Global foundation-model progress briefing
English Edition中文
Enter keywords to search ingested stories.

Today / Thursday, August 13, 2026

limbo logolimbo

Data updated

Jun 23, 05:34 PM

Live sources

17

Ingestion status

Live ingest

ResearcharXiv AI / CL

OpenThoughts-Agent: Data Recipes for Agentic Models

Summary

Agentic language models dramatically expand the applications of AI yet little is publicly known about how to curate training data for broadly capable agents.

Original Article

Captured source content or English translation, normalized into this reading format.

Read Source

Skip to main content

![](https://arxiv.org/static/base/1.0.1/images/icons/smileybones-small.svg)arXiv is now an independent nonprofit!Learn more×

Search arXiv

Press Enter to search ·Advanced search

Computer Science > Artificial Intelligence

arXiv:2606.24855v1 (cs)

[Submitted on 23 Jun 2026]

Title:OpenThoughts-Agent: Data Recipes for Agentic Models

Authors:Negin Raoof,Richard Zhuang,Marianna Nezhurina,Etash Guha,Atula Tejaswi,Ryan Marten,Charlie F. Ruan,Tyler Griggs,Alexander Glenn Shaw,Hritik Bansal,E. Kelly Buchanan,Artem Gazizov,Reinhard Heckel,Chinmay Hegde,Sankalp Jajee,Daanish Khazi,Emmanouil Koukoumidis,Xiangyi Li,Hange Liu,Shlok Natarajan,Harsh Raj,Nicholas Roberts,Ethan Shen,Nishad Singhi,Michael Siu,Ashima Suvarna,Hanwen Xing,Patrick Yubeaton,Robert Zhang,Leon Liangyu Chen,Xiaokun Chen,Steven Dillmann,Saadia Gabriel,Xunyi Jiang,Anurag Kashyap,Boxuan Li,Yein Park,Minh Pham,Sujay Sanghavi,Lin Shi,Ke Sun,Yixin Wang,Zhiwei Xu,Erica Zhang,Siyan Zhao,Wanjia Zhao,Jenia Jitsev,Alex Dimakis,Benjamin Feuer,Ludwig Schmidt

View a PDF of the paper titled OpenThoughts-Agent: Data Recipes for Agentic Models, by Negin Raoof and 49 other authors

View PDFHTML (experimental)

Abstract:Agentic language models dramatically expand the applications of AI yet little is publicly known about how to curate training data for broadly capable agents. Existing open efforts such as SWE-Smith, SERA, and Nemotron-Terminal typically target a single benchmark, leaving open the question of how to train models that generalize across diverse agentic tasks. The OpenThoughts-Agent (OT-Agent) project addresses this gap with a fully open data curation pipeline for training agentic models. We conduct more than 100 controlled ablation experiments to systematically investigate each stage of the pipeline, yielding insights on the importance of task sources and diversity. We then assemble a training set of 100K examples from our pipeline and fine-tune Qwen3-32B on this dataset, which yields an average accuracy of 44.8% across seven agentic benchmarks and a 3.9 percentage point improvement over the strongest existing open data agentic model (Nemotron-Terminal-32B, 40.9%). Moreover, our training data exhibits strong scaling properties, outperforming alternative open datasets at every training set size in compute-controlled comparisons. We publicly release our training sets, data pipeline, experimental data, and models atthis http URLto support future open research on agentic model training.

| | | | --- | --- | | Subjects: | Artificial Intelligence (cs.AI) | | Cite as: |arXiv:2606.24855[cs.AI] | | | (orarXiv:2606.24855v1[cs.AI] for this version) | | |https://doi.org/10.48550/arXiv.2606.24855<br>Focus to learn more<br>arXiv-issued DOI via DataCite |

Submission history

From: Benjamin Feuer \[view email]

[v1] Tue, 23 Jun 2026 17:34:29 UTC (2,434 KB)

Full-text links:

Access Paper:

View a PDF of the paper titled OpenThoughts-Agent: Data Recipes for Agentic Models, by Negin Raoof and 49 other authors

[view license](http://arxiv.org/licenses/nonexclusive-distrib/1.0/ "Rights to this article")

Current browse context:

cs.AI

[< prev](https://arxiv.org/prevnext?id=2606.24855&function=prev&context=cs.AI "previous in cs.AI (accesskey p)")  \|  [next >](https://arxiv.org/prevnext?id=2606.24855&function=next&context=cs.AI "next in cs.AI (accesskey n)")

new\|recent\|2026-06

Change to browse by:

cs

References & Citations

export BibTeX citation

Bookmark

![BibSonomy](http://www.bibsonomy.org/BibtexHandler?requTask=upload&url=https://arxiv.org/abs/2606.24855&description=OpenThoughts-Agent:%20Data%20Recipes%20for%20Agentic%20Models "Bookmark on BibSonomy")![Reddit](https://reddit.com/submit?url=https://arxiv.org/abs/2606.24855&title=OpenThoughts-Agent:%20Data%20Recipes%20for%20Agentic%20Models "Bookmark on Reddit")

Bibliographic Tools

Bibliographic and Citation Tools

Bibliographic Explorer Toggle

Bibliographic Explorer _(What is the Explorer?)_

Connected Papers Toggle

Connected Papers _(What is Connected Papers?)_

Litmaps Toggle

Litmaps _(What is Litmaps?)_

scite.ai Toggle

scite Smart Citations _(What are Smart Citations?)_

Code, Data, Media

Code, Data and Media Associated with this Article

alphaXiv Toggle

alphaXiv _(What is alphaXiv?)_

Links to Code Toggle

CatalyzeX Code Finder for Papers _(What is CatalyzeX?)_

DagsHub Toggle

DagsHub _(What is DagsHub?)_

GotitPub Toggle

Gotit.pub _(What is GotitPub?)_

Huggingface Toggle

Hugging Face _(What is Huggingface?)_

ScienceCast Toggle

ScienceCast _(What is ScienceCast?)_

Demos

Demos

Replicate Toggle

Replicate _(What is Replicate?)_

Spaces Toggle

Hugging Face Spaces _(What is Spaces?)_

Spaces Toggle

TXYZ.AI _(What is TXYZ.AI?)_

Related Papers

Recommenders and Search Tools

Link to Influence Flower

Influence Flower _(What are Influence Flowers?)_

Core recommender toggle

CORE Recommender _(What is CORE?)_

  • Author
  • Venue
  • Institution
  • Topic

About arXivLabs

arXivLabs: experimental projects with community collaborators

arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.

Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.

Have an idea for a project that will add value for arXiv's community?Learn more about arXivLabs.

Which authors of this paper are endorsers?\| Disable MathJax (What is MathJax?)

Region

Global

Heat Score

89

Category

Research

Language

en