OpenThoughts: Data Recipes for Reasoning Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guha, Etash, Marten, Ryan, Keh, Sedrick, Raoof, Negin, Smyrnis, Georgios, Bansal, Hritik, Nezhurina, Marianna, Mercat, Jean, Vu, Trung, Sprague, Zayne, Suvarna, Ashima, Feuer, Benjamin, Chen, Liangyu, Khan, Zaid, Frankel, Eric, Grover, Sachin, Choi, Caroline, Muennighoff, Niklas, Su, Shiye, Zhao, Wanjia, Yang, John, Pimpalgaonkar, Shreyas, Sharma, Kartik, Ji, Charlie Cheng-Jie, Deng, Yichuan, Pratt, Sarah, Ramanujan, Vivek, Saad-Falcon, Jon, Li, Jeffrey, Dave, Achal, Albalak, Alon, Arora, Kushal, Wulfe, Blake, Hegde, Chinmay, Durrett, Greg, Oh, Sewoong, Bansal, Mohit, Gabriel, Saadia, Grover, Aditya, Chang, Kai-Wei, Shankar, Vaishaal, Gokaslan, Aaron, Merrill, Mike A., Hashimoto, Tatsunori, Choi, Yejin, Jitsev, Jenia, Heckel, Reinhard, Sathiamoorthy, Maheswaran, Dimakis, Alexandros G., Schmidt, Ludwig |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Peering Through Preferences: Unraveling Feedback Acquisition for Aligning Large Language Models
von: Bansal, Hritik, et al.
Veröffentlicht: (2023)
von: Bansal, Hritik, et al.
Veröffentlicht: (2023)
Game Reasoning Arena: A Framework and Benchmark for Assessing Reasoning Capabilities of Large Language Models via Game Play
von: Cipolina-Kun, Lucia, et al.
Veröffentlicht: (2025)
von: Cipolina-Kun, Lucia, et al.
Veröffentlicht: (2025)
Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models
von: Nezhurina, Marianna, et al.
Veröffentlicht: (2024)
von: Nezhurina, Marianna, et al.
Veröffentlicht: (2024)
SkillFactory: Self-Distillation For Learning Cognitive Behaviors
von: Sprague, Zayne, et al.
Veröffentlicht: (2025)
von: Sprague, Zayne, et al.
Veröffentlicht: (2025)
MedMax: Mixed-Modal Instruction Tuning for Training Biomedical Assistants
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
Language models scale reliably with over-training and on downstream tasks
von: Gadre, Samir Yitzhak, et al.
Veröffentlicht: (2024)
von: Gadre, Samir Yitzhak, et al.
Veröffentlicht: (2024)
TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
Comparing Bad Apples to Good Oranges: Aligning Large Language Models via Joint Preference Optimization
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation
von: Bansal, Hritik, et al.
Veröffentlicht: (2025)
von: Bansal, Hritik, et al.
Veröffentlicht: (2025)
Peeking Behind Closed Doors: Risks of LLM Evaluation by Private Data Curators
von: Bansal, Hritik, et al.
Veröffentlicht: (2025)
von: Bansal, Hritik, et al.
Veröffentlicht: (2025)
Scaling Laws for Robust Comparison of Open Foundation Language-Vision Models and Datasets
von: Nezhurina, Marianna, et al.
Veröffentlicht: (2025)
von: Nezhurina, Marianna, et al.
Veröffentlicht: (2025)
Learning in Compact Spaces with Approximately Normalized Transformer
von: Franke, Jörg K. H., et al.
Veröffentlicht: (2025)
von: Franke, Jörg K. H., et al.
Veröffentlicht: (2025)
Linearizing Large Language Models
von: Mercat, Jean, et al.
Veröffentlicht: (2024)
von: Mercat, Jean, et al.
Veröffentlicht: (2024)
When To Solve, When To Verify: Compute-Optimal Problem Solving and Generative Verification for LLM Reasoning
von: Singhi, Nishad, et al.
Veröffentlicht: (2025)
von: Singhi, Nishad, et al.
Veröffentlicht: (2025)
Open-sci-ref-0.01: open and reproducible reference baselines for language model and dataset comparison
von: Nezhurina, Marianna, et al.
Veröffentlicht: (2025)
von: Nezhurina, Marianna, et al.
Veröffentlicht: (2025)
VLA Foundry: A Unified Framework for Training Vision-Language-Action Models
von: Mercat, Jean, et al.
Veröffentlicht: (2026)
von: Mercat, Jean, et al.
Veröffentlicht: (2026)
HoneyBee: Data Recipes for Vision-Language Reasoners
von: Bansal, Hritik, et al.
Veröffentlicht: (2025)
von: Bansal, Hritik, et al.
Veröffentlicht: (2025)
Scaling transformer neural networks for skillful and reliable medium-range weather forecasting
von: Nguyen, Tung, et al.
Veröffentlicht: (2023)
von: Nguyen, Tung, et al.
Veröffentlicht: (2023)
AudioToolAgent: An Agentic Framework for Audio-Language Models
von: Wijngaard, Gijs, et al.
Veröffentlicht: (2025)
von: Wijngaard, Gijs, et al.
Veröffentlicht: (2025)
DataComp-LM: In search of the next generation of training sets for language models
von: Li, Jeffrey, et al.
Veröffentlicht: (2024)
von: Li, Jeffrey, et al.
Veröffentlicht: (2024)
Word frequency and sentiment analysis of twitter messages during Coronavirus pandemic
von: Rajput, Nikhil Kumar, et al.
Veröffentlicht: (2020)
von: Rajput, Nikhil Kumar, et al.
Veröffentlicht: (2020)
LaViDa: A Large Diffusion Language Model for Multimodal Understanding
von: Li, Shufan, et al.
Veröffentlicht: (2025)
von: Li, Shufan, et al.
Veröffentlicht: (2025)
VideoPhy: Evaluating Physical Commonsense for Video Generation
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models
von: Wadhawan, Rohan, et al.
Veröffentlicht: (2024)
von: Wadhawan, Rohan, et al.
Veröffentlicht: (2024)
Resolving Discrepancies in Compute-Optimal Scaling of Language Models
von: Porian, Tomer, et al.
Veröffentlicht: (2024)
von: Porian, Tomer, et al.
Veröffentlicht: (2024)
ryangrvr/GRUT-RAI-v1.0: GRUT-RAI v2.0 — Phase IV Canonical Build
von: Ryan Grover
Veröffentlicht: (2026)
von: Ryan Grover
Veröffentlicht: (2026)
ryangrvr/GRUT-RAI-v1.0: GRUT I — Final Program Archive and GRUT-RAI v1.0
von: Ryan Grover
Veröffentlicht: (2026)
von: Ryan Grover
Veröffentlicht: (2026)
FinRLlama: A Solution to LLM-Engineered Signals Challenge at FinRL Contest 2024
von: Grover, Arnav
Veröffentlicht: (2025)
von: Grover, Arnav
Veröffentlicht: (2025)
Reliability Under Randomness: An Empirical Analysis of Sparse and Dense Language Models Across Decoding Temperatures
von: Grover, Kabir
Veröffentlicht: (2026)
von: Grover, Kabir
Veröffentlicht: (2026)
Steel City Readers
von: Grover, Mary
Veröffentlicht: (2023)
von: Grover, Mary
Veröffentlicht: (2023)
Discussion of Next Generation Models for Subsequent Sports Injuries by Wu Et Al.
von: Rhythm Grover
Veröffentlicht: (2025)
von: Rhythm Grover
Veröffentlicht: (2025)
Manual de preparación de proyectos de abastecimiento de agua y saneamiento / Brian Grover, Nicholas Burnett, Michael McGarry
von: Grover, Brian
Veröffentlicht: (1986)
von: Grover, Brian
Veröffentlicht: (1986)
The Wage Stop and Restricting Benefit Income in the United Kingdom: Discretion, Wages and Hardship
von: Chris Grover
Veröffentlicht: (2024)
von: Chris Grover
Veröffentlicht: (2024)
ABC needs more than letterman / Ronald Grover
von: Grover, Ronald
von: Grover, Ronald
Hierarchical entanglement transitions and hidden area-law sectors in quantum many-body dynamics
von: Grover, Tarun
Veröffentlicht: (2026)
von: Grover, Tarun
Veröffentlicht: (2026)
CAMBIOS FISICOQUÍMICOS POR EXPOSICIÓN A LA RADIACIÓN SOLAR EN TUBÉRCULOS DE OXALIS TUBEROSA, “OCA” CULTIVADOS EN BOLIVIA
von: Grover Castañeta
Veröffentlicht: (2022)
von: Grover Castañeta
Veröffentlicht: (2022)
MICROPLÁSTICOS: UN CONTAMINANTE QUE CRECE EN TODAS LAS ESFERAS AMBIENTALES, SUS CARACTERÍSTICAS Y POSIBLES RIESGOS PARA LA SALUD PÚBLICA POR EXPOSICIÓN
von: Grover Castañeta
Veröffentlicht: (2020)
von: Grover Castañeta
Veröffentlicht: (2020)
Assessing Information Skills Instruction.
von: Grover, Robert
Veröffentlicht: (1994)
von: Grover, Robert
Veröffentlicht: (1994)
A Proposed Model for Diagnosing Information Needs.
von: Grover, Robert
Veröffentlicht: (1993)
von: Grover, Robert
Veröffentlicht: (1993)
CHARACTERIZATION OF TERPENOIDS FROM PSEUDOGNAPHALIUM GAUDICHAUDIANUM (DC) ANDERB, WIRA-WIRA BY GC/MS, ACTIVE PRINCIPLES WITH POSSIBLE USE IN COVID-19 INFECTION PREVENTION
von: Grover Castañeta
Veröffentlicht: (2022)
von: Grover Castañeta
Veröffentlicht: (2022)
Ähnliche Einträge
-
Peering Through Preferences: Unraveling Feedback Acquisition for Aligning Large Language Models
von: Bansal, Hritik, et al.
Veröffentlicht: (2023) -
Game Reasoning Arena: A Framework and Benchmark for Assessing Reasoning Capabilities of Large Language Models via Game Play
von: Cipolina-Kun, Lucia, et al.
Veröffentlicht: (2025) -
Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models
von: Nezhurina, Marianna, et al.
Veröffentlicht: (2024) -
SkillFactory: Self-Distillation For Learning Cognitive Behaviors
von: Sprague, Zayne, et al.
Veröffentlicht: (2025) -
MedMax: Mixed-Modal Instruction Tuning for Training Biomedical Assistants
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)