Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Singh, Avi, Co-Reyes, John D., Agarwal, Rishabh, Anand, Ankesh, Patil, Piyush, Garcia, Xavier, Liu, Peter J., Harrison, James, Lee, Jaehoon, Xu, Kelvin, Parisi, Aaron, Kumar, Abhishek, Alemi, Alex, Rizkowsky, Alex, Nova, Azade, Adlam, Ben, Bohnet, Bernd, Elsayed, Gamaleldin, Sedghi, Hanie, Mordatch, Igor, Simpson, Isabelle, Gur, Izzeddin, Snoek, Jasper, Pennington, Jeffrey, Hron, Jiri, Kenealy, Kathleen, Swersky, Kevin, Mahajan, Kshiteej, Culp, Laura, Xiao, Lechao, Bileschi, Maxwell L., Constant, Noah, Novak, Roman, Liu, Rosanne, Warkentin, Tris, Qian, Yundi, Bansal, Yamini, Dyer, Ethan, Neyshabur, Behnam, Sohl-Dickstein, Jascha, Fiedel, Noah |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Training Language Models on the Knowledge Graph: Insights on Hallucinations and Their Detectability
von: Hron, Jiri, et al.
Veröffentlicht: (2024)
von: Hron, Jiri, et al.
Veröffentlicht: (2024)
Exploring and Benchmarking the Planning Capabilities of Large Language Models
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024)
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024)
Enhancing LLM Planning Capabilities through Intrinsic Self-Critique
von: Bohnet, Bernd, et al.
Veröffentlicht: (2025)
von: Bohnet, Bernd, et al.
Veröffentlicht: (2025)
A Comparative Analysis of LLM Adaptation: SFT, LoRA, and ICL in Data-Scarce Scenarios
von: Bohnet, Bernd, et al.
Veröffentlicht: (2025)
von: Bohnet, Bernd, et al.
Veröffentlicht: (2025)
Analysis of Optimality of Large Language Models on Planning Problems
von: Bohnet, Bernd, et al.
Veröffentlicht: (2026)
von: Bohnet, Bernd, et al.
Veröffentlicht: (2026)
Long-Span Question-Answering: Automatic Question Generation and QA-System Ranking via Side-by-Side Evaluation
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024)
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024)
Improving Large Language Model Planning with Action Sequence Similarity
von: Zhao, Xinran, et al.
Veröffentlicht: (2025)
von: Zhao, Xinran, et al.
Veröffentlicht: (2025)
Training LLMs over Neurally Compressed Text
von: Lester, Brian, et al.
Veröffentlicht: (2024)
von: Lester, Brian, et al.
Veröffentlicht: (2024)
Levels of AGI for Operationalizing Progress on the Path to AGI
von: Morris, Meredith Ringel, et al.
Veröffentlicht: (2023)
von: Morris, Meredith Ringel, et al.
Veröffentlicht: (2023)
Transfer Learning for Text Diffusion Models
von: Han, Kehang, et al.
Veröffentlicht: (2024)
von: Han, Kehang, et al.
Veröffentlicht: (2024)
The boundary of neural network trainability is fractal
von: Sohl-Dickstein, Jascha
Veröffentlicht: (2024)
von: Sohl-Dickstein, Jascha
Veröffentlicht: (2024)
Scaling Exponents Across Parameterizations and Optimizers
von: Everett, Katie, et al.
Veröffentlicht: (2024)
von: Everett, Katie, et al.
Veröffentlicht: (2024)
Do generative video models understand physical principles?
von: Motamed, Saman, et al.
Veröffentlicht: (2025)
von: Motamed, Saman, et al.
Veröffentlicht: (2025)
General-Purpose In-Context Learning by Meta-Learning Transformers
von: Kirsch, Louis, et al.
Veröffentlicht: (2022)
von: Kirsch, Louis, et al.
Veröffentlicht: (2022)
Because we have LLMs, we Can and Should Pursue Agentic Interpretability
von: Kim, Been, et al.
Veröffentlicht: (2025)
von: Kim, Been, et al.
Veröffentlicht: (2025)
The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?
von: Hägele, Alexander, et al.
Veröffentlicht: (2026)
von: Hägele, Alexander, et al.
Veröffentlicht: (2026)
Searching for Democratic Constraint in Donald Trump's America
von: Andrew Kenealy
Veröffentlicht: (2026)
von: Andrew Kenealy
Veröffentlicht: (2026)
Wave-Freezing and other Phenomena in Temporal Metasurfaces driven by Nonlocal Interactions
von: Deshmukh, Kshiteej J.
Veröffentlicht: (2024)
von: Deshmukh, Kshiteej J.
Veröffentlicht: (2024)
Braid group actions, Baxter polynomials, and affine quantum groups
von: Friesen, Noah, et al.
Veröffentlicht: (2024)
von: Friesen, Noah, et al.
Veröffentlicht: (2024)
Neurotoxicidad en neonatos con hiperbilirrubinemia severa. Análisis de los factores de riesgo para neurotoxicidad en neonatos con ictericia severa
von: Rasha Gamaleldin
Veröffentlicht: (2012)
von: Rasha Gamaleldin
Veröffentlicht: (2012)
Quantum Field Theory and the Limits of Reductionism
von: Adlam, Emily
Veröffentlicht: (2024)
von: Adlam, Emily
Veröffentlicht: (2024)
Wigner's Frame
von: Adlam, Emily
Veröffentlicht: (2025)
von: Adlam, Emily
Veröffentlicht: (2025)
Against Self-Location
von: Adlam, Emily
Veröffentlicht: (2024)
von: Adlam, Emily
Veröffentlicht: (2024)
How do we Observe Relational Observables?
von: Adlam, Emily
Veröffentlicht: (2024)
von: Adlam, Emily
Veröffentlicht: (2024)
Why does the wavefunction 'collapse' in relational approaches to quantum mechanics?
von: Adlam, Emily
Veröffentlicht: (2026)
von: Adlam, Emily
Veröffentlicht: (2026)
What Do Black Holes Teach Us About Wigner's Friend?
von: Adlam, Emily
Veröffentlicht: (2026)
von: Adlam, Emily
Veröffentlicht: (2026)
Moderate Physical Perspectivalism
von: Adlam, Emily
Veröffentlicht: (2024)
von: Adlam, Emily
Veröffentlicht: (2024)
Relational Observables, Quiddities, and Structural Realism
von: Adlam, Emily
Veröffentlicht: (2025)
von: Adlam, Emily
Veröffentlicht: (2025)
What Kind of Relationality does Quantum Mechanics Exhibit?
von: Adlam, Emily
Veröffentlicht: (2025)
von: Adlam, Emily
Veröffentlicht: (2025)
The Combination Problem for Relational Quantum Mechanics
von: Adlam, Emily
Veröffentlicht: (2024)
von: Adlam, Emily
Veröffentlicht: (2024)
How are Entanglement Entropies Related to Entropy Bounds?
von: Adlam, Emily
Veröffentlicht: (2024)
von: Adlam, Emily
Veröffentlicht: (2024)
O PEDIATRA E SUA FUNÇÃO APOSTÓLICA: PERCEPÇÕES DE MÉDICOS RESIDENTES SOBRE SUAS PRÁTICAS
von: Paulo Dickstein
Veröffentlicht: (2017)
von: Paulo Dickstein
Veröffentlicht: (2017)
Der metafiktionale Roman
von: Bohnet, Christine
Veröffentlicht: (2019)
von: Bohnet, Christine
Veröffentlicht: (2019)
Turning toward Edification
von: Bohnet, Adam
Veröffentlicht: (2020)
von: Bohnet, Adam
Veröffentlicht: (2020)
The Amazing Journey of Reason
von: Alemi, Mario
Veröffentlicht: (2020)
von: Alemi, Mario
Veröffentlicht: (2020)
SAIL Thomson Reuters Update
von: Culp, Kristin
Veröffentlicht: ()
von: Culp, Kristin
Veröffentlicht: ()
Exploring the Reasons Behind Iranian TEFL Graduate Students’ Academic Failure
von: Minoo Alemi
Veröffentlicht: (2021)
von: Minoo Alemi
Veröffentlicht: (2021)
Sublinear Time Low-Rank Approximation of Toeplitz Matrices
von: Musco, Cameron, et al.
Veröffentlicht: (2024)
von: Musco, Cameron, et al.
Veröffentlicht: (2024)
Gestures of neighbor‐love: Literature, philosophy, and givenness
von: Irina Hron
Veröffentlicht: (2024)
von: Irina Hron
Veröffentlicht: (2024)
Grumpy brand: Reading literary prize fiction
von: Irina Hron
Veröffentlicht: (2025)
von: Irina Hron
Veröffentlicht: (2025)
Ähnliche Einträge
-
Training Language Models on the Knowledge Graph: Insights on Hallucinations and Their Detectability
von: Hron, Jiri, et al.
Veröffentlicht: (2024) -
Exploring and Benchmarking the Planning Capabilities of Large Language Models
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024) -
Enhancing LLM Planning Capabilities through Intrinsic Self-Critique
von: Bohnet, Bernd, et al.
Veröffentlicht: (2025) -
A Comparative Analysis of LLM Adaptation: SFT, LoRA, and ICL in Data-Scarce Scenarios
von: Bohnet, Bernd, et al.
Veröffentlicht: (2025) -
Analysis of Optimality of Large Language Models on Planning Problems
von: Bohnet, Bernd, et al.
Veröffentlicht: (2026)