Salvato in:
| Autori principali: | Havrilla, Alex, Iyer, Maia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2402.04004 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Understanding Scaling Laws with Statistical and Approximation Theory for Transformer Neural Networks on Intrinsically Low-dimensional Data
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
SPARQ: Synthetic Problem Generation for Reasoning via Quality-Diversity Algorithms
di: Havrilla, Alex, et al.
Pubblicazione: (2025)
di: Havrilla, Alex, et al.
Pubblicazione: (2025)
IGDA: Interactive Graph Discovery through Large Language Model Agents
di: Havrilla, Alex, et al.
Pubblicazione: (2025)
di: Havrilla, Alex, et al.
Pubblicazione: (2025)
DFU: scale-robust diffusion model for zero-shot super-resolution image generation
di: Havrilla, Alex, et al.
Pubblicazione: (2023)
di: Havrilla, Alex, et al.
Pubblicazione: (2023)
GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
Understanding Hidden Computations in Chain-of-Thought Reasoning
di: Bharadwaj, Aryasomayajula Ram
Pubblicazione: (2024)
di: Bharadwaj, Aryasomayajula Ram
Pubblicazione: (2024)
Transformers for Learning on Noisy and Task-Level Manifolds: Approximation and Generalization Insights
di: Shen, Zhaiming, et al.
Pubblicazione: (2025)
di: Shen, Zhaiming, et al.
Pubblicazione: (2025)
Constraint-Rectified Training for Efficient Chain-of-Thought
di: Wu, Qinhang, et al.
Pubblicazione: (2026)
di: Wu, Qinhang, et al.
Pubblicazione: (2026)
Temporal Chain of Thought: Long-Video Understanding by Thinking in Frames
di: Arnab, Anurag, et al.
Pubblicazione: (2025)
di: Arnab, Anurag, et al.
Pubblicazione: (2025)
Emergence of Superposition: Unveiling the Training Dynamics of Chain of Continuous Thought
di: Zhu, Hanlin, et al.
Pubblicazione: (2025)
di: Zhu, Hanlin, et al.
Pubblicazione: (2025)
Surveying the Effects of Quality, Diversity, and Complexity in Synthetic Data From Large Language Models
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
Latent Chain-of-Thought Improves Structured-Data Transformers
di: Dudley, Carson, et al.
Pubblicazione: (2026)
di: Dudley, Carson, et al.
Pubblicazione: (2026)
Understanding Chain-of-Thought in LLMs through Information Theory
di: Ton, Jean-Francois, et al.
Pubblicazione: (2024)
di: Ton, Jean-Francois, et al.
Pubblicazione: (2024)
Understanding Reasoning in Chain-of-Thought from the Hopfieldian View
di: Hu, Lijie, et al.
Pubblicazione: (2024)
di: Hu, Lijie, et al.
Pubblicazione: (2024)
Output Supervision Can Obfuscate the Chain of Thought
di: Drori, Jacob, et al.
Pubblicazione: (2025)
di: Drori, Jacob, et al.
Pubblicazione: (2025)
Utilizing Training Data to Improve LLM Reasoning for Tabular Understanding
di: Gao, Chufan, et al.
Pubblicazione: (2025)
di: Gao, Chufan, et al.
Pubblicazione: (2025)
When More is Less: Understanding Chain-of-Thought Length in LLMs
di: Wu, Yuyang, et al.
Pubblicazione: (2025)
di: Wu, Yuyang, et al.
Pubblicazione: (2025)
Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning
di: Jia, Jinghan, et al.
Pubblicazione: (2026)
di: Jia, Jinghan, et al.
Pubblicazione: (2026)
Zipping the Thought: When and How Compressed Reasoning Data Works in LLM Post-Training
di: Matsutani, Kohsei, et al.
Pubblicazione: (2026)
di: Matsutani, Kohsei, et al.
Pubblicazione: (2026)
Understanding Silent Data Corruption in LLM Training
di: Ma, Jeffrey, et al.
Pubblicazione: (2025)
di: Ma, Jeffrey, et al.
Pubblicazione: (2025)
Make Still Further Progress: Chain of Thoughts for Tabular Data Leaderboard
di: Liu, Si-Yang, et al.
Pubblicazione: (2025)
di: Liu, Si-Yang, et al.
Pubblicazione: (2025)
Training Nonlinear Transformers for Chain-of-Thought Inference: A Theoretical Generalization Analysis
di: Li, Hongkang, et al.
Pubblicazione: (2024)
di: Li, Hongkang, et al.
Pubblicazione: (2024)
ProLLM: Protein Chain-of-Thoughts Enhanced LLM for Protein-Protein Interaction Prediction
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
Audio Flamingo Sound-CoT Technical Report: Improving Chain-of-Thought Reasoning in Sound Understanding
di: Kong, Zhifeng, et al.
Pubblicazione: (2025)
di: Kong, Zhifeng, et al.
Pubblicazione: (2025)
A Theoretical Understanding of Chain-of-Thought: Coherent Reasoning and Error-Aware Demonstration
di: Cui, Yingqian, et al.
Pubblicazione: (2024)
di: Cui, Yingqian, et al.
Pubblicazione: (2024)
Analyzing the Effect of Noise in LLM Fine-tuning
di: Li, Lingfang, et al.
Pubblicazione: (2026)
di: Li, Lingfang, et al.
Pubblicazione: (2026)
BC Protocol: Structured Dual-Expert Dialogue for Eliciting High-Quality Chain-of-Thought Post-Training Data
di: Zou, Bo, et al.
Pubblicazione: (2026)
di: Zou, Bo, et al.
Pubblicazione: (2026)
Chain-of-Thought Predictive Control
di: Jia, Zhiwei, et al.
Pubblicazione: (2023)
di: Jia, Zhiwei, et al.
Pubblicazione: (2023)
Transformers Provably Learn to Internalize Chain-of-Thought
di: Huang, Yixiao, et al.
Pubblicazione: (2026)
di: Huang, Yixiao, et al.
Pubblicazione: (2026)
Deep Thinking by Markov Chain of Continuous Thoughts
di: Liu, Jiayu, et al.
Pubblicazione: (2025)
di: Liu, Jiayu, et al.
Pubblicazione: (2025)
On Learning Verifiers and Implications to Chain-of-Thought Reasoning
di: Balcan, Maria-Florina, et al.
Pubblicazione: (2025)
di: Balcan, Maria-Florina, et al.
Pubblicazione: (2025)
Boltzmann Reinforcement Learning for Noise resilience in Analog Ising Machines
di: Choudhary, Aditya, et al.
Pubblicazione: (2026)
di: Choudhary, Aditya, et al.
Pubblicazione: (2026)
S-Chain: Structured Visual Chain-of-Thought For Medicine
di: Le-Duc, Khai, et al.
Pubblicazione: (2025)
di: Le-Duc, Khai, et al.
Pubblicazione: (2025)
Training Multimodal Large Reasoning Models Needs Better Thoughts: A Three-Stage Framework for Long Chain-of-Thought Synthesis and Selection
di: Wang, Yizhi, et al.
Pubblicazione: (2025)
di: Wang, Yizhi, et al.
Pubblicazione: (2025)
Is Chain-of-Thought Really Not Explainability? Chain-of-Thought Can Be Faithful without Hint Verbalization
di: Zaman, Kerem, et al.
Pubblicazione: (2025)
di: Zaman, Kerem, et al.
Pubblicazione: (2025)
Ehrenfeucht-Haussler Rank and Chain of Thought
di: Barceló, Pablo, et al.
Pubblicazione: (2025)
di: Barceló, Pablo, et al.
Pubblicazione: (2025)
Activation Steering for Chain-of-Thought Compression
di: Azizi, Seyedarmin, et al.
Pubblicazione: (2025)
di: Azizi, Seyedarmin, et al.
Pubblicazione: (2025)
Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs
di: Zhang, Xuan, et al.
Pubblicazione: (2024)
di: Zhang, Xuan, et al.
Pubblicazione: (2024)
Fractured Chain-of-Thought Reasoning
di: Liao, Baohao, et al.
Pubblicazione: (2025)
di: Liao, Baohao, et al.
Pubblicazione: (2025)
LLaVaOLMoBitnet1B: Ternary LLM goes Multimodal!
di: Sundaram, Jainaveen, et al.
Pubblicazione: (2024)
di: Sundaram, Jainaveen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Understanding Scaling Laws with Statistical and Approximation Theory for Transformer Neural Networks on Intrinsically Low-dimensional Data
di: Havrilla, Alex, et al.
Pubblicazione: (2024) -
SPARQ: Synthetic Problem Generation for Reasoning via Quality-Diversity Algorithms
di: Havrilla, Alex, et al.
Pubblicazione: (2025) -
IGDA: Interactive Graph Discovery through Large Language Model Agents
di: Havrilla, Alex, et al.
Pubblicazione: (2025) -
DFU: scale-robust diffusion model for zero-shot super-resolution image generation
di: Havrilla, Alex, et al.
Pubblicazione: (2023) -
GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements
di: Havrilla, Alex, et al.
Pubblicazione: (2024)