Catch Your Breath: Adaptive Computation for Self-Paced Sequence Production
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Galashov, Alexandre, Jones, Matt, Ke, Rosemary, Cao, Yuan, Nagarajan, Vaishnavh, Mozer, Michael C. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The pitfalls of next-token prediction
von: Bachmann, Gregor, et al.
Veröffentlicht: (2024)
von: Bachmann, Gregor, et al.
Veröffentlicht: (2024)
Deep sequence models tend to memorize geometrically; it is unclear why
von: Noroozizadeh, Shahriar, et al.
Veröffentlicht: (2025)
von: Noroozizadeh, Shahriar, et al.
Veröffentlicht: (2025)
Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction
von: Nagarajan, Vaishnavh, et al.
Veröffentlicht: (2025)
von: Nagarajan, Vaishnavh, et al.
Veröffentlicht: (2025)
Think before you speak: Training Language Models With Pause Tokens
von: Goyal, Sachin, et al.
Veröffentlicht: (2023)
von: Goyal, Sachin, et al.
Veröffentlicht: (2023)
Decoupling the "What" and "Where" With Polar Coordinate Positional Embeddings
von: Gopalakrishnan, Anand, et al.
Veröffentlicht: (2025)
von: Gopalakrishnan, Anand, et al.
Veröffentlicht: (2025)
Clip Your Sequences Fairly: Enforcing Length Fairness for Sequence-Level RL
von: Mao, Hanyi, et al.
Veröffentlicht: (2025)
von: Mao, Hanyi, et al.
Veröffentlicht: (2025)
Analysis of Optimality of Large Language Models on Planning Problems
von: Bohnet, Bernd, et al.
Veröffentlicht: (2026)
von: Bohnet, Bernd, et al.
Veröffentlicht: (2026)
On student-teacher deviations in distillation: does it pay to disobey?
von: Nagarajan, Vaishnavh, et al.
Veröffentlicht: (2023)
von: Nagarajan, Vaishnavh, et al.
Veröffentlicht: (2023)
CoRefine: Confidence-Guided Self-Refinement for Adaptive Test-Time Compute
von: Jin, Chen, et al.
Veröffentlicht: (2026)
von: Jin, Chen, et al.
Veröffentlicht: (2026)
SelfGoal: Your Language Agents Already Know How to Achieve High-level Goals
von: Yang, Ruihan, et al.
Veröffentlicht: (2024)
von: Yang, Ruihan, et al.
Veröffentlicht: (2024)
Think When You Need: Self-Adaptive Chain-of-Thought Learning
von: Yang, Junjie, et al.
Veröffentlicht: (2025)
von: Yang, Junjie, et al.
Veröffentlicht: (2025)
Deep MMD Gradient Flow without adversarial training
von: Galashov, Alexandre, et al.
Veröffentlicht: (2024)
von: Galashov, Alexandre, et al.
Veröffentlicht: (2024)
Put Your Money Where Your Mouth Is: Evaluating Strategic Planning and Execution of LLM Agents in an Auction Arena
von: Chen, Jiangjie, et al.
Veröffentlicht: (2023)
von: Chen, Jiangjie, et al.
Veröffentlicht: (2023)
SelfElicit: Your Language Model Secretly Knows Where is the Relevant Evidence
von: Liu, Zhining, et al.
Veröffentlicht: (2025)
von: Liu, Zhining, et al.
Veröffentlicht: (2025)
Catching Chameleons: Detecting Evolving Disinformation Generated using Large Language Models
von: Jiang, Bohan, et al.
Veröffentlicht: (2024)
von: Jiang, Bohan, et al.
Veröffentlicht: (2024)
Self-supervised Preference Optimization: Enhance Your Language Model with Preference Degree Awareness
von: Li, Jian, et al.
Veröffentlicht: (2024)
von: Li, Jian, et al.
Veröffentlicht: (2024)
Efficient Self-Evaluation for Diffusion Language Models via Sequence Regeneration
von: Zhong, Linhao, et al.
Veröffentlicht: (2026)
von: Zhong, Linhao, et al.
Veröffentlicht: (2026)
Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference
von: Feng, Yuan, et al.
Veröffentlicht: (2024)
von: Feng, Yuan, et al.
Veröffentlicht: (2024)
Luna: An Evaluation Foundation Model to Catch Language Model Hallucinations with High Accuracy and Low Cost
von: Belyi, Masha, et al.
Veröffentlicht: (2024)
von: Belyi, Masha, et al.
Veröffentlicht: (2024)
Make Every Penny Count: Difficulty-Adaptive Self-Consistency for Cost-Efficient Reasoning
von: Wang, Xinglin, et al.
Veröffentlicht: (2024)
von: Wang, Xinglin, et al.
Veröffentlicht: (2024)
Reflection Pretraining Enables Token-Level Self-Correction in Biological Sequence Models
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
Fast-weight Product Key Memory
von: Zhao, Tianyu, et al.
Veröffentlicht: (2026)
von: Zhao, Tianyu, et al.
Veröffentlicht: (2026)
SelECT-SQL: Self-correcting ensemble Chain-of-Thought for Text-to-SQL
von: Shen, Ke, et al.
Veröffentlicht: (2024)
von: Shen, Ke, et al.
Veröffentlicht: (2024)
Adaptive Rectification Sampling for Test-Time Compute Scaling
von: Tan, Zhendong, et al.
Veröffentlicht: (2025)
von: Tan, Zhendong, et al.
Veröffentlicht: (2025)
Attention Sinks: A 'Catch, Tag, Release' Mechanism for Embeddings
von: Zhang, Stephen, et al.
Veröffentlicht: (2025)
von: Zhang, Stephen, et al.
Veröffentlicht: (2025)
Self-Resource Allocation in Multi-Agent LLM Systems
von: Amayuelas, Alfonso, et al.
Veröffentlicht: (2025)
von: Amayuelas, Alfonso, et al.
Veröffentlicht: (2025)
Focus on Your Question! Interpreting and Mitigating Toxic CoT Problems in Commonsense Reasoning
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
von: Li, Jiachun, et al.
Veröffentlicht: (2024)
Catch Me If You Can? Not Yet: LLMs Still Struggle to Imitate the Implicit Writing Styles of Everyday Authors
von: Wang, Zhengxiang, et al.
Veröffentlicht: (2025)
von: Wang, Zhengxiang, et al.
Veröffentlicht: (2025)
Can AI Be as Creative as Humans?
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
Reflective Agreement: Combining Self-Mixture of Agents with a Sequence Tagger for Robust Event Extraction
von: Haji, Fatemeh, et al.
Veröffentlicht: (2025)
von: Haji, Fatemeh, et al.
Veröffentlicht: (2025)
TALEC: Teach Your LLM to Evaluate in Specific Domain with In-house Criteria by Criteria Division and Zero-shot Plus Few-shot
von: Zhang, Kaiqi, et al.
Veröffentlicht: (2024)
von: Zhang, Kaiqi, et al.
Veröffentlicht: (2024)
Shorten After You're Right: Lazy Length Penalties for Reasoning RL
von: Yuan, Danlong, et al.
Veröffentlicht: (2025)
von: Yuan, Danlong, et al.
Veröffentlicht: (2025)
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
von: Li, Zheng, et al.
Veröffentlicht: (2025)
von: Li, Zheng, et al.
Veröffentlicht: (2025)
Adaptive Graph Refinement and Label Propagation with LLMs for Cost-Effective Entity Resolution
von: Wang, Hongtao, et al.
Veröffentlicht: (2026)
von: Wang, Hongtao, et al.
Veröffentlicht: (2026)
How Persuasive is Your Context?
von: Nguyen, Tu, et al.
Veröffentlicht: (2025)
von: Nguyen, Tu, et al.
Veröffentlicht: (2025)
DSPy Assertions: Computational Constraints for Self-Refining Language Model Pipelines
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
Metaphors We Compute By: A Computational Audit of Cultural Translation vs. Thinking in LLMs
von: Chang, Yuan, et al.
Veröffentlicht: (2026)
von: Chang, Yuan, et al.
Veröffentlicht: (2026)
Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents
von: Shao, Shuai, et al.
Veröffentlicht: (2025)
von: Shao, Shuai, et al.
Veröffentlicht: (2025)
Sequence to Sequence Reward Modeling: Improving RLHF by Language Feedback
von: Zhou, Jiayi, et al.
Veröffentlicht: (2024)
von: Zhou, Jiayi, et al.
Veröffentlicht: (2024)
Are Your LLMs Capable of Stable Reasoning?
von: Liu, Junnan, et al.
Veröffentlicht: (2024)
von: Liu, Junnan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The pitfalls of next-token prediction
von: Bachmann, Gregor, et al.
Veröffentlicht: (2024) -
Deep sequence models tend to memorize geometrically; it is unclear why
von: Noroozizadeh, Shahriar, et al.
Veröffentlicht: (2025) -
Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction
von: Nagarajan, Vaishnavh, et al.
Veröffentlicht: (2025) -
Think before you speak: Training Language Models With Pause Tokens
von: Goyal, Sachin, et al.
Veröffentlicht: (2023) -
Decoupling the "What" and "Where" With Polar Coordinate Positional Embeddings
von: Gopalakrishnan, Anand, et al.
Veröffentlicht: (2025)