Concise Reasoning in the Lens of Lagrangian Optimization
Fuente:
arXiv
Guardado en:
| Autores principales: | Gao, Chengqian, Li, Haonan, Killian, Taylor W., She, Jianshu, Wang, Renxi, Ma, Liqun, Cheng, Zhoujun, Hao, Shibo, Xu, Zhiqiang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective
por: Cheng, Zhoujun, et al.
Publicado: (2025)
por: Cheng, Zhoujun, et al.
Publicado: (2025)
Hawkeye:Efficient Reasoning with Model Collaboration
por: She, Jianshu, et al.
Publicado: (2025)
por: She, Jianshu, et al.
Publicado: (2025)
SplitAgent: A Privacy-Preserving Distributed Architecture for Enterprise-Cloud Agent Collaboration
por: She, Jianshu
Publicado: (2026)
por: She, Jianshu
Publicado: (2026)
Principled Data Selection for Alignment: The Hidden Risks of Difficult Examples
por: Gao, Chengqian, et al.
Publicado: (2025)
por: Gao, Chengqian, et al.
Publicado: (2025)
Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight
por: Cui, Christopher Z., et al.
Publicado: (2026)
por: Cui, Christopher Z., et al.
Publicado: (2026)
Bi-Mamba: Towards Accurate 1-Bit State Space Models
por: Tang, Shengkun, et al.
Publicado: (2024)
por: Tang, Shengkun, et al.
Publicado: (2024)
LogicLens: Visual-Logical Co-Reasoning for Text-Centric Forgery Analysis
por: Zeng, Fanwei, et al.
Publicado: (2025)
por: Zeng, Fanwei, et al.
Publicado: (2025)
ConciseHint: Boosting Efficient Reasoning via Continuous Concise Hints during Generation
por: Tang, Siao, et al.
Publicado: (2025)
por: Tang, Siao, et al.
Publicado: (2025)
IsoCompute Playbook: Optimally Scaling Sampling Compute for LLM RL
por: Cheng, Zhoujun, et al.
Publicado: (2026)
por: Cheng, Zhoujun, et al.
Publicado: (2026)
Steering Large Reasoning Models towards Concise Reasoning via Flow Matching
por: Li, Yawei, et al.
Publicado: (2026)
por: Li, Yawei, et al.
Publicado: (2026)
Prompt Recursive Search: A Living Framework with Adaptive Growth in LLM Auto-Prompting
por: Zhao, Xiangyu, et al.
Publicado: (2024)
por: Zhao, Xiangyu, et al.
Publicado: (2024)
FBI-LLM: Scaling Up Fully Binarized LLMs from Scratch via Autoregressive Distillation
por: Ma, Liqun, et al.
Publicado: (2024)
por: Ma, Liqun, et al.
Publicado: (2024)
THOR: Tool-Integrated Hierarchical Optimization via RL for Mathematical Reasoning
por: Chang, Qikai, et al.
Publicado: (2025)
por: Chang, Qikai, et al.
Publicado: (2025)
Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey
por: Zhu, Jason, et al.
Publicado: (2025)
por: Zhu, Jason, et al.
Publicado: (2025)
Concise Reasoning, Big Gains: Pruning Long Reasoning Trace with Difficulty-Aware Prompting
por: Wu, Yifan, et al.
Publicado: (2025)
por: Wu, Yifan, et al.
Publicado: (2025)
How Does Controllability Emerge In Language Models During Pretraining?
por: She, Jianshu, et al.
Publicado: (2025)
por: She, Jianshu, et al.
Publicado: (2025)
HAPO: Training Language Models to Reason Concisely via History-Aware Policy Optimization
por: Huang, Chengyu, et al.
Publicado: (2025)
por: Huang, Chengyu, et al.
Publicado: (2025)
Concise Thoughts: Impact of Output Length on LLM Reasoning and Cost
por: Nayab, Sania, et al.
Publicado: (2024)
por: Nayab, Sania, et al.
Publicado: (2024)
Correct, Concise and Complete: Multi-stage Training For Adaptive Reasoning
por: Rakotonirina, Nathanaël Carraz, et al.
Publicado: (2026)
por: Rakotonirina, Nathanaël Carraz, et al.
Publicado: (2026)
Metacognitive Reuse: Turning Recurring LLM Reasoning Into Concise Behaviors
por: Didolkar, Aniket, et al.
Publicado: (2025)
por: Didolkar, Aniket, et al.
Publicado: (2025)
Concise and Organized Perception Facilitates Reasoning in Large Language Models
por: Liu, Junjie, et al.
Publicado: (2023)
por: Liu, Junjie, et al.
Publicado: (2023)
LeTO: Learning Constrained Visuomotor Policy with Differentiable Trajectory Optimization
por: Xu, Zhengtong, et al.
Publicado: (2024)
por: Xu, Zhengtong, et al.
Publicado: (2024)
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training
por: Wang, Chen, et al.
Publicado: (2026)
por: Wang, Chen, et al.
Publicado: (2026)
SAINT: Attention-Based Policies for Discrete Combinatorial Action Spaces
por: Landers, Matthew, et al.
Publicado: (2025)
por: Landers, Matthew, et al.
Publicado: (2025)
Intent-Enhanced Data Augmentation for Sequential Recommendation
por: Chen, Shuai, et al.
Publicado: (2024)
por: Chen, Shuai, et al.
Publicado: (2024)
C3: A Bilingual Benchmark for Spoken Dialogue Models Exploring Challenges in Complex Conversations
por: Ma, Chengqian, et al.
Publicado: (2025)
por: Ma, Chengqian, et al.
Publicado: (2025)
Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples
por: Yu, Fangxu, et al.
Publicado: (2024)
por: Yu, Fangxu, et al.
Publicado: (2024)
ConciseRL: Conciseness-Guided Reinforcement Learning for Efficient Reasoning Models
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2025)
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2025)
What Are Tools Anyway? A Survey from the Language Model Perspective
por: Wang, Zhiruo, et al.
Publicado: (2024)
por: Wang, Zhiruo, et al.
Publicado: (2024)
Beyond Size: How Gradients Shape Pruning Decisions in Large Language Models
por: Das, Rocktim Jyoti, et al.
Publicado: (2023)
por: Das, Rocktim Jyoti, et al.
Publicado: (2023)
Self-Training Elicits Concise Reasoning in Large Language Models
por: Munkhbat, Tergel, et al.
Publicado: (2025)
por: Munkhbat, Tergel, et al.
Publicado: (2025)
From Classification to Clinical Insights: Towards Analyzing and Reasoning About Mobile and Behavioral Health Data With Large Language Models
por: Englhardt, Zachary, et al.
Publicado: (2023)
por: Englhardt, Zachary, et al.
Publicado: (2023)
FairReason: Balancing Reasoning and Social Bias in MLLMs
por: Pan, Zhenyu, et al.
Publicado: (2025)
por: Pan, Zhenyu, et al.
Publicado: (2025)
Learning Modal-Mixed Chain-of-Thought Reasoning with Latent Embeddings
por: Shao, Yifei, et al.
Publicado: (2026)
por: Shao, Yifei, et al.
Publicado: (2026)
Demystifying Instruction Mixing for Fine-tuning Large Language Models
por: Wang, Renxi, et al.
Publicado: (2023)
por: Wang, Renxi, et al.
Publicado: (2023)
Explainable AI the Latest Advancements and New Trends
por: Long, Bowen, et al.
Publicado: (2025)
por: Long, Bowen, et al.
Publicado: (2025)
Efficient Agentic Reasoning Through Self-Regulated Simulative Planning
por: Deng, Mingkai, et al.
Publicado: (2026)
por: Deng, Mingkai, et al.
Publicado: (2026)
Explore the Reasoning Capability of LLMs in the Chess Testbed
por: Wang, Shu, et al.
Publicado: (2024)
por: Wang, Shu, et al.
Publicado: (2024)
LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models
por: Hao, Shibo, et al.
Publicado: (2024)
por: Hao, Shibo, et al.
Publicado: (2024)
Reasoning Matters: Mitigate Hallucination in Multimodal Large Reasoning Models via Reasoning-Conditioned Preference Optimization
por: Kong, Jiawei, et al.
Publicado: (2026)
por: Kong, Jiawei, et al.
Publicado: (2026)
Ejemplares similares
-
Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective
por: Cheng, Zhoujun, et al.
Publicado: (2025) -
Hawkeye:Efficient Reasoning with Model Collaboration
por: She, Jianshu, et al.
Publicado: (2025) -
SplitAgent: A Privacy-Preserving Distributed Architecture for Enterprise-Cloud Agent Collaboration
por: She, Jianshu
Publicado: (2026) -
Principled Data Selection for Alignment: The Hidden Risks of Difficult Examples
por: Gao, Chengqian, et al.
Publicado: (2025) -
Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight
por: Cui, Christopher Z., et al.
Publicado: (2026)