Reverse Thinking Makes LLMs Stronger Reasoners
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Justin Chih-Yao, Wang, Zifeng, Palangi, Hamid, Han, Rujun, Ebrahimi, Sayna, Le, Long, Perot, Vincent, Mishra, Swaroop, Bansal, Mohit, Lee, Chen-Yu, Pfister, Tomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PlanGEN: A Multi-Agent Framework for Generating Planning and Reasoning Trajectories for Complex Problem Solving
von: Parmar, Mihir, et al.
Veröffentlicht: (2025)
von: Parmar, Mihir, et al.
Veröffentlicht: (2025)
ReConcile: Round-Table Conference Improves Reasoning via Consensus among Diverse LLMs
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2023)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2023)
Model Swarms: Collaborative Search to Adapt LLM Experts via Swarm Intelligence
von: Feng, Shangbin, et al.
Veröffentlicht: (2024)
von: Feng, Shangbin, et al.
Veröffentlicht: (2024)
CROME: Cross-Modal Adapters for Efficient Multimodal LLM
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2024)
LANISTR: Multimodal Learning from Structured and Unstructured Data
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2023)
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2023)
Speculative RAG: Enhancing Retrieval Augmented Generation through Drafting
von: Wang, Zilong, et al.
Veröffentlicht: (2024)
von: Wang, Zilong, et al.
Veröffentlicht: (2024)
CodecLM: Aligning Language Models with Tailored Synthetic Data
von: Wang, Zifeng, et al.
Veröffentlicht: (2024)
von: Wang, Zifeng, et al.
Veröffentlicht: (2024)
ASPEST: Bridging the Gap Between Active Learning and Selective Prediction
von: Chen, Jiefeng, et al.
Veröffentlicht: (2023)
von: Chen, Jiefeng, et al.
Veröffentlicht: (2023)
In Prospect and Retrospect: Reflective Memory Management for Long-term Personalized Dialogue Agents
von: Tan, Zhen, et al.
Veröffentlicht: (2025)
von: Tan, Zhen, et al.
Veröffentlicht: (2025)
Mitigating Object Hallucination in MLLMs via Data-augmented Phrase-level Alignment
von: Sarkar, Pritam, et al.
Veröffentlicht: (2024)
von: Sarkar, Pritam, et al.
Veröffentlicht: (2024)
VeriGuard: Enhancing LLM Agent Safety via Verified Code Generation
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
Think Right: Learning to Mitigate Under-Over Thinking via Adaptive, Attentive Compression
von: Singh, Joykirat, et al.
Veröffentlicht: (2025)
von: Singh, Joykirat, et al.
Veröffentlicht: (2025)
MAGDi: Structured Distillation of Multi-Agent Interaction Graphs Improves Reasoning in Smaller Language Models
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2024)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2024)
Chain-of-Table: Evolving Tables in the Reasoning Chain for Table Understanding
von: Wang, Zilong, et al.
Veröffentlicht: (2024)
von: Wang, Zilong, et al.
Veröffentlicht: (2024)
HEART: Emotionally-Driven Test-Time Scaling of Language Models
von: Pinto, Gabriela, et al.
Veröffentlicht: (2025)
von: Pinto, Gabriela, et al.
Veröffentlicht: (2025)
TFRBench: A Reasoning Benchmark for Evaluating Forecasting Systems
von: Ahamed, Md Atik, et al.
Veröffentlicht: (2026)
von: Ahamed, Md Atik, et al.
Veröffentlicht: (2026)
Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2025)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2025)
Diversity of Thought Improves Reasoning Abilities of LLMs
von: Naik, Ranjita, et al.
Veröffentlicht: (2023)
von: Naik, Ranjita, et al.
Veröffentlicht: (2023)
ScholarPeer: A Context-Aware Multi-Agent Framework for Automated Peer Review
von: Goyal, Palash, et al.
Veröffentlicht: (2026)
von: Goyal, Palash, et al.
Veröffentlicht: (2026)
Magnet: Multi-turn Tool-use Data Synthesis and Distillation via Graph Translation
von: Yin, Fan, et al.
Veröffentlicht: (2025)
von: Yin, Fan, et al.
Veröffentlicht: (2025)
Supervised Reinforcement Learning: From Expert Trajectories to Step-wise Reasoning
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
Towards Compute-Optimal Many-Shot In-Context Learning
von: Golchin, Shahriar, et al.
Veröffentlicht: (2025)
von: Golchin, Shahriar, et al.
Veröffentlicht: (2025)
MAgICoRe: Multi-Agent, Iterative, Coarse-to-Fine Refinement for Reasoning
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2024)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2024)
Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers
von: Qi, Zhenting, et al.
Veröffentlicht: (2024)
von: Qi, Zhenting, et al.
Veröffentlicht: (2024)
RubricEM: Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards
von: Li, Gaotang, et al.
Veröffentlicht: (2026)
von: Li, Gaotang, et al.
Veröffentlicht: (2026)
Speculative Knowledge Distillation: Bridging the Teacher-Student Gap Through Interleaved Sampling
von: Xu, Wenda, et al.
Veröffentlicht: (2024)
von: Xu, Wenda, et al.
Veröffentlicht: (2024)
When Can LLMs Learn to Reason with Weak Supervision?
von: Rahman, Salman, et al.
Veröffentlicht: (2026)
von: Rahman, Salman, et al.
Veröffentlicht: (2026)
Ontology-Guided Reverse Thinking Makes Large Language Models Stronger on Knowledge Graph Question Answering
von: Liu, Runxuan, et al.
Veröffentlicht: (2025)
von: Liu, Runxuan, et al.
Veröffentlicht: (2025)
Heterogeneous Swarms: Jointly Optimizing Model Roles and Weights for Multi-LLM Systems
von: Feng, Shangbin, et al.
Veröffentlicht: (2025)
von: Feng, Shangbin, et al.
Veröffentlicht: (2025)
Exploring Group and Symmetry Principles in Large Language Models
von: Imani, Shima, et al.
Veröffentlicht: (2024)
von: Imani, Shima, et al.
Veröffentlicht: (2024)
MAMM-Refine: A Recipe for Improving Faithfulness in Generation with Multi-Agent Collaboration
von: Wan, David, et al.
Veröffentlicht: (2025)
von: Wan, David, et al.
Veröffentlicht: (2025)
ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory
von: Ouyang, Siru, et al.
Veröffentlicht: (2025)
von: Ouyang, Siru, et al.
Veröffentlicht: (2025)
Watch and Learn: Learning to Use Computers from Online Videos
von: Song, Chan Hee, et al.
Veröffentlicht: (2025)
von: Song, Chan Hee, et al.
Veröffentlicht: (2025)
SAPO: Self-Adaptive Process Optimization Makes Small Reasoners Stronger
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2026)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2026)
DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning
von: Sivakumaran, Nithin, et al.
Veröffentlicht: (2025)
von: Sivakumaran, Nithin, et al.
Veröffentlicht: (2025)
LLM-Based Multi-Agent Blackboard System for Information Discovery in Data Science
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
PLAN-TUNING: Post-Training Language Models to Learn Step-by-Step Planning for Complex Problem Solving
von: Parmar, Mihir, et al.
Veröffentlicht: (2025)
von: Parmar, Mihir, et al.
Veröffentlicht: (2025)
CaLM: Contrasting Large and Small Language Models to Verify Grounded Generation
von: Hsu, I-Hung, et al.
Veröffentlicht: (2024)
von: Hsu, I-Hung, et al.
Veröffentlicht: (2024)
Synapse: Adaptive Arbitration of Complementary Expertise in Time Series Foundational Models
von: Das, Sarkar Snigdha Sarathi, et al.
Veröffentlicht: (2025)
von: Das, Sarkar Snigdha Sarathi, et al.
Veröffentlicht: (2025)
Cog-DRIFT: Exploration on Adaptively Reformulated Instances Enables Learning from Hard Reasoning Problems
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2026)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PlanGEN: A Multi-Agent Framework for Generating Planning and Reasoning Trajectories for Complex Problem Solving
von: Parmar, Mihir, et al.
Veröffentlicht: (2025) -
ReConcile: Round-Table Conference Improves Reasoning via Consensus among Diverse LLMs
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2023) -
Model Swarms: Collaborative Search to Adapt LLM Experts via Swarm Intelligence
von: Feng, Shangbin, et al.
Veröffentlicht: (2024) -
CROME: Cross-Modal Adapters for Efficient Multimodal LLM
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2024) -
LANISTR: Multimodal Learning from Structured and Unstructured Data
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2023)