CHIMERA: Compact Synthetic Data for Generalizable LLM Reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhu, Xinyu, Feng, Yihao, Sun, Yanchao, Du, Xianzhi, Li, Pingzhi, Saarikivi, Olli, Zhu, Yun, Meng, Yu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Entropy-Gated Branching for Efficient Test-Time Reasoning
di: Li, Xianzhi, et al.
Pubblicazione: (2025)
di: Li, Xianzhi, et al.
Pubblicazione: (2025)
An Enhanced Prompt-Based LLM Reasoning Scheme via Knowledge Graph-Integrated Collaboration
di: Li, Yihao, et al.
Pubblicazione: (2024)
di: Li, Yihao, et al.
Pubblicazione: (2024)
TAT-LLM: A Specialized Language Model for Discrete Reasoning over Tabular and Textual Data
di: Zhu, Fengbin, et al.
Pubblicazione: (2024)
di: Zhu, Fengbin, et al.
Pubblicazione: (2024)
Quantization Hurts Reasoning? An Empirical Study on Quantized Reasoning Models
di: Liu, Ruikang, et al.
Pubblicazione: (2025)
di: Liu, Ruikang, et al.
Pubblicazione: (2025)
OckBench: Measuring the Efficiency of LLM Reasoning
di: Du, Zheng, et al.
Pubblicazione: (2025)
di: Du, Zheng, et al.
Pubblicazione: (2025)
Causal Abstraction in Model Interpretability: A Compact Survey
di: Zhang, Yihao
Pubblicazione: (2024)
di: Zhang, Yihao
Pubblicazione: (2024)
LLM-Driven Multi-Turn Task-Oriented Dialogue Synthesis for Realistic Reasoning
di: Zhu, Yu, et al.
Pubblicazione: (2026)
di: Zhu, Yu, et al.
Pubblicazione: (2026)
Towards Generalizable and Faithful Logic Reasoning over Natural Language via Resolution Refutation
di: Sun, Zhouhao, et al.
Pubblicazione: (2024)
di: Sun, Zhouhao, et al.
Pubblicazione: (2024)
Few-shot LLM Synthetic Data with Distribution Matching
di: Ren, Jiyuan, et al.
Pubblicazione: (2025)
di: Ren, Jiyuan, et al.
Pubblicazione: (2025)
Unleashing LLM Reasoning Capability via Scalable Question Synthesis from Scratch
di: Ding, Yuyang, et al.
Pubblicazione: (2024)
di: Ding, Yuyang, et al.
Pubblicazione: (2024)
SafeCtrl-RL: Inference-Time Adaptive Behaviour Control for LLM Dialogue via RL-Driven Prompt Optimisation
di: Orme, Michael, et al.
Pubblicazione: (2026)
di: Orme, Michael, et al.
Pubblicazione: (2026)
Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning
di: Zhou, Cai, et al.
Pubblicazione: (2026)
di: Zhou, Cai, et al.
Pubblicazione: (2026)
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models
di: Wang, Kai, et al.
Pubblicazione: (2025)
di: Wang, Kai, et al.
Pubblicazione: (2025)
Contextual Drag: How Errors in the Context Affect LLM Reasoning
di: Cheng, Yun, et al.
Pubblicazione: (2026)
di: Cheng, Yun, et al.
Pubblicazione: (2026)
The Impact of Language Mixing on Bilingual LLM Reasoning
di: Li, Yihao, et al.
Pubblicazione: (2025)
di: Li, Yihao, et al.
Pubblicazione: (2025)
Critique-GRPO: Advancing LLM Reasoning with Natural Language and Numerical Feedback
di: Zhang, Xiaoying, et al.
Pubblicazione: (2025)
di: Zhang, Xiaoying, et al.
Pubblicazione: (2025)
Is LLMs Hallucination Usable? LLM-based Negative Reasoning for Fake News Detection
di: Zhang, Chaowei, et al.
Pubblicazione: (2025)
di: Zhang, Chaowei, et al.
Pubblicazione: (2025)
From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning
di: Wang, Shaojie, et al.
Pubblicazione: (2026)
di: Wang, Shaojie, et al.
Pubblicazione: (2026)
Towards Foundation Models for Knowledge Graph Reasoning
di: Galkin, Mikhail, et al.
Pubblicazione: (2023)
di: Galkin, Mikhail, et al.
Pubblicazione: (2023)
MindMerger: Efficient Boosting LLM Reasoning in non-English Languages
di: Huang, Zixian, et al.
Pubblicazione: (2024)
di: Huang, Zixian, et al.
Pubblicazione: (2024)
Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools
di: Wu, Junde, et al.
Pubblicazione: (2025)
di: Wu, Junde, et al.
Pubblicazione: (2025)
CrowdSelect: Synthetic Instruction Data Selection with Multi-LLM Wisdom
di: Li, Yisen, et al.
Pubblicazione: (2025)
di: Li, Yisen, et al.
Pubblicazione: (2025)
Generalizable End-to-End Tool-Use RL with Synthetic CodeGym
di: Du, Weihua, et al.
Pubblicazione: (2025)
di: Du, Weihua, et al.
Pubblicazione: (2025)
Propensity Inference: Environmental Contributors to LLM Behaviour
di: Järviniemi, Olli, et al.
Pubblicazione: (2026)
di: Järviniemi, Olli, et al.
Pubblicazione: (2026)
Does Math Reasoning Improve General LLM Capabilities? Understanding Transferability of LLM Reasoning
di: Huan, Maggie, et al.
Pubblicazione: (2025)
di: Huan, Maggie, et al.
Pubblicazione: (2025)
TARGA: Targeted Synthetic Data Generation for Practical Reasoning over Structured Data
di: Huang, Xiang, et al.
Pubblicazione: (2024)
di: Huang, Xiang, et al.
Pubblicazione: (2024)
Chat-TS: Enhancing Multi-Modal Reasoning Over Time-Series and Natural Language Data
di: Quinlan, Paul, et al.
Pubblicazione: (2025)
di: Quinlan, Paul, et al.
Pubblicazione: (2025)
HGOT: Hierarchical Graph of Thoughts for Retrieval-Augmented In-Context Learning in Factuality Evaluation
di: Fang, Yihao, et al.
Pubblicazione: (2024)
di: Fang, Yihao, et al.
Pubblicazione: (2024)
Table as Thought: Exploring Structured Thoughts in LLM Reasoning
di: Sun, Zhenjie, et al.
Pubblicazione: (2025)
di: Sun, Zhenjie, et al.
Pubblicazione: (2025)
Know When To Fold 'Em: Token-Efficient LLM Synthetic Data Generation via Multi-Stage In-Flight Rejection
di: Chowdhury, Anjir Ahmed, et al.
Pubblicazione: (2026)
di: Chowdhury, Anjir Ahmed, et al.
Pubblicazione: (2026)
Mirror: A Multiple-perspective Self-Reflection Method for Knowledge-rich Reasoning
di: Yan, Hanqi, et al.
Pubblicazione: (2024)
di: Yan, Hanqi, et al.
Pubblicazione: (2024)
OpenSeeker: Democratizing Frontier Search Agents by Fully Open-Sourcing Training Data
di: Du, Yuwen, et al.
Pubblicazione: (2026)
di: Du, Yuwen, et al.
Pubblicazione: (2026)
PortLLM: Personalizing Evolving Large Language Models with Training-Free and Portable Model Patches
di: Khan, Rana Muhammad Shahroz, et al.
Pubblicazione: (2024)
di: Khan, Rana Muhammad Shahroz, et al.
Pubblicazione: (2024)
Learning from Response not Preference: A Stackelberg Approach for LLM Detoxification using Non-parallel Data
di: Xie, Xinhong, et al.
Pubblicazione: (2024)
di: Xie, Xinhong, et al.
Pubblicazione: (2024)
NCV: A Node-Wise Consistency Verification Approach for Low-Cost Structured Error Localization in LLM Reasoning
di: Zhang, Yulong, et al.
Pubblicazione: (2025)
di: Zhang, Yulong, et al.
Pubblicazione: (2025)
Confidence-Calibrated Small-Large Language Model Collaboration for Cost-Efficient Reasoning
di: Zhang, Chuang, et al.
Pubblicazione: (2026)
di: Zhang, Chuang, et al.
Pubblicazione: (2026)
Scaling Up, Speeding Up: A Benchmark of Speculative Decoding for Efficient LLM Test-Time Scaling
di: Sun, Shengyin, et al.
Pubblicazione: (2025)
di: Sun, Shengyin, et al.
Pubblicazione: (2025)
Reasoning-Driven Synthetic Data Generation and Evaluation
di: Davidson, Tim R., et al.
Pubblicazione: (2026)
di: Davidson, Tim R., et al.
Pubblicazione: (2026)
On the Step Length Confounding in LLM Reasoning Data Selection
di: Wang, Bing, et al.
Pubblicazione: (2026)
di: Wang, Bing, et al.
Pubblicazione: (2026)
X-Reasoner: Towards Generalizable Reasoning Across Modalities and Domains
di: Liu, Qianchu, et al.
Pubblicazione: (2025)
di: Liu, Qianchu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Entropy-Gated Branching for Efficient Test-Time Reasoning
di: Li, Xianzhi, et al.
Pubblicazione: (2025) -
An Enhanced Prompt-Based LLM Reasoning Scheme via Knowledge Graph-Integrated Collaboration
di: Li, Yihao, et al.
Pubblicazione: (2024) -
TAT-LLM: A Specialized Language Model for Discrete Reasoning over Tabular and Textual Data
di: Zhu, Fengbin, et al.
Pubblicazione: (2024) -
Quantization Hurts Reasoning? An Empirical Study on Quantized Reasoning Models
di: Liu, Ruikang, et al.
Pubblicazione: (2025) -
OckBench: Measuring the Efficiency of LLM Reasoning
di: Du, Zheng, et al.
Pubblicazione: (2025)