One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yinghui, Kuang, Jiayi, Huang, Haojing, Xu, Zhikun, Liang, Xinnian, Yu, Yi, Lu, Wenlian, Li, Yangning, Tan, Xiaoyu, Qu, Chao, Shen, Ying, Zheng, Hai-Tao, Yu, Philip S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Atomic Thinking of LLMs: Decoupling and Exploring Mathematical Reasoning Abilities
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025)
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025)
Refine Knowledge of Large Language Models via Adaptive Contrastive Learning
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
Mitigating Catastrophic Forgetting in Multi-domain Chinese Spelling Correction by Multi-stage Knowledge Transfer Framework
von: Xing, Peng, et al.
Veröffentlicht: (2024)
von: Xing, Peng, et al.
Veröffentlicht: (2024)
TangramPuzzle: Evaluating Multimodal Large Language Models with Compositional Spatial Reasoning
von: Liu, Daixian, et al.
Veröffentlicht: (2026)
von: Liu, Daixian, et al.
Veröffentlicht: (2026)
Process-Level Trajectory Evaluation for Environment Configuration in Software Engineering Agents
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025)
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025)
Rethinking the Roles of Large Language Models in Chinese Grammatical Error Correction
von: Li, Yinghui, et al.
Veröffentlicht: (2024)
von: Li, Yinghui, et al.
Veröffentlicht: (2024)
Correct Like Humans: Progressive Learning Framework for Chinese Text Error Correction
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
On the (In)Effectiveness of Large Language Models for Chinese Text Correction
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
Examples and Counterexamples
Veröffentlicht: (2021)
Veröffentlicht: (2021)
Let LLMs Take on the Latest Challenges! A Chinese Dynamic Question Answering Benchmark
von: Xu, Zhikun, et al.
Veröffentlicht: (2024)
von: Xu, Zhikun, et al.
Veröffentlicht: (2024)
When LLMs Meet Cunning Texts: A Fallacy Understanding Benchmark for Large Language Models
von: Li, Yinghui, et al.
Veröffentlicht: (2024)
von: Li, Yinghui, et al.
Veröffentlicht: (2024)
Embracing Ambiguity: Improving Similarity-oriented Tasks with Contextual Synonym Knowledge
von: Li, Yangning, et al.
Veröffentlicht: (2022)
von: Li, Yangning, et al.
Veröffentlicht: (2022)
EvoConfig: Self-Evolving Multi-Agent Systems for Efficient Autonomous Environment Configuration
von: Guo, Xinshuai, et al.
Veröffentlicht: (2026)
von: Guo, Xinshuai, et al.
Veröffentlicht: (2026)
DAST: Context-Aware Compression in LLMs via Dynamic Allocation of Soft Tokens
von: Chen, Shaoshen, et al.
Veröffentlicht: (2025)
von: Chen, Shaoshen, et al.
Veröffentlicht: (2025)
CORE: Concept-Oriented Reinforcement for Bridging the Definition-Application Gap in Mathematical Reasoning
von: Gao, Zijun, et al.
Veröffentlicht: (2025)
von: Gao, Zijun, et al.
Veröffentlicht: (2025)
MDIT: A Model-free Data Interpolation Method for Diverse Instruction Tuning
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
From Retrieval to Generation: Efficient and Effective Entity Set Expansion
von: Huang, Shulin, et al.
Veröffentlicht: (2023)
von: Huang, Shulin, et al.
Veröffentlicht: (2023)
Teaching According to Talents! Instruction Tuning LLMs with Competence-Aware Curriculum Learning
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
ORIGAMISPACE: Benchmarking Multimodal LLMs in Multi-Step Spatial Reasoning with Mathematical Constraints
von: Xu, Rui, et al.
Veröffentlicht: (2025)
von: Xu, Rui, et al.
Veröffentlicht: (2025)
Get Shown the Light
von: Kaler, Michael
Veröffentlicht: (2024)
von: Kaler, Michael
Veröffentlicht: (2024)
AdmTree: Compressing Lengthy Context with Adaptive Semantic Trees
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
UltraWiki: Ultra-fine-grained Entity Set Expansion with Negative Seed Entities
von: Li, Yangning, et al.
Veröffentlicht: (2024)
von: Li, Yangning, et al.
Veröffentlicht: (2024)
Examples and Counterexamples of Cost-efficiency in Incomplete Markets
von: Bernard, Carole, et al.
Veröffentlicht: (2024)
von: Bernard, Carole, et al.
Veröffentlicht: (2024)
Exploring the Implicit Semantic Ability of Multimodal Large Language Models: A Pilot Study on Entity Set Expansion
von: Wang, Hebin, et al.
Veröffentlicht: (2024)
von: Wang, Hebin, et al.
Veröffentlicht: (2024)
Cognitive Mismatch in Multimodal Large Language Models for Discrete Symbol Understanding
von: Li, Yinghui, et al.
Veröffentlicht: (2026)
von: Li, Yinghui, et al.
Veröffentlicht: (2026)
Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples
von: Yu, Fangxu, et al.
Veröffentlicht: (2024)
von: Yu, Fangxu, et al.
Veröffentlicht: (2024)
Nemotron Elastic: Towards Efficient Many-in-One Reasoning LLMs
von: Taghibakhshi, Ali, et al.
Veröffentlicht: (2025)
von: Taghibakhshi, Ali, et al.
Veröffentlicht: (2025)
Bidirectional End-to-End Learning of Retriever-Reader Paradigm for Entity Linking
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
von: Li, Yinghui, et al.
Veröffentlicht: (2023)
ProtoReasoning: Prototypes as the Foundation for Generalizable Reasoning in LLMs
von: He, Feng, et al.
Veröffentlicht: (2025)
von: He, Feng, et al.
Veröffentlicht: (2025)
Star Elastic: Many-in-One Reasoning LLMs with Efficient Budget Control
von: Taghibakhshi, Ali, et al.
Veröffentlicht: (2026)
von: Taghibakhshi, Ali, et al.
Veröffentlicht: (2026)
The Counterexample Game: Iterated Conceptual Analysis and Repair in Language Models
von: Drucker, Daniel, et al.
Veröffentlicht: (2026)
von: Drucker, Daniel, et al.
Veröffentlicht: (2026)
ProductAgent: Benchmarking Conversational Product Search Agent with Asking Clarification Questions
von: Ye, Jingheng, et al.
Veröffentlicht: (2024)
von: Ye, Jingheng, et al.
Veröffentlicht: (2024)
Mathematical Understanding and the Role of Counterexamples and Pathologies: A Case Study in Mathematical Analysis
von: Carmen Martínez-Adame
Veröffentlicht: (2018)
von: Carmen Martínez-Adame
Veröffentlicht: (2018)
Evaluating the Unseen Capabilities: How Many Theorems Do LLMs Know?
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
On the Equivalence of Synchronous Coordination Game and Asynchronous Coordination Design
von: Pan, Xinnian Kazusa
Veröffentlicht: (2024)
von: Pan, Xinnian Kazusa
Veröffentlicht: (2024)
Thought-Like-Pro: Enhancing Reasoning of Large Language Models through Self-Driven Prolog-based Chain-of-Thought
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024)
Order-preserving condition for coherence measures of projective measurements with One Example
von: Wang, Hai
Veröffentlicht: (2025)
von: Wang, Hai
Veröffentlicht: (2025)
One Does Not Simply Meme Alone: Evaluating Co-Creativity Between LLMs and Humans in the Generation of Humor
von: Wu, Zhikun, et al.
Veröffentlicht: (2025)
von: Wu, Zhikun, et al.
Veröffentlicht: (2025)
Automated Optimization Modeling via a Localizable Error-Driven Perspective
von: Liu, Weiting, et al.
Veröffentlicht: (2026)
von: Liu, Weiting, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Atomic Thinking of LLMs: Decoupling and Exploring Mathematical Reasoning Abilities
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025) -
Refine Knowledge of Large Language Models via Adaptive Contrastive Learning
von: Li, Yinghui, et al.
Veröffentlicht: (2025) -
Mitigating Catastrophic Forgetting in Multi-domain Chinese Spelling Correction by Multi-stage Knowledge Transfer Framework
von: Xing, Peng, et al.
Veröffentlicht: (2024) -
TangramPuzzle: Evaluating Multimodal Large Language Models with Compositional Spatial Reasoning
von: Liu, Daixian, et al.
Veröffentlicht: (2026) -
Process-Level Trajectory Evaluation for Environment Configuration in Software Engineering Agents
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025)