AIPO: Learning to Reason from Active Interaction
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Junnan, Luo, Linhao, Vu, Thuy-Trang, Haffari, Gholamreza |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SituatedThinker: Grounding LLM Reasoning with Real-World through Situated Thinking
by: Liu, Junnan, et al.
Published: (2025)
by: Liu, Junnan, et al.
Published: (2025)
Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
by: Luo, Linhao, et al.
Published: (2023)
by: Luo, Linhao, et al.
Published: (2023)
ChatRule: Mining Logical Rules with Large Language Models for Knowledge Graph Reasoning
by: Luo, Linhao, et al.
Published: (2023)
by: Luo, Linhao, et al.
Published: (2023)
Towards Inference-time Scaling for Continuous Space Reasoning
by: Wang, Minghan, et al.
Published: (2025)
by: Wang, Minghan, et al.
Published: (2025)
Active Continual Learning: On Balancing Knowledge Retention and Learnability
by: Vu, Thuy-Trang, et al.
Published: (2023)
by: Vu, Thuy-Trang, et al.
Published: (2023)
Continual Learning for Large Language Models: A Survey
by: Wu, Tongtong, et al.
Published: (2024)
by: Wu, Tongtong, et al.
Published: (2024)
Mixture-of-Skills: Learning to Optimize Data Usage for Fine-Tuning Large Language Models
by: Wu, Minghao, et al.
Published: (2024)
by: Wu, Minghao, et al.
Published: (2024)
TriAlign: Towards Universal Truth Consistency in Personalized LLM Alignment
by: Nguyen, Thi-Nhung, et al.
Published: (2026)
by: Nguyen, Thi-Nhung, et al.
Published: (2026)
Direct Evaluation of Chain-of-Thought in Multi-hop Reasoning with Knowledge Graphs
by: Nguyen, Minh-Vuong, et al.
Published: (2024)
by: Nguyen, Minh-Vuong, et al.
Published: (2024)
GFM-RAG: Graph Foundation Model for Retrieval Augmented Generation
by: Luo, Linhao, et al.
Published: (2025)
by: Luo, Linhao, et al.
Published: (2025)
G-reasoner: Foundation Models for Unified Reasoning over Graph-structured Knowledge
by: Luo, Linhao, et al.
Published: (2025)
by: Luo, Linhao, et al.
Published: (2025)
The Best of Both Worlds: Bridging Quality and Diversity in Data Selection with Bipartite Graph
by: Wu, Minghao, et al.
Published: (2024)
by: Wu, Minghao, et al.
Published: (2024)
Beyond Imitation: Recovering Dense Rewards from Demonstrations
by: Li, Jiangnan, et al.
Published: (2025)
by: Li, Jiangnan, et al.
Published: (2025)
GTS: Inference-Time Scaling of Latent Reasoning with a Learnable Gaussian Thought Sampler
by: Wang, Minghan, et al.
Published: (2026)
by: Wang, Minghan, et al.
Published: (2026)
Adapting Large Language Models for Document-Level Machine Translation
by: Wu, Minghao, et al.
Published: (2024)
by: Wu, Minghao, et al.
Published: (2024)
Exploring the Potential of Multimodal LLM with Knowledge-Intensive Multimodal ASR
by: Wang, Minghan, et al.
Published: (2024)
by: Wang, Minghan, et al.
Published: (2024)
Discrete Minds in a Continuous World: Do Language Models Know Time Passes?
by: Wang, Minghan, et al.
Published: (2025)
by: Wang, Minghan, et al.
Published: (2025)
Conversational SimulMT: Efficient Simultaneous Translation with Large Language Models
by: Wang, Minghan, et al.
Published: (2024)
by: Wang, Minghan, et al.
Published: (2024)
LiveCultureBench: a Multi-Agent, Multi-Cultural Benchmark for Large Language Models in Dynamic Social Simulations
by: Pham, Viet-Thanh, et al.
Published: (2026)
by: Pham, Viet-Thanh, et al.
Published: (2026)
Assistive Large Language Model Agents for Socially-Aware Negotiation Dialogues
by: Hua, Yuncheng, et al.
Published: (2024)
by: Hua, Yuncheng, et al.
Published: (2024)
Extending LLMs to New Languages: A Case Study of Llama and Persian Adaptation
by: Sani, Samin Mahdizadeh, et al.
Published: (2024)
by: Sani, Samin Mahdizadeh, et al.
Published: (2024)
Discourse Graph Guided Document Translation with Large Language Models
by: Pham, Viet-Thanh, et al.
Published: (2025)
by: Pham, Viet-Thanh, et al.
Published: (2025)
CONGRAD:Conflicting Gradient Filtering for Multilingual Preference Alignment
by: Li, Jiangnan, et al.
Published: (2025)
by: Li, Jiangnan, et al.
Published: (2025)
Simultaneous Machine Translation with Large Language Models
by: Wang, Minghan, et al.
Published: (2023)
by: Wang, Minghan, et al.
Published: (2023)
SpeechDialogueFactory: Generating High-Quality Speech Dialogue Data to Accelerate Your Speech-LLM Development
by: Wang, Minghan, et al.
Published: (2025)
by: Wang, Minghan, et al.
Published: (2025)
SCAR: Data Selection via Style Consistency-Aware Response Ranking for Efficient Instruction-Tuning of Large Language Models
by: Li, Zhuang, et al.
Published: (2024)
by: Li, Zhuang, et al.
Published: (2024)
Scalable Frame-based Construction of Sociocultural NormBases for Socially-Aware Dialogues
by: Qu, Shilin, et al.
Published: (2024)
by: Qu, Shilin, et al.
Published: (2024)
MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
by: Liu, Akide, et al.
Published: (2024)
by: Liu, Akide, et al.
Published: (2024)
IntroLM: Introspective Language Models via Prefilling-Time Self-Evaluation
by: Kasnavieh, Hossein Hosseini, et al.
Published: (2026)
by: Kasnavieh, Hossein Hosseini, et al.
Published: (2026)
Decompose, Enrich, and Extract! Schema-aware Event Extraction using LLMs
by: Shiri, Fatemeh, et al.
Published: (2024)
by: Shiri, Fatemeh, et al.
Published: (2024)
Graph-constrained Reasoning: Faithful Reasoning on Knowledge Graphs with Large Language Models
by: Luo, Linhao, et al.
Published: (2024)
by: Luo, Linhao, et al.
Published: (2024)
RIDE: Enhancing Large Language Model Alignment through Restyled In-Context Learning Demonstration Exemplars
by: Hua, Yuncheng, et al.
Published: (2025)
by: Hua, Yuncheng, et al.
Published: (2025)
SADAS: A Dialogue Assistant System Towards Remediating Norm Violations in Bilingual Socio-Cultural Conversations
by: Hua, Yuncheng, et al.
Published: (2024)
by: Hua, Yuncheng, et al.
Published: (2024)
Dissecting Tool-Integrated Reasoning: An Empirical Study and Analysis
by: Zhao, Yufeng, et al.
Published: (2025)
by: Zhao, Yufeng, et al.
Published: (2025)
Reasoning over User Preferences: Knowledge Graph-Augmented LLMs for Explainable Conversational Recommendations
by: Qiu, Zhangchi, et al.
Published: (2024)
by: Qiu, Zhangchi, et al.
Published: (2024)
Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective
by: Liu, Junnan, et al.
Published: (2025)
by: Liu, Junnan, et al.
Published: (2025)
Are Your LLMs Capable of Stable Reasoning?
by: Liu, Junnan, et al.
Published: (2024)
by: Liu, Junnan, et al.
Published: (2024)
Rectifying LLM Thought from Lens of Optimization
by: Liu, Junnan, et al.
Published: (2025)
by: Liu, Junnan, et al.
Published: (2025)
Large Language Models-guided Dynamic Adaptation for Temporal Knowledge Graph Reasoning
by: Wang, Jiapu, et al.
Published: (2024)
by: Wang, Jiapu, et al.
Published: (2024)
MAPLE: Multi-Agent Adaptive Planning with Long-Term Memory for Table Reasoning
by: Bai, Ye, et al.
Published: (2025)
by: Bai, Ye, et al.
Published: (2025)
Similar Items
-
SituatedThinker: Grounding LLM Reasoning with Real-World through Situated Thinking
by: Liu, Junnan, et al.
Published: (2025) -
Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
by: Luo, Linhao, et al.
Published: (2023) -
ChatRule: Mining Logical Rules with Large Language Models for Knowledge Graph Reasoning
by: Luo, Linhao, et al.
Published: (2023) -
Towards Inference-time Scaling for Continuous Space Reasoning
by: Wang, Minghan, et al.
Published: (2025) -
Active Continual Learning: On Balancing Knowledge Retention and Learnability
by: Vu, Thuy-Trang, et al.
Published: (2023)