Stratagem: Learning Transferable Reasoning via Trajectory-Modulated Game Self-Play
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Xiachong, Yin, Deyi, Feng, Xiaocheng, Jiang, Yi, Qin, Libo, Ye, Yangfan, Huang, Lei, Ma, Weitao, Li, Qiming, Gu, Yuxuan, Qin, Bing, Kong, Lingpeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution
by: Feng, Xiachong, et al.
Published: (2026)
by: Feng, Xiachong, et al.
Published: (2026)
ImplicitMemBench: Measuring Unconscious Behavioral Adaptation in Large Language Models
by: Qin, Chonghan, et al.
Published: (2026)
by: Qin, Chonghan, et al.
Published: (2026)
Unveiling Entity-Level Unlearning for Large Language Models: A Comprehensive Analysis
by: Ma, Weitao, et al.
Published: (2024)
by: Ma, Weitao, et al.
Published: (2024)
Reasoning Does Not Necessarily Improve Role-Playing Ability
by: Feng, Xiachong, et al.
Published: (2025)
by: Feng, Xiachong, et al.
Published: (2025)
GlobeSumm: A Challenging Benchmark Towards Unifying Multi-lingual, Cross-lingual and Multi-document News Summarization
by: Ye, Yangfan, et al.
Published: (2024)
by: Ye, Yangfan, et al.
Published: (2024)
PERSONA: Dynamic and Compositional Inference-Time Personality Control via Activation Vector Algebra
by: Feng, Xiachong, et al.
Published: (2026)
by: Feng, Xiachong, et al.
Published: (2026)
Unlocking Multilingual Reasoning Capability of LLMs and LVLMs through Representation Engineering
by: Li, Qiming, et al.
Published: (2025)
by: Li, Qiming, et al.
Published: (2025)
The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models
by: Qin, Chonghan, et al.
Published: (2026)
by: Qin, Chonghan, et al.
Published: (2026)
CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs
by: Ye, Yangfan, et al.
Published: (2026)
by: Ye, Yangfan, et al.
Published: (2026)
TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs
by: Wang, Haochuan, et al.
Published: (2024)
by: Wang, Haochuan, et al.
Published: (2024)
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
by: Li, Qiming, et al.
Published: (2025)
by: Li, Qiming, et al.
Published: (2025)
Improving Contextual Faithfulness of Large Language Models via Retrieval Heads-Induced Optimization
by: Huang, Lei, et al.
Published: (2025)
by: Huang, Lei, et al.
Published: (2025)
LangGPS: Language Separability Guided Data Pre-Selection for Joint Multilingual Instruction Tuning
by: Ye, Yangfan, et al.
Published: (2025)
by: Ye, Yangfan, et al.
Published: (2025)
Adaptive Backtracking for Privacy Protection in Large Language Models
by: Yao, Zhihao, et al.
Published: (2025)
by: Yao, Zhihao, et al.
Published: (2025)
Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges
by: Ye, Yangfan, et al.
Published: (2024)
by: Ye, Yangfan, et al.
Published: (2024)
CC-Tuning: A Cross-Lingual Connection Mechanism for Improving Joint Multilingual Supervised Fine-Tuning
by: Ye, Yangfan, et al.
Published: (2025)
by: Ye, Yangfan, et al.
Published: (2025)
x1: Learning to Think Adaptively Across Languages and Cultures
by: Ye, Yangfan, et al.
Published: (2026)
by: Ye, Yangfan, et al.
Published: (2026)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
by: Zhong, Weihong, et al.
Published: (2024)
by: Zhong, Weihong, et al.
Published: (2024)
Learning Fine-Grained Grounded Citations for Attributed Large Language Models
by: Huang, Lei, et al.
Published: (2024)
by: Huang, Lei, et al.
Published: (2024)
Adapter-based Selective Knowledge Distillation for Federated Multi-domain Meeting Summarization
by: Feng, Xiachong, et al.
Published: (2023)
by: Feng, Xiachong, et al.
Published: (2023)
Fine-Mem: Fine-Grained Feedback Alignment for Long-Horizon Memory Management
by: Ma, Weitao, et al.
Published: (2026)
by: Ma, Weitao, et al.
Published: (2026)
Context-Aware Hierarchical Taxonomy Generation for Scientific Papers via LLM-Guided Multi-Aspect Clustering
by: Zhu, Kun, et al.
Published: (2025)
by: Zhu, Kun, et al.
Published: (2025)
One for All: Update Parameterized Knowledge Across Multiple Models
by: Ma, Weitao, et al.
Published: (2025)
by: Ma, Weitao, et al.
Published: (2025)
From Hypothesis to Publication: A Comprehensive Survey of AI-Driven Research Support Systems
by: Zhou, Zekun, et al.
Published: (2025)
by: Zhou, Zekun, et al.
Published: (2025)
Proxy Compression for Language Modeling
by: Zheng, Lin, et al.
Published: (2026)
by: Zheng, Lin, et al.
Published: (2026)
Can Large Language Models Simulate Human Cognition Beyond Behavioral Imitation?
by: Gu, Yuxuan, et al.
Published: (2026)
by: Gu, Yuxuan, et al.
Published: (2026)
OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model
by: Chen, Qiguang, et al.
Published: (2026)
by: Chen, Qiguang, et al.
Published: (2026)
CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering
by: Li, Qiming, et al.
Published: (2026)
by: Li, Qiming, et al.
Published: (2026)
CAI: Caption-Sensitive Attention Intervention for Mitigating Object Hallucination in Large Vision-Language Models
by: Li, Qiming, et al.
Published: (2025)
by: Li, Qiming, et al.
Published: (2025)
A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios
by: Feng, Xiachong, et al.
Published: (2024)
by: Feng, Xiachong, et al.
Published: (2024)
MPR-GUI: Benchmarking and Enhancing Multilingual Perception and Reasoning in GUI Agents
by: Chen, Ruihan, et al.
Published: (2025)
by: Chen, Ruihan, et al.
Published: (2025)
Multimodal ArXiv: A Dataset for Improving Scientific Comprehension of Large Vision-Language Models
by: Li, Lei, et al.
Published: (2024)
by: Li, Lei, et al.
Published: (2024)
Length Extrapolation of Transformers: A Survey from the Perspective of Positional Encoding
by: Zhao, Liang, et al.
Published: (2023)
by: Zhao, Liang, et al.
Published: (2023)
Culture-Aware Machine Translation in Large Language Models: Benchmarking and Investigation
by: Yuan, Zekun, et al.
Published: (2026)
by: Yuan, Zekun, et al.
Published: (2026)
Discrete Modeling via Boundary Conditional Diffusion Processes
by: Gu, Yuxuan, et al.
Published: (2024)
by: Gu, Yuxuan, et al.
Published: (2024)
Advancing Large Language Model Attribution through Self-Improving
by: Huang, Lei, et al.
Published: (2024)
by: Huang, Lei, et al.
Published: (2024)
Improving Language Model Reasoning with Self-motivated Learning
by: Feng, Yunlong, et al.
Published: (2024)
by: Feng, Yunlong, et al.
Published: (2024)
Length Controlled Generation for Black-box LLMs
by: Gu, Yuxuan, et al.
Published: (2024)
by: Gu, Yuxuan, et al.
Published: (2024)
CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
by: Ye, Zekai, et al.
Published: (2025)
by: Ye, Zekai, et al.
Published: (2025)
Reasoning Path Divergence: A New Metric and Curation Strategy to Unlock LLM Diverse Thinking
by: Ju, Feng, et al.
Published: (2025)
by: Ju, Feng, et al.
Published: (2025)
Similar Items
-
SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution
by: Feng, Xiachong, et al.
Published: (2026) -
ImplicitMemBench: Measuring Unconscious Behavioral Adaptation in Large Language Models
by: Qin, Chonghan, et al.
Published: (2026) -
Unveiling Entity-Level Unlearning for Large Language Models: A Comprehensive Analysis
by: Ma, Weitao, et al.
Published: (2024) -
Reasoning Does Not Necessarily Improve Role-Playing Ability
by: Feng, Xiachong, et al.
Published: (2025) -
GlobeSumm: A Challenging Benchmark Towards Unifying Multi-lingual, Cross-lingual and Multi-document News Summarization
by: Ye, Yangfan, et al.
Published: (2024)