TCPO: Thought-Centric Preference Optimization for Effective Embodied Decision-making
Fuente:
arXiv
Saved in:
| Main Authors: | Jiao, Kechen, Fang, Zhirui, Liu, Jiahao, Li, Bei, Wang, Qifan, Liu, Xinyu, Ruan, Junhao, Qiao, Zhongjian, Zhu, Yifan, Xu, Yaxin, Wang, Jingang, Li, Xiu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IIET: Efficient Numerical Transformer via Implicit Iterative Euler Method
by: Liu, Xinyu, et al.
Published: (2025)
by: Liu, Xinyu, et al.
Published: (2025)
SUMO: Search-Based Uncertainty Estimation for Model-Based Offline Reinforcement Learning
by: Qiao, Zhongjian, et al.
Published: (2024)
by: Qiao, Zhongjian, et al.
Published: (2024)
MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks
by: Ruan, Junhao, et al.
Published: (2026)
by: Ruan, Junhao, et al.
Published: (2026)
Speculative Decoding via Early-exiting for Faster LLM Inference with Thompson Sampling Control Mechanism
by: Liu, Jiahao, et al.
Published: (2024)
by: Liu, Jiahao, et al.
Published: (2024)
Mind the Model, Not the Agent: The Primacy Bias in Model-based RL
by: Qiao, Zhongjian, et al.
Published: (2023)
by: Qiao, Zhongjian, et al.
Published: (2023)
Numerical Approach for On-the-Fly Active Flow Control via Flow Map Learning Method
by: Liu, Xinyu, et al.
Published: (2026)
by: Liu, Xinyu, et al.
Published: (2026)
Dynamic Fisher-weighted Model Merging via Bayesian Optimization
by: Lee, Sanwoo, et al.
Published: (2025)
by: Lee, Sanwoo, et al.
Published: (2025)
C-ICL: Contrastive In-context Learning for Information Extraction
by: Mo, Ying, et al.
Published: (2024)
by: Mo, Ying, et al.
Published: (2024)
mCL-NER: Cross-Lingual Named Entity Recognition via Multi-view Contrastive Learning
by: Mo, Ying, et al.
Published: (2023)
by: Mo, Ying, et al.
Published: (2023)
VLP: Vision-Language Preference Learning for Embodied Manipulation
by: Liu, Runze, et al.
Published: (2025)
by: Liu, Runze, et al.
Published: (2025)
WHALE: Towards Generalizable and Scalable World Models for Embodied Decision-making
by: Zhang, Zhilong, et al.
Published: (2024)
by: Zhang, Zhilong, et al.
Published: (2024)
Parallel Decoding via Hidden Transfer for Lossless Large Language Model Acceleration
by: Wu, Pengfei, et al.
Published: (2024)
by: Wu, Pengfei, et al.
Published: (2024)
FIRP: Faster LLM inference via future intermediate representation prediction
by: Wu, Pengfei, et al.
Published: (2024)
by: Wu, Pengfei, et al.
Published: (2024)
MiniDisc: Minimal Distillation Schedule for Language Model Compression
by: Zhang, Chen, et al.
Published: (2022)
by: Zhang, Chen, et al.
Published: (2022)
Earlier Tokens Contribute More: Learning Direct Preference Optimization From Temporal Decay Perspective
by: Shao, Ruichen, et al.
Published: (2025)
by: Shao, Ruichen, et al.
Published: (2025)
DecisionNCE: Embodied Multimodal Representations via Implicit Preference Learning
by: Li, Jianxiong, et al.
Published: (2024)
by: Li, Jianxiong, et al.
Published: (2024)
When Transformers Meet Recommenders: Integrating Self-Attentive Sequential Recommendation with Fine-Tuned LLMs
by: Liu, Kechen
Published: (2025)
by: Liu, Kechen
Published: (2025)
CGoT: A Novel Inference Mechanism for Embodied Multi-Agent Systems Using Composable Graphs of Thoughts
by: Nie, Yixiao, et al.
Published: (2025)
by: Nie, Yixiao, et al.
Published: (2025)
ReMamba: Equip Mamba with Effective Long-Sequence Modeling
by: Yuan, Danlong, et al.
Published: (2024)
by: Yuan, Danlong, et al.
Published: (2024)
Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
by: Li, Manling, et al.
Published: (2024)
by: Li, Manling, et al.
Published: (2024)
StreetviewLLM: Extracting Geographic Information Using a Chain-of-Thought Multimodal Large Language Model
by: Li, Zongrong, et al.
Published: (2024)
by: Li, Zongrong, et al.
Published: (2024)
EAVE: Efficient Product Attribute Value Extraction via Lightweight Sparse-layer Interaction
by: Yang, Li, et al.
Published: (2024)
by: Yang, Li, et al.
Published: (2024)
Causal Autoregressive Diffusion Language Model
by: Ruan, Junhao, et al.
Published: (2026)
by: Ruan, Junhao, et al.
Published: (2026)
Libra: Assessing and Improving Reward Model by Learning to Think
by: Zhou, Meng, et al.
Published: (2025)
by: Zhou, Meng, et al.
Published: (2025)
Assistant-Guided Mitigation of Teacher Preference Bias in LLM-as-a-Judge
by: Liu, Zhuo, et al.
Published: (2025)
by: Liu, Zhuo, et al.
Published: (2025)
Chain of Event-Centric Causal Thought for Physically Plausible Video Generation
by: Wang, Zixuan, et al.
Published: (2026)
by: Wang, Zixuan, et al.
Published: (2026)
Forgetting Curve: A Reliable Method for Evaluating Memorization Capability for Long-context Models
by: Liu, Xinyu, et al.
Published: (2024)
by: Liu, Xinyu, et al.
Published: (2024)
Last-Layer-Centric Feature Recombination: Unleashing 3D Geometric Knowledge in DINOv3 for Monocular Depth Estimation
by: Wang, Gongshu, et al.
Published: (2026)
by: Wang, Gongshu, et al.
Published: (2026)
NDP: Next Distribution Prediction as a More Broad Target
by: Ruan, Junhao, et al.
Published: (2024)
by: Ruan, Junhao, et al.
Published: (2024)
Embodied Tree of Thoughts: Deliberate Manipulation Planning with Embodied World Model
by: Xu, Wenjiang, et al.
Published: (2025)
by: Xu, Wenjiang, et al.
Published: (2025)
Temporal Difference Learning with Constrained Initial Representations
by: Lyu, Jiafei, et al.
Published: (2026)
by: Lyu, Jiafei, et al.
Published: (2026)
VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI
by: Cheng, Sijie, et al.
Published: (2024)
by: Cheng, Sijie, et al.
Published: (2024)
Prescribing Decision Conservativeness in Two-Stage Power Markets: A Distributionally Robust End-to-End Approach
by: Liang, Zhirui, et al.
Published: (2024)
by: Liang, Zhirui, et al.
Published: (2024)
Exploring the Robustness of Decision-Level Through Adversarial Attacks on LLM-Based Embodied Models
by: Liu, Shuyuan, et al.
Published: (2024)
by: Liu, Shuyuan, et al.
Published: (2024)
Argus: Vision-Centric Reasoning with Grounded Chain-of-Thought
by: Man, Yunze, et al.
Published: (2025)
by: Man, Yunze, et al.
Published: (2025)
Unified Predefined‐Time Stability Theorem and Sliding Mode Control for Fractional‐Order Nonlinear Systems
by: Jingang Liu, et al.
Published: (2025)
by: Jingang Liu, et al.
Published: (2025)
Refine Thought: A Test-Time Inference Method for Embedding Model Reasoning
by: Wang, Guangzhi, et al.
Published: (2025)
by: Wang, Guangzhi, et al.
Published: (2025)
MemoSight: Unifying Context Compression and Multi Token Prediction for Reasoning Acceleration
by: Liu, Xinyu, et al.
Published: (2026)
by: Liu, Xinyu, et al.
Published: (2026)
Event-Centric World Modeling with Memory-Augmented Retrieval for Embodied Decision-Making
by: Fan, Zhaowen, et al.
Published: (2026)
by: Fan, Zhaowen, et al.
Published: (2026)
Towards Anatomically Plausible Human Image Generation via Synthetic Localized Preferences
by: Li, Bao, et al.
Published: (2026)
by: Li, Bao, et al.
Published: (2026)
Similar Items
-
IIET: Efficient Numerical Transformer via Implicit Iterative Euler Method
by: Liu, Xinyu, et al.
Published: (2025) -
SUMO: Search-Based Uncertainty Estimation for Model-Based Offline Reinforcement Learning
by: Qiao, Zhongjian, et al.
Published: (2024) -
MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks
by: Ruan, Junhao, et al.
Published: (2026) -
Speculative Decoding via Early-exiting for Faster LLM Inference with Thompson Sampling Control Mechanism
by: Liu, Jiahao, et al.
Published: (2024) -
Mind the Model, Not the Agent: The Primacy Bias in Model-based RL
by: Qiao, Zhongjian, et al.
Published: (2023)