Structured Preference Optimization for Vision-Language Long-Horizon Task Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Xiwen, Lin, Min, Ruan, Weiqi, Xu, Rongtao, Liu, Yuecheng, Chen, Jiaqi, Lin, Bingqian, Zhuang, Yuzheng, Liang, Xiaodan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
Unseen from Seen: Rewriting Observation-Instruction Using Foundation Models for Augmenting Vision-Language Navigation
von: Wei, Ziming, et al.
Veröffentlicht: (2025)
von: Wei, Ziming, et al.
Veröffentlicht: (2025)
CorNav: Autonomous Agent with Self-Corrected Planning for Zero-Shot Vision-and-Language Navigation
von: Liang, Xiwen, et al.
Veröffentlicht: (2023)
von: Liang, Xiwen, et al.
Veröffentlicht: (2023)
EvolveNav: Empowering LLM-Based Vision-Language Navigation via Self-Improving Embodied Reasoning
von: Lin, Bingqian, et al.
Veröffentlicht: (2025)
von: Lin, Bingqian, et al.
Veröffentlicht: (2025)
ActionSink: Toward Precise Robot Manipulation with Dynamic Integration of Action Flow
von: Guo, Shanshan, et al.
Veröffentlicht: (2025)
von: Guo, Shanshan, et al.
Veröffentlicht: (2025)
NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning
von: Lin, Bingqian, et al.
Veröffentlicht: (2024)
von: Lin, Bingqian, et al.
Veröffentlicht: (2024)
Correctable Landmark Discovery via Large Models for Vision-Language Navigation
von: Lin, Bingqian, et al.
Veröffentlicht: (2024)
von: Lin, Bingqian, et al.
Veröffentlicht: (2024)
MapGPT: Map-Guided Prompting with Adaptive Path Planning for Vision-and-Language Navigation
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
Actional Atomic-Concept Learning for Demystifying Vision-Language Navigation
von: Lin, Bingqian, et al.
Veröffentlicht: (2023)
von: Lin, Bingqian, et al.
Veröffentlicht: (2023)
HorizonBench: Long-Horizon Personalization with Evolving Preferences
von: Li, Shuyue Stella, et al.
Veröffentlicht: (2026)
von: Li, Shuyue Stella, et al.
Veröffentlicht: (2026)
Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks
von: Erdogan, Lutfi Eren, et al.
Veröffentlicht: (2025)
von: Erdogan, Lutfi Eren, et al.
Veröffentlicht: (2025)
RAT: Retrieval Augmented Thoughts Elicit Context-Aware Reasoning in Long-Horizon Generation
von: Wang, Zihao, et al.
Veröffentlicht: (2024)
von: Wang, Zihao, et al.
Veröffentlicht: (2024)
MineAnyBuild: Benchmarking Spatial Planning for Open-world AI Agents
von: Wei, Ziming, et al.
Veröffentlicht: (2025)
von: Wei, Ziming, et al.
Veröffentlicht: (2025)
The Cognitive Bandwidth Bottleneck: Shifting Long-Horizon Agent from Planning with Actions to Planning with Schemas
von: Xu, Baixuan, et al.
Veröffentlicht: (2025)
von: Xu, Baixuan, et al.
Veröffentlicht: (2025)
PhyBlock: A Progressive Benchmark for Physical Understanding and Planning via 3D Block Assembly
von: Ma, Liang, et al.
Veröffentlicht: (2025)
von: Ma, Liang, et al.
Veröffentlicht: (2025)
Temporal Preferences in Language Models for Long-Horizon Assistance
von: Mazyaki, Ali, et al.
Veröffentlicht: (2025)
von: Mazyaki, Ali, et al.
Veröffentlicht: (2025)
Psy-Chronicle:A Structured Pipeline for Synthesizing Long-Horizon Campus Psychological Counseling Dialogues
von: Gou, Chaogui, et al.
Veröffentlicht: (2026)
von: Gou, Chaogui, et al.
Veröffentlicht: (2026)
Empowering LLMs with Parameterized Skills for Adversarial Long-Horizon Planning
von: Cui, Sijia, et al.
Veröffentlicht: (2025)
von: Cui, Sijia, et al.
Veröffentlicht: (2025)
Shopping Companion: Benchmarking and Training LLM Agents for Long-Horizon Preference-Grounded E-Commerce Tasks
von: Yu, Zijian, et al.
Veröffentlicht: (2026)
von: Yu, Zijian, et al.
Veröffentlicht: (2026)
DeepPlanning: Benchmarking Long-Horizon Agentic Planning with Verifiable Constraints
von: Zhang, Yinger, et al.
Veröffentlicht: (2026)
von: Zhang, Yinger, et al.
Veröffentlicht: (2026)
ColorBench: Benchmarking Mobile Agents with Graph-Structured Framework for Complex Long-Horizon Tasks
von: Song, Yuanyi, et al.
Veröffentlicht: (2025)
von: Song, Yuanyi, et al.
Veröffentlicht: (2025)
Intrinsic Mutual Information as a Modulator for Preference Optimization
von: Liao, Peng, et al.
Veröffentlicht: (2026)
von: Liao, Peng, et al.
Veröffentlicht: (2026)
Self-Steering Optimization: Autonomous Preference Optimization for Large Language Models
von: Xiang, Hao, et al.
Veröffentlicht: (2024)
von: Xiang, Hao, et al.
Veröffentlicht: (2024)
3D-MoE: A Mixture-of-Experts Multi-modal LLM for 3D Vision and Pose Diffusion via Rectified Flow
von: Ma, Yueen, et al.
Veröffentlicht: (2025)
von: Ma, Yueen, et al.
Veröffentlicht: (2025)
MiniLongBench: The Low-cost Long Context Understanding Benchmark for Large Language Models
von: Huang, Zhongzhan, et al.
Veröffentlicht: (2025)
von: Huang, Zhongzhan, et al.
Veröffentlicht: (2025)
RADAR: Revealing Asymmetric Development of Abilities in MLLM Pre-training
von: Nie, Yunshuang, et al.
Veröffentlicht: (2026)
von: Nie, Yunshuang, et al.
Veröffentlicht: (2026)
Can We Predict Performance of Large Models across Vision-Language Tasks?
von: Zhao, Qinyu, et al.
Veröffentlicht: (2024)
von: Zhao, Qinyu, et al.
Veröffentlicht: (2024)
Localize-and-Stitch: Efficient Model Merging via Sparse Task Arithmetic
von: He, Yifei, et al.
Veröffentlicht: (2024)
von: He, Yifei, et al.
Veröffentlicht: (2024)
Agentic Aggregation for Parallel Scaling of Long-Horizon Agentic Tasks
von: Lee, Yoonsang, et al.
Veröffentlicht: (2026)
von: Lee, Yoonsang, et al.
Veröffentlicht: (2026)
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences
von: Shrivastava, Vaishnavi, et al.
Veröffentlicht: (2025)
von: Shrivastava, Vaishnavi, et al.
Veröffentlicht: (2025)
SeaPO: Strategic Error Amplification for Robust Preference Optimization of Large Language Models
von: Rao, Jun, et al.
Veröffentlicht: (2025)
von: Rao, Jun, et al.
Veröffentlicht: (2025)
LiteCoder-Terminal: Scaling Long-Horizon Terminal Environments for Learning Language Agents
von: Peng, Xiaoxuan, et al.
Veröffentlicht: (2026)
von: Peng, Xiaoxuan, et al.
Veröffentlicht: (2026)
ExpertLongBench: Benchmarking Language Models on Expert-Level Long-Form Generation Tasks with Structured Checklists
von: Ruan, Jie, et al.
Veröffentlicht: (2025)
von: Ruan, Jie, et al.
Veröffentlicht: (2025)
LoMo: Local Modality Substitution for Deeper Vision-Language Fusion
von: Han, Feng, et al.
Veröffentlicht: (2026)
von: Han, Feng, et al.
Veröffentlicht: (2026)
Mitigating Hallucinations in Large Vision-Language Models via Entity-Centric Multimodal Preference Optimization
von: Wu, Jiulong, et al.
Veröffentlicht: (2025)
von: Wu, Jiulong, et al.
Veröffentlicht: (2025)
WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences
von: Lu, Yujie, et al.
Veröffentlicht: (2024)
von: Lu, Yujie, et al.
Veröffentlicht: (2024)
A Hybrid GA LLM Framework for Structured Task Optimization
von: Shum, William, et al.
Veröffentlicht: (2025)
von: Shum, William, et al.
Veröffentlicht: (2025)
BPO: Revisiting Preference Modeling in Direct Preference Optimization
von: Sun, Lin, et al.
Veröffentlicht: (2025)
von: Sun, Lin, et al.
Veröffentlicht: (2025)
Milestone-Guided Policy Learning for Long-Horizon Language Agents
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
Vision-Language Interpreter for Robot Task Planning
von: Shirai, Keisuke, et al.
Veröffentlicht: (2023)
von: Shirai, Keisuke, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024) -
Unseen from Seen: Rewriting Observation-Instruction Using Foundation Models for Augmenting Vision-Language Navigation
von: Wei, Ziming, et al.
Veröffentlicht: (2025) -
CorNav: Autonomous Agent with Self-Corrected Planning for Zero-Shot Vision-and-Language Navigation
von: Liang, Xiwen, et al.
Veröffentlicht: (2023) -
EvolveNav: Empowering LLM-Based Vision-Language Navigation via Self-Improving Embodied Reasoning
von: Lin, Bingqian, et al.
Veröffentlicht: (2025) -
ActionSink: Toward Precise Robot Manipulation with Dynamic Integration of Action Flow
von: Guo, Shanshan, et al.
Veröffentlicht: (2025)