Watch Before You Answer: Learning from Visually Grounded Post-Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yuxuan, Hwang, EunJeong, Zhang, Huaisong, Du, Penghui, Jia, Yiming, Jiang, Dongfu, He, Xuan, Zhang, Shenhui, Nie, Ping, West, Peter, Allen, Kelsey R. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RewardHarness: Self-Evolving Agentic Post-Training
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2026)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2026)
BottleHumor: Self-Informed Humor Explanation using the Information Bottleneck Principle
von: Hwang, EunJeong, et al.
Veröffentlicht: (2025)
von: Hwang, EunJeong, et al.
Veröffentlicht: (2025)
Infusing Theory of Mind into Socially Intelligent LLM Agents
von: Hwang, EunJeong, et al.
Veröffentlicht: (2025)
von: Hwang, EunJeong, et al.
Veröffentlicht: (2025)
SWI: Speaking with Intent in Large Language Models
von: Yin, Yuwei, et al.
Veröffentlicht: (2025)
von: Yin, Yuwei, et al.
Veröffentlicht: (2025)
Fast-Food Intimacy: How Chinese Women Navigate Soul's AI Boyfriend
von: Lai, Huiqian, et al.
Veröffentlicht: (2026)
von: Lai, Huiqian, et al.
Veröffentlicht: (2026)
Fulfillment of the Work Games: Warehouse Workers' Experiences with Algorithmic Management
von: Cheon, EunJeong, et al.
Veröffentlicht: (2025)
von: Cheon, EunJeong, et al.
Veröffentlicht: (2025)
Is Robot Labor Labor? Delivery Robots and the Politics of Work in Public Space
von: Cheon, EunJeong, et al.
Veröffentlicht: (2026)
von: Cheon, EunJeong, et al.
Veröffentlicht: (2026)
Labor, Capital, and Machine: Toward a Labor Process Theory for HCI
von: Qin, Yigang, et al.
Veröffentlicht: (2026)
von: Qin, Yigang, et al.
Veröffentlicht: (2026)
"Hello, I'm Delivering. Let Me Pass By": Navigating Public Pathways with Walk-along with Robots in Crowded City Streets
von: Cheon, EunJeong, et al.
Veröffentlicht: (2026)
von: Cheon, EunJeong, et al.
Veröffentlicht: (2026)
Organization Matters: A Qualitative Study of Organizational Dynamics in Red Teaming Practices for Generative AI
von: Ren, Bixuan, et al.
Veröffentlicht: (2025)
von: Ren, Bixuan, et al.
Veröffentlicht: (2025)
Encountering Robotic Art: The Social, Material, and Temporal Processes of Creation with Machines
von: Qin, Yigang, et al.
Veröffentlicht: (2025)
von: Qin, Yigang, et al.
Veröffentlicht: (2025)
ClawBench: Can AI Agents Complete Everyday Online Tasks?
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2026)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2026)
Centralized Trust in Decentralized Systems: Unveiling Hidden Contradictions in Blockchain and Cryptocurrency
von: Bappy, Faisal Haque, et al.
Veröffentlicht: (2025)
von: Bappy, Faisal Haque, et al.
Veröffentlicht: (2025)
Enhancing Incremental Summarization with Structured Representations
von: Hwang, EunJeong, et al.
Veröffentlicht: (2024)
von: Hwang, EunJeong, et al.
Veröffentlicht: (2024)
VideoScore2: Think before You Score in Generative Video Evaluation
von: He, Xuan, et al.
Veröffentlicht: (2025)
von: He, Xuan, et al.
Veröffentlicht: (2025)
WAT: Online Video Understanding Needs Watching Before Thinking
von: Han, Zifan, et al.
Veröffentlicht: (2026)
von: Han, Zifan, et al.
Veröffentlicht: (2026)
Explain Before You Answer: A Survey on Compositional Visual Reasoning
von: Ke, Fucai, et al.
Veröffentlicht: (2025)
von: Ke, Fucai, et al.
Veröffentlicht: (2025)
EvolveCoder: Evolving Test Cases via Adversarial Verification for Code Reinforcement Learning
von: Ruan, Chi, et al.
Veröffentlicht: (2026)
von: Ruan, Chi, et al.
Veröffentlicht: (2026)
What Happens Before Decoding? Prefill Determines GUI Grounding in VLMs
von: Lin, Jiaping, et al.
Veröffentlicht: (2026)
von: Lin, Jiaping, et al.
Veröffentlicht: (2026)
Understand Before You Generate: Self-Guided Training for Autoregressive Image Generation
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2025)
Self-Train Before You Transcribe
von: Flynn, Robert, et al.
Veröffentlicht: (2024)
von: Flynn, Robert, et al.
Veröffentlicht: (2024)
When Safety Fails Before the Answer: Benchmarking Harmful Behavior Detection in Reasoning Chains
von: Kakkar, Ishita, et al.
Veröffentlicht: (2026)
von: Kakkar, Ishita, et al.
Veröffentlicht: (2026)
GraphDancer: Training LLMs to Explore and Reason over Graphs via Two-Stage Curriculum Post-Training
von: Bai, Yuyang, et al.
Veröffentlicht: (2026)
von: Bai, Yuyang, et al.
Veröffentlicht: (2026)
"My body is not your Porn": Identifying Trends of Harm and Oppression through a Sociotechnical Genealogy of Digital Sexual Violence in South Korea
von: Cha, Inha, et al.
Veröffentlicht: (2026)
von: Cha, Inha, et al.
Veröffentlicht: (2026)
D.Va: Validate Your Demonstration First Before You Use It
von: Zhang, Qi, et al.
Veröffentlicht: (2025)
von: Zhang, Qi, et al.
Veröffentlicht: (2025)
Listen and Chant Before You Read: The Ladder of Beauty in LM Pre-Training
von: Nomura, Yoshinori
Veröffentlicht: (2026)
von: Nomura, Yoshinori
Veröffentlicht: (2026)
Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models
von: Zou, Xin, et al.
Veröffentlicht: (2024)
von: Zou, Xin, et al.
Veröffentlicht: (2024)
Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation
von: Luo, Jiani, et al.
Veröffentlicht: (2026)
von: Luo, Jiani, et al.
Veröffentlicht: (2026)
Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM
von: Zhang, Shaoqing, et al.
Veröffentlicht: (2024)
von: Zhang, Shaoqing, et al.
Veröffentlicht: (2024)
Diffusion Language Models Know the Answer Before Decoding
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
ACECODER: Acing Coder RL via Automated Test-Case Synthesis
von: Zeng, Huaye, et al.
Veröffentlicht: (2025)
von: Zeng, Huaye, et al.
Veröffentlicht: (2025)
PEO: Improving Bi-Factorial Preference Alignment with Post-Training Policy Extrapolation
von: Liu, Yuxuan
Veröffentlicht: (2025)
von: Liu, Yuxuan
Veröffentlicht: (2025)
Post-Hoc Answer Attribution for Grounded and Trustworthy Long Document Comprehension: Task, Insights, and Challenges
von: Sancheti, Abhilasha, et al.
Veröffentlicht: (2024)
von: Sancheti, Abhilasha, et al.
Veröffentlicht: (2024)
Think Before You Diffuse: Infusing Physical Rules into Video Diffusion
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
Dr. Bench: A Multidimensional Evaluation for Deep Research Agents, from Answers to Reports
von: Yao, Yang, et al.
Veröffentlicht: (2025)
von: Yao, Yang, et al.
Veröffentlicht: (2025)
OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis
von: Li, Zhuofeng, et al.
Veröffentlicht: (2026)
von: Li, Zhuofeng, et al.
Veröffentlicht: (2026)
Think Before You Prune: Selective Self-Generated Calibration for Pruning Large Reasoning Models
von: Xiang, Yang, et al.
Veröffentlicht: (2025)
von: Xiang, Yang, et al.
Veröffentlicht: (2025)
Think Before You Drive: World Model-Inspired Multimodal Grounding for Autonomous Vehicles
von: Liao, Haicheng, et al.
Veröffentlicht: (2025)
von: Liao, Haicheng, et al.
Veröffentlicht: (2025)
PISA Experiments: Exploring Physics Post-Training for Video Diffusion Models by Watching Stuff Drop
von: Li, Chenyu, et al.
Veröffentlicht: (2025)
von: Li, Chenyu, et al.
Veröffentlicht: (2025)
General Table Question Answering via Answer-Formula Joint Generation
von: Wang, Zhongyuan, et al.
Veröffentlicht: (2025)
von: Wang, Zhongyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RewardHarness: Self-Evolving Agentic Post-Training
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2026) -
BottleHumor: Self-Informed Humor Explanation using the Information Bottleneck Principle
von: Hwang, EunJeong, et al.
Veröffentlicht: (2025) -
Infusing Theory of Mind into Socially Intelligent LLM Agents
von: Hwang, EunJeong, et al.
Veröffentlicht: (2025) -
SWI: Speaking with Intent in Large Language Models
von: Yin, Yuwei, et al.
Veröffentlicht: (2025) -
Fast-Food Intimacy: How Chinese Women Navigate Soul's AI Boyfriend
von: Lai, Huiqian, et al.
Veröffentlicht: (2026)