Temporal Evidence Routing with Structured Visual Evidence for TimeLogicQA
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Yuyang, Wu, Yongliang, Zhu, Xingyu, Chen, Yuxia, Jiang, Zhenxiang, Ji, Yangguang, Zhu, Wenbo, Shi, Yanxi, Wu, Jay, Wang, Shuo, Yang, Xu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Dense Evidence Refinement for Video Relational Reasoning for VRR-QA Challenge
von: Sun, Yuyang, et al.
Veröffentlicht: (2026)
von: Sun, Yuyang, et al.
Veröffentlicht: (2026)
RSVP: Reasoning Segmentation via Visual Prompting and Multi-modal Chain-of-Thought
von: Lu, Yi, et al.
Veröffentlicht: (2025)
von: Lu, Yi, et al.
Veröffentlicht: (2025)
Number it: Temporal Grounding Videos like Flipping Manga
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
VEU-Bench: Towards Comprehensive Understanding of Video Editing
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
Reframe Anything: LLM Agent for Open World Video Reframing
von: Cao, Jiawang, et al.
Veröffentlicht: (2024)
von: Cao, Jiawang, et al.
Veröffentlicht: (2024)
TimeLogic: A Temporal Logic Benchmark for Video QA
von: Swetha, Sirnam, et al.
Veröffentlicht: (2025)
von: Swetha, Sirnam, et al.
Veröffentlicht: (2025)
OpusAnimation: Code-Based Dynamic Chart Generation
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
TimeLogic Challenge @ CVPR 2026: Strong MLLMs Meet Evidence-Seeking Agents for Temporal-Logic Video Question Answering
von: Xu, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Xu, Zhaoyang, et al.
Veröffentlicht: (2026)
Adapting Point Cloud Analysis via Multimodal Bayesian Distribution Learning
von: Zhu, Xingyu, et al.
Veröffentlicht: (2026)
von: Zhu, Xingyu, et al.
Veröffentlicht: (2026)
Performance pressure and annual report text manipulation: Evidence from China
von: Yanxi Li, et al.
Veröffentlicht: (2024)
von: Yanxi Li, et al.
Veröffentlicht: (2024)
Economic Growth Expectations and Corporate Innovation: Evidence From China
von: Lujun Wang, et al.
Veröffentlicht: (2024)
von: Lujun Wang, et al.
Veröffentlicht: (2024)
Social Collaboration Does Not Shape the Effects Caused by Self‐Encoding: Evidence From Ongoing and Enduring Collaboration
von: Aiqing Nie, et al.
Veröffentlicht: (2025)
von: Aiqing Nie, et al.
Veröffentlicht: (2025)
Firm Digital Transformation and Investment Efficiency: Evidence From China
von: Maoguo Wu, et al.
Veröffentlicht: (2026)
von: Maoguo Wu, et al.
Veröffentlicht: (2026)
Med-CMR: A Fine-Grained Benchmark Integrating Visual Evidence and Clinical Logic for Medical Complex Multimodal Reasoning
von: Gong, Haozhen, et al.
Veröffentlicht: (2025)
von: Gong, Haozhen, et al.
Veröffentlicht: (2025)
VoQA: Visual-only Question Answering
von: An, Jianing, et al.
Veröffentlicht: (2025)
von: An, Jianing, et al.
Veröffentlicht: (2025)
EVIDENT: Routing MLLM Adaptation through Entity-Grounded Visual Evidence for Cross-Domain Video Temporal Grounding
von: Ahn, Geo, et al.
Veröffentlicht: (2026)
von: Ahn, Geo, et al.
Veröffentlicht: (2026)
Multifunctional agents based on 3‐dicycanovinylindan‐1‐one acceptor: Molecular design and phototheranostic application
von: Najia Zhu, et al.
Veröffentlicht: (2024)
von: Najia Zhu, et al.
Veröffentlicht: (2024)
Vibroseis Vehicle Routing Problem with Spatio-Temporal Coupled Constraints
von: Zhu, Kexin, et al.
Veröffentlicht: (2025)
von: Zhu, Kexin, et al.
Veröffentlicht: (2025)
NeoQA: Evidence-based Question Answering with Generated News Events
von: Glockner, Max, et al.
Veröffentlicht: (2025)
von: Glockner, Max, et al.
Veröffentlicht: (2025)
Evidence for a Damped Millisecond Quasi-Periodic Structure in a Fast Radio Burst
von: Xiao, Shuo, et al.
Veröffentlicht: (2026)
von: Xiao, Shuo, et al.
Veröffentlicht: (2026)
DateLogicQA: Benchmarking Temporal Biases in Large Language Models
von: Bhatia, Gagan, et al.
Veröffentlicht: (2024)
von: Bhatia, Gagan, et al.
Veröffentlicht: (2024)
MarkIt: Training-Free Visual Markers for Precise Video Temporal Grounding
von: Fang, Pengcheng, et al.
Veröffentlicht: (2026)
von: Fang, Pengcheng, et al.
Veröffentlicht: (2026)
TimeToM: Temporal Space is the Key to Unlocking the Door of Large Language Models' Theory-of-Mind
von: Hou, Guiyang, et al.
Veröffentlicht: (2024)
von: Hou, Guiyang, et al.
Veröffentlicht: (2024)
The Impact of Customer Environmental Pressure Perception on Supplier Greenwashing: Evidence From Environmental Tone in Annual Reports
von: Yunge Hu, et al.
Veröffentlicht: (2026)
von: Yunge Hu, et al.
Veröffentlicht: (2026)
Error-Mitigated Quantum Routing on Noisy Devices
von: Shi, Wenbo, et al.
Veröffentlicht: (2023)
von: Shi, Wenbo, et al.
Veröffentlicht: (2023)
Rumor Detection on Social Media with Temporal Propagation Structure Optimization
von: Peng, Xingyu, et al.
Veröffentlicht: (2024)
von: Peng, Xingyu, et al.
Veröffentlicht: (2024)
HulluEdit: Single-Pass Evidence-Consistent Subspace Editing for Mitigating Hallucinations in Large Vision-Language Models
von: Lin, Yangguang, et al.
Veröffentlicht: (2026)
von: Lin, Yangguang, et al.
Veröffentlicht: (2026)
"The Roller Conduction Effect" from the A-share Data Evidence
von: Lyu, Wenbo
Veröffentlicht: (2023)
von: Lyu, Wenbo
Veröffentlicht: (2023)
ROVER: Routing Object-Centric Visual Evidence for Grounded Multi-Image Reasoning
von: Lv, Guannan, et al.
Veröffentlicht: (2026)
von: Lv, Guannan, et al.
Veröffentlicht: (2026)
CircuitProbe: Tracing Visual Temporal Evidence Flow in Video Language Models
von: Zhang, Yiming, et al.
Veröffentlicht: (2025)
von: Zhang, Yiming, et al.
Veröffentlicht: (2025)
DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding
von: Yan, Hao, et al.
Veröffentlicht: (2026)
von: Yan, Hao, et al.
Veröffentlicht: (2026)
Auctioning Time to Mitigate Latency Races: Theory and Evidence from Blockchains
von: Capponi, Agostino, et al.
Veröffentlicht: (2025)
von: Capponi, Agostino, et al.
Veröffentlicht: (2025)
Look on Demand: A Cognitive Scheduling Framework for Visual Evidence Acquisition in Multimodal Reasoning
von: Zhang, Yang, et al.
Veröffentlicht: (2026)
von: Zhang, Yang, et al.
Veröffentlicht: (2026)
Video Repurposing from User Generated Content: A Large-scale Dataset and Benchmark
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
Zero-Shot Long-Form Video Understanding through Screenplay
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
A Multi-Expert Structural-Semantic Hybrid Framework for Unveiling Historical Patterns in Temporal Knowledge Graphs
von: Deng, Yimin, et al.
Veröffentlicht: (2025)
von: Deng, Yimin, et al.
Veröffentlicht: (2025)
Nearly invariant subspaces and kernels of Toeplitz operators on the Hardy space over the bidisk
von: Zhu, Senhua, et al.
Veröffentlicht: (2024)
von: Zhu, Senhua, et al.
Veröffentlicht: (2024)
DeferMem: Query-Time Evidence Distillation via Reinforcement Learning for Long-Term Memory QA
von: Yin, Jianing, et al.
Veröffentlicht: (2026)
von: Yin, Jianing, et al.
Veröffentlicht: (2026)
Prediction-Augmented Mechanism Design for Weighted Facility Location
von: Shi, Yangguang, et al.
Veröffentlicht: (2025)
von: Shi, Yangguang, et al.
Veröffentlicht: (2025)
Pure Exciton UV Emission of Colloidal ZnO Nanocrystals and Evidence of Surface Origin of Green Luminescence
von: Jiahuan Zhu, et al.
Veröffentlicht: (2025)
von: Jiahuan Zhu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Adaptive Dense Evidence Refinement for Video Relational Reasoning for VRR-QA Challenge
von: Sun, Yuyang, et al.
Veröffentlicht: (2026) -
RSVP: Reasoning Segmentation via Visual Prompting and Multi-modal Chain-of-Thought
von: Lu, Yi, et al.
Veröffentlicht: (2025) -
Number it: Temporal Grounding Videos like Flipping Manga
von: Wu, Yongliang, et al.
Veröffentlicht: (2024) -
VEU-Bench: Towards Comprehensive Understanding of Video Editing
von: Li, Bozheng, et al.
Veröffentlicht: (2025) -
Reframe Anything: LLM Agent for Open World Video Reframing
von: Cao, Jiawang, et al.
Veröffentlicht: (2024)