MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chowdhury, Sanjoy, Elmoghany, Mohamed, Abeysinghe, Yohan, Fei, Junjie, Nag, Sayan, Khan, Salman, Elhoseiny, Mohamed, Manocha, Dinesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Aurelia: Test-time Reasoning Distillation in Audio-Visual LLMs
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
Meerkat: Audio-Visual Large Language Model for Grounding in Space and Time
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2024)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2024)
AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
Uncovering the Representation Geometry of Minimal Cores in Overcomplete Reasoning Traces
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2026)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2026)
Document Haystacks: Vision-Language Reasoning Over Piles of 1000+ Documents
von: Chen, Jun, et al.
Veröffentlicht: (2024)
von: Chen, Jun, et al.
Veröffentlicht: (2024)
MeLFusion: Synthesizing Music from Image and Language Cues using Diffusion Models
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2024)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2024)
AURA: A Fine-Grained Benchmark and Decomposed Metric for Audio-Visual Reasoning
von: Galougah, Siminfar Samakoush, et al.
Veröffentlicht: (2025)
von: Galougah, Siminfar Samakoush, et al.
Veröffentlicht: (2025)
Can LLMs Generate Human-Like Wayfinding Instructions? Towards Platform-Agnostic Embodied Instruction Synthesis
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2024)
von: Dorbala, Vishnu Sashank, et al.
Veröffentlicht: (2024)
Finding BSM Needles in Electromagnetic Haystacks at DUNE
von: Brdar, Vedran, et al.
Veröffentlicht: (2025)
von: Brdar, Vedran, et al.
Veröffentlicht: (2025)
Reasoning on Multiple Needles In A Haystack
von: Wang, Yidong
Veröffentlicht: (2025)
von: Wang, Yidong
Veröffentlicht: (2025)
Hidden in the Haystack: Smaller Needles are More Difficult for LLMs to Find
von: Bianchi, Owen, et al.
Veröffentlicht: (2025)
von: Bianchi, Owen, et al.
Veröffentlicht: (2025)
AMUSE: Audio-Visual Benchmark and Alignment Framework for Agentic Multi-Speaker Understanding
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
Two Causally Related Needles in a Video Haystack
von: Li, Miaoyu, et al.
Veröffentlicht: (2025)
von: Li, Miaoyu, et al.
Veröffentlicht: (2025)
EgoAdapt: Adaptive Multisensory Distillation and Policy Learning for Efficient Egocentric Perception
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
Finding Diamonds in Conversation Haystacks: A Benchmark for Conversational Data Retrieval
von: Lee, Yohan, et al.
Veröffentlicht: (2025)
von: Lee, Yohan, et al.
Veröffentlicht: (2025)
Artificial Intelligence for Quantum Matter: Finding a Needle in a Haystack
von: Nazaryan, Khachatur, et al.
Veröffentlicht: (2025)
von: Nazaryan, Khachatur, et al.
Veröffentlicht: (2025)
Inst4DGS: Instance-Decomposed 4D Gaussian Splatting with Multi-Video Label Permutation Learning
von: Lee, Yonghan, et al.
Veröffentlicht: (2026)
von: Lee, Yonghan, et al.
Veröffentlicht: (2026)
Needle In A Multimodal Haystack
von: Wang, Weiyun, et al.
Veröffentlicht: (2024)
von: Wang, Weiyun, et al.
Veröffentlicht: (2024)
Kestrel: 3D Multimodal LLM for Part-Aware Grounded Description
von: Ahmed, Mahmoud, et al.
Veröffentlicht: (2024)
von: Ahmed, Mahmoud, et al.
Veröffentlicht: (2024)
GOLD PANNING: Strategic Context Shuffling for Needle-in-Haystack Reasoning
von: Byerly, Adam, et al.
Veröffentlicht: (2025)
von: Byerly, Adam, et al.
Veröffentlicht: (2025)
DMCA: Dense Multi-agent Navigation using Attention and Communication
von: Arul, Senthil Hariharan, et al.
Veröffentlicht: (2022)
von: Arul, Senthil Hariharan, et al.
Veröffentlicht: (2022)
Finding Interest Needle in Popularity Haystack: Improving Retrieval by Modeling Item Exposure
von: Agarwal, Rahul, et al.
Veröffentlicht: (2025)
von: Agarwal, Rahul, et al.
Veröffentlicht: (2025)
MultiHaystack: Benchmarking Multimodal Retrieval and Reasoning over 40K Images, Videos, and Documents
von: Xu, Dannong, et al.
Veröffentlicht: (2026)
von: Xu, Dannong, et al.
Veröffentlicht: (2026)
CalibFree: Self-Supervised View Feature Separation for Calibration-Free Multi-Camera Multi-Object Tracking
von: Xian, Ruiqi, et al.
Veröffentlicht: (2026)
von: Xian, Ruiqi, et al.
Veröffentlicht: (2026)
Vgent: Graph-based Retrieval-Reasoning-Augmented Generation For Long Video Understanding
von: Shen, Xiaoqian, et al.
Veröffentlicht: (2025)
von: Shen, Xiaoqian, et al.
Veröffentlicht: (2025)
Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark
von: Wu, Tsung-Han, et al.
Veröffentlicht: (2024)
von: Wu, Tsung-Han, et al.
Veröffentlicht: (2024)
Needle In A Video Haystack: A Scalable Synthetic Evaluator for Video MLLMs
von: Zhao, Zijia, et al.
Veröffentlicht: (2024)
von: Zhao, Zijia, et al.
Veröffentlicht: (2024)
Looking for the Information Needle in the Internet Haystack.
von: Clausen, Helge
Veröffentlicht: (1996)
von: Clausen, Helge
Veröffentlicht: (1996)
An Infinite Needle in a Finite Haystack: Finding Infinite Counter-Models in Deductive Verification
von: Elad, Neta, et al.
Veröffentlicht: (2023)
von: Elad, Neta, et al.
Veröffentlicht: (2023)
MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark
von: Sakshi, S, et al.
Veröffentlicht: (2024)
von: Sakshi, S, et al.
Veröffentlicht: (2024)
In Search of Needles in a 11M Haystack: Recurrent Memory Finds What LLMs Miss
von: Kuratov, Yuri, et al.
Veröffentlicht: (2024)
von: Kuratov, Yuri, et al.
Veröffentlicht: (2024)
Synergistic Neural Forecasting of Air Pollution with Stochastic Sampling
von: Abeysinghe, Yohan, et al.
Veröffentlicht: (2025)
von: Abeysinghe, Yohan, et al.
Veröffentlicht: (2025)
Looking for (Genomic) Needles in a Haystack: Sparsity-Driven Search for Identifying Correlated Genetic Mutations in Cancer
von: Prabhu, Ritvik, et al.
Veröffentlicht: (2026)
von: Prabhu, Ritvik, et al.
Veröffentlicht: (2026)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
von: Baliah, Sanoojan, et al.
Veröffentlicht: (2026)
von: Baliah, Sanoojan, et al.
Veröffentlicht: (2026)
StoryGPT-V: Large Language Models as Consistent Story Visualizers
von: Shen, Xiaoqian, et al.
Veröffentlicht: (2023)
von: Shen, Xiaoqian, et al.
Veröffentlicht: (2023)
WikiAutoGen: Towards Multi-Modal Wikipedia-Style Article Generation
von: Yang, Zhongyu, et al.
Veröffentlicht: (2025)
von: Yang, Zhongyu, et al.
Veröffentlicht: (2025)
Finding Needles in Emb(a)dding Haystacks: Legal Document Retrieval via Bagging and SVR Ensembles
von: Bönisch, Kevin, et al.
Veröffentlicht: (2025)
von: Bönisch, Kevin, et al.
Veröffentlicht: (2025)
Beyond Needle(s) in the Embodied Haystack: Environment, Architecture, and Training Considerations for Long Context Reasoning
von: Kim, Bosung, et al.
Veröffentlicht: (2025)
von: Kim, Bosung, et al.
Veröffentlicht: (2025)
Needle in a Poissonian Haystack—An X‐Ray Astronomer's Guide to QPE Fishing
von: Erwan Quintin, et al.
Veröffentlicht: (2026)
von: Erwan Quintin, et al.
Veröffentlicht: (2026)
Needle in the Haystack for Memory Based Large Language Models
von: Nelson, Elliot, et al.
Veröffentlicht: (2024)
von: Nelson, Elliot, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Aurelia: Test-time Reasoning Distillation in Audio-Visual LLMs
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025) -
Meerkat: Audio-Visual Large Language Model for Grounding in Space and Time
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2024) -
AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025) -
Uncovering the Representation Geometry of Minimal Cores in Overcomplete Reasoning Traces
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2026) -
Document Haystacks: Vision-Language Reasoning Over Piles of 1000+ Documents
von: Chen, Jun, et al.
Veröffentlicht: (2024)