CourseTimeQA: A Lecture-Video Benchmark and a Latency-Constrained Cross-Modal Fusion Method for Timestamped QA
Fuente:
arXiv
Saved in:
| Main Authors: | Kovalev, Vsevolod, Kumar, Parteek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
KARMA-MV: A Benchmark for Causal Question Answering on Music Videos
by: Ghosh, Archishman, et al.
Published: (2026)
by: Ghosh, Archishman, et al.
Published: (2026)
Evaluating Perspectival Biases in Cross-Modal Retrieval
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
ReCoVR: Closing the Loop in Interactive Composed Video Retrieval
by: Zhang, Bingqing, et al.
Published: (2026)
by: Zhang, Bingqing, et al.
Published: (2026)
Playing telephone with generative models: "verification disability," "compelled reliance," and accessibility in data visualization
by: Elavsky, Frank, et al.
Published: (2025)
by: Elavsky, Frank, et al.
Published: (2025)
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
by: Park, Sungho, et al.
Published: (2026)
by: Park, Sungho, et al.
Published: (2026)
A Grounded Memory System For Smart Personal Assistants
by: Ocker, Felix, et al.
Published: (2025)
by: Ocker, Felix, et al.
Published: (2025)
EdgeJury: Cross-Reviewed Small-Model Ensembles for Truthful Question Answering on Serverless Edge Inference
by: Kumar, Aayush
Published: (2025)
by: Kumar, Aayush
Published: (2025)
Large Language Model for Qualitative Research -- A Systematic Mapping Study
by: Barros, Cauã Ferreira, et al.
Published: (2024)
by: Barros, Cauã Ferreira, et al.
Published: (2024)
Leveraging OpenFlamingo for Multimodal Embedding Analysis of C2C Car Parts Data
by: Rashid, Maisha Binte, et al.
Published: (2025)
by: Rashid, Maisha Binte, et al.
Published: (2025)
Bottleneck-based Encoder-decoder ARchitecture (BEAR) for Learning Unbiased Consumer-to-Consumer Image Representations
by: Rivas, Pablo, et al.
Published: (2024)
by: Rivas, Pablo, et al.
Published: (2024)
TriAlignGR: Triangular Multitask Alignment with Multimodal Deep Interest Mining for Generative Recommendation
by: Zeng, Yangchen, et al.
Published: (2026)
by: Zeng, Yangchen, et al.
Published: (2026)
Drawing on Memory: Dual-Trace Encoding Improves Cross-Session Recall in LLM Agents
by: Stern, Benjamin, et al.
Published: (2026)
by: Stern, Benjamin, et al.
Published: (2026)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
by: Pradhan, Anu, et al.
Published: (2025)
by: Pradhan, Anu, et al.
Published: (2025)
Retrieval Is Not Enough: Why Organizational AI Needs Epistemic Infrastructure
by: Bottino, Federico, et al.
Published: (2026)
by: Bottino, Federico, et al.
Published: (2026)
What Matters in LLM-Based Feature Extractor for Recommender? A Systematic Analysis of Prompts, Models, and Adaptation
by: Shi, Kainan, et al.
Published: (2025)
by: Shi, Kainan, et al.
Published: (2025)
When to Forget: A Memory Governance Primitive
by: Simsek, Baris
Published: (2026)
by: Simsek, Baris
Published: (2026)
Mind the Gap: A Generalized Approach for Cross-Modal Embedding Alignment
by: Yadav, Arihan, et al.
Published: (2024)
by: Yadav, Arihan, et al.
Published: (2024)
Bi-View Embedding Fusion: A Hybrid Learning Approach for Knowledge Graph's Nodes Classification Addressing Problems with Limited Data
by: Napoli, Rosario, et al.
Published: (2025)
by: Napoli, Rosario, et al.
Published: (2025)
Local Hybrid Retrieval-Augmented Document QA
by: Astrino, Paolo
Published: (2025)
by: Astrino, Paolo
Published: (2025)
NCTB-QA: A Large-Scale Bangla Educational Question Answering Dataset and Benchmarking Performance
by: Eyasir, Abrar, et al.
Published: (2026)
by: Eyasir, Abrar, et al.
Published: (2026)
Beyond Vision: Contextually Enriched Image Captioning with Multi-Modal Retrieval
by: Quy, Nguyen Lam Phu, et al.
Published: (2025)
by: Quy, Nguyen Lam Phu, et al.
Published: (2025)
ORPHEAS: A Cross-Lingual Greek-English Embedding Model for Retrieval-Augmented Generation
by: Livieris, Ioannis E., et al.
Published: (2026)
by: Livieris, Ioannis E., et al.
Published: (2026)
Experimenting active and sequential learning in a medieval music manuscript
by: Sharma, Sachin, et al.
Published: (2025)
by: Sharma, Sachin, et al.
Published: (2025)
Vision-Language Cross-Attention for Real-Time Autonomous Driving
by: Patapati, Santosh, et al.
Published: (2025)
by: Patapati, Santosh, et al.
Published: (2025)
PRECEPT: Planning Resilience via Experience, Context Engineering & Probing Trajectories A Unified Framework for Test-Time Adaptation with Compositional Rule Learning and Pareto-Guided Prompt Evolution
by: Shahmansoori, Arash
Published: (2026)
by: Shahmansoori, Arash
Published: (2026)
Tulip Agent -- Enabling LLM-Based Agents to Solve Tasks Using Large Tool Libraries
by: Ocker, Felix, et al.
Published: (2024)
by: Ocker, Felix, et al.
Published: (2024)
LENS: A Staged Design for Interaction Granularityin Sequential CTR Prediction
by: Wang, Yuan, et al.
Published: (2026)
by: Wang, Yuan, et al.
Published: (2026)
Adaptive$^2$: Adaptive Domain Mining for Fine-grained Domain Adaptation Modeling
by: Sun, Wenxuan, et al.
Published: (2024)
by: Sun, Wenxuan, et al.
Published: (2024)
Bayesian Coreset Optimization for Personalized Federated Learning
by: Chanda, Prateek, et al.
Published: (2025)
by: Chanda, Prateek, et al.
Published: (2025)
Beyond the Flat Sequence: Hierarchical and Preference-Aware Generative Recommendations
by: Chen, Zerui, et al.
Published: (2026)
by: Chen, Zerui, et al.
Published: (2026)
QoSGMAA: A Robust Multi-Order Graph Attention and Adversarial Framework for Sparse QoS Prediction
by: Du, Guanchen, et al.
Published: (2025)
by: Du, Guanchen, et al.
Published: (2025)
Experimentation Accelerator: Interpretable Insights and Creative Recommendations for A/B Testing with Content-Aware ranking
by: Hu, Zhengmian, et al.
Published: (2026)
by: Hu, Zhengmian, et al.
Published: (2026)
DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark
by: Hu, Ruofan, et al.
Published: (2026)
by: Hu, Ruofan, et al.
Published: (2026)
Quantifying and Narrowing the Unknown: Interactive Text-to-Video Retrieval via Uncertainty Minimization
by: Zhang, Bingqing, et al.
Published: (2025)
by: Zhang, Bingqing, et al.
Published: (2025)
Optimizing Retrieval-Augmented Generation (RAG) for Colloquial Cantonese: A LoRA-Based Systematic Review
by: Calonge, David Santandreu, et al.
Published: (2025)
by: Calonge, David Santandreu, et al.
Published: (2025)
Evolve: A Persistent Knowledge Lifecycle for Small Language Models
by: Hovagimian, Dikran
Published: (2026)
by: Hovagimian, Dikran
Published: (2026)
Qtok: A Comprehensive Framework for Evaluating Multilingual Tokenizer Quality in Large Language Models
by: Chelombitko, Iaroslav, et al.
Published: (2024)
by: Chelombitko, Iaroslav, et al.
Published: (2024)
Task Memory Engine: Spatial Memory for Robust Multi-Step LLM Agents
by: Ye, Ye
Published: (2025)
by: Ye, Ye
Published: (2025)
Adaptive Data Flywheel: Applying MAPE Control Loops to AI Agent Improvement
by: Shukla, Aaditya, et al.
Published: (2025)
by: Shukla, Aaditya, et al.
Published: (2025)
STAR : Bridging Statistical and Agentic Reasoning for Large Model Performance Prediction
by: Wang, Xiaoxiao, et al.
Published: (2026)
by: Wang, Xiaoxiao, et al.
Published: (2026)
Similar Items
-
KARMA-MV: A Benchmark for Causal Question Answering on Music Videos
by: Ghosh, Archishman, et al.
Published: (2026) -
Evaluating Perspectival Biases in Cross-Modal Retrieval
by: Saengsukhiran, Teerapol, et al.
Published: (2025) -
ReCoVR: Closing the Loop in Interactive Composed Video Retrieval
by: Zhang, Bingqing, et al.
Published: (2026) -
Playing telephone with generative models: "verification disability," "compelled reliance," and accessibility in data visualization
by: Elavsky, Frank, et al.
Published: (2025) -
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
by: Park, Sungho, et al.
Published: (2026)