CoVR-R:Reason-Aware Composed Video Retrieval
Fuente:
arXiv
Guardado en:
| Autores principales: | Thawakar, Omkar, Demidov, Dmitry, Potlapalli, Vaishnav, Bogireddy, Sai Prasanna Teja Reddy, Gajjala, Viswanatha Reddy, Lasheen, Alaa Mostafa, Anwer, Rao Muhammad, Khan, Fahad |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Neural at ArchEHR-QA 2025: Agentic Prompt Optimization for Evidence-Grounded Clinical Question Answering
por: Bogireddy, Sai Prasanna Teja Reddy, et al.
Publicado: (2025)
por: Bogireddy, Sai Prasanna Teja Reddy, et al.
Publicado: (2025)
Neural at ArchEHR-QA 2026: One Method Fits All: Unified Prompt Optimization for Clinical QA over EHRs
por: Majeedi, Abrar, et al.
Publicado: (2026)
por: Majeedi, Abrar, et al.
Publicado: (2026)
Beyond Simple Edits: Composed Video Retrieval with Dense Modifications
por: Thawakar, Omkar, et al.
Publicado: (2025)
por: Thawakar, Omkar, et al.
Publicado: (2025)
Composed Video Retrieval via Enriched Context and Discriminative Embeddings
por: Thawakar, Omkar, et al.
Publicado: (2024)
por: Thawakar, Omkar, et al.
Publicado: (2024)
Thinking Beyond Labels: Vocabulary-Free Fine-Grained Recognition using Reasoning-Augmented LMMs
por: Demidov, Dmitry, et al.
Publicado: (2025)
por: Demidov, Dmitry, et al.
Publicado: (2025)
Vocabulary-free Fine-grained Visual Recognition via Enriched Contextually Grounded Vision-Language Model
por: Demidov, Dmitry, et al.
Publicado: (2025)
por: Demidov, Dmitry, et al.
Publicado: (2025)
CoVR-2: Automatic Data Construction for Composed Video Retrieval
por: Ventura, Lucas, et al.
Publicado: (2023)
por: Ventura, Lucas, et al.
Publicado: (2023)
Predicting Estimated Times of Restoration for Electrical Outages Using Longitudinal Tabular Transformers
por: Teja, Bogireddy Sai Prasanna, et al.
Publicado: (2025)
por: Teja, Bogireddy Sai Prasanna, et al.
Publicado: (2025)
RICA2: Rubric-Informed, Calibrated Assessment of Actions
por: Majeedi, Abrar, et al.
Publicado: (2024)
por: Majeedi, Abrar, et al.
Publicado: (2024)
EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards
por: Thawakar, Omkar, et al.
Publicado: (2025)
por: Thawakar, Omkar, et al.
Publicado: (2025)
XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models
por: Thawakar, Omkar, et al.
Publicado: (2023)
por: Thawakar, Omkar, et al.
Publicado: (2023)
LETS Forecast: Learning Embedology for Time Series Forecasting
por: Majeedi, Abrar, et al.
Publicado: (2025)
por: Majeedi, Abrar, et al.
Publicado: (2025)
Time Travel: A Comprehensive Benchmark to Evaluate LMMs on Historical and Cultural Artifacts
por: Ghaboura, Sara, et al.
Publicado: (2025)
por: Ghaboura, Sara, et al.
Publicado: (2025)
Cognitive Load Limits in Large Language Models: Benchmarking Multi-Hop Reasoning
por: Adapala, Sai Teja Reddy
Publicado: (2025)
por: Adapala, Sai Teja Reddy
Publicado: (2025)
The Anti-Ouroboros Effect: Emergent Resilience in Large Language Models from Recursive Selective Feedback
por: Adapala, Sai Teja Reddy
Publicado: (2025)
por: Adapala, Sai Teja Reddy
Publicado: (2025)
Reason-Then-Retrieve for CoVR-R with Structured Edit Prompts and Dense-Sparse Fusion
por: Liu, DongQing, et al.
Publicado: (2026)
por: Liu, DongQing, et al.
Publicado: (2026)
DuwatBench: Bridging Language and Visual Heritage through an Arabic Calligraphy Benchmark for Multimodal Understanding
por: Patle, Shubham, et al.
Publicado: (2026)
por: Patle, Shubham, et al.
Publicado: (2026)
MobiLlama: Towards Accurate and Lightweight Fully Transparent GPT
por: Thawakar, Omkar, et al.
Publicado: (2024)
por: Thawakar, Omkar, et al.
Publicado: (2024)
CAMEL-Bench: A Comprehensive Arabic LMM Benchmark
por: Ghaboura, Sara, et al.
Publicado: (2024)
por: Ghaboura, Sara, et al.
Publicado: (2024)
Graph Neural Networks (GNNs) in Intelligent Transportation Systems
por: Reddy, Sai Teja Reddy, et al.
Publicado: (2026)
por: Reddy, Sai Teja Reddy, et al.
Publicado: (2026)
The Aegis Protocol: A Foundational Security Framework for Autonomous AI Agents
por: Adapala, Sai Teja Reddy, et al.
Publicado: (2025)
por: Adapala, Sai Teja Reddy, et al.
Publicado: (2025)
ARB: A Comprehensive Arabic Multimodal Reasoning Benchmark
por: Ghaboura, Sara, et al.
Publicado: (2025)
por: Ghaboura, Sara, et al.
Publicado: (2025)
Fann or Flop: A Multigenre, Multiera Benchmark for Arabic Poetry Understanding in LLMs
por: Alghallabi, Wafa, et al.
Publicado: (2025)
por: Alghallabi, Wafa, et al.
Publicado: (2025)
LLM Post-Training: A Deep Dive into Reasoning Large Language Models
por: Kumar, Komal, et al.
Publicado: (2025)
por: Kumar, Komal, et al.
Publicado: (2025)
Composed Object Retrieval: Object-level Retrieval via Composed Expressions
por: Wang, Tong, et al.
Publicado: (2025)
por: Wang, Tong, et al.
Publicado: (2025)
How Good are Foundation Models in Step-by-Step Embodied Reasoning?
por: Dissanayake, Dinura, et al.
Publicado: (2025)
por: Dissanayake, Dinura, et al.
Publicado: (2025)
Distilling Local Texture Features for Colorectal Tissue Classification in Low Data Regimes
por: Demidov, Dmitry, et al.
Publicado: (2024)
por: Demidov, Dmitry, et al.
Publicado: (2024)
Dynamic Pre-training: Towards Efficient and Scalable All-in-One Image Restoration
por: Dudhane, Akshay, et al.
Publicado: (2024)
por: Dudhane, Akshay, et al.
Publicado: (2024)
AIN: The Arabic INclusive Large Multimodal Model
por: Heakl, Ahmed, et al.
Publicado: (2025)
por: Heakl, Ahmed, et al.
Publicado: (2025)
DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding
por: Ishaq, Ayesha, et al.
Publicado: (2025)
por: Ishaq, Ayesha, et al.
Publicado: (2025)
SketchYourSeg: Mask-Free Subjective Image Segmentation via Freehand Sketches
por: Koley, Subhadeep, et al.
Publicado: (2025)
por: Koley, Subhadeep, et al.
Publicado: (2025)
Designing Scalable SAP-Centered Enterprise Systems For High-Volume Transaction Environments
por: Somasekharreddy Bogireddy
Publicado: (2026)
por: Somasekharreddy Bogireddy
Publicado: (2026)
Improving Data Accuracy and Financial Integrity in SAP-Centered Enterprise Transaction Systems
por: Somasekharreddy Bogireddy
Publicado: (2026)
por: Somasekharreddy Bogireddy
Publicado: (2026)
Integrating Agile Governance in Compliance-Driven Enterprises: A Hybrid Delivery Model for Regulated Environments
por: Tarun Teja Reddy Palyam
Publicado: (2026)
por: Tarun Teja Reddy Palyam
Publicado: (2026)
Empowering Enterprise Development by Building and Deploying Admin Dashboard using Refine Framework
por: Gajjala, Sai Teja, et al.
Publicado: (2024)
por: Gajjala, Sai Teja, et al.
Publicado: (2024)
(Table 3) Chemical analyses from DSDP Holes 22-214, 22-215 and 22-216 in the Indian Ocean
por: Reddy, V Viswanatha, et al.
Publicado: (1978)
por: Reddy, V Viswanatha, et al.
Publicado: (1978)
Fourier-Based GAN Fingerprint Detection using ResNet50
por: Erukude, Sai Teja, et al.
Publicado: (2025)
por: Erukude, Sai Teja, et al.
Publicado: (2025)
FedOnco-Bench: A Reproducible Benchmark for Privacy-Aware Federated Tumor Segmentation with Synthetic CT Data
por: Marella, Viswa Chaitanya, et al.
Publicado: (2025)
por: Marella, Viswa Chaitanya, et al.
Publicado: (2025)
AI-Driven Cybersecurity Threats: A Survey of Emerging Risks and Defensive Strategies
por: Erukude, Sai Teja, et al.
Publicado: (2026)
por: Erukude, Sai Teja, et al.
Publicado: (2026)
Multimodal Detection of Fake Reviews using BERT and ResNet-50
por: Veluru, Suhasnadh Reddy, et al.
Publicado: (2025)
por: Veluru, Suhasnadh Reddy, et al.
Publicado: (2025)
Ejemplares similares
-
Neural at ArchEHR-QA 2025: Agentic Prompt Optimization for Evidence-Grounded Clinical Question Answering
por: Bogireddy, Sai Prasanna Teja Reddy, et al.
Publicado: (2025) -
Neural at ArchEHR-QA 2026: One Method Fits All: Unified Prompt Optimization for Clinical QA over EHRs
por: Majeedi, Abrar, et al.
Publicado: (2026) -
Beyond Simple Edits: Composed Video Retrieval with Dense Modifications
por: Thawakar, Omkar, et al.
Publicado: (2025) -
Composed Video Retrieval via Enriched Context and Discriminative Embeddings
por: Thawakar, Omkar, et al.
Publicado: (2024) -
Thinking Beyond Labels: Vocabulary-Free Fine-Grained Recognition using Reasoning-Augmented LMMs
por: Demidov, Dmitry, et al.
Publicado: (2025)