Test-Time Adaptation for Combating Missing Modalities in Egocentric Videos
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ramazanova, Merey, Pardo, Alejandro, Ghanem, Bernard, Alfarra, Motasem |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring Missing Modality in Multimodal Egocentric Datasets
von: Ramazanova, Merey, et al.
Veröffentlicht: (2024)
von: Ramazanova, Merey, et al.
Veröffentlicht: (2024)
ADVMEM: Adversarial Memory Initialization for Realistic Test-Time Adaptation via Tracklet-Based Benchmarking
von: Alhuwaider, Shyma, et al.
Veröffentlicht: (2025)
von: Alhuwaider, Shyma, et al.
Veröffentlicht: (2025)
Evaluation of Test-Time Adaptation Under Computational Time Constraints
von: Alfarra, Motasem, et al.
Veröffentlicht: (2023)
von: Alfarra, Motasem, et al.
Veröffentlicht: (2023)
GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation
von: Alhuwaider, Shyma, et al.
Veröffentlicht: (2026)
von: Alhuwaider, Shyma, et al.
Veröffentlicht: (2026)
Pixels or Positions? Benchmarking Modalities in Group Activity Recognition
von: Karki, Drishya, et al.
Veröffentlicht: (2025)
von: Karki, Drishya, et al.
Veröffentlicht: (2025)
SimCS: Simulation for Domain Incremental Online Continual Segmentation
von: Alfarra, Motasem, et al.
Veröffentlicht: (2022)
von: Alfarra, Motasem, et al.
Veröffentlicht: (2022)
Skill-Aligned Annotation for Reliable Evaluation in Text-to-Image Generation
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2026)
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models
von: Villa, Andrés, et al.
Veröffentlicht: (2025)
von: Villa, Andrés, et al.
Veröffentlicht: (2025)
Deep Learning at the Intersection: Certified Robustness as a Tool for 3D Vision
von: S, Gabriel Pérez, et al.
Veröffentlicht: (2024)
von: S, Gabriel Pérez, et al.
Veröffentlicht: (2024)
Forget Less, Retain More: A Lightweight Regularizer for Rehearsal-Based Continual Learning
von: Alssum, Lama, et al.
Veröffentlicht: (2025)
von: Alssum, Lama, et al.
Veröffentlicht: (2025)
Online Distillation with Continual Learning for Cyclic Domain Shifts
von: Houyon, Joachim, et al.
Veröffentlicht: (2023)
von: Houyon, Joachim, et al.
Veröffentlicht: (2023)
OpenTAD: A Unified Framework and Comprehensive Study of Temporal Action Detection
von: Liu, Shuming, et al.
Veröffentlicht: (2025)
von: Liu, Shuming, et al.
Veröffentlicht: (2025)
FedMedICL: Towards Holistic Evaluation of Distribution Shifts in Federated Medical Imaging
von: Alhamoud, Kumail, et al.
Veröffentlicht: (2024)
von: Alhamoud, Kumail, et al.
Veröffentlicht: (2024)
Multimodal Knowledge Distillation for Egocentric Action Recognition Robust to Missing Modalities
von: Santos-Villafranca, Maria, et al.
Veröffentlicht: (2025)
von: Santos-Villafranca, Maria, et al.
Veröffentlicht: (2025)
Hybrid Structure-from-Motion and Camera Relocalization for Enhanced Egocentric Localization
von: Mai, Jinjie, et al.
Veröffentlicht: (2024)
von: Mai, Jinjie, et al.
Veröffentlicht: (2024)
Video Self-Stitching Graph Network for Temporal Action Localization
von: Zhao, Chen, et al.
Veröffentlicht: (2020)
von: Zhao, Chen, et al.
Veröffentlicht: (2020)
EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation
von: Pei, Baoqi, et al.
Veröffentlicht: (2024)
von: Pei, Baoqi, et al.
Veröffentlicht: (2024)
Test-Time Adaptation for Video Highlight Detection Using Meta-Auxiliary Learning and Cross-Modality Hallucinations
von: Islam, Zahidul, et al.
Veröffentlicht: (2025)
von: Islam, Zahidul, et al.
Veröffentlicht: (2025)
Efficient Prompting for Continual Adaptation to Missing Modalities
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
EgoAdapt: Enhancing Robustness in Egocentric Interactive Speaker Detection Under Missing Modalities
von: Qian, Xinyuan, et al.
Veröffentlicht: (2026)
von: Qian, Xinyuan, et al.
Veröffentlicht: (2026)
MoRA: Missing Modality Low-Rank Adaptation for Visual Recognition
von: Zhao, Shu, et al.
Veröffentlicht: (2025)
von: Zhao, Shu, et al.
Veröffentlicht: (2025)
BOLT: Boost Large Vision-Language Model Without Training for Long-form Video Understanding
von: Liu, Shuming, et al.
Veröffentlicht: (2025)
von: Liu, Shuming, et al.
Veröffentlicht: (2025)
Compressed-Language Models for Understanding Compressed File Formats: a JPEG Exploration
von: Pérez, Juan C., et al.
Veröffentlicht: (2024)
von: Pérez, Juan C., et al.
Veröffentlicht: (2024)
SMILE: Infusing Spatial and Motion Semantics in Masked Video Learning
von: Thoker, Fida Mohammad, et al.
Veröffentlicht: (2025)
von: Thoker, Fida Mohammad, et al.
Veröffentlicht: (2025)
Robust Egocentric Referring Video Object Segmentation via Dual-Modal Causal Intervention
von: Liu, Haijing, et al.
Veröffentlicht: (2025)
von: Liu, Haijing, et al.
Veröffentlicht: (2025)
Low-Cost Test-Time Adaptation for Robust Video Editing
von: Wang, Jianhui, et al.
Veröffentlicht: (2025)
von: Wang, Jianhui, et al.
Veröffentlicht: (2025)
TTA-Vid: Generalized Test-Time Adaptation for Video Reasoning
von: Jahagirdar, Soumya Shamarao, et al.
Veröffentlicht: (2026)
von: Jahagirdar, Soumya Shamarao, et al.
Veröffentlicht: (2026)
Ego-VPA: Egocentric Video Understanding with Parameter-efficient Adaptation
von: Wu, Tz-Ying, et al.
Veröffentlicht: (2024)
von: Wu, Tz-Ying, et al.
Veröffentlicht: (2024)
TrackMAE: Video Representation Learning via Track Mask and Predict
von: Vandeghen, Renaud, et al.
Veröffentlicht: (2026)
von: Vandeghen, Renaud, et al.
Veröffentlicht: (2026)
Interactive Test-Time Adaptation with Reliable Spatial-Temporal Voxels for Multi-Modal Segmentation
von: Cao, Haozhi, et al.
Veröffentlicht: (2024)
von: Cao, Haozhi, et al.
Veröffentlicht: (2024)
Omnia de EgoTempo: Benchmarking Temporal Understanding of Multi-Modal LLMs in Egocentric Videos
von: Plizzari, Chiara, et al.
Veröffentlicht: (2025)
von: Plizzari, Chiara, et al.
Veröffentlicht: (2025)
Decoupling Stability and Plasticity for Multi-Modal Test-Time Adaptation
von: He, Yongbo, et al.
Veröffentlicht: (2026)
von: He, Yongbo, et al.
Veröffentlicht: (2026)
TimeLoc: A Unified End-to-End Framework for Precise Timestamp Localization in Long Videos
von: Zhang, Chen-Lin, et al.
Veröffentlicht: (2025)
von: Zhang, Chen-Lin, et al.
Veröffentlicht: (2025)
Investigating Event-Based Cameras for Video Frame Interpolation in Sports
von: Deckyvere, Antoine, et al.
Veröffentlicht: (2024)
von: Deckyvere, Antoine, et al.
Veröffentlicht: (2024)
Multi-Stream Cellular Test-Time Adaptation of Real-Time Models Evolving in Dynamic Environments
von: Gérin, Benoît, et al.
Veröffentlicht: (2024)
von: Gérin, Benoît, et al.
Veröffentlicht: (2024)
X-VARS: Introducing Explainability in Football Refereeing with Multi-Modal Large Language Model
von: Held, Jan, et al.
Veröffentlicht: (2024)
von: Held, Jan, et al.
Veröffentlicht: (2024)
SoccerNet-Caption: Dense Video Captioning for Soccer Broadcasts Commentaries
von: Mkhallati, Hassan, et al.
Veröffentlicht: (2023)
von: Mkhallati, Hassan, et al.
Veröffentlicht: (2023)
Majorization-Guided Test-Time Adaptation for Vision-Language Models under Modality-Specific Shift
von: Chen, Lixian, et al.
Veröffentlicht: (2026)
von: Chen, Lixian, et al.
Veröffentlicht: (2026)
DMC$^3$: Dual-Modal Counterfactual Contrastive Construction for Egocentric Video Question Answering
von: Zou, Jiayi, et al.
Veröffentlicht: (2025)
von: Zou, Jiayi, et al.
Veröffentlicht: (2025)
Robust Multi-Modal Face Anti-Spoofing with Domain Adaptation: Tackling Missing Modalities, Noisy Pseudo-Labels, and Model Degradation
von: Hsu, Ming-Tsung, et al.
Veröffentlicht: (2025)
von: Hsu, Ming-Tsung, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Exploring Missing Modality in Multimodal Egocentric Datasets
von: Ramazanova, Merey, et al.
Veröffentlicht: (2024) -
ADVMEM: Adversarial Memory Initialization for Realistic Test-Time Adaptation via Tracklet-Based Benchmarking
von: Alhuwaider, Shyma, et al.
Veröffentlicht: (2025) -
Evaluation of Test-Time Adaptation Under Computational Time Constraints
von: Alfarra, Motasem, et al.
Veröffentlicht: (2023) -
GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation
von: Alhuwaider, Shyma, et al.
Veröffentlicht: (2026) -
Pixels or Positions? Benchmarking Modalities in Group Activity Recognition
von: Karki, Drishya, et al.
Veröffentlicht: (2025)