A Probabilistic Jump-Diffusion Framework for Open-World Egocentric Activity Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kundu, Sanjoy, Vellamcheti, Shanmukha, Aakur, Sathyanarayanan N. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ProbRes: Probabilistic Jump Diffusion for Open-World Egocentric Activity Recognition
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2025)
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2025)
Hallucinate, Ground, Repeat: A Framework for Generalized Visual Relationship Detection
von: Vellamcheti, Shanmukha, et al.
Veröffentlicht: (2025)
von: Vellamcheti, Shanmukha, et al.
Veröffentlicht: (2025)
ALGO: Object-Grounded Visual Commonsense Reasoning for Open-World Egocentric Action Recognition
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2024)
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2024)
Discovering Novel Actions from Open World Egocentric Videos with Object-Grounded Visual Commonsense Reasoning
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2023)
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2023)
CVT-Bench: Counterfactual Viewpoint Transformations Reveal Unstable Spatial Representations in Multimodal LLMs
von: Vellamcheti, Shanmukha, et al.
Veröffentlicht: (2026)
von: Vellamcheti, Shanmukha, et al.
Veröffentlicht: (2026)
EASE: Embodied Active Event Perception via Self-Supervised Energy Minimization
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
Self-supervised Multi-actor Social Activity Understanding in Streaming Videos
von: Trehan, Shubham, et al.
Veröffentlicht: (2024)
von: Trehan, Shubham, et al.
Veröffentlicht: (2024)
CRAFT: A Neuro-Symbolic Framework for Visual Functional Affordance Grounding
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
Generalized Event Partonomy Inference with Structured Hierarchical Predictive Learning
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
STaTS: Structure-Aware Temporal Sequence Summarization via Statistical Window Merging
von: Bhowmick, Disharee, et al.
Veröffentlicht: (2025)
von: Bhowmick, Disharee, et al.
Veröffentlicht: (2025)
Capturing Temporal Components for Time Series Classification
von: Vavilthota, Venkata Ragavendra, et al.
Veröffentlicht: (2024)
von: Vavilthota, Venkata Ragavendra, et al.
Veröffentlicht: (2024)
FSP-DETR: Few-Shot Prototypical Parasitic Ova Detection
von: Trehan, Shubham, et al.
Veröffentlicht: (2025)
von: Trehan, Shubham, et al.
Veröffentlicht: (2025)
Towards Continual Egocentric Activity Recognition: A Multi-modal Egocentric Activity Dataset for Continual Learning
von: Xu, Linfeng, et al.
Veröffentlicht: (2023)
von: Xu, Linfeng, et al.
Veröffentlicht: (2023)
Human Activity Recognition in an Open World
von: Prijatelj, Derek S., et al.
Veröffentlicht: (2022)
von: Prijatelj, Derek S., et al.
Veröffentlicht: (2022)
WEAR: An Outdoor Sports Dataset for Wearable and Egocentric Activity Recognition
von: Bock, Marius, et al.
Veröffentlicht: (2023)
von: Bock, Marius, et al.
Veröffentlicht: (2023)
Towards Open-World Gesture Recognition
von: Shen, Junxiao, et al.
Veröffentlicht: (2024)
von: Shen, Junxiao, et al.
Veröffentlicht: (2024)
PlayerOne: Egocentric World Simulator
von: Tu, Yuanpeng, et al.
Veröffentlicht: (2025)
von: Tu, Yuanpeng, et al.
Veröffentlicht: (2025)
Continual Multimodal Egocentric Activity Recognition via Modality-Aware Novel Detection
von: Lim, Wonseon, et al.
Veröffentlicht: (2026)
von: Lim, Wonseon, et al.
Veröffentlicht: (2026)
SimpleEgo: Predicting Probabilistic Body Pose from Egocentric Cameras
von: Cuevas-Velasquez, Hanz, et al.
Veröffentlicht: (2024)
von: Cuevas-Velasquez, Hanz, et al.
Veröffentlicht: (2024)
WorldWander: Bridging Egocentric and Exocentric Worlds in Video Generation
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
EgoPrompt: Prompt Learning for Egocentric Action Recognition
von: Lyu, Huaihai, et al.
Veröffentlicht: (2025)
von: Lyu, Huaihai, et al.
Veröffentlicht: (2025)
Motion Focus Recognition in Fast-Moving Egocentric Video
von: Hong, Si-En, et al.
Veröffentlicht: (2026)
von: Hong, Si-En, et al.
Veröffentlicht: (2026)
EgoCHARM: Resource-Efficient Hierarchical Activity Recognition using an Egocentric IMU Sensor
von: Padmanabha, Akhil, et al.
Veröffentlicht: (2025)
von: Padmanabha, Akhil, et al.
Veröffentlicht: (2025)
Learning for Transductive Threshold Calibration in Open-World Recognition
von: Zhang, Qin, et al.
Veröffentlicht: (2023)
von: Zhang, Qin, et al.
Veröffentlicht: (2023)
Domain Generalization using Action Sequences for Egocentric Action Recognition
von: Nasirimajd, Amirshayan, et al.
Veröffentlicht: (2025)
von: Nasirimajd, Amirshayan, et al.
Veröffentlicht: (2025)
LookOut: Real-World Humanoid Egocentric Navigation
von: Pan, Boxiao, et al.
Veröffentlicht: (2025)
von: Pan, Boxiao, et al.
Veröffentlicht: (2025)
egoEMOTION: Egocentric Vision and Physiological Signals for Emotion and Personality Recognition in Real-World Tasks
von: Jammot, Matthias, et al.
Veröffentlicht: (2025)
von: Jammot, Matthias, et al.
Veröffentlicht: (2025)
Boosting Open Set Recognition Performance through Modulated Representation Learning
von: Kundu, Amit Kumar, et al.
Veröffentlicht: (2025)
von: Kundu, Amit Kumar, et al.
Veröffentlicht: (2025)
Multimodal Knowledge Distillation for Egocentric Action Recognition Robust to Missing Modalities
von: Santos-Villafranca, Maria, et al.
Veröffentlicht: (2025)
von: Santos-Villafranca, Maria, et al.
Veröffentlicht: (2025)
Multimodal Cross-Domain Few-Shot Learning for Egocentric Action Recognition
von: Hatano, Masashi, et al.
Veröffentlicht: (2024)
von: Hatano, Masashi, et al.
Veröffentlicht: (2024)
Masked Video and Body-worn IMU Autoencoder for Egocentric Action Recognition
von: Zhang, Mingfang, et al.
Veröffentlicht: (2024)
von: Zhang, Mingfang, et al.
Veröffentlicht: (2024)
Cross-view Action Recognition Understanding From Exocentric to Egocentric Perspective
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
Spherical World-Locking for Audio-Visual Localization in Egocentric Videos
von: Yun, Heeseung, et al.
Veröffentlicht: (2024)
von: Yun, Heeseung, et al.
Veröffentlicht: (2024)
DP-DeGauss: Dynamic Probabilistic Gaussian Decomposition for Egocentric 4D Scene Reconstruction
von: Chen, Tingxi, et al.
Veröffentlicht: (2026)
von: Chen, Tingxi, et al.
Veröffentlicht: (2026)
EgoLifter: Open-world 3D Segmentation for Egocentric Perception
von: Gu, Qiao, et al.
Veröffentlicht: (2024)
von: Gu, Qiao, et al.
Veröffentlicht: (2024)
Efficient Egocentric Action Recognition with Multimodal Data
von: Calzavara, Marco, et al.
Veröffentlicht: (2025)
von: Calzavara, Marco, et al.
Veröffentlicht: (2025)
EgoLCD: Egocentric Video Generation with Long Context Diffusion
von: Zhang, Liuzhou, et al.
Veröffentlicht: (2025)
von: Zhang, Liuzhou, et al.
Veröffentlicht: (2025)
CalibFree: Self-Supervised View Feature Separation for Calibration-Free Multi-Camera Multi-Object Tracking
von: Xian, Ruiqi, et al.
Veröffentlicht: (2026)
von: Xian, Ruiqi, et al.
Veröffentlicht: (2026)
WHOLE: World-Grounded Hand-Object Lifted from Egocentric Videos
von: Ye, Yufei, et al.
Veröffentlicht: (2026)
von: Ye, Yufei, et al.
Veröffentlicht: (2026)
Walk through Paintings: Egocentric World Models from Internet Priors
von: Bagchi, Anurag, et al.
Veröffentlicht: (2026)
von: Bagchi, Anurag, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ProbRes: Probabilistic Jump Diffusion for Open-World Egocentric Activity Recognition
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2025) -
Hallucinate, Ground, Repeat: A Framework for Generalized Visual Relationship Detection
von: Vellamcheti, Shanmukha, et al.
Veröffentlicht: (2025) -
ALGO: Object-Grounded Visual Commonsense Reasoning for Open-World Egocentric Action Recognition
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2024) -
Discovering Novel Actions from Open World Egocentric Videos with Object-Grounded Visual Commonsense Reasoning
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2023) -
CVT-Bench: Counterfactual Viewpoint Transformations Reveal Unstable Spatial Representations in Multimodal LLMs
von: Vellamcheti, Shanmukha, et al.
Veröffentlicht: (2026)