EASE: Embodied Active Event Perception via Self-Supervised Energy Minimization
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Zhou, Kundu, Sanjoy, Baweja, Harsimran S., Aakur, Sathyanarayanan N. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Probabilistic Jump-Diffusion Framework for Open-World Egocentric Activity Recognition
by: Kundu, Sanjoy, et al.
Published: (2025)
by: Kundu, Sanjoy, et al.
Published: (2025)
Hallucinate, Ground, Repeat: A Framework for Generalized Visual Relationship Detection
by: Vellamcheti, Shanmukha, et al.
Published: (2025)
by: Vellamcheti, Shanmukha, et al.
Published: (2025)
ProbRes: Probabilistic Jump Diffusion for Open-World Egocentric Activity Recognition
by: Kundu, Sanjoy, et al.
Published: (2025)
by: Kundu, Sanjoy, et al.
Published: (2025)
Discovering Novel Actions from Open World Egocentric Videos with Object-Grounded Visual Commonsense Reasoning
by: Kundu, Sanjoy, et al.
Published: (2023)
by: Kundu, Sanjoy, et al.
Published: (2023)
ALGO: Object-Grounded Visual Commonsense Reasoning for Open-World Egocentric Action Recognition
by: Kundu, Sanjoy, et al.
Published: (2024)
by: Kundu, Sanjoy, et al.
Published: (2024)
Generalized Event Partonomy Inference with Structured Hierarchical Predictive Learning
by: Chen, Zhou, et al.
Published: (2025)
by: Chen, Zhou, et al.
Published: (2025)
Self-supervised Multi-actor Social Activity Understanding in Streaming Videos
by: Trehan, Shubham, et al.
Published: (2024)
by: Trehan, Shubham, et al.
Published: (2024)
CRAFT: A Neuro-Symbolic Framework for Visual Functional Affordance Grounding
by: Chen, Zhou, et al.
Published: (2025)
by: Chen, Zhou, et al.
Published: (2025)
Event Camera Meets Mobile Embodied Perception: Abstraction, Algorithm, Acceleration, Application
by: Wang, Haoyang, et al.
Published: (2025)
by: Wang, Haoyang, et al.
Published: (2025)
STaTS: Structure-Aware Temporal Sequence Summarization via Statistical Window Merging
by: Bhowmick, Disharee, et al.
Published: (2025)
by: Bhowmick, Disharee, et al.
Published: (2025)
Strategy-Supervised Autonomous Laparoscopic Camera Control via Event-Driven Graph Mining
by: Zhou, Keyu, et al.
Published: (2026)
by: Zhou, Keyu, et al.
Published: (2026)
Look, Zoom, Understand: The Robotic Eyeball for Embodied Perception
by: Yang, Jiashu, et al.
Published: (2025)
by: Yang, Jiashu, et al.
Published: (2025)
Capturing Temporal Components for Time Series Classification
by: Vavilthota, Venkata Ragavendra, et al.
Published: (2024)
by: Vavilthota, Venkata Ragavendra, et al.
Published: (2024)
CVT-Bench: Counterfactual Viewpoint Transformations Reveal Unstable Spatial Representations in Multimodal LLMs
by: Vellamcheti, Shanmukha, et al.
Published: (2026)
by: Vellamcheti, Shanmukha, et al.
Published: (2026)
On-Device Self-Supervised Learning of Low-Latency Monocular Depth from Only Events
by: Hagenaars, Jesse, et al.
Published: (2024)
by: Hagenaars, Jesse, et al.
Published: (2024)
Perception Matters: Enhancing Embodied AI with Uncertainty-Aware Semantic Segmentation
by: Prasanna, Sai, et al.
Published: (2024)
by: Prasanna, Sai, et al.
Published: (2024)
ActiveUMI: Robotic Manipulation with Active Perception from Robot-Free Human Demonstrations
by: Zeng, Qiyuan, et al.
Published: (2025)
by: Zeng, Qiyuan, et al.
Published: (2025)
EventFly: Event Camera Perception from Ground to the Sky
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
by: Shang, Yu, et al.
Published: (2026)
by: Shang, Yu, et al.
Published: (2026)
Learning Underwater Active Perception in Simulation
by: Cardaillac, Alexandre, et al.
Published: (2025)
by: Cardaillac, Alexandre, et al.
Published: (2025)
FSP-DETR: Few-Shot Prototypical Parasitic Ova Detection
by: Trehan, Shubham, et al.
Published: (2025)
by: Trehan, Shubham, et al.
Published: (2025)
Self-Supervised Bootstrapping of Action-Predictive Embodied Reasoning
by: Ganai, Milan, et al.
Published: (2026)
by: Ganai, Milan, et al.
Published: (2026)
Embodied Scene Understanding for Vision Language Models via MetaVQA
by: Wang, Weizhen, et al.
Published: (2025)
by: Wang, Weizhen, et al.
Published: (2025)
Robo-Cortex: A Self-Evolving Embodied Agent via Dual-Grain Cognitive Memory and Autonomous Knowledge Induction
by: Chan, Nga Teng, et al.
Published: (2026)
by: Chan, Nga Teng, et al.
Published: (2026)
SaPaVe: Towards Active Perception and Manipulation in Vision-Language-Action Models for Robotics
by: Liu, Mengzhen, et al.
Published: (2026)
by: Liu, Mengzhen, et al.
Published: (2026)
Test-Time Certifiable Self-Supervision to Bridge the Sim2Real Gap in Event-Based Satellite Pose Estimation
by: Jawaid, Mohsi, et al.
Published: (2024)
by: Jawaid, Mohsi, et al.
Published: (2024)
RobotPan: A 360$^\circ$ Surround-View Robotic Vision System for Embodied Perception
by: Ma, Jiahao, et al.
Published: (2026)
by: Ma, Jiahao, et al.
Published: (2026)
FreeOcc: Training-Free Embodied Open-Vocabulary Occupancy Prediction
by: Jiang, Zeyu, et al.
Published: (2026)
by: Jiang, Zeyu, et al.
Published: (2026)
Ego to World: Collaborative Spatial Reasoning in Embodied Systems via Reinforcement Learning
by: Zhou, Heng, et al.
Published: (2026)
by: Zhou, Heng, et al.
Published: (2026)
SemanticFlow: A Self-Supervised Framework for Joint Scene Flow Prediction and Instance Segmentation in Dynamic Environments
by: Chen, Yinqi, et al.
Published: (2025)
by: Chen, Yinqi, et al.
Published: (2025)
Learning 3D Persistent Embodied World Models
by: Zhou, Siyuan, et al.
Published: (2025)
by: Zhou, Siyuan, et al.
Published: (2025)
An Event-Based Perception Pipeline for a Table Tennis Robot
by: Ziegler, Andreas, et al.
Published: (2025)
by: Ziegler, Andreas, et al.
Published: (2025)
DST-Calib: A Dual-Path, Self-Supervised, Target-Free LiDAR-Camera Extrinsic Calibration Network
by: Huang, Zhiwei, et al.
Published: (2026)
by: Huang, Zhiwei, et al.
Published: (2026)
HERE: Hierarchical Active Exploration of Radiance Field with Epistemic Uncertainty Minimization
by: Lee, Taekbeom, et al.
Published: (2026)
by: Lee, Taekbeom, et al.
Published: (2026)
Emergent Active Perception and Dexterity of Simulated Humanoids from Visual Reinforcement Learning
by: Luo, Zhengyi, et al.
Published: (2025)
by: Luo, Zhengyi, et al.
Published: (2025)
EmbodiedBrain: Expanding Performance Boundaries of Task Planning for Embodied Intelligence
by: Zou, Ding, et al.
Published: (2025)
by: Zou, Ding, et al.
Published: (2025)
Towards Physically Realizable Adversarial Attacks in Embodied Vision Navigation
by: Chen, Meng, et al.
Published: (2024)
by: Chen, Meng, et al.
Published: (2024)
DualVLA: Building a Generalizable Embodied Agent via Partial Decoupling of Reasoning and Action
by: Fang, Zhen, et al.
Published: (2025)
by: Fang, Zhen, et al.
Published: (2025)
EnerVerse-AC: Envisioning Embodied Environments with Action Condition
by: Jiang, Yuxin, et al.
Published: (2025)
by: Jiang, Yuxin, et al.
Published: (2025)
Embodied Image Captioning: Self-supervised Learning Agents for Spatially Coherent Image Descriptions
by: Galliena, Tommaso, et al.
Published: (2025)
by: Galliena, Tommaso, et al.
Published: (2025)
Similar Items
-
A Probabilistic Jump-Diffusion Framework for Open-World Egocentric Activity Recognition
by: Kundu, Sanjoy, et al.
Published: (2025) -
Hallucinate, Ground, Repeat: A Framework for Generalized Visual Relationship Detection
by: Vellamcheti, Shanmukha, et al.
Published: (2025) -
ProbRes: Probabilistic Jump Diffusion for Open-World Egocentric Activity Recognition
by: Kundu, Sanjoy, et al.
Published: (2025) -
Discovering Novel Actions from Open World Egocentric Videos with Object-Grounded Visual Commonsense Reasoning
by: Kundu, Sanjoy, et al.
Published: (2023) -
ALGO: Object-Grounded Visual Commonsense Reasoning for Open-World Egocentric Action Recognition
by: Kundu, Sanjoy, et al.
Published: (2024)