Generalized Event Partonomy Inference with Structured Hierarchical Predictive Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Zhou, Lin, Joe, Aakur\\, Sathyanarayanan N. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CRAFT: A Neuro-Symbolic Framework for Visual Functional Affordance Grounding
by: Chen, Zhou, et al.
Published: (2025)
by: Chen, Zhou, et al.
Published: (2025)
EASE: Embodied Active Event Perception via Self-Supervised Energy Minimization
by: Chen, Zhou, et al.
Published: (2025)
by: Chen, Zhou, et al.
Published: (2025)
Self-supervised Multi-actor Social Activity Understanding in Streaming Videos
by: Trehan, Shubham, et al.
Published: (2024)
by: Trehan, Shubham, et al.
Published: (2024)
Hallucinate, Ground, Repeat: A Framework for Generalized Visual Relationship Detection
by: Vellamcheti, Shanmukha, et al.
Published: (2025)
by: Vellamcheti, Shanmukha, et al.
Published: (2025)
STaTS: Structure-Aware Temporal Sequence Summarization via Statistical Window Merging
by: Bhowmick, Disharee, et al.
Published: (2025)
by: Bhowmick, Disharee, et al.
Published: (2025)
A Probabilistic Jump-Diffusion Framework for Open-World Egocentric Activity Recognition
by: Kundu, Sanjoy, et al.
Published: (2025)
by: Kundu, Sanjoy, et al.
Published: (2025)
ProbRes: Probabilistic Jump Diffusion for Open-World Egocentric Activity Recognition
by: Kundu, Sanjoy, et al.
Published: (2025)
by: Kundu, Sanjoy, et al.
Published: (2025)
Discovering Novel Actions from Open World Egocentric Videos with Object-Grounded Visual Commonsense Reasoning
by: Kundu, Sanjoy, et al.
Published: (2023)
by: Kundu, Sanjoy, et al.
Published: (2023)
ALGO: Object-Grounded Visual Commonsense Reasoning for Open-World Egocentric Action Recognition
by: Kundu, Sanjoy, et al.
Published: (2024)
by: Kundu, Sanjoy, et al.
Published: (2024)
Capturing Temporal Components for Time Series Classification
by: Vavilthota, Venkata Ragavendra, et al.
Published: (2024)
by: Vavilthota, Venkata Ragavendra, et al.
Published: (2024)
CVT-Bench: Counterfactual Viewpoint Transformations Reveal Unstable Spatial Representations in Multimodal LLMs
by: Vellamcheti, Shanmukha, et al.
Published: (2026)
by: Vellamcheti, Shanmukha, et al.
Published: (2026)
FSP-DETR: Few-Shot Prototypical Parasitic Ova Detection
by: Trehan, Shubham, et al.
Published: (2025)
by: Trehan, Shubham, et al.
Published: (2025)
Learning to Generate Diverse Pedestrian Movements from Web Videos with Noisy Labels
by: Liu, Zhizheng, et al.
Published: (2024)
by: Liu, Zhizheng, et al.
Published: (2024)
Structured Context Learning for Generic Event Boundary Detection
by: Gu, Xin, et al.
Published: (2025)
by: Gu, Xin, et al.
Published: (2025)
HyperST: Hierarchical Hyperbolic Learning for Spatial Transcriptomics Prediction
by: Zhang, Chen, et al.
Published: (2025)
by: Zhang, Chen, et al.
Published: (2025)
Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation
by: Chen, Gordon, et al.
Published: (2026)
by: Chen, Gordon, et al.
Published: (2026)
Revisit Event Generation Model: Self-Supervised Learning of Event-to-Video Reconstruction with Implicit Neural Representations
by: Wang, Zipeng, et al.
Published: (2024)
by: Wang, Zipeng, et al.
Published: (2024)
SuperEIO: Self-Supervised Event Feature Learning for Event Inertial Odometry
by: Chen, Peiyu, et al.
Published: (2025)
by: Chen, Peiyu, et al.
Published: (2025)
OmniEvent: Unified Event Representation Learning
by: Yan, Weiqi, et al.
Published: (2025)
by: Yan, Weiqi, et al.
Published: (2025)
EventBind: Learning a Unified Representation to Bind Them All for Event-based Open-world Understanding
by: Zhou, Jiazhou, et al.
Published: (2023)
by: Zhou, Jiazhou, et al.
Published: (2023)
Elite-EvGS: Learning Event-based 3D Gaussian Splatting by Distilling Event-to-Video Priors
by: Zhang, Zixin, et al.
Published: (2024)
by: Zhang, Zixin, et al.
Published: (2024)
OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning
by: Liang, Zhijia, et al.
Published: (2026)
by: Liang, Zhijia, et al.
Published: (2026)
Temporal-contextual Event Learning for Pedestrian Crossing Intent Prediction
by: Liang, Hongbin, et al.
Published: (2025)
by: Liang, Hongbin, et al.
Published: (2025)
An image speaks a thousand words, but can everyone listen? On image transcreation for cultural relevance
by: Khanuja, Simran, et al.
Published: (2024)
by: Khanuja, Simran, et al.
Published: (2024)
EchoFoley: Event-Centric Hierarchical Control for Video Grounded Creative Sound Generation
by: Li, Bingxuan, et al.
Published: (2025)
by: Li, Bingxuan, et al.
Published: (2025)
Bring Event into RGB and LiDAR: Hierarchical Visual-Motion Fusion for Scene Flow
by: Zhou, Hanyu, et al.
Published: (2024)
by: Zhou, Hanyu, et al.
Published: (2024)
Energy-Aware Imitation Learning for Steering Prediction Using Events and Frames
by: Cao, Hu, et al.
Published: (2026)
by: Cao, Hu, et al.
Published: (2026)
Hierarchical Brain Structure Modeling for Predicting Genotype of Glioma
by: Tang, Haotian, et al.
Published: (2025)
by: Tang, Haotian, et al.
Published: (2025)
Feudal Steering: Hierarchical Learning for Steering Angle Prediction
by: Johnson, Faith, et al.
Published: (2020)
by: Johnson, Faith, et al.
Published: (2020)
PEPR: Privileged Event-based Predictive Regularization for Domain Generalization
by: Magrini, Gabriele, et al.
Published: (2026)
by: Magrini, Gabriele, et al.
Published: (2026)
Hierarchical Banzhaf Interaction for General Video-Language Representation Learning
by: Jin, Peng, et al.
Published: (2024)
by: Jin, Peng, et al.
Published: (2024)
Self-supervised Representation Learning for Cell Event Recognition through Time Arrow Prediction
by: Chen, Cangxiong, et al.
Published: (2024)
by: Chen, Cangxiong, et al.
Published: (2024)
PASS: Path-selective State Space Model for Event-based Recognition
by: Zhou, Jiazhou, et al.
Published: (2024)
by: Zhou, Jiazhou, et al.
Published: (2024)
State-Change Learning for Prediction of Future Events in Endoscopic Videos
by: Sharma, Saurav, et al.
Published: (2025)
by: Sharma, Saurav, et al.
Published: (2025)
Video-as-Answer: Predict and Generate Next Video Event with Joint-GRPO
by: Cheng, Junhao, et al.
Published: (2025)
by: Cheng, Junhao, et al.
Published: (2025)
HVIS: A Human-like Vision and Inference System for Human Motion Prediction
by: Lyu, Kedi, et al.
Published: (2025)
by: Lyu, Kedi, et al.
Published: (2025)
Event-Customized Image Generation
by: Wang, Zhen, et al.
Published: (2024)
by: Wang, Zhen, et al.
Published: (2024)
Autobiasing Event Cameras for Flickering Mitigation
by: Dilmaghani, Mehdi Sefidgar, et al.
Published: (2025)
by: Dilmaghani, Mehdi Sefidgar, et al.
Published: (2025)
UniHDSA: A Unified Relation Prediction Approach for Hierarchical Document Structure Analysis
by: Wang, Jiawei, et al.
Published: (2025)
by: Wang, Jiawei, et al.
Published: (2025)
EventMemAgent: Hierarchical Event-Centric Memory for Online Video Understanding with Adaptive Tool Use
by: Wen, Siwei, et al.
Published: (2026)
by: Wen, Siwei, et al.
Published: (2026)
Similar Items
-
CRAFT: A Neuro-Symbolic Framework for Visual Functional Affordance Grounding
by: Chen, Zhou, et al.
Published: (2025) -
EASE: Embodied Active Event Perception via Self-Supervised Energy Minimization
by: Chen, Zhou, et al.
Published: (2025) -
Self-supervised Multi-actor Social Activity Understanding in Streaming Videos
by: Trehan, Shubham, et al.
Published: (2024) -
Hallucinate, Ground, Repeat: A Framework for Generalized Visual Relationship Detection
by: Vellamcheti, Shanmukha, et al.
Published: (2025) -
STaTS: Structure-Aware Temporal Sequence Summarization via Statistical Window Merging
by: Bhowmick, Disharee, et al.
Published: (2025)