What Changed and What Could Have Changed? State-Change Counterfactuals for Procedure-Aware Video Representation Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Kung, Chi-Hsi, Ramirez, Frangil, Ha, Juhyung, Chen, Yi-Ting, Crandall, David, Tsai, Yi-Hsuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
EgoVIS@CVPR: What Changed and What Could Have Changed? State-Change Counterfactuals for Procedure-Aware Video Representation Learning
di: Kung, Chi-Hsi, et al.
Pubblicazione: (2025)
di: Kung, Chi-Hsi, et al.
Pubblicazione: (2025)
GateFusion: Hierarchical Gated Cross-Modal Fusion for Active Speaker Detection
di: Wang, Yu, et al.
Pubblicazione: (2025)
di: Wang, Yu, et al.
Pubblicazione: (2025)
Action-slot: Visual Action-centric Representations for Multi-label Atomic Activity Recognition in Traffic Scenes
di: Kung, Chi-Hsi, et al.
Pubblicazione: (2023)
di: Kung, Chi-Hsi, et al.
Pubblicazione: (2023)
EgoVIS@CVPR: PAIR-Net: Enhancing Egocentric Speaker Detection via Pretrained Audio-Visual Fusion and Alignment Loss
di: Wang, Yu, et al.
Pubblicazione: (2025)
di: Wang, Yu, et al.
Pubblicazione: (2025)
A solution to generalized learning from small training sets found in infant repeated visual experiences of individual objects
di: Ramirez, Frangil, et al.
Pubblicazione: (2025)
di: Ramirez, Frangil, et al.
Pubblicazione: (2025)
Anticipating Object State Changes in Long Procedural Videos
di: Manousaki, Victoria, et al.
Pubblicazione: (2024)
di: Manousaki, Victoria, et al.
Pubblicazione: (2024)
Multi-resolution Guided 3D GANs for Medical Image Translation
di: Ha, Juhyung, et al.
Pubblicazione: (2024)
di: Ha, Juhyung, et al.
Pubblicazione: (2024)
Imagine How To Change: Explicit Procedure Modeling for Change Captioning
di: Sun, Jiayang, et al.
Pubblicazione: (2026)
di: Sun, Jiayang, et al.
Pubblicazione: (2026)
Show Me What and Where has Changed? Question Answering and Grounding for Remote Sensing Change Detection
di: Li, Ke, et al.
Pubblicazione: (2024)
di: Li, Ke, et al.
Pubblicazione: (2024)
Learning Object State Changes in Videos: An Open-World Perspective
di: Xue, Zihui, et al.
Pubblicazione: (2023)
di: Xue, Zihui, et al.
Pubblicazione: (2023)
State-Change Learning for Prediction of Future Events in Endoscopic Videos
di: Sharma, Saurav, et al.
Pubblicazione: (2025)
di: Sharma, Saurav, et al.
Pubblicazione: (2025)
ATARS: An Aerial Traffic Atomic Activity Recognition and Temporal Segmentation Dataset
di: Chen, Zihao, et al.
Pubblicazione: (2025)
di: Chen, Zihao, et al.
Pubblicazione: (2025)
Toward Real-world BEV Perception: Depth Uncertainty Estimation via Gaussian Splatting
di: Lu, Shu-Wei, et al.
Pubblicazione: (2025)
di: Lu, Shu-Wei, et al.
Pubblicazione: (2025)
Motion Modes: What Could Happen Next?
di: Pandey, Karran, et al.
Pubblicazione: (2024)
di: Pandey, Karran, et al.
Pubblicazione: (2024)
SPOC: Spatially-Progressing Object State Change Segmentation in Video
di: Mandikal, Priyanka, et al.
Pubblicazione: (2025)
di: Mandikal, Priyanka, et al.
Pubblicazione: (2025)
Scene Change Detection with Vision-Language Representation Learning
di: Sheng, Diwei, et al.
Pubblicazione: (2026)
di: Sheng, Diwei, et al.
Pubblicazione: (2026)
OSCaR: Object State Captioning and State Change Representation
di: Nguyen, Nguyen, et al.
Pubblicazione: (2024)
di: Nguyen, Nguyen, et al.
Pubblicazione: (2024)
IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video Generation
di: Lin, Yuanze, et al.
Pubblicazione: (2025)
di: Lin, Yuanze, et al.
Pubblicazione: (2025)
The Modern Public Library and Melvil Dewey: What He Changed, What We've Changed, and What Hasn't Changed.
di: Salmon, Laura E.
Pubblicazione: (2000)
di: Salmon, Laura E.
Pubblicazione: (2000)
Efficient Remote Sensing Change Detection with Change State Space Models
di: Ghazaei, Elman, et al.
Pubblicazione: (2025)
di: Ghazaei, Elman, et al.
Pubblicazione: (2025)
What If Trojan Horse Nanoparticles Could Change the Game for HPV Gene‐Targeted Therapies?
di: Trairong Chokwassanasakulkit, et al.
Pubblicazione: (2026)
di: Trairong Chokwassanasakulkit, et al.
Pubblicazione: (2026)
Attention-based Shape and Gait Representations Learning for Video-based Cloth-Changing Person Re-Identification
di: Nguyen, Vuong D., et al.
Pubblicazione: (2024)
di: Nguyen, Vuong D., et al.
Pubblicazione: (2024)
Learning Procedural-aware Video Representations through State-Grounded Hierarchy Unfolding
di: Zhao, Jinghan, et al.
Pubblicazione: (2025)
di: Zhao, Jinghan, et al.
Pubblicazione: (2025)
See What You Seek: Semantic Contextual Integration for Cloth-Changing Person Re-Identification
di: Han, Xiyu, et al.
Pubblicazione: (2024)
di: Han, Xiyu, et al.
Pubblicazione: (2024)
OSCBench: Benchmarking Object State Change in Text-to-Video Generation
di: Han, Xianjing, et al.
Pubblicazione: (2026)
di: Han, Xianjing, et al.
Pubblicazione: (2026)
What You Have is What You Track: Adaptive and Robust Multimodal Tracking
di: Tan, Yuedong, et al.
Pubblicazione: (2025)
di: Tan, Yuedong, et al.
Pubblicazione: (2025)
Change3D: Revisiting Change Detection and Captioning from A Video Modeling Perspective
di: Zhu, Duowang, et al.
Pubblicazione: (2025)
di: Zhu, Duowang, et al.
Pubblicazione: (2025)
Potential Field as Scene Affordance for Behavior Change-Based Visual Risk Object Identification
di: Pao, Pang-Yuan, et al.
Pubblicazione: (2024)
di: Pao, Pang-Yuan, et al.
Pubblicazione: (2024)
What has Changed? Has Anything Changed? What Needs to Change? Does Anything Need to Change? Late‐Career Reflections 1
di: Andrew Samuels
Pubblicazione: (2026)
di: Andrew Samuels
Pubblicazione: (2026)
Label What Matters: Modality-Balanced and Difficulty-Aware Multimodal Active Learning
di: Zeng, Yuqiao, et al.
Pubblicazione: (2026)
di: Zeng, Yuqiao, et al.
Pubblicazione: (2026)
PROFUSEme: PROstate Cancer Biochemical Recurrence Prediction via FUSEd Multi-modal Embeddings
di: You, Suhang, et al.
Pubblicazione: (2025)
di: You, Suhang, et al.
Pubblicazione: (2025)
How Video Meetings Change Your Expression
di: Sarin, Sumit, et al.
Pubblicazione: (2024)
di: Sarin, Sumit, et al.
Pubblicazione: (2024)
ChangeBind: A Hybrid Change Encoder for Remote Sensing Change Detection
di: Noman, Mubashir, et al.
Pubblicazione: (2024)
di: Noman, Mubashir, et al.
Pubblicazione: (2024)
SAM-Based Building Change Detection with Distribution-Aware Fourier Adaptation and Edge-Constrained Warping
di: Li, Yun-Cheng, et al.
Pubblicazione: (2025)
di: Li, Yun-Cheng, et al.
Pubblicazione: (2025)
Weakly Supervised 3D Object Detection via Multi-Level Visual Guidance
di: Huang, Kuan-Chih, et al.
Pubblicazione: (2023)
di: Huang, Kuan-Chih, et al.
Pubblicazione: (2023)
See More, Change Less: Anatomy-Aware Diffusion for Contrast Enhancement
di: Liu, Junqi, et al.
Pubblicazione: (2025)
di: Liu, Junqi, et al.
Pubblicazione: (2025)
Exemplar Masking for Multimodal Incremental Learning
di: Lee, Yi-Lun, et al.
Pubblicazione: (2024)
di: Lee, Yi-Lun, et al.
Pubblicazione: (2024)
Hierarchical Dual-Change Collaborative Learning for UAV Scene Change Captioning
di: Chen, Fuhai, et al.
Pubblicazione: (2026)
di: Chen, Fuhai, et al.
Pubblicazione: (2026)
ChangeDiff: A Multi-Temporal Change Detection Data Generator with Flexible Text Prompts via Diffusion Model
di: Zang, Qi, et al.
Pubblicazione: (2024)
di: Zang, Qi, et al.
Pubblicazione: (2024)
ChangeMamba: Remote Sensing Change Detection With Spatiotemporal State Space Model
di: Chen, Hongruixuan, et al.
Pubblicazione: (2024)
di: Chen, Hongruixuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
EgoVIS@CVPR: What Changed and What Could Have Changed? State-Change Counterfactuals for Procedure-Aware Video Representation Learning
di: Kung, Chi-Hsi, et al.
Pubblicazione: (2025) -
GateFusion: Hierarchical Gated Cross-Modal Fusion for Active Speaker Detection
di: Wang, Yu, et al.
Pubblicazione: (2025) -
Action-slot: Visual Action-centric Representations for Multi-label Atomic Activity Recognition in Traffic Scenes
di: Kung, Chi-Hsi, et al.
Pubblicazione: (2023) -
EgoVIS@CVPR: PAIR-Net: Enhancing Egocentric Speaker Detection via Pretrained Audio-Visual Fusion and Alignment Loss
di: Wang, Yu, et al.
Pubblicazione: (2025) -
A solution to generalized learning from small training sets found in infant repeated visual experiences of individual objects
di: Ramirez, Frangil, et al.
Pubblicazione: (2025)