ViPro-2: Unsupervised State Estimation via Integrated Dynamics for Guiding Video Prediction
Fuente:
arXiv
Salvato in:
| Autori principali: | Takenaka, Patrick, Maucher, Johannes, Huber, Marco F. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ViPro: Enabling and Controlling Video Prediction for Complex Dynamical Scenarios using Procedural Knowledge
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
Guiding Video Prediction with Explicit Procedural Knowledge
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
Anonymization of Documents for Law Enforcement with Machine Learning
di: Eberhardinger, Manuel, et al.
Pubblicazione: (2025)
di: Eberhardinger, Manuel, et al.
Pubblicazione: (2025)
Classification of Inkjet Printers based on Droplet Statistics
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
di: Takenaka, Patrick, et al.
Pubblicazione: (2024)
Automatic Odometry-Less OpenDRIVE Generation From Sparse Point Clouds
di: Eisemann, Leon, et al.
Pubblicazione: (2024)
di: Eisemann, Leon, et al.
Pubblicazione: (2024)
ViSMaP: Unsupervised Hour-long Video Summarisation by Meta-Prompting
di: Hu, Jian, et al.
Pubblicazione: (2025)
di: Hu, Jian, et al.
Pubblicazione: (2025)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
di: Siméoni, Oriane, et al.
Pubblicazione: (2023)
di: Siméoni, Oriane, et al.
Pubblicazione: (2023)
Integrating Disparity Confidence Estimation into Relative Depth Prior-Guided Unsupervised Stereo Matching
di: Liu, Chuang-Wei, et al.
Pubblicazione: (2025)
di: Liu, Chuang-Wei, et al.
Pubblicazione: (2025)
ViASNet: A Video Ad Saliency Network for Predicting Dynamic Saliency and Viewer Engagement
di: Ye, Jianping, et al.
Pubblicazione: (2026)
di: Ye, Jianping, et al.
Pubblicazione: (2026)
ViSAGE @ NTIRE 2026 Challenge on Video Saliency Prediction
di: Wang, Kun, et al.
Pubblicazione: (2026)
di: Wang, Kun, et al.
Pubblicazione: (2026)
ConViS-Bench: Estimating Video Similarity Through Semantic Concepts
di: Liberatori, Benedetta, et al.
Pubblicazione: (2025)
di: Liberatori, Benedetta, et al.
Pubblicazione: (2025)
Guided Slot Attention for Unsupervised Video Object Segmentation
di: Lee, Minhyeok, et al.
Pubblicazione: (2023)
di: Lee, Minhyeok, et al.
Pubblicazione: (2023)
DiViD: Disentangled Video Diffusion for Static-Dynamic Factorization
di: Gheisari, Marzieh, et al.
Pubblicazione: (2025)
di: Gheisari, Marzieh, et al.
Pubblicazione: (2025)
CoProU-VO: Combining Projected Uncertainty for End-to-End Unsupervised Monocular Visual Odometry
di: Xie, Jingchao, et al.
Pubblicazione: (2025)
di: Xie, Jingchao, et al.
Pubblicazione: (2025)
OCK: Unsupervised Dynamic Video Prediction with Object-Centric Kinematics
di: Song, Yeon-Ji, et al.
Pubblicazione: (2024)
di: Song, Yeon-Ji, et al.
Pubblicazione: (2024)
ViViD: Video Virtual Try-on using Diffusion Models
di: Fang, Zixun, et al.
Pubblicazione: (2024)
di: Fang, Zixun, et al.
Pubblicazione: (2024)
Learning Physics From Video: Unsupervised Physical Parameter Estimation for Continuous Dynamical Systems
di: Garcia, Alejandro Castañeda, et al.
Pubblicazione: (2024)
di: Garcia, Alejandro Castañeda, et al.
Pubblicazione: (2024)
Exploring Plain ViT Reconstruction for Multi-class Unsupervised Anomaly Detection
di: Zhang, Jiangning, et al.
Pubblicazione: (2023)
di: Zhang, Jiangning, et al.
Pubblicazione: (2023)
ViKey: Enhancing Temporal Understanding in Videos via Visual Prompting
di: Lee, Yeonkyung, et al.
Pubblicazione: (2026)
di: Lee, Yeonkyung, et al.
Pubblicazione: (2026)
ProDyG: Progressive Dynamic Scene Reconstruction via Gaussian Splatting from Monocular Videos
di: Chen, Shi, et al.
Pubblicazione: (2025)
di: Chen, Shi, et al.
Pubblicazione: (2025)
Unsupervised Monocular Depth Estimation Based on Hierarchical Feature-Guided Diffusion
di: Liu, Runze, et al.
Pubblicazione: (2024)
di: Liu, Runze, et al.
Pubblicazione: (2024)
Future Slot Prediction for Unsupervised Object Discovery in Surgical Video
di: Liao, Guiqiu, et al.
Pubblicazione: (2025)
di: Liao, Guiqiu, et al.
Pubblicazione: (2025)
VersaViT: Enhancing MLLM Vision Backbones via Task-Guided Optimization
di: Liu, Yikun, et al.
Pubblicazione: (2026)
di: Liu, Yikun, et al.
Pubblicazione: (2026)
Boosting Unsupervised Video Instance Segmentation with Automatic Quality-Guided Self-Training
di: Lu, Kaixuan, et al.
Pubblicazione: (2025)
di: Lu, Kaixuan, et al.
Pubblicazione: (2025)
Unsupervised Cross-Domain 3D Human Pose Estimation via Pseudo-Label-Guided Global Transforms
di: Liu, Jingjing, et al.
Pubblicazione: (2025)
di: Liu, Jingjing, et al.
Pubblicazione: (2025)
AdaViPro: Region-based Adaptive Visual Prompt for Large-Scale Models Adapting
di: Yang, Mengyu, et al.
Pubblicazione: (2024)
di: Yang, Mengyu, et al.
Pubblicazione: (2024)
ViDiC: Video Difference Captioning
di: Wu, Jiangtao, et al.
Pubblicazione: (2025)
di: Wu, Jiangtao, et al.
Pubblicazione: (2025)
DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles
di: Daniel, Tal, et al.
Pubblicazione: (2023)
di: Daniel, Tal, et al.
Pubblicazione: (2023)
Vid2Avatar-Pro: Authentic Avatar from Videos in the Wild via Universal Prior
di: Guo, Chen, et al.
Pubblicazione: (2025)
di: Guo, Chen, et al.
Pubblicazione: (2025)
ViLA: Efficient Video-Language Alignment for Video Question Answering
di: Wang, Xijun, et al.
Pubblicazione: (2023)
di: Wang, Xijun, et al.
Pubblicazione: (2023)
ViSpeak: Visual Instruction Feedback in Streaming Videos
di: Fu, Shenghao, et al.
Pubblicazione: (2025)
di: Fu, Shenghao, et al.
Pubblicazione: (2025)
ViLL-E: Video LLM Embeddings for Retrieval
di: Gupta, Rohit, et al.
Pubblicazione: (2026)
di: Gupta, Rohit, et al.
Pubblicazione: (2026)
MoViE: Mobile Diffusion for Video Editing
di: Karjauv, Adil, et al.
Pubblicazione: (2024)
di: Karjauv, Adil, et al.
Pubblicazione: (2024)
LAMM-ViT: AI Face Detection via Layer-Aware Modulation of Region-Guided Attention
di: Zhang, Jiangling, et al.
Pubblicazione: (2025)
di: Zhang, Jiangling, et al.
Pubblicazione: (2025)
Self-supervised Optimization of Hand Pose Estimation using Anatomical Features and Iterative Learning
di: Jauch, Christian, et al.
Pubblicazione: (2023)
di: Jauch, Christian, et al.
Pubblicazione: (2023)
TiViBench: Benchmarking Think-in-Video Reasoning for Video Generative Models
di: Chen, Harold Haodong, et al.
Pubblicazione: (2025)
di: Chen, Harold Haodong, et al.
Pubblicazione: (2025)
VideoPro: Adaptive Program Reasoning for Long Video Understanding
di: Li, Chenglin, et al.
Pubblicazione: (2025)
di: Li, Chenglin, et al.
Pubblicazione: (2025)
DynaGuide: A Generalizable Dynamic Guidance Framework for Unsupervised Semantic Segmentation
di: Guermazi, Boujemaa, et al.
Pubblicazione: (2026)
di: Guermazi, Boujemaa, et al.
Pubblicazione: (2026)
ViThinker: Active Vision-Language Reasoning via Dynamic Perceptual Querying
di: You, Weihang, et al.
Pubblicazione: (2026)
di: You, Weihang, et al.
Pubblicazione: (2026)
ViMU: Benchmarking Video Metaphorical Understanding
di: Li, Qi, et al.
Pubblicazione: (2026)
di: Li, Qi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
ViPro: Enabling and Controlling Video Prediction for Complex Dynamical Scenarios using Procedural Knowledge
di: Takenaka, Patrick, et al.
Pubblicazione: (2024) -
Guiding Video Prediction with Explicit Procedural Knowledge
di: Takenaka, Patrick, et al.
Pubblicazione: (2024) -
Anonymization of Documents for Law Enforcement with Machine Learning
di: Eberhardinger, Manuel, et al.
Pubblicazione: (2025) -
Classification of Inkjet Printers based on Droplet Statistics
di: Takenaka, Patrick, et al.
Pubblicazione: (2024) -
Automatic Odometry-Less OpenDRIVE Generation From Sparse Point Clouds
di: Eisemann, Leon, et al.
Pubblicazione: (2024)