Gespeichert in:
| Hauptverfasser: | Etaat, Daniel, Kalaria, Dvij, Rahmanian, Nima, Sastry, Shankar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2503.20936 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TT4D: A Pipeline and Dataset for Table Tennis 4D Reconstruction From Monocular Videos
von: Rahmanian, Nima, et al.
Veröffentlicht: (2026)
von: Rahmanian, Nima, et al.
Veröffentlicht: (2026)
Real-time Accident Anticipation for Autonomous Driving Through Monocular Depth-Enhanced 3D Modeling
von: Liao, Haicheng, et al.
Veröffentlicht: (2024)
von: Liao, Haicheng, et al.
Veröffentlicht: (2024)
Towards Ball Spin and Trajectory Analysis in Table Tennis Broadcast Videos via Physically Grounded Synthetic-to-Real Transfer
von: Kienzle, Daniel, et al.
Veröffentlicht: (2025)
von: Kienzle, Daniel, et al.
Veröffentlicht: (2025)
On Moving Object Segmentation from Monocular Video with Transformers
von: Homeyer, Christian, et al.
Veröffentlicht: (2024)
von: Homeyer, Christian, et al.
Veröffentlicht: (2024)
α-RACER: Real-Time Algorithm for Game-Theoretic Motion Planning and Control in Autonomous Racing using Near-Potential Function
von: Kalaria, Dvij, et al.
Veröffentlicht: (2024)
von: Kalaria, Dvij, et al.
Veröffentlicht: (2024)
GFlow: Recovering 4D World from Monocular Video
von: Wang, Shizun, et al.
Veröffentlicht: (2024)
von: Wang, Shizun, et al.
Veröffentlicht: (2024)
Towards Scene Graph Anticipation
von: Peddi, Rohith, et al.
Veröffentlicht: (2024)
von: Peddi, Rohith, et al.
Veröffentlicht: (2024)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
von: Zuo, Qi, et al.
Veröffentlicht: (2024)
von: Zuo, Qi, et al.
Veröffentlicht: (2024)
MV-S2V: Multi-View Subject-Consistent Video Generation
von: Song, Ziyang, et al.
Veröffentlicht: (2026)
von: Song, Ziyang, et al.
Veröffentlicht: (2026)
Rethinking Video Human-Object Interaction: Set Prediction over Time for Unified Detection and Anticipation
von: Luo, Yuanhao, et al.
Veröffentlicht: (2026)
von: Luo, Yuanhao, et al.
Veröffentlicht: (2026)
SpatialDreamer: Self-supervised Stereo Video Synthesis from Monocular Input
von: Lv, Zhen, et al.
Veröffentlicht: (2024)
von: Lv, Zhen, et al.
Veröffentlicht: (2024)
iHuman: Instant Animatable Digital Humans From Monocular Videos
von: Paudel, Pramish, et al.
Veröffentlicht: (2024)
von: Paudel, Pramish, et al.
Veröffentlicht: (2024)
Automated Tennis Player and Ball Tracking with Court Keypoints Detection (Hawk Eye System)
von: Desu, Venkata Manikanta, et al.
Veröffentlicht: (2025)
von: Desu, Venkata Manikanta, et al.
Veröffentlicht: (2025)
Semantically Guided Representation Learning For Action Anticipation
von: Diko, Anxhelo, et al.
Veröffentlicht: (2024)
von: Diko, Anxhelo, et al.
Veröffentlicht: (2024)
VideoArtGS: Building Digital Twins of Articulated Objects from Monocular Video
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
Endo3R: Unified Online Reconstruction from Dynamic Monocular Endoscopic Video
von: Guo, Jiaxin, et al.
Veröffentlicht: (2025)
von: Guo, Jiaxin, et al.
Veröffentlicht: (2025)
MV-RAG: Retrieval Augmented Multiview Diffusion
von: Dayani, Yosef, et al.
Veröffentlicht: (2025)
von: Dayani, Yosef, et al.
Veröffentlicht: (2025)
Short-term Object Interaction Anticipation with Disentangled Object Detection @ Ego4D Short Term Object Interaction Anticipation Challenge
von: Cho, Hyunjin, et al.
Veröffentlicht: (2024)
von: Cho, Hyunjin, et al.
Veröffentlicht: (2024)
Feed-Forward Bullet-Time Reconstruction of Dynamic Scenes from Monocular Videos
von: Liang, Hanxue, et al.
Veröffentlicht: (2024)
von: Liang, Hanxue, et al.
Veröffentlicht: (2024)
Semantically Guided Action Anticipation
von: Diko, Anxhelo, et al.
Veröffentlicht: (2024)
von: Diko, Anxhelo, et al.
Veröffentlicht: (2024)
MV-Swin-T: Mammogram Classification with Multi-view Swin Transformer
von: Sarker, Sushmita, et al.
Veröffentlicht: (2024)
von: Sarker, Sushmita, et al.
Veröffentlicht: (2024)
4D3R: Motion-Aware Neural Reconstruction and Rendering of Dynamic Scenes from Monocular Videos
von: Guo, Mengqi, et al.
Veröffentlicht: (2025)
von: Guo, Mengqi, et al.
Veröffentlicht: (2025)
Tamaththul3D: High-Fidelity 3D Saudi Sign Language Avatars from Monocular Video
von: Alghamdi, Eyad, et al.
Veröffentlicht: (2026)
von: Alghamdi, Eyad, et al.
Veröffentlicht: (2026)
Predict and Resist: Long-Term Accident Anticipation under Sensor Noise
von: Liu, Xingcheng, et al.
Veröffentlicht: (2025)
von: Liu, Xingcheng, et al.
Veröffentlicht: (2025)
Intention-Guided Cognitive Reasoning for Egocentric Long-Term Action Anticipation
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
GAF: Gaussian Avatar Reconstruction from Monocular Videos via Multi-view Diffusion
von: Tang, Jiapeng, et al.
Veröffentlicht: (2024)
von: Tang, Jiapeng, et al.
Veröffentlicht: (2024)
The Crystal Ball Hypothesis in diffusion models: Anticipating object positions from initial noise
von: Ban, Yuanhao, et al.
Veröffentlicht: (2024)
von: Ban, Yuanhao, et al.
Veröffentlicht: (2024)
EQ-TAA: Equivariant Traffic Accident Anticipation via Diffusion-Based Accident Video Synthesis
von: Fang, Jianwu, et al.
Veröffentlicht: (2025)
von: Fang, Jianwu, et al.
Veröffentlicht: (2025)
LATTE: Latent Trajectory Embedding for Diffusion-Generated Image Detection
von: Vasilcoiu, Ana, et al.
Veröffentlicht: (2025)
von: Vasilcoiu, Ana, et al.
Veröffentlicht: (2025)
TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
von: YU, Mark, et al.
Veröffentlicht: (2025)
von: YU, Mark, et al.
Veröffentlicht: (2025)
MsFIN: Multi-scale Feature Interaction Network for Traffic Accident Anticipation
von: Wu, Tongshuai, et al.
Veröffentlicht: (2025)
von: Wu, Tongshuai, et al.
Veröffentlicht: (2025)
Focusable Monocular Depth Estimation
von: Du, Yuxin, et al.
Veröffentlicht: (2026)
von: Du, Yuxin, et al.
Veröffentlicht: (2026)
GGAvatar: Reconstructing Garment-Separated 3D Gaussian Splatting Avatars from Monocular Video
von: Chen, Jingxuan
Veröffentlicht: (2024)
von: Chen, Jingxuan
Veröffentlicht: (2024)
DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments
von: Sarkar, Anindya, et al.
Veröffentlicht: (2026)
von: Sarkar, Anindya, et al.
Veröffentlicht: (2026)
GeoSAM-3D: Geodesic Prompt Propagation for Open-Vocabulary 3D Scene Segmentation from Monocular Video
von: Sharma, Arun
Veröffentlicht: (2026)
von: Sharma, Arun
Veröffentlicht: (2026)
LATTE: Learning to Think with Vision Specialists
von: Ma, Zixian, et al.
Veröffentlicht: (2024)
von: Ma, Zixian, et al.
Veröffentlicht: (2024)
Comparing Learning Paradigms for Egocentric Video Summarization
von: Wen, Daniel
Veröffentlicht: (2025)
von: Wen, Daniel
Veröffentlicht: (2025)
Technical Report for Ego4D Long-Term Action Anticipation Challenge 2025
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
MV-DUSt3R+: Single-Stage Scene Reconstruction from Sparse Views In 2 Seconds
von: Tang, Zhenggang, et al.
Veröffentlicht: (2024)
von: Tang, Zhenggang, et al.
Veröffentlicht: (2024)
MV-MOS: Multi-View Feature Fusion for 3D Moving Object Segmentation
von: Cheng, Jintao, et al.
Veröffentlicht: (2024)
von: Cheng, Jintao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
TT4D: A Pipeline and Dataset for Table Tennis 4D Reconstruction From Monocular Videos
von: Rahmanian, Nima, et al.
Veröffentlicht: (2026) -
Real-time Accident Anticipation for Autonomous Driving Through Monocular Depth-Enhanced 3D Modeling
von: Liao, Haicheng, et al.
Veröffentlicht: (2024) -
Towards Ball Spin and Trajectory Analysis in Table Tennis Broadcast Videos via Physically Grounded Synthetic-to-Real Transfer
von: Kienzle, Daniel, et al.
Veröffentlicht: (2025) -
On Moving Object Segmentation from Monocular Video with Transformers
von: Homeyer, Christian, et al.
Veröffentlicht: (2024) -
α-RACER: Real-Time Algorithm for Game-Theoretic Motion Planning and Control in Autonomous Racing using Near-Potential Function
von: Kalaria, Dvij, et al.
Veröffentlicht: (2024)