Lester: rotoscope animation through video object segmentation and tracking
Fuente:
arXiv
Salvato in:
| Autore principale: | Tous, Ruben |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Extreme Compression of Adaptive Neural Images
di: Hoshikawa, Leo, et al.
Pubblicazione: (2024)
di: Hoshikawa, Leo, et al.
Pubblicazione: (2024)
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
di: Shen, Qiuhong, et al.
Pubblicazione: (2024)
di: Shen, Qiuhong, et al.
Pubblicazione: (2024)
Freehand Sketch Generation from Mechanical Components
di: Liao, Zhichao, et al.
Pubblicazione: (2024)
di: Liao, Zhichao, et al.
Pubblicazione: (2024)
HiSC4D: Human-centered interaction and 4D Scene Capture in Large-scale Space Using Wearable IMUs and LiDAR
di: Dai, Yudi, et al.
Pubblicazione: (2024)
di: Dai, Yudi, et al.
Pubblicazione: (2024)
A Survey on 3D Gaussian Splatting
di: Chen, Guikun, et al.
Pubblicazione: (2024)
di: Chen, Guikun, et al.
Pubblicazione: (2024)
Zero-Shot Visual Deepfake Detection: Can AI Predict and Prevent Fake Content Before It's Created?
di: Sar, Ayan, et al.
Pubblicazione: (2025)
di: Sar, Ayan, et al.
Pubblicazione: (2025)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
di: Lyu, Tianle, et al.
Pubblicazione: (2025)
di: Lyu, Tianle, et al.
Pubblicazione: (2025)
Seeing World Dynamics in a Nutshell
di: Shen, Qiuhong, et al.
Pubblicazione: (2025)
di: Shen, Qiuhong, et al.
Pubblicazione: (2025)
ChoreoMuse: Robust Music-to-Dance Video Generation with Style Transfer and Beat-Adherent Motion
di: Wang, Xuanchen, et al.
Pubblicazione: (2025)
di: Wang, Xuanchen, et al.
Pubblicazione: (2025)
Towards Unified Co-Speech Gesture Generation via Hierarchical Implicit Periodicity Learning
di: Guo, Xin, et al.
Pubblicazione: (2025)
di: Guo, Xin, et al.
Pubblicazione: (2025)
Multimodal Cinematic Video Synthesis Using Text-to-Image and Audio Generation Models
di: S, Sridhar, et al.
Pubblicazione: (2025)
di: S, Sridhar, et al.
Pubblicazione: (2025)
ToonAging: Face Re-Aging upon Artistic Portrait Style Transfer
di: Kim, Bumsoo, et al.
Pubblicazione: (2024)
di: Kim, Bumsoo, et al.
Pubblicazione: (2024)
SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation
di: Xu, Xiangyu, et al.
Pubblicazione: (2024)
di: Xu, Xiangyu, et al.
Pubblicazione: (2024)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
di: Kim, Bumsoo, et al.
Pubblicazione: (2024)
di: Kim, Bumsoo, et al.
Pubblicazione: (2024)
GoodDrag: Towards Good Practices for Drag Editing with Diffusion Models
di: Zhang, Zewei, et al.
Pubblicazione: (2024)
di: Zhang, Zewei, et al.
Pubblicazione: (2024)
Identity Preserving 3D Head Stylization with Multiview Score Distillation
di: Bilecen, Bahri Batuhan, et al.
Pubblicazione: (2024)
di: Bilecen, Bahri Batuhan, et al.
Pubblicazione: (2024)
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
di: Sun, Zeyi, et al.
Pubblicazione: (2024)
di: Sun, Zeyi, et al.
Pubblicazione: (2024)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
di: Park, Inkyu, et al.
Pubblicazione: (2023)
di: Park, Inkyu, et al.
Pubblicazione: (2023)
Instant3D: Instant Text-to-3D Generation
di: Li, Ming, et al.
Pubblicazione: (2023)
di: Li, Ming, et al.
Pubblicazione: (2023)
Time-to-Move: Training-Free Motion Controlled Video Generation via Dual-Clock Denoising
di: Singer, Assaf, et al.
Pubblicazione: (2025)
di: Singer, Assaf, et al.
Pubblicazione: (2025)
Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters
di: Chiu, Pin-Yen, et al.
Pubblicazione: (2025)
di: Chiu, Pin-Yen, et al.
Pubblicazione: (2025)
Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
di: Girdhar, Rohit, et al.
Pubblicazione: (2023)
di: Girdhar, Rohit, et al.
Pubblicazione: (2023)
DesignAsCode: Bridging Structural Editability and Visual Fidelity in Graphic Design Generation
di: Liu, Ziyuan, et al.
Pubblicazione: (2026)
di: Liu, Ziyuan, et al.
Pubblicazione: (2026)
Instruction-Driven 3D Facial Expression Generation and Transition
di: Vo, Anh H., et al.
Pubblicazione: (2026)
di: Vo, Anh H., et al.
Pubblicazione: (2026)
Cross-Scenario Deraining Adaptation with Unpaired Data: Superpixel Structural Priors and Multi-Stage Pseudo-Rain Synthesis
di: Zhao, Kangbo, et al.
Pubblicazione: (2026)
di: Zhao, Kangbo, et al.
Pubblicazione: (2026)
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
di: Zheng, Shuhong, et al.
Pubblicazione: (2026)
di: Zheng, Shuhong, et al.
Pubblicazione: (2026)
Coral Model Generation from Single Images for Virtual Reality Applications
di: Fu, Jie, et al.
Pubblicazione: (2024)
di: Fu, Jie, et al.
Pubblicazione: (2024)
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
di: Rakesh, Vineet Kumar, et al.
Pubblicazione: (2025)
di: Rakesh, Vineet Kumar, et al.
Pubblicazione: (2025)
Photoreal Scene Reconstruction from an Egocentric Device
di: Lv, Zhaoyang, et al.
Pubblicazione: (2025)
di: Lv, Zhaoyang, et al.
Pubblicazione: (2025)
d-Sketch: Improving Visual Fidelity of Sketch-to-Image Translation with Pretrained Latent Diffusion Models without Retraining
di: Roy, Prasun, et al.
Pubblicazione: (2025)
di: Roy, Prasun, et al.
Pubblicazione: (2025)
DrawVideo: Generating Long Video from Storyboard Keyframe Sketches
di: Xu, Chuanzhi, et al.
Pubblicazione: (2026)
di: Xu, Chuanzhi, et al.
Pubblicazione: (2026)
Neuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical Prognosis
di: Aiersilan, Aizierjiang, et al.
Pubblicazione: (2026)
di: Aiersilan, Aizierjiang, et al.
Pubblicazione: (2026)
DreamCinema: Cinematic Transfer with Free Camera and 3D Character
di: Chen, Weiliang, et al.
Pubblicazione: (2024)
di: Chen, Weiliang, et al.
Pubblicazione: (2024)
Unveiling Deep Shadows: A Survey and Benchmark on Image and Video Shadow Detection, Removal, and Generation in the Deep Learning Era
di: Hu, Xiaowei, et al.
Pubblicazione: (2024)
di: Hu, Xiaowei, et al.
Pubblicazione: (2024)
Representing Long Volumetric Video with Temporal Gaussian Hierarchy
di: Xu, Zhen, et al.
Pubblicazione: (2024)
di: Xu, Zhen, et al.
Pubblicazione: (2024)
Break-for-Make: Modular Low-Rank Adaptations for Composable Content-Style Customization
di: Xu, Yu, et al.
Pubblicazione: (2024)
di: Xu, Yu, et al.
Pubblicazione: (2024)
Neural Network-Based Tracking and 3D Reconstruction of Baseball Pitch Trajectories from Single-View 2D Video
di: Hsieh, Jhen
Pubblicazione: (2024)
di: Hsieh, Jhen
Pubblicazione: (2024)
Real-Time Position-Aware View Synthesis from Single-View Input
di: Gond, Manu, et al.
Pubblicazione: (2024)
di: Gond, Manu, et al.
Pubblicazione: (2024)
ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer
di: Guan, Jiazhi, et al.
Pubblicazione: (2024)
di: Guan, Jiazhi, et al.
Pubblicazione: (2024)
SVGS: Enhancing Gaussian Splatting Using Primitives with Spatially Varying Colors
di: Xu, Rui, et al.
Pubblicazione: (2024)
di: Xu, Rui, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Extreme Compression of Adaptive Neural Images
di: Hoshikawa, Leo, et al.
Pubblicazione: (2024) -
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
di: Shen, Qiuhong, et al.
Pubblicazione: (2024) -
Freehand Sketch Generation from Mechanical Components
di: Liao, Zhichao, et al.
Pubblicazione: (2024) -
HiSC4D: Human-centered interaction and 4D Scene Capture in Large-scale Space Using Wearable IMUs and LiDAR
di: Dai, Yudi, et al.
Pubblicazione: (2024) -
A Survey on 3D Gaussian Splatting
di: Chen, Guikun, et al.
Pubblicazione: (2024)