Guardado en:
| Autores principales: | Henriques, Joao F., Campbell, Dylan, Han, Tengda |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2404.01079 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Character-Centric Understanding of Animated Movies
por: Gui, Zhongrui, et al.
Publicado: (2025)
por: Gui, Zhongrui, et al.
Publicado: (2025)
Flash3D: Feed-Forward Generalisable 3D Scene Reconstruction from a Single Image
por: Szymanowicz, Stanislaw, et al.
Publicado: (2024)
por: Szymanowicz, Stanislaw, et al.
Publicado: (2024)
SCENES: Subpixel Correspondence Estimation With Epipolar Supervision
por: Kloepfer, Dominik A., et al.
Publicado: (2024)
por: Kloepfer, Dominik A., et al.
Publicado: (2024)
PoseDreamer: Scalable and Photorealistic Human Data Generation Pipeline with Diffusion Models
por: Prospero, Lorenza, et al.
Publicado: (2026)
por: Prospero, Lorenza, et al.
Publicado: (2026)
GMOS: Grounding Moving Object Segmentation in 3D Space and Time
por: Xie, Junyu, et al.
Publicado: (2026)
por: Xie, Junyu, et al.
Publicado: (2026)
HandCraft: Anatomically Correct Restoration of Malformed Hands in Diffusion Generated Images
por: Qin, Zhenyue, et al.
Publicado: (2024)
por: Qin, Zhenyue, et al.
Publicado: (2024)
PlückeRF: A Line-based 3D Representation for Few-view Reconstruction
por: Bahrami, Sam, et al.
Publicado: (2025)
por: Bahrami, Sam, et al.
Publicado: (2025)
Text2Loc++: Generalizing 3D Point Cloud Localization from Natural Language
por: Xia, Yan, et al.
Publicado: (2025)
por: Xia, Yan, et al.
Publicado: (2025)
HyperHuman: Hyper-Realistic Human Generation with Latent Structural Diffusion
por: Liu, Xian, et al.
Publicado: (2023)
por: Liu, Xian, et al.
Publicado: (2023)
MPDS: A Movie Posters Dataset for Image Generation with Diffusion Model
por: Xu, Meng, et al.
Publicado: (2024)
por: Xu, Meng, et al.
Publicado: (2024)
Prompt Generation Networks for Input-Space Adaptation of Frozen Vision Transformers
por: Loedeman, Jochem, et al.
Publicado: (2022)
por: Loedeman, Jochem, et al.
Publicado: (2022)
CountGD: Multi-Modal Open-World Counting
por: Amini-Naieni, Niki, et al.
Publicado: (2024)
por: Amini-Naieni, Niki, et al.
Publicado: (2024)
Learning to Count without Annotations
por: Knobel, Lukas, et al.
Publicado: (2023)
por: Knobel, Lukas, et al.
Publicado: (2023)
HyperDiff: Hypergraph Guided Diffusion Model for 3D Human Pose Estimation
por: Han, Bing, et al.
Publicado: (2025)
por: Han, Bing, et al.
Publicado: (2025)
MovieBench: A Hierarchical Movie Level Dataset for Long Video Generation
por: Wu, Weijia, et al.
Publicado: (2024)
por: Wu, Weijia, et al.
Publicado: (2024)
Refracting Reality: Generating Images with Realistic Transparent Objects
por: Yin, Yue, et al.
Publicado: (2025)
por: Yin, Yue, et al.
Publicado: (2025)
Unsupervised Object Detection with Theoretical Guarantees
por: Longa, Marian, et al.
Publicado: (2024)
por: Longa, Marian, et al.
Publicado: (2024)
Interpretable Representation Learning from Videos using Nonlinear Priors
por: Longa, Marian, et al.
Publicado: (2024)
por: Longa, Marian, et al.
Publicado: (2024)
It's Just Another Day: Unique Video Captioning by Discriminative Prompting
por: Perrett, Toby, et al.
Publicado: (2024)
por: Perrett, Toby, et al.
Publicado: (2024)
Towards Automated Movie Trailer Generation
por: Argaw, Dawit Mureja, et al.
Publicado: (2024)
por: Argaw, Dawit Mureja, et al.
Publicado: (2024)
Room Envelopes: A Synthetic Dataset for Indoor Layout Reconstruction from Images
por: Bahrami, Sam, et al.
Publicado: (2025)
por: Bahrami, Sam, et al.
Publicado: (2025)
Text2Loc: 3D Point Cloud Localization from Natural Language
por: Xia, Yan, et al.
Publicado: (2023)
por: Xia, Yan, et al.
Publicado: (2023)
IMPUS: Image Morphing with Perceptually-Uniform Sampling Using Diffusion Models
por: Yang, Zhaoyuan, et al.
Publicado: (2023)
por: Yang, Zhaoyuan, et al.
Publicado: (2023)
3D Arena: An Open Platform for Generative 3D Evaluation
por: Ebert, Dylan
Publicado: (2025)
por: Ebert, Dylan
Publicado: (2025)
CRAM: Large-scale Video Continual Learning with Bootstrapped Compression
por: Mall, Shivani, et al.
Publicado: (2025)
por: Mall, Shivani, et al.
Publicado: (2025)
Captain Cinema: Towards Short Movie Generation
por: Xiao, Junfei, et al.
Publicado: (2025)
por: Xiao, Junfei, et al.
Publicado: (2025)
GST: Precise 3D Human Body from a Single Image with Gaussian Splatting Transformers
por: Prospero, Lorenza, et al.
Publicado: (2024)
por: Prospero, Lorenza, et al.
Publicado: (2024)
TrafficLoc: Localizing Traffic Surveillance Cameras in 3D Scenes
por: Xia, Yan, et al.
Publicado: (2024)
por: Xia, Yan, et al.
Publicado: (2024)
HyperDiT: Hyper-Connected Transformers for High-Fidelity Pixel-Space Diffusion
por: He, Yu, et al.
Publicado: (2026)
por: He, Yu, et al.
Publicado: (2026)
Dynamic Reflections: Probing Video Representations with Text Alignment
por: Zhu, Tyler, et al.
Publicado: (2025)
por: Zhu, Tyler, et al.
Publicado: (2025)
Seeing without Pixels: Perception from Camera Trajectories
por: Xue, Zihui, et al.
Publicado: (2025)
por: Xue, Zihui, et al.
Publicado: (2025)
SEED4D: A Synthetic Ego--Exo Dynamic 4D Data Generator, Driving Dataset and Benchmark
por: Kästingschäfer, Marius, et al.
Publicado: (2024)
por: Kästingschäfer, Marius, et al.
Publicado: (2024)
MovieCORE: COgnitive REasoning in Movies
por: Faure, Gueter Josmy, et al.
Publicado: (2025)
por: Faure, Gueter Josmy, et al.
Publicado: (2025)
D-Cube: Exploiting Hyper-Features of Diffusion Model for Robust Medical Classification
por: Jang, Minhee, et al.
Publicado: (2024)
por: Jang, Minhee, et al.
Publicado: (2024)
MovieDreamer: Hierarchical Generation for Coherent Long Visual Sequence
por: Zhao, Canyu, et al.
Publicado: (2024)
por: Zhao, Canyu, et al.
Publicado: (2024)
Automated Movie Generation via Multi-Agent CoT Planning
por: Wu, Weijia, et al.
Publicado: (2025)
por: Wu, Weijia, et al.
Publicado: (2025)
Unpaired Photo-realistic Image Deraining with Energy-informed Diffusion Model
por: Wen, Yuanbo, et al.
Publicado: (2024)
por: Wen, Yuanbo, et al.
Publicado: (2024)
Shot-by-Shot: Film-Grammar-Aware Training-Free Audio Description Generation
por: Xie, Junyu, et al.
Publicado: (2025)
por: Xie, Junyu, et al.
Publicado: (2025)
Believing is Seeing: Unobserved Object Detection using Generative Models
por: Bhattacharjee, Subhransu S., et al.
Publicado: (2024)
por: Bhattacharjee, Subhransu S., et al.
Publicado: (2024)
Learning from Streaming Video with Orthogonal Gradients
por: Han, Tengda, et al.
Publicado: (2025)
por: Han, Tengda, et al.
Publicado: (2025)
Ejemplares similares
-
Character-Centric Understanding of Animated Movies
por: Gui, Zhongrui, et al.
Publicado: (2025) -
Flash3D: Feed-Forward Generalisable 3D Scene Reconstruction from a Single Image
por: Szymanowicz, Stanislaw, et al.
Publicado: (2024) -
SCENES: Subpixel Correspondence Estimation With Epipolar Supervision
por: Kloepfer, Dominik A., et al.
Publicado: (2024) -
PoseDreamer: Scalable and Photorealistic Human Data Generation Pipeline with Diffusion Models
por: Prospero, Lorenza, et al.
Publicado: (2026) -
GMOS: Grounding Moving Object Segmentation in 3D Space and Time
por: Xie, Junyu, et al.
Publicado: (2026)