ImagenHub: Standardizing the evaluation of conditional image generation models
Fuente:
arXiv
Salvato in:
| Autori principali: | Ku, Max, Li, Tianle, Zhang, Kai, Lu, Yujie, Fu, Xingyu, Zhuang, Wenwen, Chen, Wenhu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
di: Lyu, Tianle, et al.
Pubblicazione: (2025)
di: Lyu, Tianle, et al.
Pubblicazione: (2025)
altiro3D: Scene representation from single image and novel view synthesis
di: Canessa, E., et al.
Pubblicazione: (2023)
di: Canessa, E., et al.
Pubblicazione: (2023)
Unveiling Deep Shadows: A Survey and Benchmark on Image and Video Shadow Detection, Removal, and Generation in the Deep Learning Era
di: Hu, Xiaowei, et al.
Pubblicazione: (2024)
di: Hu, Xiaowei, et al.
Pubblicazione: (2024)
DreamCinema: Cinematic Transfer with Free Camera and 3D Character
di: Chen, Weiliang, et al.
Pubblicazione: (2024)
di: Chen, Weiliang, et al.
Pubblicazione: (2024)
VerbDiff: Text-Only Diffusion Models with Enhanced Interaction Awareness
di: Cha, SeungJu, et al.
Pubblicazione: (2025)
di: Cha, SeungJu, et al.
Pubblicazione: (2025)
Kiss3DGen: Repurposing Image Diffusion Models for 3D Asset Generation
di: Lin, Jiantao, et al.
Pubblicazione: (2025)
di: Lin, Jiantao, et al.
Pubblicazione: (2025)
HiScene: Creating Hierarchical 3D Scenes with Isometric View Generation
di: Dong, Wenqi, et al.
Pubblicazione: (2025)
di: Dong, Wenqi, et al.
Pubblicazione: (2025)
FairyGen: Storied Cartoon Video from a Single Child-Drawn Character
di: Zheng, Jiayi, et al.
Pubblicazione: (2025)
di: Zheng, Jiayi, et al.
Pubblicazione: (2025)
Perceive-Sample-Compress: Towards Real-Time 3D Gaussian Splatting
di: Wang, Zijian, et al.
Pubblicazione: (2025)
di: Wang, Zijian, et al.
Pubblicazione: (2025)
Splatography: Sparse multi-view dynamic Gaussian Splatting for filmmaking challenges
di: Azzarelli, Adrian, et al.
Pubblicazione: (2025)
di: Azzarelli, Adrian, et al.
Pubblicazione: (2025)
MesonGS++: Post-training Compression of 3D Gaussian Splatting with Hyperparameter Searching
di: Xie, Shuzhao, et al.
Pubblicazione: (2026)
di: Xie, Shuzhao, et al.
Pubblicazione: (2026)
Representing Long Volumetric Video with Temporal Gaussian Hierarchy
di: Xu, Zhen, et al.
Pubblicazione: (2024)
di: Xu, Zhen, et al.
Pubblicazione: (2024)
Break-for-Make: Modular Low-Rank Adaptations for Composable Content-Style Customization
di: Xu, Yu, et al.
Pubblicazione: (2024)
di: Xu, Yu, et al.
Pubblicazione: (2024)
Neural Network-Based Tracking and 3D Reconstruction of Baseball Pitch Trajectories from Single-View 2D Video
di: Hsieh, Jhen
Pubblicazione: (2024)
di: Hsieh, Jhen
Pubblicazione: (2024)
Exploring Palette based Color Guidance in Diffusion Models
di: Qiu, Qianru, et al.
Pubblicazione: (2025)
di: Qiu, Qianru, et al.
Pubblicazione: (2025)
Real-Time Position-Aware View Synthesis from Single-View Input
di: Gond, Manu, et al.
Pubblicazione: (2024)
di: Gond, Manu, et al.
Pubblicazione: (2024)
ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer
di: Guan, Jiazhi, et al.
Pubblicazione: (2024)
di: Guan, Jiazhi, et al.
Pubblicazione: (2024)
Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation
di: Wu, Yuheng, et al.
Pubblicazione: (2026)
di: Wu, Yuheng, et al.
Pubblicazione: (2026)
SAGE: Semantic-Driven Adaptive Gaussian Splatting in Extended Reality
di: Schiavo, Chiara, et al.
Pubblicazione: (2025)
di: Schiavo, Chiara, et al.
Pubblicazione: (2025)
SVGS: Enhancing Gaussian Splatting Using Primitives with Spatially Varying Colors
di: Xu, Rui, et al.
Pubblicazione: (2024)
di: Xu, Rui, et al.
Pubblicazione: (2024)
Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation
di: He, Liu, et al.
Pubblicazione: (2024)
di: He, Liu, et al.
Pubblicazione: (2024)
Casual3DHDR: Deblurring High Dynamic Range 3D Gaussian Splatting from Casually Captured Videos
di: Gong, Shucheng, et al.
Pubblicazione: (2025)
di: Gong, Shucheng, et al.
Pubblicazione: (2025)
TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree Sequencing
di: Lionar, Stefan, et al.
Pubblicazione: (2025)
di: Lionar, Stefan, et al.
Pubblicazione: (2025)
GS-ProCams: Gaussian Splatting-based Projector-Camera Systems
di: Deng, Qingyue, et al.
Pubblicazione: (2024)
di: Deng, Qingyue, et al.
Pubblicazione: (2024)
Textured mesh Quality Assessment using Geometry and Color Field Similarity
di: Yang, Kaifa, et al.
Pubblicazione: (2025)
di: Yang, Kaifa, et al.
Pubblicazione: (2025)
AudCast: Audio-Driven Human Video Generation by Cascaded Diffusion Transformers
di: Guan, Jiazhi, et al.
Pubblicazione: (2025)
di: Guan, Jiazhi, et al.
Pubblicazione: (2025)
ArchGPT: Understanding the World's Architectures with Large Multimodal Models
di: Wang, Yuze, et al.
Pubblicazione: (2025)
di: Wang, Yuze, et al.
Pubblicazione: (2025)
Laplacian Analysis Meets Dynamics Modelling: Gaussian Splatting for 4D Reconstruction
di: Zhou, Yifan, et al.
Pubblicazione: (2025)
di: Zhou, Yifan, et al.
Pubblicazione: (2025)
Perceptual Visual Quality Assessment: Principles, Methods, and Future Directions
di: Zhou, Wei, et al.
Pubblicazione: (2025)
di: Zhou, Wei, et al.
Pubblicazione: (2025)
PersonaGest: Personalized Co-Speech Gesture Generation with Semantic-Guided Hierarchical Motion Representation
di: Zhao, Junchuan, et al.
Pubblicazione: (2026)
di: Zhao, Junchuan, et al.
Pubblicazione: (2026)
InteractEdit: Zero-Shot Editing of Human-Object Interactions in Images
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2025)
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2025)
COutfitGAN: Learning to Synthesize Compatible Outfits Supervised by Silhouette Masks and Fashion Styles
di: Zhou, Dongliang, et al.
Pubblicazione: (2025)
di: Zhou, Dongliang, et al.
Pubblicazione: (2025)
InteractDiffusion: Interaction Control in Text-to-Image Diffusion Models
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2023)
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2023)
STAR: Skeleton-aware Text-based 4D Avatar Generation with In-Network Motion Retargeting
di: Chai, Zenghao, et al.
Pubblicazione: (2024)
di: Chai, Zenghao, et al.
Pubblicazione: (2024)
Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation
di: Cheng, Shihao, et al.
Pubblicazione: (2026)
di: Cheng, Shihao, et al.
Pubblicazione: (2026)
MDD: A Dataset for Text-and-Music Conditioned Duet Dance Generation
di: Gupta, Prerit, et al.
Pubblicazione: (2025)
di: Gupta, Prerit, et al.
Pubblicazione: (2025)
DanceEditor: Towards Iterative Editable Music-driven Dance Generation with Open-Vocabulary Descriptions
di: Zhang, Hengyuan, et al.
Pubblicazione: (2025)
di: Zhang, Hengyuan, et al.
Pubblicazione: (2025)
Sound Sparks Motion: Audio and Text Tuning for Video Editing
di: Razlighi, AmirHossein Naghi, et al.
Pubblicazione: (2026)
di: Razlighi, AmirHossein Naghi, et al.
Pubblicazione: (2026)
Lester: rotoscope animation through video object segmentation and tracking
di: Tous, Ruben
Pubblicazione: (2024)
di: Tous, Ruben
Pubblicazione: (2024)
ReFiNe: Recursive Field Networks for Cross-modal Multi-scene Representation
di: Zakharov, Sergey, et al.
Pubblicazione: (2024)
di: Zakharov, Sergey, et al.
Pubblicazione: (2024)
Documenti analoghi
-
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
di: Lyu, Tianle, et al.
Pubblicazione: (2025) -
altiro3D: Scene representation from single image and novel view synthesis
di: Canessa, E., et al.
Pubblicazione: (2023) -
Unveiling Deep Shadows: A Survey and Benchmark on Image and Video Shadow Detection, Removal, and Generation in the Deep Learning Era
di: Hu, Xiaowei, et al.
Pubblicazione: (2024) -
DreamCinema: Cinematic Transfer with Free Camera and 3D Character
di: Chen, Weiliang, et al.
Pubblicazione: (2024) -
VerbDiff: Text-Only Diffusion Models with Enhanced Interaction Awareness
di: Cha, SeungJu, et al.
Pubblicazione: (2025)