On the Content Bias in Fréchet Video Distance
Fuente:
arXiv
Guardado en:
| Autores principales: | Ge, Songwei, Mahapatra, Aniruddha, Parmar, Gaurav, Zhu, Jun-Yan, Huang, Jia-Bin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Expressive Text-to-Image Generation with Rich Text
por: Ge, Songwei, et al.
Publicado: (2023)
por: Ge, Songwei, et al.
Publicado: (2023)
One-Step Image Translation with Text-to-Image Models
por: Parmar, Gaurav, et al.
Publicado: (2024)
por: Parmar, Gaurav, et al.
Publicado: (2024)
Preserve Your Own Correlation: A Noise Prior for Video Diffusion Models
por: Ge, Songwei, et al.
Publicado: (2023)
por: Ge, Songwei, et al.
Publicado: (2023)
Scaling Group Inference for Diverse and High-Quality Generation
por: Parmar, Gaurav, et al.
Publicado: (2025)
por: Parmar, Gaurav, et al.
Publicado: (2025)
Rethinking Score Distillation as a Bridge Between Image Distributions
por: McAllister, David, et al.
Publicado: (2024)
por: McAllister, David, et al.
Publicado: (2024)
MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation
por: Liu, Dongxia, et al.
Publicado: (2026)
por: Liu, Dongxia, et al.
Publicado: (2026)
Depth-supervised NeRF: Fewer Views and Faster Training for Free
por: Deng, Kangle, et al.
Publicado: (2021)
por: Deng, Kangle, et al.
Publicado: (2021)
Energy-Based Sliced Wasserstein Distance
por: Nguyen, Khai, et al.
Publicado: (2023)
por: Nguyen, Khai, et al.
Publicado: (2023)
Generating Multi-Image Synthetic Data for Text-to-Image Customization
por: Kumari, Nupur, et al.
Publicado: (2025)
por: Kumari, Nupur, et al.
Publicado: (2025)
Customizing Text-to-Image Models with a Single Image Pair
por: Jones, Maxwell, et al.
Publicado: (2024)
por: Jones, Maxwell, et al.
Publicado: (2024)
Consolidating Attention Features for Multi-view Image Editing
por: Patashnik, Or, et al.
Publicado: (2024)
por: Patashnik, Or, et al.
Publicado: (2024)
FlashTex: Fast Relightable Mesh Texturing with LightControlNet
por: Deng, Kangle, et al.
Publicado: (2024)
por: Deng, Kangle, et al.
Publicado: (2024)
LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias
por: Jin, Haian, et al.
Publicado: (2024)
por: Jin, Haian, et al.
Publicado: (2024)
Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
por: Cui, Jiahao, et al.
Publicado: (2024)
por: Cui, Jiahao, et al.
Publicado: (2024)
ReSWD: ReSTIR'd, not shaken. Combining Reservoir Sampling and Sliced Wasserstein Distance for Variance Reduction
por: Boss, Mark, et al.
Publicado: (2025)
por: Boss, Mark, et al.
Publicado: (2025)
Distilling Diffusion Models into Conditional GANs
por: Kang, Minguk, et al.
Publicado: (2024)
por: Kang, Minguk, et al.
Publicado: (2024)
Approximating Signed Distance Fields With Sparse Ellipsoidal Radial Basis Function Networks: A Dynamic Multi-Objective Optimization Strategy
por: Lian, Bobo, et al.
Publicado: (2025)
por: Lian, Bobo, et al.
Publicado: (2025)
CookingDiffusion: Cooking Procedural Image Generation with Stable Diffusion
por: Wang, Yuan, et al.
Publicado: (2025)
por: Wang, Yuan, et al.
Publicado: (2025)
Dynamic Concepts Personalization from Single Videos
por: Abdal, Rameen, et al.
Publicado: (2025)
por: Abdal, Rameen, et al.
Publicado: (2025)
From Competition to Synergy: Unlocking Reinforcement Learning for Subject-Driven Image Generation
por: Huang, Ziwei, et al.
Publicado: (2025)
por: Huang, Ziwei, et al.
Publicado: (2025)
VoteSplat: Hough Voting Gaussian Splatting for 3D Scene Understanding
por: Jiang, Minchao, et al.
Publicado: (2025)
por: Jiang, Minchao, et al.
Publicado: (2025)
BioHuman: Learning Biomechanical Human Representations from Video
por: Huo, Yujun, et al.
Publicado: (2026)
por: Huo, Yujun, et al.
Publicado: (2026)
Adaptive Hybrid Caching for Efficient Text-to-Video Diffusion Model Acceleration
por: Wei, Yuanxin, et al.
Publicado: (2025)
por: Wei, Yuanxin, et al.
Publicado: (2025)
DreamWaltz-G: Expressive 3D Gaussian Avatars from Skeleton-Guided 2D Diffusion
por: Huang, Yukun, et al.
Publicado: (2024)
por: Huang, Yukun, et al.
Publicado: (2024)
ArtGS: Building Interactable Replicas of Complex Articulated Objects via Gaussian Splatting
por: Liu, Yu, et al.
Publicado: (2025)
por: Liu, Yu, et al.
Publicado: (2025)
FLex: Joint Pose and Dynamic Radiance Fields Optimization for Stereo Endoscopic Videos
por: Stilz, Florian Philipp, et al.
Publicado: (2024)
por: Stilz, Florian Philipp, et al.
Publicado: (2024)
Movie Weaver: Tuning-Free Multi-Concept Video Personalization with Anchored Prompts
por: Liang, Feng, et al.
Publicado: (2025)
por: Liang, Feng, et al.
Publicado: (2025)
Alice v1: Distillation-Enhanced Video Generation Surpassing Closed-Source Models
por: Xiaoyu, Wang, et al.
Publicado: (2026)
por: Xiaoyu, Wang, et al.
Publicado: (2026)
R-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flow
por: Wu, Zijie, et al.
Publicado: (2026)
por: Wu, Zijie, et al.
Publicado: (2026)
Neural Material Adaptor for Visual Grounding of Intrinsic Dynamics
por: Cao, Junyi, et al.
Publicado: (2024)
por: Cao, Junyi, et al.
Publicado: (2024)
VFusion3D: Learning Scalable 3D Generative Models from Video Diffusion Models
por: Han, Junlin, et al.
Publicado: (2024)
por: Han, Junlin, et al.
Publicado: (2024)
D-NPC: Dynamic Neural Point Clouds for Non-Rigid View Synthesis from Monocular Video
por: Kappel, Moritz, et al.
Publicado: (2024)
por: Kappel, Moritz, et al.
Publicado: (2024)
SViM3D: Stable Video Material Diffusion for Single Image 3D Generation
por: Engelhardt, Andreas, et al.
Publicado: (2025)
por: Engelhardt, Andreas, et al.
Publicado: (2025)
Multimodal Latent Diffusion Model for Complex Sewing Pattern Generation
por: Liu, Shengqi, et al.
Publicado: (2024)
por: Liu, Shengqi, et al.
Publicado: (2024)
Continuous Edit Distance, Geodesics and Barycenters of Time-varying Persistence Diagrams
por: Tchitchek, Sebastien, et al.
Publicado: (2025)
por: Tchitchek, Sebastien, et al.
Publicado: (2025)
Nabla-R2D3: Effective and Efficient 3D Diffusion Alignment with 2D Rewards
por: Liu, Qingming, et al.
Publicado: (2025)
por: Liu, Qingming, et al.
Publicado: (2025)
SceneWeaver: All-in-One 3D Scene Synthesis with an Extensible and Self-Reflective Agent
por: Yang, Yandan, et al.
Publicado: (2025)
por: Yang, Yandan, et al.
Publicado: (2025)
Spatial and Surface Correspondence Field for Interaction Transfer
por: Huang, Zeyu, et al.
Publicado: (2024)
por: Huang, Zeyu, et al.
Publicado: (2024)
Flexible Motion In-betweening with Diffusion Models
por: Cohan, Setareh, et al.
Publicado: (2024)
por: Cohan, Setareh, et al.
Publicado: (2024)
DreamCube: 3D Panorama Generation via Multi-plane Synchronization
por: Huang, Yukun, et al.
Publicado: (2025)
por: Huang, Yukun, et al.
Publicado: (2025)
Ejemplares similares
-
Expressive Text-to-Image Generation with Rich Text
por: Ge, Songwei, et al.
Publicado: (2023) -
One-Step Image Translation with Text-to-Image Models
por: Parmar, Gaurav, et al.
Publicado: (2024) -
Preserve Your Own Correlation: A Noise Prior for Video Diffusion Models
por: Ge, Songwei, et al.
Publicado: (2023) -
Scaling Group Inference for Diverse and High-Quality Generation
por: Parmar, Gaurav, et al.
Publicado: (2025) -
Rethinking Score Distillation as a Bridge Between Image Distributions
por: McAllister, David, et al.
Publicado: (2024)