GazeFusion: Saliency-Guided Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Yunxiang, Wu, Nan, Lin, Connor Z., Wetzstein, Gordon, Sun, Qi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
WonderPlay: Dynamic 3D Scene Generation from a Single Image and Actions
di: Li, Zizhang, et al.
Pubblicazione: (2025)
di: Li, Zizhang, et al.
Pubblicazione: (2025)
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
di: Cai, Shengqu, et al.
Pubblicazione: (2024)
di: Cai, Shengqu, et al.
Pubblicazione: (2024)
MultiGen: Level-Design for Editable Multiplayer Worlds in Diffusion Game Engines
di: Po, Ryan, et al.
Pubblicazione: (2026)
di: Po, Ryan, et al.
Pubblicazione: (2026)
Robust Symmetry Detection via Riemannian Langevin Dynamics
di: Je, Jihyeon, et al.
Pubblicazione: (2024)
di: Je, Jihyeon, et al.
Pubblicazione: (2024)
Mixture of Contexts for Long Video Generation
di: Cai, Shengqu, et al.
Pubblicazione: (2025)
di: Cai, Shengqu, et al.
Pubblicazione: (2025)
Res2NetFuse: A Novel Res2Net-based Fusion Method for Infrared and Visible Images
di: Song, Xu, et al.
Pubblicazione: (2021)
di: Song, Xu, et al.
Pubblicazione: (2021)
BlockFusion: Expandable 3D Scene Generation using Latent Tri-plane Extrapolation
di: Wu, Zhennan, et al.
Pubblicazione: (2024)
di: Wu, Zhennan, et al.
Pubblicazione: (2024)
SyncDreamer: Generating Multiview-consistent Images from a Single-view Image
di: Liu, Yuan, et al.
Pubblicazione: (2023)
di: Liu, Yuan, et al.
Pubblicazione: (2023)
ScanGAN360: A Generative Model of Realistic Scanpaths for 360$^{\circ}$ Images
di: Martin, Daniel, et al.
Pubblicazione: (2021)
di: Martin, Daniel, et al.
Pubblicazione: (2021)
AMG: Avatar Motion Guided Video Generation
di: Yang, Zhangsihao, et al.
Pubblicazione: (2024)
di: Yang, Zhangsihao, et al.
Pubblicazione: (2024)
Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
di: Liu, Bingchen, et al.
Pubblicazione: (2024)
di: Liu, Bingchen, et al.
Pubblicazione: (2024)
FabricDiffusion: High-Fidelity Texture Transfer for 3D Garments Generation from In-The-Wild Clothing Images
di: Zhang, Cheng, et al.
Pubblicazione: (2024)
di: Zhang, Cheng, et al.
Pubblicazione: (2024)
InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models
di: Lu, Yifan, et al.
Pubblicazione: (2024)
di: Lu, Yifan, et al.
Pubblicazione: (2024)
WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting
di: Wang, Lezhong, et al.
Pubblicazione: (2026)
di: Wang, Lezhong, et al.
Pubblicazione: (2026)
SeqTex: Generate Mesh Textures in Video Sequence
di: Yuan, Ze, et al.
Pubblicazione: (2025)
di: Yuan, Ze, et al.
Pubblicazione: (2025)
SVGBuilder: Component-Based Colored SVG Generation with Text-Guided Autoregressive Transformers
di: Chen, Zehao, et al.
Pubblicazione: (2024)
di: Chen, Zehao, et al.
Pubblicazione: (2024)
ObjectMover: Generative Object Movement with Video Prior
di: Yu, Xin, et al.
Pubblicazione: (2025)
di: Yu, Xin, et al.
Pubblicazione: (2025)
Object-level Visual Prompts for Compositional Image Generation
di: Parmar, Gaurav, et al.
Pubblicazione: (2025)
di: Parmar, Gaurav, et al.
Pubblicazione: (2025)
A Survey on Quality Metrics for Text-to-Image Generation
di: Hartwig, Sebastian, et al.
Pubblicazione: (2024)
di: Hartwig, Sebastian, et al.
Pubblicazione: (2024)
Lazy Diffusion Transformer for Interactive Image Editing
di: Nitzan, Yotam, et al.
Pubblicazione: (2024)
di: Nitzan, Yotam, et al.
Pubblicazione: (2024)
SkeletonGaussian: Editable 4D Generation through Gaussian Skeletonization
di: Wu, Lifan, et al.
Pubblicazione: (2026)
di: Wu, Lifan, et al.
Pubblicazione: (2026)
Make It Count: Text-to-Image Generation with an Accurate Number of Objects
di: Binyamin, Lital, et al.
Pubblicazione: (2024)
di: Binyamin, Lital, et al.
Pubblicazione: (2024)
GraphicsDreamer: Image to 3D Generation with Physical Consistency
di: Chen, Pei, et al.
Pubblicazione: (2024)
di: Chen, Pei, et al.
Pubblicazione: (2024)
Uncovering Conceptual Blindspots in Generative Image Models Using Sparse Autoencoders
di: Bohacek, Matyas, et al.
Pubblicazione: (2025)
di: Bohacek, Matyas, et al.
Pubblicazione: (2025)
FlexCAD: Unified and Versatile Controllable CAD Generation with Fine-tuned Large Language Models
di: Zhang, Zhanwei, et al.
Pubblicazione: (2024)
di: Zhang, Zhanwei, et al.
Pubblicazione: (2024)
GuideFlow3D: Optimization-Guided Rectified Flow For Appearance Transfer
di: Sarkar, Sayan Deb, et al.
Pubblicazione: (2025)
di: Sarkar, Sayan Deb, et al.
Pubblicazione: (2025)
Template-Guided Reconstruction of Pulmonary Segments with Neural Implicit Functions
di: Xie, Kangxian, et al.
Pubblicazione: (2025)
di: Xie, Kangxian, et al.
Pubblicazione: (2025)
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
di: Wang, Kuan-Chieh, et al.
Pubblicazione: (2024)
di: Wang, Kuan-Chieh, et al.
Pubblicazione: (2024)
Vector Grimoire: Codebook-based Shape Generation under Raster Image Supervision
di: Feuerpfeil, Moritz, et al.
Pubblicazione: (2024)
di: Feuerpfeil, Moritz, et al.
Pubblicazione: (2024)
PFAvatar: Pose-Fusion 3D Personalized Avatar Reconstruction from Real-World Outfit-of-the-Day Photos
di: Xi, Dianbing, et al.
Pubblicazione: (2025)
di: Xi, Dianbing, et al.
Pubblicazione: (2025)
3DPhysVideo: Consistency-Guided Flow SDE for Video Generation via 3D Scene Reconstruction and Physical Simulation
di: Kim, Hwidong, et al.
Pubblicazione: (2026)
di: Kim, Hwidong, et al.
Pubblicazione: (2026)
UniLat3D: Geometry-Appearance Unified Latents for Single-Stage 3D Generation
di: Wu, Guanjun, et al.
Pubblicazione: (2025)
di: Wu, Guanjun, et al.
Pubblicazione: (2025)
OmniHands: Towards Robust 4D Hand Mesh Recovery via A Versatile Transformer
di: Lin, Dixuan, et al.
Pubblicazione: (2024)
di: Lin, Dixuan, et al.
Pubblicazione: (2024)
DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion
di: Sun, Wenqiang, et al.
Pubblicazione: (2024)
di: Sun, Wenqiang, et al.
Pubblicazione: (2024)
HY-Motion 1.0: Scaling Flow Matching Models for Text-To-Motion Generation
di: Wen, Yuxin, et al.
Pubblicazione: (2025)
di: Wen, Yuxin, et al.
Pubblicazione: (2025)
DreamDrive: Generative 4D Scene Modeling from Street View Images
di: Mao, Jiageng, et al.
Pubblicazione: (2024)
di: Mao, Jiageng, et al.
Pubblicazione: (2024)
Learning to Synthesize Graphics Programs for Geometric Artworks
di: Bing, Qi, et al.
Pubblicazione: (2024)
di: Bing, Qi, et al.
Pubblicazione: (2024)
BootPIG: Bootstrapping Zero-shot Personalized Image Generation Capabilities in Pretrained Diffusion Models
di: Purushwalkam, Senthil, et al.
Pubblicazione: (2024)
di: Purushwalkam, Senthil, et al.
Pubblicazione: (2024)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
di: Zuo, Qi, et al.
Pubblicazione: (2024)
di: Zuo, Qi, et al.
Pubblicazione: (2024)
MeshAnything V2: Artist-Created Mesh Generation With Adjacent Mesh Tokenization
di: Chen, Yiwen, et al.
Pubblicazione: (2024)
di: Chen, Yiwen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
WonderPlay: Dynamic 3D Scene Generation from a Single Image and Actions
di: Li, Zizhang, et al.
Pubblicazione: (2025) -
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
di: Cai, Shengqu, et al.
Pubblicazione: (2024) -
MultiGen: Level-Design for Editable Multiplayer Worlds in Diffusion Game Engines
di: Po, Ryan, et al.
Pubblicazione: (2026) -
Robust Symmetry Detection via Riemannian Langevin Dynamics
di: Je, Jihyeon, et al.
Pubblicazione: (2024) -
Mixture of Contexts for Long Video Generation
di: Cai, Shengqu, et al.
Pubblicazione: (2025)