Scaling Group Inference for Diverse and High-Quality Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Parmar, Gaurav, Patashnik, Or, Ostashev, Daniil, Wang, Kuan-Chieh, Aberman, Kfir, Narasimhan, Srinivasa, Zhu, Jun-Yan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Object-level Visual Prompts for Compositional Image Generation
by: Parmar, Gaurav, et al.
Published: (2025)
by: Parmar, Gaurav, et al.
Published: (2025)
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
by: Wang, Kuan-Chieh, et al.
Published: (2024)
by: Wang, Kuan-Chieh, et al.
Published: (2024)
Nested Attention: Semantic-aware Attention Values for Concept Personalization
by: Patashnik, Or, et al.
Published: (2025)
by: Patashnik, Or, et al.
Published: (2025)
Continuous Control of Editing Models via Adaptive-Origin Guidance
by: Wolf, Alon, et al.
Published: (2026)
by: Wolf, Alon, et al.
Published: (2026)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
by: Dahary, Omer, et al.
Published: (2024)
by: Dahary, Omer, et al.
Published: (2024)
One-Step Image Translation with Text-to-Image Models
by: Parmar, Gaurav, et al.
Published: (2024)
by: Parmar, Gaurav, et al.
Published: (2024)
Be Decisive: Noise-Induced Layouts for Multi-Subject Generation
by: Dahary, Omer, et al.
Published: (2025)
by: Dahary, Omer, et al.
Published: (2025)
Omni-ID: Holistic Identity Representation Designed for Generative Tasks
by: Qian, Guocheng, et al.
Published: (2024)
by: Qian, Guocheng, et al.
Published: (2024)
Stable Flow: Vital Layers for Training-Free Image Editing
by: Avrahami, Omri, et al.
Published: (2024)
by: Avrahami, Omri, et al.
Published: (2024)
3D PixBrush: Image-Guided Local Texture Synthesis
by: Decatur, Dale, et al.
Published: (2025)
by: Decatur, Dale, et al.
Published: (2025)
Dynamic Concepts Personalization from Single Videos
by: Abdal, Rameen, et al.
Published: (2025)
by: Abdal, Rameen, et al.
Published: (2025)
Zero-Shot Dynamic Concept Personalization with Grid-Based LoRA
by: Abdal, Rameen, et al.
Published: (2025)
by: Abdal, Rameen, et al.
Published: (2025)
Interpreting the Weight Space of Customized Diffusion Models
by: Dravid, Amil, et al.
Published: (2024)
by: Dravid, Amil, et al.
Published: (2024)
ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation
by: Qian, Guocheng Gordon, et al.
Published: (2025)
by: Qian, Guocheng Gordon, et al.
Published: (2025)
VLM-Guided Adaptive Negative Prompting for Creative Generation
by: Golan, Shelly, et al.
Published: (2025)
by: Golan, Shelly, et al.
Published: (2025)
On the Content Bias in Fréchet Video Distance
by: Ge, Songwei, et al.
Published: (2024)
by: Ge, Songwei, et al.
Published: (2024)
Consolidating Attention Features for Multi-view Image Editing
by: Patashnik, Or, et al.
Published: (2024)
by: Patashnik, Or, et al.
Published: (2024)
Kontinuous Kontext: Continuous Strength Control for Instruction-based Image Editing
by: Parihar, Rishubh, et al.
Published: (2025)
by: Parihar, Rishubh, et al.
Published: (2025)
TurboEdit: Text-Based Image Editing Using Few-Step Diffusion Models
by: Deutch, Gilad, et al.
Published: (2024)
by: Deutch, Gilad, et al.
Published: (2024)
Untwisting RoPE: Frequency Control for Shared Attention in DiTs
by: Mikaeili, Aryan, et al.
Published: (2026)
by: Mikaeili, Aryan, et al.
Published: (2026)
Iterative Motion Editing with Natural Language
by: Goel, Purvi, et al.
Published: (2023)
by: Goel, Purvi, et al.
Published: (2023)
Incorporating dense metric depth into neural 3D representations for view synthesis and relighting
by: Chaudhury, Arkadeep Narayan, et al.
Published: (2024)
by: Chaudhury, Arkadeep Narayan, et al.
Published: (2024)
LCM-Lookahead for Encoder-based Text-to-Image Personalization
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
MotioNet: 3D Human Motion Reconstruction from Monocular Video with Skeleton Consistency
by: Shi, Mingyi, et al.
Published: (2020)
by: Shi, Mingyi, et al.
Published: (2020)
Generative Photomontage
by: Liu, Sean J., et al.
Published: (2024)
by: Liu, Sean J., et al.
Published: (2024)
In-Context Sync-LoRA for Portrait Video Editing
by: Polaczek, Sagi, et al.
Published: (2025)
by: Polaczek, Sagi, et al.
Published: (2025)
RealFill: Reference-Driven Generation for Authentic Image Completion
by: Tang, Luming, et al.
Published: (2023)
by: Tang, Luming, et al.
Published: (2023)
LATTICE: Democratize High-Fidelity 3D Generation at Scale
by: Lai, Zeqiang, et al.
Published: (2025)
by: Lai, Zeqiang, et al.
Published: (2025)
HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models
by: Ruiz, Nataniel, et al.
Published: (2023)
by: Ruiz, Nataniel, et al.
Published: (2023)
Tactile DreamFusion: Exploiting Tactile Sensing for 3D Generation
by: Gao, Ruihan, et al.
Published: (2024)
by: Gao, Ruihan, et al.
Published: (2024)
Tight Inversion: Image-Conditioned Inversion for Real Image Editing
by: Kadosh, Edo, et al.
Published: (2025)
by: Kadosh, Edo, et al.
Published: (2025)
Surf-D: Generating High-Quality Surfaces of Arbitrary Topologies Using Diffusion Models
by: Yu, Zhengming, et al.
Published: (2023)
by: Yu, Zhengming, et al.
Published: (2023)
MeshFormer: High-Quality Mesh Generation with 3D-Guided Reconstruction Model
by: Liu, Minghua, et al.
Published: (2024)
by: Liu, Minghua, et al.
Published: (2024)
Image Generation from Contextually-Contradictory Prompts
by: Huberman, Saar, et al.
Published: (2025)
by: Huberman, Saar, et al.
Published: (2025)
Meshtron: High-Fidelity, Artist-Like 3D Mesh Generation at Scale
by: Hao, Zekun, et al.
Published: (2024)
by: Hao, Zekun, et al.
Published: (2024)
GenUDC: High Quality 3D Mesh Generation with Unsigned Dual Contouring Representation
by: Wang, Ruowei, et al.
Published: (2024)
by: Wang, Ruowei, et al.
Published: (2024)
High-Quality Mesh Blendshape Generation from Face Videos via Neural Inverse Rendering
by: Ming, Xin, et al.
Published: (2024)
by: Ming, Xin, et al.
Published: (2024)
HumanEdit: A High-Quality Human-Rewarded Dataset for Instruction-based Image Editing
by: Bai, Jinbin, et al.
Published: (2024)
by: Bai, Jinbin, et al.
Published: (2024)
Sketch2Motion: Text-driven 2D Sketch to 3D Animation via Diffusion-guided Skeleton Optimization
by: Rai, Gaurav, et al.
Published: (2026)
by: Rai, Gaurav, et al.
Published: (2026)
Enhancing Sketch Animation: Text-to-Video Diffusion Models with Temporal Consistency and Rigidity Constraints
by: Rai, Gaurav, et al.
Published: (2024)
by: Rai, Gaurav, et al.
Published: (2024)
Similar Items
-
Object-level Visual Prompts for Compositional Image Generation
by: Parmar, Gaurav, et al.
Published: (2025) -
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
by: Wang, Kuan-Chieh, et al.
Published: (2024) -
Nested Attention: Semantic-aware Attention Values for Concept Personalization
by: Patashnik, Or, et al.
Published: (2025) -
Continuous Control of Editing Models via Adaptive-Origin Guidance
by: Wolf, Alon, et al.
Published: (2026) -
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
by: Dahary, Omer, et al.
Published: (2024)