Does FLUX Already Know How to Perform Physically Plausible Image Composition?
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Shilin, Lian, Zhuming, Zhou, Zihan, Zhang, Shaocong, Zhao, Chen, Kong, Adams Wai-Kin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DragFlow: Unleashing DiT Priors with Region Based Supervision for Drag Editing
by: Zhou, Zihan, et al.
Published: (2025)
by: Zhou, Zihan, et al.
Published: (2025)
Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances
by: Lu, Shilin, et al.
Published: (2024)
by: Lu, Shilin, et al.
Published: (2024)
All That Glitters Is Not Gold: Key-Secured 3D Secrets within 3D Gaussian Splatting
by: Ren, Yan, et al.
Published: (2025)
by: Ren, Yan, et al.
Published: (2025)
Set You Straight: Auto-Steering Denoising Trajectories to Sidestep Unwanted Concepts
by: Li, Leyang, et al.
Published: (2025)
by: Li, Leyang, et al.
Published: (2025)
MACE: Mass Concept Erasure in Diffusion Models
by: Lu, Shilin, et al.
Published: (2024)
by: Lu, Shilin, et al.
Published: (2024)
SparseMamba-PCL: Scribble-Supervised Medical Image Segmentation via SAM-Guided Progressive Collaborative Learning
by: Qiu, Luyi, et al.
Published: (2025)
by: Qiu, Luyi, et al.
Published: (2025)
Flux Already Knows -- Activating Subject-Driven Image Generation without Training
by: Kang, Hao, et al.
Published: (2025)
by: Kang, Hao, et al.
Published: (2025)
Improving Concept Alignment in Vision-Language Concept Bottleneck Models
by: Selvaraj, Nithish Muthuchamy, et al.
Published: (2024)
by: Selvaraj, Nithish Muthuchamy, et al.
Published: (2024)
DiffPop: Plausibility-Guided Object Placement Diffusion for Image Composition
by: Liu, Jiacheng, et al.
Published: (2024)
by: Liu, Jiacheng, et al.
Published: (2024)
MACS: Multi-source Audio-to-image Generation with Contextual Significance and Semantic Alignment
by: Zhou, Hao, et al.
Published: (2025)
by: Zhou, Hao, et al.
Published: (2025)
PhyCAGE: Physically Plausible Compositional 3D Asset Generation from a Single Image
by: Yan, Han, et al.
Published: (2024)
by: Yan, Han, et al.
Published: (2024)
UniLumos: Fast and Unified Image and Video Relighting with Physics-Plausible Feedback
by: Liu, Ropeway, et al.
Published: (2025)
by: Liu, Ropeway, et al.
Published: (2025)
Mixture of Low-rank Experts for Transferable AI-Generated Image Detection
by: Liu, Zihan, et al.
Published: (2024)
by: Liu, Zihan, et al.
Published: (2024)
3DIS-FLUX: simple and efficient multi-instance generation with DiT rendering
by: Zhou, Dewei, et al.
Published: (2025)
by: Zhou, Dewei, et al.
Published: (2025)
Model Already Knows the Best Noise: Bayesian Active Noise Selection via Attention in Video Diffusion Model
by: Kim, Kwanyoung, et al.
Published: (2025)
by: Kim, Kwanyoung, et al.
Published: (2025)
FLUX that Plays Music
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
1.58-bit FLUX
by: Yang, Chenglin, et al.
Published: (2024)
by: Yang, Chenglin, et al.
Published: (2024)
Video Self-Distillation for Single-Image Encoders: A Step Toward Physically Plausible Perception
by: Simon, Marcel, et al.
Published: (2025)
by: Simon, Marcel, et al.
Published: (2025)
PhyRecon: Physically Plausible Neural Scene Reconstruction
by: Ni, Junfeng, et al.
Published: (2024)
by: Ni, Junfeng, et al.
Published: (2024)
Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility
by: Hao, Yutong, et al.
Published: (2025)
by: Hao, Yutong, et al.
Published: (2025)
SpectralTrain: A Universal Framework for Hyperspectral Image Classification
by: Zhou, Meihua, et al.
Published: (2025)
by: Zhou, Meihua, et al.
Published: (2025)
ColorFLUX: A Structure-Color Decoupling Framework for Old Photo Colorization
by: Li, Bingchen, et al.
Published: (2026)
by: Li, Bingchen, et al.
Published: (2026)
Real2Sim in HOI: Toward Physically Plausible HOI Reconstruction from Monocular Videos
by: Zhao, Yubo, et al.
Published: (2026)
by: Zhao, Yubo, et al.
Published: (2026)
NanoFLUX: Distillation-Driven Compression of Large Text-to-Image Generation Models for Mobile Devices
by: Chavhan, Ruchika, et al.
Published: (2026)
by: Chavhan, Ruchika, et al.
Published: (2026)
Tempered Self-Similarity Alignment for Physically Plausible Video Generation
by: Kim, Manjin, et al.
Published: (2026)
by: Kim, Manjin, et al.
Published: (2026)
Physical Plausibility-aware Trajectory Prediction via Locomotion Embodiment
by: Taketsugu, Hiromu, et al.
Published: (2025)
by: Taketsugu, Hiromu, et al.
Published: (2025)
Measuring Physical Plausibility of 3D Human Poses Using Physics Simulation
by: Louis, Nathan, et al.
Published: (2025)
by: Louis, Nathan, et al.
Published: (2025)
FLUX-Text: A Simple and Advanced Diffusion Transformer Baseline for Scene Text Editing
by: Lan, Rui, et al.
Published: (2025)
by: Lan, Rui, et al.
Published: (2025)
PhySIC: Physically Plausible 3D Human-Scene Interaction and Contact from a Single Image
by: Muralidhar, Pradyumna Yalandur, et al.
Published: (2025)
by: Muralidhar, Pradyumna Yalandur, et al.
Published: (2025)
What Does DALL-E 2 Know About Radiology?
by: Adams, Lisa C., et al.
Published: (2022)
by: Adams, Lisa C., et al.
Published: (2022)
Hierarchical Fine-grained Preference Optimization for Physically Plausible Video Generation
by: Chen, Harold Haodong, et al.
Published: (2025)
by: Chen, Harold Haodong, et al.
Published: (2025)
ReinDiffuse: Crafting Physically Plausible Motions with Reinforced Diffusion Model
by: Han, Gaoge, et al.
Published: (2024)
by: Han, Gaoge, et al.
Published: (2024)
Chain of Event-Centric Causal Thought for Physically Plausible Video Generation
by: Wang, Zixuan, et al.
Published: (2026)
by: Wang, Zixuan, et al.
Published: (2026)
THOM: Generating Physically Plausible Hand-Object Meshes From Text
by: Jeong, Uyoung, et al.
Published: (2026)
by: Jeong, Uyoung, et al.
Published: (2026)
FLUX-Makeup: High-Fidelity, Identity-Consistent, and Robust Makeup Transfer via Diffusion Transformer
by: Zhu, Jian, et al.
Published: (2025)
by: Zhu, Jian, et al.
Published: (2025)
Diffusion Knows Transparency: Repurposing Video Diffusion for Transparent Object Depth and Normal Estimation
by: Xu, Shaocong, et al.
Published: (2025)
by: Xu, Shaocong, et al.
Published: (2025)
Towards Anatomically Plausible Human Image Generation via Synthetic Localized Preferences
by: Li, Bao, et al.
Published: (2026)
by: Li, Bao, et al.
Published: (2026)
From Generated Human Videos to Physically Plausible Robot Trajectories
by: Ni, James, et al.
Published: (2025)
by: Ni, James, et al.
Published: (2025)
PhysPart: Physically Plausible Part Completion for Interactable Objects
by: Luo, Rundong, et al.
Published: (2024)
by: Luo, Rundong, et al.
Published: (2024)
MM-CondChain: A Programmatically Verified Benchmark for Visually Grounded Deep Compositional Reasoning
by: Shen, Haozhan, et al.
Published: (2026)
by: Shen, Haozhan, et al.
Published: (2026)
Similar Items
-
DragFlow: Unleashing DiT Priors with Region Based Supervision for Drag Editing
by: Zhou, Zihan, et al.
Published: (2025) -
Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances
by: Lu, Shilin, et al.
Published: (2024) -
All That Glitters Is Not Gold: Key-Secured 3D Secrets within 3D Gaussian Splatting
by: Ren, Yan, et al.
Published: (2025) -
Set You Straight: Auto-Steering Denoising Trajectories to Sidestep Unwanted Concepts
by: Li, Leyang, et al.
Published: (2025) -
MACE: Mass Concept Erasure in Diffusion Models
by: Lu, Shilin, et al.
Published: (2024)