FlexGen: Flexible Multi-View Generation from Text and Image Inputs
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Xinli, Ge, Wenhang, Lin, Jiantao, Feng, Jiawei, Xu, Lie, Zhao, HanFeng, Zhang, Shunsi, Chen, Ying-Cong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FlexPainter: Flexible and Multi-View Consistent Texture Generation
by: Yan, Dongyu, et al.
Published: (2025)
by: Yan, Dongyu, et al.
Published: (2025)
PRM: Photometric Stereo based Large Reconstruction Model
by: Ge, Wenhang, et al.
Published: (2024)
by: Ge, Wenhang, et al.
Published: (2024)
Kiss3DGen: Repurposing Image Diffusion Models for 3D Asset Generation
by: Lin, Jiantao, et al.
Published: (2025)
by: Lin, Jiantao, et al.
Published: (2025)
GaussianProperty: Integrating Physical Properties to 3D Gaussians with LMMs
by: Xu, Xinli, et al.
Published: (2024)
by: Xu, Xinli, et al.
Published: (2024)
SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
by: Shen, Guibao, et al.
Published: (2024)
by: Shen, Guibao, et al.
Published: (2024)
EmoGen: Emotional Image Content Generation with Text-to-Image Diffusion Models
by: Yang, Jingyuan, et al.
Published: (2024)
by: Yang, Jingyuan, et al.
Published: (2024)
Flex3D: Feed-Forward 3D Generation with Flexible Reconstruction Model and Input View Curation
by: Han, Junlin, et al.
Published: (2024)
by: Han, Junlin, et al.
Published: (2024)
A Mechanistic View on Video Generation as World Models: State and Dynamics
by: Wang, Luozhou, et al.
Published: (2026)
by: Wang, Luozhou, et al.
Published: (2026)
FlexEControl: Flexible and Efficient Multimodal Control for Text-to-Image Generation
by: He, Xuehai, et al.
Published: (2024)
by: He, Xuehai, et al.
Published: (2024)
Human Multi-View Synthesis from a Single-View Model:Transferred Body and Face Representations
by: Feng, Yu, et al.
Published: (2024)
by: Feng, Yu, et al.
Published: (2024)
ComfyMind: Toward General-Purpose Generation via Tree-Based Planning and Reactive Feedback
by: Guo, Litao, et al.
Published: (2025)
by: Guo, Litao, et al.
Published: (2025)
FlexAM: Flexible Appearance-Motion Decomposition for Versatile Video Generation Control
by: Sheng, Mingzhi, et al.
Published: (2026)
by: Sheng, Mingzhi, et al.
Published: (2026)
FlexRAG: A Flexible and Comprehensive Framework for Retrieval-Augmented Generation
by: Zhang, Zhuocheng, et al.
Published: (2025)
by: Zhang, Zhuocheng, et al.
Published: (2025)
Text-Anchored Score Composition: Tackling Condition Misalignment in Text-to-Image Diffusion Models
by: Wang, Luozhou, et al.
Published: (2023)
by: Wang, Luozhou, et al.
Published: (2023)
CamPilot: Improving Camera Control in Video Diffusion Model with Efficient Camera Reward Feedback
by: Ge, Wenhang, et al.
Published: (2026)
by: Ge, Wenhang, et al.
Published: (2026)
RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data
by: Chen, Harold Haodong, et al.
Published: (2026)
by: Chen, Harold Haodong, et al.
Published: (2026)
LLM-Optic: Unveiling the Capabilities of Large Language Models for Universal Visual Grounding
by: Zhao, Haoyu, et al.
Published: (2024)
by: Zhao, Haoyu, et al.
Published: (2024)
FlexPara: Flexible Neural Surface Parameterization
by: Zhao, Yuming, et al.
Published: (2025)
by: Zhao, Yuming, et al.
Published: (2025)
FlexControl: Computation-Aware ControlNet with Differentiable Router for Text-to-Image Generation
by: Fang, Zheng, et al.
Published: (2025)
by: Fang, Zheng, et al.
Published: (2025)
FlexSQL: Flexible Exploration and Execution Make Better Text-to-SQL Agents
by: Pham, Quang Hieu, et al.
Published: (2026)
by: Pham, Quang Hieu, et al.
Published: (2026)
Uni-Renderer: Unifying Rendering and Inverse Rendering Via Dual Stream Diffusion
by: Chen, Zhifei, et al.
Published: (2024)
by: Chen, Zhifei, et al.
Published: (2024)
FlexID: Training-Free Flexible Identity Injection via Intent-Aware Modulation for Text-to-Image Generation
by: Li, Guandong, et al.
Published: (2026)
by: Li, Guandong, et al.
Published: (2026)
FlexTraj: Image-to-Video Generation with Flexible Point Trajectory Control
by: Zhang, Zhiyuan, et al.
Published: (2025)
by: Zhang, Zhiyuan, et al.
Published: (2025)
FlexMap: Generalized HD Map Construction from Flexible Camera Configurations
by: Wang, Run, et al.
Published: (2026)
by: Wang, Run, et al.
Published: (2026)
FlexProofs: A Vector Commitment with Flexible Linear Time for Computing All Proofs
by: Liu, Jing, et al.
Published: (2026)
by: Liu, Jing, et al.
Published: (2026)
GenTron: Diffusion Transformers for Image and Video Generation
by: Chen, Shoufa, et al.
Published: (2023)
by: Chen, Shoufa, et al.
Published: (2023)
PreGenie: An Agentic Framework for High-quality Visual Presentation Generation
by: Xu, Xiaojie, et al.
Published: (2025)
by: Xu, Xiaojie, et al.
Published: (2025)
FlexCare: Leveraging Cross-Task Synergy for Flexible Multimodal Healthcare Prediction
by: Xu, Muhao, et al.
Published: (2024)
by: Xu, Muhao, et al.
Published: (2024)
X-Ray: A Sequential 3D Representation For Generation
by: Hu, Tao, et al.
Published: (2024)
by: Hu, Tao, et al.
Published: (2024)
Forward Kinematics of Object Transporting by a Multi-Robot System with a Deformable Sheet
by: Hu, Jiawei, et al.
Published: (2023)
by: Hu, Jiawei, et al.
Published: (2023)
FlexIO: Flexible Single- and Multi-Channel Speech Separation and Enhancement
by: Masuyama, Yoshiki, et al.
Published: (2025)
by: Masuyama, Yoshiki, et al.
Published: (2025)
SketchFlex: Facilitating Spatial-Semantic Coherence in Text-to-Image Generation with Region-Based Sketches
by: Lin, Haichuan, et al.
Published: (2025)
by: Lin, Haichuan, et al.
Published: (2025)
MathGen: Revealing the Illusion of Mathematical Competence through Text-to-Image Generation
by: Liu, Ruiyao, et al.
Published: (2026)
by: Liu, Ruiyao, et al.
Published: (2026)
Rethinking Image-to-3D Generation with Sparse Queries: Efficiency, Capacity, and Input-View Bias
by: Xu, Zhiyuan, et al.
Published: (2026)
by: Xu, Zhiyuan, et al.
Published: (2026)
Towards Robust Text-to-Image Person Retrieval: Multi-View Reformulation for Semantic Compensation
by: Yuan, Chao, et al.
Published: (2026)
by: Yuan, Chao, et al.
Published: (2026)
FlexEdit: Marrying Free-Shape Masks to VLLM for Flexible Image Editing
by: Yuan, Tianshuo, et al.
Published: (2024)
by: Yuan, Tianshuo, et al.
Published: (2024)
FlexTSF: A Flexible Forecasting Model for Time Series with Variable Regularities
by: Xiao, Jingge, et al.
Published: (2024)
by: Xiao, Jingge, et al.
Published: (2024)
DiMeR: Disentangled Mesh Reconstruction Model
by: Jiang, Lutao, et al.
Published: (2025)
by: Jiang, Lutao, et al.
Published: (2025)
FlexLLM: Composable HLS Library for Flexible Hybrid LLM Accelerator Design
by: Zhang, Jiahao, et al.
Published: (2026)
by: Zhang, Jiahao, et al.
Published: (2026)
FlexWorld: Progressively Expanding 3D Scenes for Flexiable-View Synthesis
by: Chen, Luxi, et al.
Published: (2025)
by: Chen, Luxi, et al.
Published: (2025)
Similar Items
-
FlexPainter: Flexible and Multi-View Consistent Texture Generation
by: Yan, Dongyu, et al.
Published: (2025) -
PRM: Photometric Stereo based Large Reconstruction Model
by: Ge, Wenhang, et al.
Published: (2024) -
Kiss3DGen: Repurposing Image Diffusion Models for 3D Asset Generation
by: Lin, Jiantao, et al.
Published: (2025) -
GaussianProperty: Integrating Physical Properties to 3D Gaussians with LMMs
by: Xu, Xinli, et al.
Published: (2024) -
SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
by: Shen, Guibao, et al.
Published: (2024)