Agentic Design of Compositional Machines
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Wenqian, Liu, Weiyang, Liu, Zhen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Controlling Text-to-Image Diffusion by Orthogonal Finetuning
by: Qiu, Zeju, et al.
Published: (2023)
by: Qiu, Zeju, et al.
Published: (2023)
Multi-LoRA Composition for Image Generation
by: Zhong, Ming, et al.
Published: (2024)
by: Zhong, Ming, et al.
Published: (2024)
Neuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical Prognosis
by: Aiersilan, Aizierjiang, et al.
Published: (2026)
by: Aiersilan, Aizierjiang, et al.
Published: (2026)
JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation
by: Song, Lin, et al.
Published: (2026)
by: Song, Lin, et al.
Published: (2026)
StyleMotif: Multi-Modal Motion Stylization using Style-Content Cross Fusion
by: Guo, Ziyu, et al.
Published: (2025)
by: Guo, Ziyu, et al.
Published: (2025)
A Solver-Aided Hierarchical Language for LLM-Driven CAD Design
by: Jones, Benjamin T., et al.
Published: (2025)
by: Jones, Benjamin T., et al.
Published: (2025)
Gesture2Text: A Generalizable Decoder for Word-Gesture Keyboards in XR Through Trajectory Coarse Discretization and Pre-training
by: Shen, Junxiao, et al.
Published: (2024)
by: Shen, Junxiao, et al.
Published: (2024)
Stylus: Automatic Adapter Selection for Diffusion Models
by: Luo, Michael, et al.
Published: (2024)
by: Luo, Michael, et al.
Published: (2024)
An Image is Worth Multiple Words: Discovering Object Level Concepts using Multi-Concept Prompt Learning
by: Jin, Chen, et al.
Published: (2023)
by: Jin, Chen, et al.
Published: (2023)
Text-guided Controllable Mesh Refinement for Interactive 3D Modeling
by: Chen, Yun-Chun, et al.
Published: (2024)
by: Chen, Yun-Chun, et al.
Published: (2024)
Unbounded: A Generative Infinite Game of Character Life Simulation
by: Li, Jialu, et al.
Published: (2024)
by: Li, Jialu, et al.
Published: (2024)
GraphDreamer: Compositional 3D Scene Synthesis from Scene Graphs
by: Gao, Gege, et al.
Published: (2023)
by: Gao, Gege, et al.
Published: (2023)
3DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code
by: Gao, Yipeng, et al.
Published: (2026)
by: Gao, Yipeng, et al.
Published: (2026)
Compositional Neural Textures
by: Tu, Peihan, et al.
Published: (2024)
by: Tu, Peihan, et al.
Published: (2024)
DesignAsCode: Bridging Structural Editability and Visual Fidelity in Graphic Design Generation
by: Liu, Ziyuan, et al.
Published: (2026)
by: Liu, Ziyuan, et al.
Published: (2026)
Machine Learning for Scientific Visualization: Ensemble Data Analysis
by: Gadirov, Hamid
Published: (2025)
by: Gadirov, Hamid
Published: (2025)
GAF-FusionNet: Multimodal ECG Analysis via Gramian Angular Fields and Split Attention
by: Qin, Jiahao, et al.
Published: (2024)
by: Qin, Jiahao, et al.
Published: (2024)
GeneOH Diffusion: Towards Generalizable Hand-Object Interaction Denoising via Denoising Diffusion
by: Liu, Xueyi, et al.
Published: (2024)
by: Liu, Xueyi, et al.
Published: (2024)
Improving Video Generation with Human Feedback
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
LRM: Large Reconstruction Model for Single Image to 3D
by: Hong, Yicong, et al.
Published: (2023)
by: Hong, Yicong, et al.
Published: (2023)
Ghost on the Shell: An Expressive Representation of General 3D Shapes
by: Liu, Zhen, et al.
Published: (2023)
by: Liu, Zhen, et al.
Published: (2023)
GaussianVAE: Adaptive Learning Dynamics of 3D Gaussians for High-Fidelity Super-Resolution
by: Khalid, Shuja, et al.
Published: (2025)
by: Khalid, Shuja, et al.
Published: (2025)
ArchComplete: Autoregressive 3D Architectural Design Generation with Hierarchical Diffusion-Based Upsampling
by: Rasoulzadeh, S., et al.
Published: (2024)
by: Rasoulzadeh, S., et al.
Published: (2024)
RealmDreamer: Text-Driven 3D Scene Generation with Inpainting and Depth Diffusion
by: Shriram, Jaidev, et al.
Published: (2024)
by: Shriram, Jaidev, et al.
Published: (2024)
ComboStoc: Combinatorial Stochasticity for Diffusion Generative Models
by: Xu, Rui, et al.
Published: (2024)
by: Xu, Rui, et al.
Published: (2024)
An objective comparison of methods for augmented reality in laparoscopic liver resection by preoperative-to-intraoperative image fusion
by: Ali, Sharib, et al.
Published: (2024)
by: Ali, Sharib, et al.
Published: (2024)
DGS-LRM: Real-Time Deformable 3D Gaussian Reconstruction From Monocular Videos
by: Lin, Chieh Hubert, et al.
Published: (2025)
by: Lin, Chieh Hubert, et al.
Published: (2025)
GoodDrag: Towards Good Practices for Drag Editing with Diffusion Models
by: Zhang, Zewei, et al.
Published: (2024)
by: Zhang, Zewei, et al.
Published: (2024)
DynamicGTR: Leveraging Graph Topology Representation Preferences to Boost VLM Capabilities on Graph QAs
by: Wei, Yanbin, et al.
Published: (2026)
by: Wei, Yanbin, et al.
Published: (2026)
NeuSDFusion: A Spatial-Aware Generative Model for 3D Shape Completion, Reconstruction, and Generation
by: Cui, Ruikai, et al.
Published: (2024)
by: Cui, Ruikai, et al.
Published: (2024)
PAPR in Motion: Seamless Point-level 3D Scene Interpolation
by: Peng, Shichong, et al.
Published: (2024)
by: Peng, Shichong, et al.
Published: (2024)
Asset Harvester: Extracting 3D Assets from Autonomous Driving Logs for Simulation
by: Cao, Tianshi, et al.
Published: (2026)
by: Cao, Tianshi, et al.
Published: (2026)
DiffusionBrowser: Interactive Diffusion Previews via Multi-Branch Decoders
by: Hong, Susung, et al.
Published: (2025)
by: Hong, Susung, et al.
Published: (2025)
Learning to Edit Visual Programs with Self-Supervision
by: Jones, R. Kenny, et al.
Published: (2024)
by: Jones, R. Kenny, et al.
Published: (2024)
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
by: Cai, Shengqu, et al.
Published: (2024)
by: Cai, Shengqu, et al.
Published: (2024)
Improved Baselines with Representation Autoencoders
by: Singh, Jaskirat, et al.
Published: (2026)
by: Singh, Jaskirat, et al.
Published: (2026)
What matters for Representation Alignment: Global Information or Spatial Structure?
by: Singh, Jaskirat, et al.
Published: (2025)
by: Singh, Jaskirat, et al.
Published: (2025)
One Trajectory, One Token: Grounded Video Tokenization via Panoptic Sub-object Trajectory
by: Zheng, Chenhao, et al.
Published: (2025)
by: Zheng, Chenhao, et al.
Published: (2025)
MeshSplat: Generalizable Sparse-View Surface Reconstruction via Gaussian Splatting
by: Chang, Hanzhi, et al.
Published: (2025)
by: Chang, Hanzhi, et al.
Published: (2025)
End-to-End Training for Unified Tokenization and Latent Denoising
by: Duggal, Shivam, et al.
Published: (2026)
by: Duggal, Shivam, et al.
Published: (2026)
Similar Items
-
Controlling Text-to-Image Diffusion by Orthogonal Finetuning
by: Qiu, Zeju, et al.
Published: (2023) -
Multi-LoRA Composition for Image Generation
by: Zhong, Ming, et al.
Published: (2024) -
Neuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical Prognosis
by: Aiersilan, Aizierjiang, et al.
Published: (2026) -
JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation
by: Song, Lin, et al.
Published: (2026) -
StyleMotif: Multi-Modal Motion Stylization using Style-Content Cross Fusion
by: Guo, Ziyu, et al.
Published: (2025)