ComfyMind: Toward General-Purpose Generation via Tree-Based Planning and Reactive Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Litao, Xu, Xinli, Wang, Luozhou, Lin, Jiantao, Zhou, Jinsong, Zhang, Zixin, Su, Bolan, Chen, Ying-Cong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VideoMemory: Toward Consistent Video Generation via Memory Integration
by: Zhou, Jinsong, et al.
Published: (2026)
by: Zhou, Jinsong, et al.
Published: (2026)
PresentCoach: Dual-Agent Presentation Coaching through Exemplars and Interactive Feedback
by: Chen, Sirui, et al.
Published: (2025)
by: Chen, Sirui, et al.
Published: (2025)
ComfySearch: Autonomous Exploration and Reasoning for ComfyUI Workflows
by: Su, Jinwei, et al.
Published: (2026)
by: Su, Jinwei, et al.
Published: (2026)
ComfyGPT: A Self-Optimizing Multi-Agent System for Comprehensive ComfyUI Workflow Generation
by: Huang, Oucheng, et al.
Published: (2025)
by: Huang, Oucheng, et al.
Published: (2025)
ComfyGI: Automatic Improvement of Image Generation Workflows
by: Sobania, Dominik, et al.
Published: (2024)
by: Sobania, Dominik, et al.
Published: (2024)
FlexGen: Flexible Multi-View Generation from Text and Image Inputs
by: Xu, Xinli, et al.
Published: (2024)
by: Xu, Xinli, et al.
Published: (2024)
ComfyGen: Prompt-Adaptive Workflows for Text-to-Image Generation
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
ComfyUI-R1: Exploring Reasoning Models for Workflow Generation
by: Xu, Zhenran, et al.
Published: (2025)
by: Xu, Zhenran, et al.
Published: (2025)
SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
by: Shen, Guibao, et al.
Published: (2024)
by: Shen, Guibao, et al.
Published: (2024)
Kiss3DGen: Repurposing Image Diffusion Models for 3D Asset Generation
by: Lin, Jiantao, et al.
Published: (2025)
by: Lin, Jiantao, et al.
Published: (2025)
Comprehensive Bioinformatic Analysis of TONSL Expression in Pan‐Cancer
by: Ying Yang, et al.
Published: (2026)
by: Ying Yang, et al.
Published: (2026)
Show, Don't Tell: Morphing Latent Reasoning into Image Generation
by: Chen, Harold Haodong, et al.
Published: (2026)
by: Chen, Harold Haodong, et al.
Published: (2026)
ComfyBench: Benchmarking LLM-based Agents in ComfyUI for Autonomously Designing Collaborative AI Systems
by: Xue, Xiangyuan, et al.
Published: (2024)
by: Xue, Xiangyuan, et al.
Published: (2024)
PRM: Photometric Stereo based Large Reconstruction Model
by: Ge, Wenhang, et al.
Published: (2024)
by: Ge, Wenhang, et al.
Published: (2024)
StereoPilot: Learning Unified and Efficient Stereo Conversion via Generative Priors
by: Shen, Guibao, et al.
Published: (2025)
by: Shen, Guibao, et al.
Published: (2025)
FlexPainter: Flexible and Multi-View Consistent Texture Generation
by: Yan, Dongyu, et al.
Published: (2025)
by: Yan, Dongyu, et al.
Published: (2025)
Sentiment Analysis Based on RoBERTa for Amazon Review: An Empirical Study on Decision Making
by: Guo, Xinli
Published: (2024)
by: Guo, Xinli
Published: (2024)
FLIP: Flow-Centric Generative Planning as General-Purpose Manipulation World Model
by: Gao, Chongkai, et al.
Published: (2024)
by: Gao, Chongkai, et al.
Published: (2024)
CamPilot: Improving Camera Control in Video Diffusion Model with Efficient Camera Reward Feedback
by: Ge, Wenhang, et al.
Published: (2026)
by: Ge, Wenhang, et al.
Published: (2026)
A Mechanistic View on Video Generation as World Models: State and Dynamics
by: Wang, Luozhou, et al.
Published: (2026)
by: Wang, Luozhou, et al.
Published: (2026)
DuoGen: Towards General Purpose Interleaved Multimodal Generation
by: Shi, Min, et al.
Published: (2026)
by: Shi, Min, et al.
Published: (2026)
Towards A General-Purpose Motion Planning for Autonomous Vehicles Using Fluid Dynamics
by: Sormoli, MReza Alipour, et al.
Published: (2024)
by: Sormoli, MReza Alipour, et al.
Published: (2024)
PhysToolBench: Benchmarking Physical Tool Understanding for MLLMs
by: Zhang, Zixin, et al.
Published: (2025)
by: Zhang, Zixin, et al.
Published: (2025)
Uncovering Modality Discrepancy and Generalization Illusion for General-Purpose 3D Medical Segmentation
by: Zhang, Yichi, et al.
Published: (2026)
by: Zhang, Yichi, et al.
Published: (2026)
ProToM: Promoting Prosocial Behaviour via Theory of Mind-Informed Feedback
by: Bortoletto, Matteo, et al.
Published: (2025)
by: Bortoletto, Matteo, et al.
Published: (2025)
PreGenie: An Agentic Framework for High-quality Visual Presentation Generation
by: Xu, Xiaojie, et al.
Published: (2025)
by: Xu, Xiaojie, et al.
Published: (2025)
Generalized Pseudo-Relevance Feedback
by: Tu, Yiteng, et al.
Published: (2025)
by: Tu, Yiteng, et al.
Published: (2025)
Towards General-Purpose Text-Instruction-Guided Voice Conversion
by: Kuan, Chun-Yi, et al.
Published: (2023)
by: Kuan, Chun-Yi, et al.
Published: (2023)
ComfyUI-Copilot: An Intelligent Assistant for Automated Workflow Development
by: Xu, Zhenran, et al.
Published: (2025)
by: Xu, Zhenran, et al.
Published: (2025)
Towards Building General Purpose Embedding Models for Industry 4.0 Agents
by: Constantinides, Christodoulos, et al.
Published: (2025)
by: Constantinides, Christodoulos, et al.
Published: (2025)
Bailouts by Representation: A Minimal TLC Theory with Weighted Consent
by: Guo, Xinli
Published: (2025)
by: Guo, Xinli
Published: (2025)
Optimal Transfer Mechanism for Municipal Soft-Budget Constraints in Newfoundland
by: Guo, Xinli
Published: (2025)
by: Guo, Xinli
Published: (2025)
Two-Instrument Screening under Soft Budget Constraints
by: Guo, Xinli
Published: (2025)
by: Guo, Xinli
Published: (2025)
A4-Agent: An Agentic Framework for Zero-Shot Affordance Reasoning
by: Zhang, Zixin, et al.
Published: (2025)
by: Zhang, Zixin, et al.
Published: (2025)
Beyond Task and Motion Planning: Hierarchical Robot Planning with General-Purpose Skills
by: Hedegaard, Benned, et al.
Published: (2025)
by: Hedegaard, Benned, et al.
Published: (2025)
LucidFusion: Reconstructing 3D Gaussians with Arbitrary Unposed Images
by: He, Hao, et al.
Published: (2024)
by: He, Hao, et al.
Published: (2024)
Astra: Toward General-Purpose Mobile Robots via Hierarchical Multimodal Learning
by: Chen, Sheng, et al.
Published: (2025)
by: Chen, Sheng, et al.
Published: (2025)
Curvature-Informed SGD via General Purpose Lie-Group Preconditioners
by: Pooladzandi, Omead, et al.
Published: (2024)
by: Pooladzandi, Omead, et al.
Published: (2024)
Mind to Hand: Purposeful Robotic Control via Embodied Reasoning
by: Tang, Peijun, et al.
Published: (2025)
by: Tang, Peijun, et al.
Published: (2025)
Towards General-Purpose Model-Free Reinforcement Learning
by: Fujimoto, Scott, et al.
Published: (2025)
by: Fujimoto, Scott, et al.
Published: (2025)
Similar Items
-
VideoMemory: Toward Consistent Video Generation via Memory Integration
by: Zhou, Jinsong, et al.
Published: (2026) -
PresentCoach: Dual-Agent Presentation Coaching through Exemplars and Interactive Feedback
by: Chen, Sirui, et al.
Published: (2025) -
ComfySearch: Autonomous Exploration and Reasoning for ComfyUI Workflows
by: Su, Jinwei, et al.
Published: (2026) -
ComfyGPT: A Self-Optimizing Multi-Agent System for Comprehensive ComfyUI Workflow Generation
by: Huang, Oucheng, et al.
Published: (2025) -
ComfyGI: Automatic Improvement of Image Generation Workflows
by: Sobania, Dominik, et al.
Published: (2024)