DiffGraph: An Automated Agent-driven Model Merging Framework for In-the-Wild Text-to-Image Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Zhuoling, Rahmani, Hossein, Zhang, Jiarui, Xue, Yu, Mirmehdi, Majid, Kuen, Jason, Gu, Jiuxiang, Liu, Jun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Automatic Method Illustration Generation for AI Scientific Papers via Drawing Middleware Creation, Evolution, and Orchestration
por: Li, Zhuoling, et al.
Publicado: (2026)
por: Li, Zhuoling, et al.
Publicado: (2026)
LongDiff: Training-Free Long Video Generation in One Go
por: Li, Zhuoling, et al.
Publicado: (2025)
por: Li, Zhuoling, et al.
Publicado: (2025)
DiffGraph: Heterogeneous Graph Diffusion Model
por: Li, Zongwei, et al.
Publicado: (2025)
por: Li, Zongwei, et al.
Publicado: (2025)
Diff-Tracker: Text-to-Image Diffusion Models are Unsupervised Trackers
por: Zhang, Zhengbo, et al.
Publicado: (2024)
por: Zhang, Zhengbo, et al.
Publicado: (2024)
Learning to Generate Cross-Task Unexploitable Examples
por: Qu, Haoxuan, et al.
Publicado: (2025)
por: Qu, Haoxuan, et al.
Publicado: (2025)
When Visual Privacy Protection Meets Multimodal Large Language Models
por: Hui, Xiaofei, et al.
Publicado: (2026)
por: Hui, Xiaofei, et al.
Publicado: (2026)
SNCE: Geometry-Aware Supervision for Scalable Discrete Image Generation
por: Li, Shufan, et al.
Publicado: (2026)
por: Li, Shufan, et al.
Publicado: (2026)
ToolFG: Towards Well-Grounded Fine-Grained Image Classification
por: Xue, Yu, et al.
Publicado: (2026)
por: Xue, Yu, et al.
Publicado: (2026)
ImageFolder: Autoregressive Image Generation with Folded Tokens
por: Li, Xiang, et al.
Publicado: (2024)
por: Li, Xiang, et al.
Publicado: (2024)
XQ-GAN: An Open-source Image Tokenization Framework for Autoregressive Generation
por: Li, Xiang, et al.
Publicado: (2024)
por: Li, Xiang, et al.
Publicado: (2024)
Automated Radiology Report Generation: A Review of Recent Advances
por: Sloan, Phillip, et al.
Publicado: (2024)
por: Sloan, Phillip, et al.
Publicado: (2024)
DisC-GS: Discontinuity-aware Gaussian Splatting
por: Qu, Haoxuan, et al.
Publicado: (2024)
por: Qu, Haoxuan, et al.
Publicado: (2024)
Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation
por: Li, Shufan, et al.
Publicado: (2025)
por: Li, Shufan, et al.
Publicado: (2025)
SceneLLM: Implicit Language Reasoning in LLM for Dynamic Scene Graph Generation
por: Zhang, Hang, et al.
Publicado: (2024)
por: Zhang, Hang, et al.
Publicado: (2024)
Robust Latent Matters: Boosting Image Generation with Sampling Error Synthesis
por: Qiu, Kai, et al.
Publicado: (2025)
por: Qiu, Kai, et al.
Publicado: (2025)
Customization Assistant for Text-to-image Generation
por: Zhou, Yufan, et al.
Publicado: (2023)
por: Zhou, Yufan, et al.
Publicado: (2023)
Sparse-LaViDa: Sparse Multimodal Discrete Diffusion Language Models
por: Li, Shufan, et al.
Publicado: (2025)
por: Li, Shufan, et al.
Publicado: (2025)
Visual-textual Dermatoglyphic Animal Biometrics: A First Case Study on Panthera tigris
por: Li, Wenshuo, et al.
Publicado: (2025)
por: Li, Wenshuo, et al.
Publicado: (2025)
Co-STAR: Collaborative Curriculum Self-Training with Adaptive Regularization for Source-Free Video Domain Adaptation
por: Dadashzadeh, Amirhossein, et al.
Publicado: (2025)
por: Dadashzadeh, Amirhossein, et al.
Publicado: (2025)
Unsupervised View-Invariant Human Posture Representation
por: Sardari, Faegheh, et al.
Publicado: (2021)
por: Sardari, Faegheh, et al.
Publicado: (2021)
Clinically-aligned Multi-modal Chest X-ray Classification
por: Sloan, Phillip, et al.
Publicado: (2025)
por: Sloan, Phillip, et al.
Publicado: (2025)
KG-FairDiff: Knowledge Graph-Guided Prompt Refinement for Demographically Fair Text-to-Image Generation
por: Davoodi, Farbod, et al.
Publicado: (2026)
por: Davoodi, Farbod, et al.
Publicado: (2026)
METAL: A Multi-Agent Framework for Chart Generation with Test-Time Scaling
por: Li, Bingxuan, et al.
Publicado: (2025)
por: Li, Bingxuan, et al.
Publicado: (2025)
Image Tokenizer Needs Post-Training
por: Qiu, Kai, et al.
Publicado: (2025)
por: Qiu, Kai, et al.
Publicado: (2025)
Is Monitoring Enough? Strategic Agent Selection For Stealthy Attack in Multi-Agent Discussions
por: Xiang, Qiuchi, et al.
Publicado: (2026)
por: Xiang, Qiuchi, et al.
Publicado: (2026)
Prediction of Thrombectomy Functional Outcomes using Multimodal Data
por: Samak, Zeynel A., et al.
Publicado: (2020)
por: Samak, Zeynel A., et al.
Publicado: (2020)
TranSOP: Transformer-based Multimodal Classification for Stroke Treatment Outcome Prediction
por: Samak, Zeynel A., et al.
Publicado: (2023)
por: Samak, Zeynel A., et al.
Publicado: (2023)
Automatic Prediction of Stroke Treatment Outcomes: Latest Advances and Perspectives
por: Samak, Zeynel A., et al.
Publicado: (2024)
por: Samak, Zeynel A., et al.
Publicado: (2024)
Multimodal LLMs as Customized Reward Models for Text-to-Image Generation
por: Zhou, Shijie, et al.
Publicado: (2025)
por: Zhou, Shijie, et al.
Publicado: (2025)
ARTIST: Improving the Generation of Text-rich Images with Disentangled Diffusion Models and Large Language Models
por: Zhang, Jianyi, et al.
Publicado: (2024)
por: Zhang, Jianyi, et al.
Publicado: (2024)
High‐Velocity Impact Response and Comfort Properties of Discrete‐Droplet‐Coated Cushioning Composite Fabrics
por: Zhuoling Yu, et al.
Publicado: (2026)
por: Zhuoling Yu, et al.
Publicado: (2026)
XGRAG: A Graph-Native Framework for Explaining KG-based Retrieval-Augmented Generation
por: Li, Zhuoling, et al.
Publicado: (2026)
por: Li, Zhuoling, et al.
Publicado: (2026)
ChimpVLM: Ethogram-Enhanced Chimpanzee Behaviour Recognition
por: Brookes, Otto, et al.
Publicado: (2024)
por: Brookes, Otto, et al.
Publicado: (2024)
Trajectory-guided Motion Perception for Facial Expression Quality Assessment in Neurological Disorders
por: Duan, Shuchao, et al.
Publicado: (2025)
por: Duan, Shuchao, et al.
Publicado: (2025)
AI-Generated Content (AIGC) for Various Data Modalities: A Survey
por: Foo, Lin Geng, et al.
Publicado: (2023)
por: Foo, Lin Geng, et al.
Publicado: (2023)
An Image-like Diffusion Method for Human-Object Interaction Detection
por: Hui, Xiaofei, et al.
Publicado: (2025)
por: Hui, Xiaofei, et al.
Publicado: (2025)
TSTMotion: Training-free Scene-aware Text-to-motion Generation
por: Guo, Ziyan, et al.
Publicado: (2025)
por: Guo, Ziyan, et al.
Publicado: (2025)
PFB-Diff: Progressive Feature Blending Diffusion for Text-driven Image Editing
por: Huang, Wenjing, et al.
Publicado: (2023)
por: Huang, Wenjing, et al.
Publicado: (2023)
ViewDiff: 3D-Consistent Image Generation with Text-to-Image Models
por: Höllein, Lukas, et al.
Publicado: (2024)
por: Höllein, Lukas, et al.
Publicado: (2024)
Fuzz-Testing Meets LLM-Based Agents: An Automated and Efficient Framework for Jailbreaking Text-To-Image Generation Models
por: Dong, Yingkai, et al.
Publicado: (2024)
por: Dong, Yingkai, et al.
Publicado: (2024)
Ejemplares similares
-
Automatic Method Illustration Generation for AI Scientific Papers via Drawing Middleware Creation, Evolution, and Orchestration
por: Li, Zhuoling, et al.
Publicado: (2026) -
LongDiff: Training-Free Long Video Generation in One Go
por: Li, Zhuoling, et al.
Publicado: (2025) -
DiffGraph: Heterogeneous Graph Diffusion Model
por: Li, Zongwei, et al.
Publicado: (2025) -
Diff-Tracker: Text-to-Image Diffusion Models are Unsupervised Trackers
por: Zhang, Zhengbo, et al.
Publicado: (2024) -
Learning to Generate Cross-Task Unexploitable Examples
por: Qu, Haoxuan, et al.
Publicado: (2025)