CountLoop: Training-Free High-Instance Image Generation via Iterative Agent Guidance
Fuente:
arXiv
Guardado en:
| Autores principales: | Mondal, Anindya, Banerjee, Ayan, Nag, Sauradip, Llados, Josep, Zhu, Xiatian, Dutta, Anjan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
OmniCount: Multi-label Object Counting with Semantic-Geometric Priors
por: Mondal, Anindya, et al.
Publicado: (2024)
por: Mondal, Anindya, et al.
Publicado: (2024)
Actor-agnostic Multi-label Action Recognition with Multi-modal Query
por: Mondal, Anindya, et al.
Publicado: (2023)
por: Mondal, Anindya, et al.
Publicado: (2023)
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
por: Banerjee, Ayan, et al.
Publicado: (2025)
por: Banerjee, Ayan, et al.
Publicado: (2025)
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
por: Banerjee, Ayan, et al.
Publicado: (2024)
por: Banerjee, Ayan, et al.
Publicado: (2024)
CraftGraffiti: Exploring Human Identity with Custom Graffiti Art via Facial-Preserving Diffusion Models
por: Banerjee, Ayan, et al.
Publicado: (2025)
por: Banerjee, Ayan, et al.
Publicado: (2025)
RespoDiff: Dual-Module Bottleneck Transformation for Responsible & Faithful T2I Generation
por: Sreelatha, Silpa Vadakkeeveetil, et al.
Publicado: (2025)
por: Sreelatha, Silpa Vadakkeeveetil, et al.
Publicado: (2025)
GraphKD: Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation
por: Banerjee, Ayan, et al.
Publicado: (2024)
por: Banerjee, Ayan, et al.
Publicado: (2024)
DocRevive: A Unified Pipeline for Document Text Restoration
por: Purkayastha, Kunal, et al.
Publicado: (2026)
por: Purkayastha, Kunal, et al.
Publicado: (2026)
In-2-4D: Inbetweening from Two Single-View Images to 4D Generation
por: Nag, Sauradip, et al.
Publicado: (2025)
por: Nag, Sauradip, et al.
Publicado: (2025)
La cadena global de valor en la industria electrónica
por: Josep Lladós Masllorens
Publicado: (2018)
por: Josep Lladós Masllorens
Publicado: (2018)
Articulate That Object Part (ATOP): 3D Part Articulation via Text and Motion Personalization
por: Vora, Aditya, et al.
Publicado: (2025)
por: Vora, Aditya, et al.
Publicado: (2025)
CountDiffusion: Text-to-Image Synthesis with Training-Free Counting-Guidance Diffusion
por: Li, Yanyu, et al.
Publicado: (2025)
por: Li, Yanyu, et al.
Publicado: (2025)
A Fixed Point Iteration Technique for Proving Correctness of Slicing for Probabilistic Programs
por: Amtoft, Torben, et al.
Publicado: (2024)
por: Amtoft, Torben, et al.
Publicado: (2024)
SketchGPT: Autoregressive Modeling for Sketch Generation and Recognition
por: Tiwari, Adarsh, et al.
Publicado: (2024)
por: Tiwari, Adarsh, et al.
Publicado: (2024)
Fetch-A-Set: A Large-Scale OCR-Free Benchmark for Historical Document Retrieval
por: Molina, Adrià, et al.
Publicado: (2024)
por: Molina, Adrià, et al.
Publicado: (2024)
Visual Model Checking: Graph-Based Inference of Visual Routines for Image Retrieval
por: Molina, Adrià, et al.
Publicado: (2026)
por: Molina, Adrià, et al.
Publicado: (2026)
DistilDoc: Knowledge Distillation for Visually-Rich Document Applications
por: Van Landeghem, Jordy, et al.
Publicado: (2024)
por: Van Landeghem, Jordy, et al.
Publicado: (2024)
CountCluster: Training-Free Object Quantity Guidance with Cross-Attention Map Clustering for Text-to-Image Generation
por: Lee, Joohyeon, et al.
Publicado: (2025)
por: Lee, Joohyeon, et al.
Publicado: (2025)
ASIA: Adaptive 3D Segmentation using Few Image Annotations
por: Perla, Sai Raj Kishore, et al.
Publicado: (2025)
por: Perla, Sai Raj Kishore, et al.
Publicado: (2025)
Strategic Advances in the Total Syntheses of Opioids: Codeine, Morphine, and Related Alkaloids
por: Ayan Mondal, et al.
Publicado: (2025)
por: Ayan Mondal, et al.
Publicado: (2025)
Towards Generative Class Prompt Learning for Fine-grained Visual Recognition
por: Chattopadhyay, Soumitri, et al.
Publicado: (2024)
por: Chattopadhyay, Soumitri, et al.
Publicado: (2024)
In Quest of an Efficient Positive Electrode Material for Aqueous Al‐Ion Hybrid Capacitor: Investigation of a High Entropy Prussian Blue Analog
por: Arijit Dey, et al.
Publicado: (2025)
por: Arijit Dey, et al.
Publicado: (2025)
The OCR Quest for Generalization: Learning to recognize low-resource alphabets with model editing
por: Rodríguez, Adrià Molina, et al.
Publicado: (2025)
por: Rodríguez, Adrià Molina, et al.
Publicado: (2025)
Investigating the Impact of Copper and Zinc Doping in High‐Entropy Prussian Blue Analogues for Na‐Ion Batteries: From Material Analysis to Device Fabrication
por: Pappu Naskar, et al.
Publicado: (2024)
por: Pappu Naskar, et al.
Publicado: (2024)
Resistance hysteresis in twisted bilayer graphene: Intrinsic versus extrinsic effects
por: Dutta, Ranit, et al.
Publicado: (2025)
por: Dutta, Ranit, et al.
Publicado: (2025)
Counting Guidance for High Fidelity Text-to-Image Synthesis
por: Kang, Wonjun, et al.
Publicado: (2023)
por: Kang, Wonjun, et al.
Publicado: (2023)
Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement
por: Gao, Jiayi, et al.
Publicado: (2025)
por: Gao, Jiayi, et al.
Publicado: (2025)
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
por: Das, Alloy, et al.
Publicado: (2024)
por: Das, Alloy, et al.
Publicado: (2024)
GAN-based Content-Conditioned Generation of Handwritten Musical Symbols
por: Asbert, Gerard, et al.
Publicado: (2025)
por: Asbert, Gerard, et al.
Publicado: (2025)
AgentPose: Progressive Distribution Alignment via Feature Agent for Human Pose Distillation
por: Zhang, Feng, et al.
Publicado: (2025)
por: Zhang, Feng, et al.
Publicado: (2025)
A note on the Penrose process in rotating regular black holes
por: Kar, Anjan, et al.
Publicado: (2025)
por: Kar, Anjan, et al.
Publicado: (2025)
Surface Plasmon Mediated Giant Goos-Hanchen and Imbert-Fedorov Shifts on a Corrugated Metal Surface
por: Maiti, Arani, et al.
Publicado: (2025)
por: Maiti, Arani, et al.
Publicado: (2025)
Entropy as a Design Principle for Boosting Water Splitting: A Case Study on 1,3,5‐Benzenetricarboxylic Acid‐Based Metal–Organic‐Frameworks
por: Subarna Mandal, et al.
Publicado: (2025)
por: Subarna Mandal, et al.
Publicado: (2025)
GeoContrastNet: Contrastive Key-Value Edge Learning for Language-Agnostic Document Understanding
por: Biescas, Nil, et al.
Publicado: (2024)
por: Biescas, Nil, et al.
Publicado: (2024)
Diving into the Depths of Spotting Text in Multi-Domain Noisy Scenes
por: Das, Alloy, et al.
Publicado: (2023)
por: Das, Alloy, et al.
Publicado: (2023)
TIMI: Training-Free Image-to-3D Multi-Instance Generation with Spatial Fidelity
por: Cai, Xiao, et al.
Publicado: (2026)
por: Cai, Xiao, et al.
Publicado: (2026)
On relative Ulrich bundles and generalized Clifford algebras
por: Mondal, Soham, et al.
Publicado: (2026)
por: Mondal, Soham, et al.
Publicado: (2026)
Experience with Single Domain Generalization in Real World Medical Imaging Deployments
por: Banerjee, Ayan, et al.
Publicado: (2026)
por: Banerjee, Ayan, et al.
Publicado: (2026)
Identification of microstructure from macroscopic measurement using inverse multiscale analysis
por: Mukherjee, Anjan, et al.
Publicado: (2024)
por: Mukherjee, Anjan, et al.
Publicado: (2024)
Elastic-gap free strain gradient crystal plasticity model that effectively account for plastic slip gradient and grain boundary dissipation
por: Mukherjee, Anjan, et al.
Publicado: (2024)
por: Mukherjee, Anjan, et al.
Publicado: (2024)
Ejemplares similares
-
OmniCount: Multi-label Object Counting with Semantic-Geometric Priors
por: Mondal, Anindya, et al.
Publicado: (2024) -
Actor-agnostic Multi-label Action Recognition with Multi-modal Query
por: Mondal, Anindya, et al.
Publicado: (2023) -
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
por: Banerjee, Ayan, et al.
Publicado: (2025) -
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
por: Banerjee, Ayan, et al.
Publicado: (2024) -
CraftGraffiti: Exploring Human Identity with Custom Graffiti Art via Facial-Preserving Diffusion Models
por: Banerjee, Ayan, et al.
Publicado: (2025)