$\infty$-Brush: Controllable Large Image Synthesis with Diffusion Models in Infinite Dimensions
Fuente:
arXiv
Guardado en:
| Autores principales: | Le, Minh-Quan, Graikos, Alexandros, Yellapragada, Srikar, Gupta, Rajarsi, Saltz, Joel, Samaras, Dimitris |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ZoomLDM: Latent Diffusion Model for multi-scale image generation
por: Yellapragada, Srikar, et al.
Publicado: (2024)
por: Yellapragada, Srikar, et al.
Publicado: (2024)
Learned representation-guided diffusion models for large-image generation
por: Graikos, Alexandros, et al.
Publicado: (2023)
por: Graikos, Alexandros, et al.
Publicado: (2023)
PathSegDiff: Pathology Segmentation using Diffusion model representations
por: Danisetty, Sachin Kumar, et al.
Publicado: (2025)
por: Danisetty, Sachin Kumar, et al.
Publicado: (2025)
GECKO: Gigapixel Vision-Concept Contrastive Pretraining in Histopathology
por: Kapse, Saarthak, et al.
Publicado: (2025)
por: Kapse, Saarthak, et al.
Publicado: (2025)
Pathology Image Compression with Pre-trained Autoencoders
por: Yellapragada, Srikar, et al.
Publicado: (2025)
por: Yellapragada, Srikar, et al.
Publicado: (2025)
Gen-SIS: Generative Self-augmentation Improves Self-supervised Learning
por: Belagali, Varun, et al.
Publicado: (2024)
por: Belagali, Varun, et al.
Publicado: (2024)
Fast constrained sampling in pre-trained diffusion models
por: Graikos, Alexandros, et al.
Publicado: (2024)
por: Graikos, Alexandros, et al.
Publicado: (2024)
PixCell: A generative foundation model for digital histopathology images
por: Yellapragada, Srikar, et al.
Publicado: (2025)
por: Yellapragada, Srikar, et al.
Publicado: (2025)
Diffusion-Refined VQA Annotations for Semi-Supervised Gaze Following
por: Miao, Qiaomu, et al.
Publicado: (2024)
por: Miao, Qiaomu, et al.
Publicado: (2024)
GriDiT: Factorized Grid-Based Diffusion for Efficient Long Image Sequence Generation
por: Tomar, Snehal Singh, et al.
Publicado: (2025)
por: Tomar, Snehal Singh, et al.
Publicado: (2025)
Mitigating Diffusion Model Hallucinations with Dynamic Guidance
por: Triaridis, Kostas, et al.
Publicado: (2025)
por: Triaridis, Kostas, et al.
Publicado: (2025)
CDG-MAE: Learning Correspondences from Diffusion Generated Views
por: Belagali, Varun, et al.
Publicado: (2025)
por: Belagali, Varun, et al.
Publicado: (2025)
Poppy: Polarization-based Plug-and-Play Guidance for Enhancing Monocular Normal Estimation
por: Kim, Irene, et al.
Publicado: (2026)
por: Kim, Irene, et al.
Publicado: (2026)
SI-MIL: Taming Deep MIL for Self-Interpretability in Gigapixel Histopathology
por: Kapse, Saarthak, et al.
Publicado: (2023)
por: Kapse, Saarthak, et al.
Publicado: (2023)
TopoDiffusionNet: A Topology-aware Diffusion Model
por: Gupta, Saumya, et al.
Publicado: (2024)
por: Gupta, Saumya, et al.
Publicado: (2024)
Generating metamers of human scene understanding
por: Raina, Ritik, et al.
Publicado: (2026)
por: Raina, Ritik, et al.
Publicado: (2026)
What about gravity in video generation? Post-Training Newton's Laws with Verifiable Rewards
por: Le, Minh-Quan, et al.
Publicado: (2025)
por: Le, Minh-Quan, et al.
Publicado: (2025)
Decoding the visual attention of pathologists to reveal their level of expertise
por: Chakraborty, Souradeep, et al.
Publicado: (2024)
por: Chakraborty, Souradeep, et al.
Publicado: (2024)
Leveraging Registers in Vision Transformers for Robust Adaptation
por: Yellapragada, Srikar, et al.
Publicado: (2025)
por: Yellapragada, Srikar, et al.
Publicado: (2025)
TICON: A Slide-Level Tile Contextualizer for Histopathology Representation Learning
por: Belagali, Varun, et al.
Publicado: (2025)
por: Belagali, Varun, et al.
Publicado: (2025)
Measuring and Predicting Where and When Pathologists Focus their Visual Attention while Grading Whole Slide Images of Cancer
por: Chakraborty, Souradeep, et al.
Publicado: (2025)
por: Chakraborty, Souradeep, et al.
Publicado: (2025)
Assessing Sample Quality via the Latent Space of Generative Models
por: Xu, Jingyi, et al.
Publicado: (2024)
por: Xu, Jingyi, et al.
Publicado: (2024)
PISCES: Annotation-free Text-to-Video Post-Training via Optimal Transport-Aligned Rewards
por: Le, Minh-Quan, et al.
Publicado: (2026)
por: Le, Minh-Quan, et al.
Publicado: (2026)
Importance-Based Token Merging for Efficient Image and Video Generation
por: Wu, Haoyu, et al.
Publicado: (2024)
por: Wu, Haoyu, et al.
Publicado: (2024)
Weighting Pseudo-Labels via High-Activation Feature Index Similarity and Object Detection for Semi-Supervised Segmentation
por: Howlader, Prantik, et al.
Publicado: (2024)
por: Howlader, Prantik, et al.
Publicado: (2024)
Hummingbird: High Fidelity Image Generation via Multimodal Context Alignment
por: Le, Minh-Quan, et al.
Publicado: (2025)
por: Le, Minh-Quan, et al.
Publicado: (2025)
Embedding Physical Reasoning into Diffusion-Based Shadow Generation
por: Hu, Shilin, et al.
Publicado: (2025)
por: Hu, Shilin, et al.
Publicado: (2025)
Self-supervised co-salient object detection via feature correspondence at multiple scales
por: Chakraborty, Souradeep, et al.
Publicado: (2024)
por: Chakraborty, Souradeep, et al.
Publicado: (2024)
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
por: Wu, Haoyu, et al.
Publicado: (2025)
por: Wu, Haoyu, et al.
Publicado: (2025)
Personalized Image Descriptions from Attention Sequences
por: Xue, Ruoyu, et al.
Publicado: (2025)
por: Xue, Ruoyu, et al.
Publicado: (2025)
Open and reusable deep learning for pathology with WSInfer and QuPath
por: Kaczmarzyk, Jakub R., et al.
Publicado: (2023)
por: Kaczmarzyk, Jakub R., et al.
Publicado: (2023)
Beyond Pixels: Semi-Supervised Semantic Segmentation with a Multi-scale Patch-based Multi-Label Classifier
por: Howlader, Prantik, et al.
Publicado: (2024)
por: Howlader, Prantik, et al.
Publicado: (2024)
Streamlining Image Editing with Layered Diffusion Brushes
por: Gholami, Peyman, et al.
Publicado: (2024)
por: Gholami, Peyman, et al.
Publicado: (2024)
Few-shot Personalized Scanpath Prediction
por: Xue, Ruoyu, et al.
Publicado: (2025)
por: Xue, Ruoyu, et al.
Publicado: (2025)
Talking Head Generation via AU-Guided Landmark Prediction
por: Chang, Shao-Yu, et al.
Publicado: (2025)
por: Chang, Shao-Yu, et al.
Publicado: (2025)
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model
por: Dao, Quan, et al.
Publicado: (2026)
por: Dao, Quan, et al.
Publicado: (2026)
Rig3DGS: Creating Controllable Portraits from Casual Monocular Videos
por: Rivero, Alfredo, et al.
Publicado: (2024)
por: Rivero, Alfredo, et al.
Publicado: (2024)
Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning
por: Hu, Shilin, et al.
Publicado: (2025)
por: Hu, Shilin, et al.
Publicado: (2025)
$\infty$-Diff: Infinite Resolution Diffusion with Subsampled Mollified States
por: Bond-Taylor, Sam, et al.
Publicado: (2023)
por: Bond-Taylor, Sam, et al.
Publicado: (2023)
MaskDiff: Modeling Mask Distribution with Diffusion Probabilistic Model for Few-Shot Instance Segmentation
por: Le, Minh-Quan, et al.
Publicado: (2023)
por: Le, Minh-Quan, et al.
Publicado: (2023)
Ejemplares similares
-
ZoomLDM: Latent Diffusion Model for multi-scale image generation
por: Yellapragada, Srikar, et al.
Publicado: (2024) -
Learned representation-guided diffusion models for large-image generation
por: Graikos, Alexandros, et al.
Publicado: (2023) -
PathSegDiff: Pathology Segmentation using Diffusion model representations
por: Danisetty, Sachin Kumar, et al.
Publicado: (2025) -
GECKO: Gigapixel Vision-Concept Contrastive Pretraining in Histopathology
por: Kapse, Saarthak, et al.
Publicado: (2025) -
Pathology Image Compression with Pre-trained Autoencoders
por: Yellapragada, Srikar, et al.
Publicado: (2025)