GriDiT: Factorized Grid-Based Diffusion for Efficient Long Image Sequence Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tomar, Snehal Singh, Graikos, Alexandros, Krishna, Arjun, Samaras, Dimitris, Mueller, Klaus |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fast constrained sampling in pre-trained diffusion models
von: Graikos, Alexandros, et al.
Veröffentlicht: (2024)
von: Graikos, Alexandros, et al.
Veröffentlicht: (2024)
PathSegDiff: Pathology Segmentation using Diffusion model representations
von: Danisetty, Sachin Kumar, et al.
Veröffentlicht: (2025)
von: Danisetty, Sachin Kumar, et al.
Veröffentlicht: (2025)
$\infty$-Brush: Controllable Large Image Synthesis with Diffusion Models in Infinite Dimensions
von: Le, Minh-Quan, et al.
Veröffentlicht: (2024)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2024)
Mitigating Diffusion Model Hallucinations with Dynamic Guidance
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
Diffusion-Refined VQA Annotations for Semi-Supervised Gaze Following
von: Miao, Qiaomu, et al.
Veröffentlicht: (2024)
von: Miao, Qiaomu, et al.
Veröffentlicht: (2024)
Generating metamers of human scene understanding
von: Raina, Ritik, et al.
Veröffentlicht: (2026)
von: Raina, Ritik, et al.
Veröffentlicht: (2026)
Poppy: Polarization-based Plug-and-Play Guidance for Enhancing Monocular Normal Estimation
von: Kim, Irene, et al.
Veröffentlicht: (2026)
von: Kim, Irene, et al.
Veröffentlicht: (2026)
ZoomLDM: Latent Diffusion Model for multi-scale image generation
von: Yellapragada, Srikar, et al.
Veröffentlicht: (2024)
von: Yellapragada, Srikar, et al.
Veröffentlicht: (2024)
Multi-Conditioned Denoising Diffusion Probabilistic Model (mDDPM) for Medical Image Synthesis
von: Krishna, Arjun, et al.
Veröffentlicht: (2024)
von: Krishna, Arjun, et al.
Veröffentlicht: (2024)
Importance-Based Token Merging for Efficient Image and Video Generation
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
Learned representation-guided diffusion models for large-image generation
von: Graikos, Alexandros, et al.
Veröffentlicht: (2023)
von: Graikos, Alexandros, et al.
Veröffentlicht: (2023)
Pathology Image Compression with Pre-trained Autoencoders
von: Yellapragada, Srikar, et al.
Veröffentlicht: (2025)
von: Yellapragada, Srikar, et al.
Veröffentlicht: (2025)
Embedding Physical Reasoning into Diffusion-Based Shadow Generation
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
Gen-SIS: Generative Self-augmentation Improves Self-supervised Learning
von: Belagali, Varun, et al.
Veröffentlicht: (2024)
von: Belagali, Varun, et al.
Veröffentlicht: (2024)
TopoDiffusionNet: A Topology-aware Diffusion Model
von: Gupta, Saumya, et al.
Veröffentlicht: (2024)
von: Gupta, Saumya, et al.
Veröffentlicht: (2024)
Self-supervised co-salient object detection via feature correspondence at multiple scales
von: Chakraborty, Souradeep, et al.
Veröffentlicht: (2024)
von: Chakraborty, Souradeep, et al.
Veröffentlicht: (2024)
Assessing Sample Quality via the Latent Space of Generative Models
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
Personalized Image Descriptions from Attention Sequences
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
JEAN: Joint Expression and Audio-guided NeRF-based Talking Face Generation
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
Weighting Pseudo-Labels via High-Activation Feature Index Similarity and Object Detection for Semi-Supervised Segmentation
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
Talking Head Generation via AU-Guided Landmark Prediction
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
DiMSUM: Diffusion Mamba -- A Scalable and Unified Spatial-Frequency Method for Image Generation
von: Phung, Hao, et al.
Veröffentlicht: (2024)
von: Phung, Hao, et al.
Veröffentlicht: (2024)
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
DiSA: Diffusion Step Annealing in Autoregressive Image Generation
von: Zhao, Qinyu, et al.
Veröffentlicht: (2025)
von: Zhao, Qinyu, et al.
Veröffentlicht: (2025)
MI-NeRF: Learning a Single Face NeRF from Multiple Identities
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
MIGS: Multi-Identity Gaussian Splatting via Tensor Decomposition
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
MaDiS: Taming Masked Diffusion Language Models for Sign Language Generation
von: Zuo, Ronglai, et al.
Veröffentlicht: (2026)
von: Zuo, Ronglai, et al.
Veröffentlicht: (2026)
ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
PNeRV: A Polynomial Neural Representation for Videos
von: Gupta, Sonam, et al.
Veröffentlicht: (2024)
von: Gupta, Sonam, et al.
Veröffentlicht: (2024)
Direct and Explicit 3D Generation from a Single Image
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
PixelDiT: Pixel Diffusion Transformers for Image Generation
von: Yu, Yongsheng, et al.
Veröffentlicht: (2025)
von: Yu, Yongsheng, et al.
Veröffentlicht: (2025)
2DMamba: Efficient State Space Model for Image Representation with Applications on Giga-Pixel Whole Slide Image Classification
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
EdgeDiT: Hardware-Aware Diffusion Transformers for Efficient On-Device Image Generation
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
Learning 3D Reconstruction with Priors in Test Time
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
Beyond Pixels: Semi-Supervised Semantic Segmentation with a Multi-scale Patch-based Multi-Label Classifier
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
DiM: Diffusion Mamba for Efficient High-Resolution Image Synthesis
von: Teng, Yao, et al.
Veröffentlicht: (2024)
von: Teng, Yao, et al.
Veröffentlicht: (2024)
Modeling Deep Learning Based Privacy Attacks on Physical Mail
von: Huang, Bingyao, et al.
Veröffentlicht: (2020)
von: Huang, Bingyao, et al.
Veröffentlicht: (2020)
Phrase-Instance Alignment for Generalized Referring Segmentation
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2024)
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2024)
DiVE: Efficient Multi-View Driving Scenes Generation Based on Video Diffusion Transformer
von: Jiang, Junpeng, et al.
Veröffentlicht: (2025)
von: Jiang, Junpeng, et al.
Veröffentlicht: (2025)
Hummingbird: High Fidelity Image Generation via Multimodal Context Alignment
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Fast constrained sampling in pre-trained diffusion models
von: Graikos, Alexandros, et al.
Veröffentlicht: (2024) -
PathSegDiff: Pathology Segmentation using Diffusion model representations
von: Danisetty, Sachin Kumar, et al.
Veröffentlicht: (2025) -
$\infty$-Brush: Controllable Large Image Synthesis with Diffusion Models in Infinite Dimensions
von: Le, Minh-Quan, et al.
Veröffentlicht: (2024) -
Mitigating Diffusion Model Hallucinations with Dynamic Guidance
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025) -
Diffusion-Refined VQA Annotations for Semi-Supervised Gaze Following
von: Miao, Qiaomu, et al.
Veröffentlicht: (2024)