Token Painter: Training-Free Text-Guided Image Inpainting via Mask Autoregressive Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Longtao, Huang, Jie, Han, Mingfei, Chen, Lei, Yu, Yongqiang, Zhao, Feng, Chang, Xiaojun, Li, Zhihui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PainterNet: Adaptive Image Inpainting with Actual-Token Attention and Diverse Mask Control
von: Wang, Ruichen, et al.
Veröffentlicht: (2024)
von: Wang, Ruichen, et al.
Veröffentlicht: (2024)
HD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion Models
von: Manukyan, Hayk, et al.
Veröffentlicht: (2023)
von: Manukyan, Hayk, et al.
Veröffentlicht: (2023)
DreamPainter: Image Background Inpainting for E-commerce Scenarios
von: Zhao, Sijie, et al.
Veröffentlicht: (2025)
von: Zhao, Sijie, et al.
Veröffentlicht: (2025)
Mettle: Meta-Token Learning for Memory-Efficient Audio-Visual Adaptation
von: Zhou, Jinxing, et al.
Veröffentlicht: (2025)
von: Zhou, Jinxing, et al.
Veröffentlicht: (2025)
Zero-Painter: Training-Free Layout Control for Text-to-Image Synthesis
von: Ohanyan, Marianna, et al.
Veröffentlicht: (2024)
von: Ohanyan, Marianna, et al.
Veröffentlicht: (2024)
Frequency Autoregressive Image Generation with Continuous Tokens
von: Yu, Hu, et al.
Veröffentlicht: (2025)
von: Yu, Hu, et al.
Veröffentlicht: (2025)
Training-Free Text-Guided Image Editing with Visual Autoregressive Model
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
Beyond Dense Futures: World Models as Structured Planners for Robotic Manipulation
von: Jin, Minghao, et al.
Veröffentlicht: (2026)
von: Jin, Minghao, et al.
Veröffentlicht: (2026)
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation
von: Hao, Haihong, et al.
Veröffentlicht: (2025)
von: Hao, Haihong, et al.
Veröffentlicht: (2025)
Self-Consistency as a Free Lunch: Reducing Hallucinations in Vision-Language Models via Self-Reflection
von: Han, Mingfei, et al.
Veröffentlicht: (2025)
von: Han, Mingfei, et al.
Veröffentlicht: (2025)
FreeCond: Free Lunch in the Input Conditions of Text-Guided Inpainting
von: Hsiao, Teng-Fang, et al.
Veröffentlicht: (2024)
von: Hsiao, Teng-Fang, et al.
Veröffentlicht: (2024)
PathoPainter: Augmenting Histopathology Segmentation via Tumor-aware Inpainting
von: Liu, Hong, et al.
Veröffentlicht: (2025)
von: Liu, Hong, et al.
Veröffentlicht: (2025)
Text Image Inpainting via Global Structure-Guided Diffusion Models
von: Zhu, Shipeng, et al.
Veröffentlicht: (2024)
von: Zhu, Shipeng, et al.
Veröffentlicht: (2024)
Optimised ProPainter for Video Diminished Reality Inpainting
von: Li, Pengze, et al.
Veröffentlicht: (2024)
von: Li, Pengze, et al.
Veröffentlicht: (2024)
Diffree: Text-Guided Shape Free Object Inpainting with Diffusion Model
von: Zhao, Lirui, et al.
Veröffentlicht: (2024)
von: Zhao, Lirui, et al.
Veröffentlicht: (2024)
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking
von: Zheng, Zirui, et al.
Veröffentlicht: (2025)
von: Zheng, Zirui, et al.
Veröffentlicht: (2025)
Training-and-Prompt-Free General Painterly Harmonization via Zero-Shot Disentenglement on Style and Content References
von: Hsiao, Teng-Fang, et al.
Veröffentlicht: (2024)
von: Hsiao, Teng-Fang, et al.
Veröffentlicht: (2024)
LatentPilot: Scene-Aware Vision-and-Language Navigation by Dreaming Ahead with Latent Visual Reasoning
von: Hao, Haihong, et al.
Veröffentlicht: (2026)
von: Hao, Haihong, et al.
Veröffentlicht: (2026)
MTADiffusion: Mask Text Alignment Diffusion Model for Object Inpainting
von: Huang, Jun, et al.
Veröffentlicht: (2025)
von: Huang, Jun, et al.
Veröffentlicht: (2025)
Training-Free Watermarking for Autoregressive Image Generation
von: Tong, Yu, et al.
Veröffentlicht: (2025)
von: Tong, Yu, et al.
Veröffentlicht: (2025)
ALTo: Adaptive-Length Tokenizer for Autoregressive Mask Generation
von: Wang, Lingfeng, et al.
Veröffentlicht: (2025)
von: Wang, Lingfeng, et al.
Veröffentlicht: (2025)
InsightTok: Improving Text and Face Fidelity in Discrete Tokenization for Autoregressive Image Generation
von: Yue, Yang, et al.
Veröffentlicht: (2026)
von: Yue, Yang, et al.
Veröffentlicht: (2026)
RedVTP: Training-Free Acceleration of Diffusion Vision-Language Models Inference via Masked Token-Guided Visual Token Pruning
von: Xu, Jingqi, et al.
Veröffentlicht: (2025)
von: Xu, Jingqi, et al.
Veröffentlicht: (2025)
Enabling Autoregressive Models to Fill In Masked Tokens
von: Israel, Daniel, et al.
Veröffentlicht: (2025)
von: Israel, Daniel, et al.
Veröffentlicht: (2025)
FreeGraftor: Training-Free Cross-Image Feature Grafting for Subject-Driven Text-to-Image Generation
von: Yao, Zebin, et al.
Veröffentlicht: (2025)
von: Yao, Zebin, et al.
Veröffentlicht: (2025)
LlamaSeg: Image Segmentation via Autoregressive Mask Generation
von: Deng, Jiru, et al.
Veröffentlicht: (2025)
von: Deng, Jiru, et al.
Veröffentlicht: (2025)
LongVLM: Efficient Long Video Understanding via Large Language Models
von: Weng, Yuetian, et al.
Veröffentlicht: (2024)
von: Weng, Yuetian, et al.
Veröffentlicht: (2024)
VideoPainter: Any-length Video Inpainting and Editing with Plug-and-Play Context Control
von: Bian, Yuxuan, et al.
Veröffentlicht: (2025)
von: Bian, Yuxuan, et al.
Veröffentlicht: (2025)
From Inpainting to Editing: Unlocking Robust Mask-Free Visual Dubbing via Generative Bootstrapping
von: He, Xu, et al.
Veröffentlicht: (2025)
von: He, Xu, et al.
Veröffentlicht: (2025)
Revolutionizing Text-to-Image Retrieval as Autoregressive Token-to-Voken Generation
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
Token Merging for Training-Free Semantic Binding in Text-to-Image Synthesis
von: Hu, Taihang, et al.
Veröffentlicht: (2024)
von: Hu, Taihang, et al.
Veröffentlicht: (2024)
NEP: Autoregressive Image Editing via Next Editing Token Prediction
von: Wu, Huimin, et al.
Veröffentlicht: (2025)
von: Wu, Huimin, et al.
Veröffentlicht: (2025)
General Phrase Debiaser: Debiasing Masked Language Models at a Multi-Token Level
von: Shi, Bingkang, et al.
Veröffentlicht: (2023)
von: Shi, Bingkang, et al.
Veröffentlicht: (2023)
I Dream My Painting: Connecting MLLMs and Diffusion Models via Prompt Generation for Text-Guided Multi-Mask Inpainting
von: Fanelli, Nicola, et al.
Veröffentlicht: (2024)
von: Fanelli, Nicola, et al.
Veröffentlicht: (2024)
Entropy-Guided Token Dropout: Training Autoregressive Language Models with Limited Domain Data
von: Wang, Jiapeng, et al.
Veröffentlicht: (2025)
von: Wang, Jiapeng, et al.
Veröffentlicht: (2025)
User-Feedback-Driven Adaptation for Vision-and-Language Navigation
von: Yu, Yongqiang, et al.
Veröffentlicht: (2025)
von: Yu, Yongqiang, et al.
Veröffentlicht: (2025)
VCE: Safe Autoregressive Image Generation via Visual Contrast Exploitation
von: Han, Feng, et al.
Veröffentlicht: (2025)
von: Han, Feng, et al.
Veröffentlicht: (2025)
MARché: Fast Masked Autoregressive Image Generation with Cache-Aware Attention
von: Jiang, Chaoyi, et al.
Veröffentlicht: (2025)
von: Jiang, Chaoyi, et al.
Veröffentlicht: (2025)
Do Inpainting Yourself: Generative Facial Inpainting Guided by Exemplars
von: Lu, Wanglong, et al.
Veröffentlicht: (2022)
von: Lu, Wanglong, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
PainterNet: Adaptive Image Inpainting with Actual-Token Attention and Diverse Mask Control
von: Wang, Ruichen, et al.
Veröffentlicht: (2024) -
HD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion Models
von: Manukyan, Hayk, et al.
Veröffentlicht: (2023) -
DreamPainter: Image Background Inpainting for E-commerce Scenarios
von: Zhao, Sijie, et al.
Veröffentlicht: (2025) -
Mettle: Meta-Token Learning for Memory-Efficient Audio-Visual Adaptation
von: Zhou, Jinxing, et al.
Veröffentlicht: (2025) -
Zero-Painter: Training-Free Layout Control for Text-to-Image Synthesis
von: Ohanyan, Marianna, et al.
Veröffentlicht: (2024)