Layton: Latent Consistency Tokenizer for 1024-pixel Image Reconstruction and Generation by 256 Tokens
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xie, Qingsong, Zhang, Zhao, Huang, Zhe, Zhang, Yanhao, Lu, Haonan, Yang, Zhenyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MergeTok: Unified Continuous and Discrete Visual Tokenization via Token Merging
von: Zhang, Luyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Luyuan, et al.
Veröffentlicht: (2026)
PainterNet: Adaptive Image Inpainting with Actual-Token Attention and Diverse Mask Control
von: Wang, Ruichen, et al.
Veröffentlicht: (2024)
von: Wang, Ruichen, et al.
Veröffentlicht: (2024)
TLCM: Training-efficient Latent Consistency Model for Image Generation with 2-8 Steps
von: Xie, Qingsong, et al.
Veröffentlicht: (2024)
von: Xie, Qingsong, et al.
Veröffentlicht: (2024)
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM
von: Liu, Peng, et al.
Veröffentlicht: (2025)
von: Liu, Peng, et al.
Veröffentlicht: (2025)
Improved Visual-Spatial Reasoning via R1-Zero-Like Training
von: Liao, Zhenyi, et al.
Veröffentlicht: (2025)
von: Liao, Zhenyi, et al.
Veröffentlicht: (2025)
GloTok: Global Perspective Tokenizer for Image Reconstruction and Generation
von: Zhao, Xuan, et al.
Veröffentlicht: (2025)
von: Zhao, Xuan, et al.
Veröffentlicht: (2025)
Semantic One-Dimensional Tokenizer for Image Reconstruction and Generation
von: Qu, Yunpeng, et al.
Veröffentlicht: (2026)
von: Qu, Yunpeng, et al.
Veröffentlicht: (2026)
An Image is Worth 32 Tokens for Reconstruction and Generation
von: Yu, Qihang, et al.
Veröffentlicht: (2024)
von: Yu, Qihang, et al.
Veröffentlicht: (2024)
EvoTok: A Unified Image Tokenizer via Residual Latent Evolution for Visual Understanding and Generation
von: Li, Yan, et al.
Veröffentlicht: (2026)
von: Li, Yan, et al.
Veröffentlicht: (2026)
Switchable Token-Specific Codebook Quantization For Face Image Compression
von: Wang, Yongbo, et al.
Veröffentlicht: (2025)
von: Wang, Yongbo, et al.
Veröffentlicht: (2025)
Image Understanding Makes for A Good Tokenizer for Image Generation
von: Wang, Luting, et al.
Veröffentlicht: (2024)
von: Wang, Luting, et al.
Veröffentlicht: (2024)
PEA-Diffusion: Parameter-Efficient Adapter with Knowledge Distillation in non-English Text-to-Image Generation
von: Ma, Jian, et al.
Veröffentlicht: (2023)
von: Ma, Jian, et al.
Veröffentlicht: (2023)
MergeVQ: A Unified Framework for Visual Generation and Representation with Disentangled Token Merging and Quantization
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
ImageFolder: Autoregressive Image Generation with Folded Tokens
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Latent Denoising Makes Good Tokenizers
von: Yang, Jiawei, et al.
Veröffentlicht: (2025)
von: Yang, Jiawei, et al.
Veröffentlicht: (2025)
Guiding a Diffusion Model by Swapping Its Tokens
von: Zhang, Weijia, et al.
Veröffentlicht: (2026)
von: Zhang, Weijia, et al.
Veröffentlicht: (2026)
OptiPrune: Boosting Prompt-Image Consistency with Attention-Guided Noise and Dynamic Token Selection
von: Lu, Ziji
Veröffentlicht: (2025)
von: Lu, Ziji
Veröffentlicht: (2025)
Difference Inversion: Interpolate and Isolate the Difference with Token Consistency for Image Analogy Generation
von: Kim, Hyunsoo, et al.
Veröffentlicht: (2025)
von: Kim, Hyunsoo, et al.
Veröffentlicht: (2025)
CETCAM: Camera-Controllable Video Generation via Consistent and Extensible Tokenization
von: Zhao, Zelin, et al.
Veröffentlicht: (2025)
von: Zhao, Zelin, et al.
Veröffentlicht: (2025)
TokenSplat: Token-aligned 3D Gaussian Splatting for Feed-forward Pose-free Reconstruction
von: Li, Yihui, et al.
Veröffentlicht: (2026)
von: Li, Yihui, et al.
Veröffentlicht: (2026)
H2VU-Benchmark: A Comprehensive Benchmark for Hierarchical Holistic Video Understanding
von: Wu, Qi, et al.
Veröffentlicht: (2025)
von: Wu, Qi, et al.
Veröffentlicht: (2025)
OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation
von: Wang, Junke, et al.
Veröffentlicht: (2024)
von: Wang, Junke, et al.
Veröffentlicht: (2024)
ToDRE: Effective Visual Token Pruning via Token Diversity and Task Relevance
von: Li, Duo, et al.
Veröffentlicht: (2025)
von: Li, Duo, et al.
Veröffentlicht: (2025)
TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
von: Qu, Liao, et al.
Veröffentlicht: (2024)
von: Qu, Liao, et al.
Veröffentlicht: (2024)
ResTok: Learning Hierarchical Residuals in 1D Visual Tokenizers for Autoregressive Image Generation
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
TokensGen: Harnessing Condensed Tokens for Long Video Generation
von: Ouyang, Wenqi, et al.
Veröffentlicht: (2025)
von: Ouyang, Wenqi, et al.
Veröffentlicht: (2025)
TokenTrace: Multi-Concept Attribution through Watermarked Token Recovery
von: Zhang, Li, et al.
Veröffentlicht: (2026)
von: Zhang, Li, et al.
Veröffentlicht: (2026)
GPSToken: Gaussian Parameterized Spatially-adaptive Tokenization for Image Representation and Generation
von: Zhang, Zhengqiang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhengqiang, et al.
Veröffentlicht: (2025)
Hita: Holistic Tokenizer for Autoregressive Image Generation
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
von: Zheng, Anlin, et al.
Veröffentlicht: (2025)
Frequency Autoregressive Image Generation with Continuous Tokens
von: Yu, Hu, et al.
Veröffentlicht: (2025)
von: Yu, Hu, et al.
Veröffentlicht: (2025)
DaMo: Data Mixing Optimizer in Fine-tuning Multimodal LLMs for Mobile Phone Agents
von: Shi, Kai, et al.
Veröffentlicht: (2025)
von: Shi, Kai, et al.
Veröffentlicht: (2025)
LAPTOP-Diff: Layer Pruning and Normalized Distillation for Compressing Diffusion Models
von: Zhang, Dingkun, et al.
Veröffentlicht: (2024)
von: Zhang, Dingkun, et al.
Veröffentlicht: (2024)
Advancing Text-to-3D Generation with Linearized Lookahead Variational Score Distillation
von: Lei, Yu, et al.
Veröffentlicht: (2025)
von: Lei, Yu, et al.
Veröffentlicht: (2025)
XQ-GAN: An Open-source Image Tokenization Framework for Autoregressive Generation
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Vision Foundation Models as Generalist Tokenizers for Image Generation
von: Zheng, Anlin, et al.
Veröffentlicht: (2026)
von: Zheng, Anlin, et al.
Veröffentlicht: (2026)
Image Tokenizer Needs Post-Training
von: Qiu, Kai, et al.
Veröffentlicht: (2025)
von: Qiu, Kai, et al.
Veröffentlicht: (2025)
ActVAR: Activating Mixtures of Weights and Tokens for Efficient Visual Autoregressive Generation
von: Zhang, Kaixin, et al.
Veröffentlicht: (2025)
von: Zhang, Kaixin, et al.
Veröffentlicht: (2025)
PIXART-δ: Fast and Controllable Image Generation with Latent Consistency Models
von: Chen, Junsong, et al.
Veröffentlicht: (2024)
von: Chen, Junsong, et al.
Veröffentlicht: (2024)
Improved Masked Image Generation with Knowledge-Augmented Token Representations
von: Liang, Guotao, et al.
Veröffentlicht: (2025)
von: Liang, Guotao, et al.
Veröffentlicht: (2025)
AdaNAT: Exploring Adaptive Policy for Token-Based Image Generation
von: Ni, Zanlin, et al.
Veröffentlicht: (2024)
von: Ni, Zanlin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MergeTok: Unified Continuous and Discrete Visual Tokenization via Token Merging
von: Zhang, Luyuan, et al.
Veröffentlicht: (2026) -
PainterNet: Adaptive Image Inpainting with Actual-Token Attention and Diverse Mask Control
von: Wang, Ruichen, et al.
Veröffentlicht: (2024) -
TLCM: Training-efficient Latent Consistency Model for Image Generation with 2-8 Steps
von: Xie, Qingsong, et al.
Veröffentlicht: (2024) -
Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM
von: Liu, Peng, et al.
Veröffentlicht: (2025) -
Improved Visual-Spatial Reasoning via R1-Zero-Like Training
von: Liao, Zhenyi, et al.
Veröffentlicht: (2025)