LGQ: Learning Discretization Geometry for Scalable and Stable Image Tokenization
Fuente:
arXiv
Saved in:
| Main Authors: | Altun, Idil Bilge, Cakiroglu, Mert Onur, Buxton, Elham, Dalkilic, Mehmet, Kurban, Hasan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multivariate de Bruijn Graphs: A Symbolic Graph Framework for Time Series Forecasting
by: Cakiroglu, Mert Onur, et al.
Published: (2025)
by: Cakiroglu, Mert Onur, et al.
Published: (2025)
Temporal Realism Evaluation of Generated Videos Using Compressed-Domain Motion Vectors
by: Cakiroglu, Mert Onur, et al.
Published: (2025)
by: Cakiroglu, Mert Onur, et al.
Published: (2025)
Stress-Testing Multimodal Foundation Models for Crystallographic Reasoning
by: Polat, Can, et al.
Published: (2025)
by: Polat, Can, et al.
Published: (2025)
IRIS: A Real-World Benchmark for Inverse Recovery and Identification of Physical Dynamic Systems from Monocular Video
by: Khanbayov, Rasul, et al.
Published: (2026)
by: Khanbayov, Rasul, et al.
Published: (2026)
On the Role of Discrete Tokenization in Visual Representation Learning
by: Du, Tianqi, et al.
Published: (2024)
by: Du, Tianqi, et al.
Published: (2024)
AdaCL:Adaptive Continual Learning
by: Yildirim, Elif Ceren Gok, et al.
Published: (2023)
by: Yildirim, Elif Ceren Gok, et al.
Published: (2023)
Diffusion Autoencoders are Scalable Image Tokenizers
by: Chen, Yinbo, et al.
Published: (2025)
by: Chen, Yinbo, et al.
Published: (2025)
SFTok: Bridging the Performance Gap in Discrete Tokenizers
by: Rao, Qihang, et al.
Published: (2025)
by: Rao, Qihang, et al.
Published: (2025)
Beyond Single Tokens: Distilling Discrete Diffusion Models via Discrete MMD
by: Hoogeboom, Emiel, et al.
Published: (2026)
by: Hoogeboom, Emiel, et al.
Published: (2026)
Semantic Prompting with Image-Token for Continual Learning
by: Han, Jisu, et al.
Published: (2024)
by: Han, Jisu, et al.
Published: (2024)
A Personalized Zero-Shot ECG Arrhythmia Monitoring System: From Sparse Representation Based Domain Adaption to Energy Efficient Abnormal Beat Detection for Practical ECG Surveillance
by: Yamaç, Mehmet, et al.
Published: (2022)
by: Yamaç, Mehmet, et al.
Published: (2022)
FuseLIP: Multimodal Embeddings via Early Fusion of Discrete Tokens
by: Schlarmann, Christian, et al.
Published: (2025)
by: Schlarmann, Christian, et al.
Published: (2025)
Spectral Image Tokenizer
by: Esteves, Carlos, et al.
Published: (2024)
by: Esteves, Carlos, et al.
Published: (2024)
Investigating Permutation-Invariant Discrete Representation Learning for Spatially Aligned Images
by: Stirling, Jamie S. J., et al.
Published: (2026)
by: Stirling, Jamie S. J., et al.
Published: (2026)
Geometric-k-means: A Bound Free Approach to Fast and Eco-Friendly k-means
by: Sharma, Parichit, et al.
Published: (2025)
by: Sharma, Parichit, et al.
Published: (2025)
Geometry Fidelity for Spherical Images
by: Christensen, Anders, et al.
Published: (2024)
by: Christensen, Anders, et al.
Published: (2024)
Latent Geometry of Taste: Scalable Low-Rank Matrix Factorization for Recommender Systems
by: Salako, Joshua
Published: (2026)
by: Salako, Joshua
Published: (2026)
SLNet: A Super-Lightweight Geometry-Adaptive Network for 3D Point Cloud Recognition
by: Saeid, Mohammad, et al.
Published: (2026)
by: Saeid, Mohammad, et al.
Published: (2026)
Student Capacity Moderates Knowledge Distillation Effectiveness: A Systematic Study Across ResNet Teacher-Student Pairs on CIFAR-10
by: Yasar, Umut Onur
Published: (2026)
by: Yasar, Umut Onur
Published: (2026)
VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations
by: Patel, Maitreya, et al.
Published: (2026)
by: Patel, Maitreya, et al.
Published: (2026)
Token Pruning for Caching Better: 9 Times Acceleration on Stable Diffusion for Free
by: Zhang, Evelyn, et al.
Published: (2024)
by: Zhang, Evelyn, et al.
Published: (2024)
A Comparison of Object Detection and Phrase Grounding Models in Chest X-ray Abnormality Localization using Eye-tracking Data
by: Ghelichkhan, Elham, et al.
Published: (2025)
by: Ghelichkhan, Elham, et al.
Published: (2025)
Controllable Image Generation with Composed Parallel Token Prediction
by: Stirling, Jamie, et al.
Published: (2024)
by: Stirling, Jamie, et al.
Published: (2024)
Variational Self-Supervised Learning
by: Yavuz, Mehmet Can, et al.
Published: (2025)
by: Yavuz, Mehmet Can, et al.
Published: (2025)
Continual Learning on a Data Diet
by: Yildirim, Elif Ceren Gok, et al.
Published: (2024)
by: Yildirim, Elif Ceren Gok, et al.
Published: (2024)
Discretizing Group-Convolutional Neural Networks for 3D Geometry in Feature Space
by: Franzen, Daniel, et al.
Published: (2026)
by: Franzen, Daniel, et al.
Published: (2026)
EasyARC: Evaluating Vision Language Models on True Visual Reasoning
by: Unsal, Mert, et al.
Published: (2025)
by: Unsal, Mert, et al.
Published: (2025)
The Compression Gap: Why Discrete Tokenization Limits Vision-Language-Action Model Scaling
by: Shiba, Takuya
Published: (2026)
by: Shiba, Takuya
Published: (2026)
Variational Self-Supervised Contrastive Learning Using Beta Divergence
by: Yavuz, Mehmet Can, et al.
Published: (2023)
by: Yavuz, Mehmet Can, et al.
Published: (2023)
UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models
by: Wang, Jiaqi, et al.
Published: (2026)
by: Wang, Jiaqi, et al.
Published: (2026)
Multimodal Deep Learning for Diabetic Foot Ulcer Staging Using Integrated RGB and Thermal Imaging
by: Mermer, Gulengul, et al.
Published: (2026)
by: Mermer, Gulengul, et al.
Published: (2026)
Scalable Evaluation of the Realism of Synthetic Environmental Augmentations in Images
by: Ruck, Damian J., et al.
Published: (2026)
by: Ruck, Damian J., et al.
Published: (2026)
Grouped Discrete Representation for Object-Centric Learning
by: Zhao, Rongzhen, et al.
Published: (2024)
by: Zhao, Rongzhen, et al.
Published: (2024)
MaskBit: Embedding-free Image Generation via Bit Tokens
by: Weber, Mark, et al.
Published: (2024)
by: Weber, Mark, et al.
Published: (2024)
End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer
by: Chu, Wenda, et al.
Published: (2026)
by: Chu, Wenda, et al.
Published: (2026)
One-D-Piece: Image Tokenizer Meets Quality-Controllable Compression
by: Miwa, Keita, et al.
Published: (2025)
by: Miwa, Keita, et al.
Published: (2025)
Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation
by: Krishnan, Akshay, et al.
Published: (2025)
by: Krishnan, Akshay, et al.
Published: (2025)
Global Context with Discrete Diffusion in Vector Quantised Modelling for Image Generation
by: Hu, Minghui, et al.
Published: (2021)
by: Hu, Minghui, et al.
Published: (2021)
LoFi: Neural Local Fields for Scalable Image Reconstruction
by: Khorashadizadeh, AmirEhsan, et al.
Published: (2024)
by: Khorashadizadeh, AmirEhsan, et al.
Published: (2024)
DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation
by: Gu, Jiatao, et al.
Published: (2024)
by: Gu, Jiatao, et al.
Published: (2024)
Similar Items
-
Multivariate de Bruijn Graphs: A Symbolic Graph Framework for Time Series Forecasting
by: Cakiroglu, Mert Onur, et al.
Published: (2025) -
Temporal Realism Evaluation of Generated Videos Using Compressed-Domain Motion Vectors
by: Cakiroglu, Mert Onur, et al.
Published: (2025) -
Stress-Testing Multimodal Foundation Models for Crystallographic Reasoning
by: Polat, Can, et al.
Published: (2025) -
IRIS: A Real-World Benchmark for Inverse Recovery and Identification of Physical Dynamic Systems from Monocular Video
by: Khanbayov, Rasul, et al.
Published: (2026) -
On the Role of Discrete Tokenization in Visual Representation Learning
by: Du, Tianqi, et al.
Published: (2024)