Token Reduction Should Go Beyond Efficiency in Generative Models -- From Vision, Language to Multimodality
Fuente:
arXiv
Saved in:
| Main Authors: | Kong, Zhenglun, Li, Yize, Zeng, Fanhu, Xin, Lei, Messica, Shvat, Lin, Xue, Zhao, Pu, Kellis, Manolis, Tang, Hao, Zitnik, Marinka |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multimodal Medical Code Tokenizer
by: Su, Xiaorui, et al.
Published: (2025)
by: Su, Xiaorui, et al.
Published: (2025)
Token Transforming: A Unified and Training-Free Token Compression Framework for Vision Transformer Acceleration
by: Zeng, Fanhu, et al.
Published: (2025)
by: Zeng, Fanhu, et al.
Published: (2025)
Adaptive Time Series Reasoning via Segment Selection
by: Messica, Shvat, et al.
Published: (2026)
by: Messica, Shvat, et al.
Published: (2026)
SPATIA: Multimodal Generation and Prediction of Spatial Cell Phenotypes
by: Kong, Zhenglun, et al.
Published: (2025)
by: Kong, Zhenglun, et al.
Published: (2025)
Greater than the Sum of Its Parts: Building Substructure into Protein Encoding Models
by: Calef, Robert, et al.
Published: (2025)
by: Calef, Robert, et al.
Published: (2025)
Controllable Sequence Editing for Biological and Clinical Trajectories
by: Li, Michelle M., et al.
Published: (2025)
by: Li, Michelle M., et al.
Published: (2025)
Rethinking Token Reduction for State Space Models
by: Zhan, Zheng, et al.
Published: (2024)
by: Zhan, Zheng, et al.
Published: (2024)
ProteinRPN: Towards Accurate Protein Function Prediction with Graph-Based Region Proposals
by: Mitra, Shania, et al.
Published: (2024)
by: Mitra, Shania, et al.
Published: (2024)
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies
by: Li, Kevin, et al.
Published: (2025)
by: Li, Kevin, et al.
Published: (2025)
Ken Utilization Layer: Hebbian Replay Within a Student's Ken for Adaptive Exercise Recommendation
by: Kuling, Grey, et al.
Published: (2025)
by: Kuling, Grey, et al.
Published: (2025)
TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools
by: Gao, Shanghua, et al.
Published: (2025)
by: Gao, Shanghua, et al.
Published: (2025)
PPT: Token Pruning and Pooling for Efficient Vision Transformers
by: Wu, Xinjian, et al.
Published: (2023)
by: Wu, Xinjian, et al.
Published: (2023)
Protein Structure Tokenization via Geometric Byte Pair Encoding
by: Sun, Michael, et al.
Published: (2025)
by: Sun, Michael, et al.
Published: (2025)
NAST: Noise Aware Speech Tokenization for Speech Language Models
by: Messica, Shoval, et al.
Published: (2024)
by: Messica, Shoval, et al.
Published: (2024)
AutoScientists: Self-Organizing Agent Teams for Long-Running Scientific Experimentation
by: Gao, Shanghua, et al.
Published: (2026)
by: Gao, Shanghua, et al.
Published: (2026)
Generalized Protein Pocket Generation with Prior-Informed Flow Matching
by: Zhang, Zaixi, et al.
Published: (2024)
by: Zhang, Zaixi, et al.
Published: (2024)
PyTDC: A multimodal machine learning training, evaluation, and inference platform for biomedical foundation models
by: Velez-Arce, Alejandro, et al.
Published: (2025)
by: Velez-Arce, Alejandro, et al.
Published: (2025)
Graph Representation Learning in Biomedicine
by: Li, Michelle M., et al.
Published: (2021)
by: Li, Michelle M., et al.
Published: (2021)
Exploring Token Pruning in Vision State Space Models
by: Zhan, Zheng, et al.
Published: (2024)
by: Zhan, Zheng, et al.
Published: (2024)
Democratizing AI scientists using ToolUniverse
by: Gao, Shanghua, et al.
Published: (2025)
by: Gao, Shanghua, et al.
Published: (2025)
Learning Generalized Medical Image Representations through Image-Graph Contrastive Pretraining
by: Khanna, Sameer, et al.
Published: (2024)
by: Khanna, Sameer, et al.
Published: (2024)
A Causality-aware Paradigm for Evaluating Creativity of Multimodal Large Language Models
by: Huang, Zhongzhan, et al.
Published: (2025)
by: Huang, Zhongzhan, et al.
Published: (2025)
A versatile informative diffusion model for single-cell ATAC-seq data generation and analysis
by: Huang, Lei, et al.
Published: (2024)
by: Huang, Lei, et al.
Published: (2024)
Learning Compact Vision Tokens for Efficient Large Multimodal Models
by: Tang, Hao, et al.
Published: (2025)
by: Tang, Hao, et al.
Published: (2025)
When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making
by: Liu, Jun, et al.
Published: (2026)
by: Liu, Jun, et al.
Published: (2026)
PhylaFlow: Hybrid Flow Matching in Billera-Holmes-Vogtmann Tree Space for Phylogenetic Inference
by: Ektefaie, Yasha, et al.
Published: (2026)
by: Ektefaie, Yasha, et al.
Published: (2026)
Repurposing Foundation Model for Generalizable Medical Time Series Classification
by: Huang, Nan, et al.
Published: (2024)
by: Huang, Nan, et al.
Published: (2024)
Qworld: Question-Specific Evaluation Criteria for LLMs
by: Gao, Shanghua, et al.
Published: (2026)
by: Gao, Shanghua, et al.
Published: (2026)
Wearable Foundation Models Should Go Beyond Static Encoders
by: Wu, Yu Yvonne, et al.
Published: (2026)
by: Wu, Yu Yvonne, et al.
Published: (2026)
RobustMerge: Parameter-Efficient Model Merging for MLLMs with Direction Robustness
by: Zeng, Fanhu, et al.
Published: (2025)
by: Zeng, Fanhu, et al.
Published: (2025)
Pruning Foundation Models for High Accuracy without Retraining
by: Zhao, Pu, et al.
Published: (2024)
by: Zhao, Pu, et al.
Published: (2024)
Topic Modelling: Going Beyond Token Outputs
by: Williams, Lowri, et al.
Published: (2024)
by: Williams, Lowri, et al.
Published: (2024)
Towards Efficient and General-Purpose Few-Shot Misclassification Detection for Vision-Language Models
by: Zeng, Fanhu, et al.
Published: (2025)
by: Zeng, Fanhu, et al.
Published: (2025)
Evaluating Relational Reasoning in LLMs with REL
by: Fesser, Lukas, et al.
Published: (2026)
by: Fesser, Lukas, et al.
Published: (2026)
Agent Skills Should Go Beyond Text: The Case for Visual Skills
by: Xu, Binxiao, et al.
Published: (2026)
by: Xu, Binxiao, et al.
Published: (2026)
LIVRO DIDÁTICO DE LÍNGUA INGLESA E O QUE OS DISCURSOS ESCRITOS REVELAM SOBRE IDENTIDADE RACIAL
by: Kellis Coelho Farias
Published: (2014)
by: Kellis Coelho Farias
Published: (2014)
TSLA: A Task-Specific Learning Adaptation for Semantic Segmentation on Autonomous Vehicles Platform
by: Liu, Jun, et al.
Published: (2025)
by: Liu, Jun, et al.
Published: (2025)
Invariant Tokenization of Crystalline Materials for Language Model Enabled Generation
by: Yan, Keqiang, et al.
Published: (2025)
by: Yan, Keqiang, et al.
Published: (2025)
Quasar-ViT: Hardware-Oriented Quantization-Aware Architecture Search for Vision Transformers
by: Li, Zhengang, et al.
Published: (2024)
by: Li, Zhengang, et al.
Published: (2024)
GNN101: Visual Learning of Graph Neural Networks in Your Web Browser
by: Lu, Yilin, et al.
Published: (2024)
by: Lu, Yilin, et al.
Published: (2024)
Similar Items
-
Multimodal Medical Code Tokenizer
by: Su, Xiaorui, et al.
Published: (2025) -
Token Transforming: A Unified and Training-Free Token Compression Framework for Vision Transformer Acceleration
by: Zeng, Fanhu, et al.
Published: (2025) -
Adaptive Time Series Reasoning via Segment Selection
by: Messica, Shvat, et al.
Published: (2026) -
SPATIA: Multimodal Generation and Prediction of Spatial Cell Phenotypes
by: Kong, Zhenglun, et al.
Published: (2025) -
Greater than the Sum of Its Parts: Building Substructure into Protein Encoding Models
by: Calef, Robert, et al.
Published: (2025)