OmniDrop: Layer-wise Token Pruning for Omni-modal LLMs via Query-Guidance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, Yeo Jeong, Jang, Hyemi, Choi, Minseo, Lee, Jongsun, Choi, Jooyoung, Jeon, Yongkweon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Two-Stage Grid Optimization for Group-wise Quantization of LLMs
von: Kim, Junhan, et al.
Veröffentlicht: (2026)
von: Kim, Junhan, et al.
Veröffentlicht: (2026)
LookaheadKV: Fast and Accurate KV Cache Eviction by Glimpsing into the Future without Generation
von: Ahn, Jinwoo, et al.
Veröffentlicht: (2026)
von: Ahn, Jinwoo, et al.
Veröffentlicht: (2026)
Keep What Audio Cannot Say: Context-Preserving Token Pruning for Omni-LLMs
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2026)
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2026)
On the Importance of a Multi-Scale Calibration for Quantization
von: Son, Seungwoo, et al.
Veröffentlicht: (2026)
von: Son, Seungwoo, et al.
Veröffentlicht: (2026)
OmniRetrieval: Unified Retrieval across Heterogeneous Knowledge Sources
von: Baek, Jinheon, et al.
Veröffentlicht: (2026)
von: Baek, Jinheon, et al.
Veröffentlicht: (2026)
OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation
von: Zhang, Guohui, et al.
Veröffentlicht: (2026)
von: Zhang, Guohui, et al.
Veröffentlicht: (2026)
ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding
von: Guan, Yiran, et al.
Veröffentlicht: (2026)
von: Guan, Yiran, et al.
Veröffentlicht: (2026)
OmniEval: A Benchmark for Evaluating Omni-modal Models with Visual, Auditory, and Textual Inputs
von: Zhang, Yiman, et al.
Veröffentlicht: (2025)
von: Zhang, Yiman, et al.
Veröffentlicht: (2025)
UNO-Bench: A Unified Benchmark for Exploring the Compositional Law Between Uni-modal and Omni-modal in Omni Models
von: Chen, Chen, et al.
Veröffentlicht: (2025)
von: Chen, Chen, et al.
Veröffentlicht: (2025)
ViT-Lens: Towards Omni-modal Representations
von: Lei, Weixian, et al.
Veröffentlicht: (2023)
von: Lei, Weixian, et al.
Veröffentlicht: (2023)
OmniPlay: Benchmarking Omni-Modal Models on Omni-Modal Game Playing
von: Bie, Fuqing, et al.
Veröffentlicht: (2025)
von: Bie, Fuqing, et al.
Veröffentlicht: (2025)
MGM-Omni: Scaling Omni LLMs to Personalized Long-Horizon Speech
von: Wang, Chengyao, et al.
Veröffentlicht: (2025)
von: Wang, Chengyao, et al.
Veröffentlicht: (2025)
Explore the Limits of Omni-modal Pretraining at Scale
von: Zhang, Yiyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yiyuan, et al.
Veröffentlicht: (2024)
OmniTrace: A Unified Framework for Generation-Time Attribution in Omni-Modal LLMs
von: Yan, Qianqi, et al.
Veröffentlicht: (2026)
von: Yan, Qianqi, et al.
Veröffentlicht: (2026)
Ex-Omni: Enabling 3D Facial Animation Generation for Omni-modal Large Language Models
von: Zhang, Haoyu, et al.
Veröffentlicht: (2026)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2026)
Context-Aware Wireless Token Communication via Joint Token Masking and Detection
von: Shin, Junyong, et al.
Veröffentlicht: (2026)
von: Shin, Junyong, et al.
Veröffentlicht: (2026)
OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens
von: Willeke, Konstantin F., et al.
Veröffentlicht: (2026)
von: Willeke, Konstantin F., et al.
Veröffentlicht: (2026)
OmniSIFT: Modality-Asymmetric Token Compression for Efficient Omni-modal Large Language Models
von: Ding, Yue, et al.
Veröffentlicht: (2026)
von: Ding, Yue, et al.
Veröffentlicht: (2026)
e5-omni: Explicit Cross-modal Alignment for Omni-modal Embeddings
von: Chen, Haonan, et al.
Veröffentlicht: (2026)
von: Chen, Haonan, et al.
Veröffentlicht: (2026)
Context-Aware Iterative Token Detection and Masked Transmission for Wireless Token Communication
von: Shin, Junyong, et al.
Veröffentlicht: (2026)
von: Shin, Junyong, et al.
Veröffentlicht: (2026)
SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models
von: Xie, Tianyu, et al.
Veröffentlicht: (2026)
von: Xie, Tianyu, et al.
Veröffentlicht: (2026)
Personalized Federated Learning via Sequential Layer Expansion in Representation Learning
von: Jang, Jaewon, et al.
Veröffentlicht: (2024)
von: Jang, Jaewon, et al.
Veröffentlicht: (2024)
Multimodal Prompt Injection Attacks: Risks and Defenses for Modern LLMs
von: Yeo, Andrew, et al.
Veröffentlicht: (2025)
von: Yeo, Andrew, et al.
Veröffentlicht: (2025)
CLIP Tricks You: Training-free Token Pruning for Efficient Pixel Grounding in Large VIsion-Language Models
von: Lee, Sangin, et al.
Veröffentlicht: (2026)
von: Lee, Sangin, et al.
Veröffentlicht: (2026)
Not Like Transformers: Drop the Beat Representation for Dance Generation with Mamba-Based Diffusion Model
von: Park, Sangjune, et al.
Veröffentlicht: (2026)
von: Park, Sangjune, et al.
Veröffentlicht: (2026)
DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs
von: Choi, Nayoung, et al.
Veröffentlicht: (2026)
von: Choi, Nayoung, et al.
Veröffentlicht: (2026)
OmniSelect: Dynamic Modality-Aware Token Compression for Efficient Omni-modal Large Language Models
von: Yang, Morunliu, et al.
Veröffentlicht: (2026)
von: Yang, Morunliu, et al.
Veröffentlicht: (2026)
OmniVideoBench: Towards Audio-Visual Understanding Evaluation for Omni MLLMs
von: Li, Caorui, et al.
Veröffentlicht: (2025)
von: Li, Caorui, et al.
Veröffentlicht: (2025)
OmniDiT: Extending Diffusion Transformer to Omni-VTON Framework
von: Zeng, Weixuan, et al.
Veröffentlicht: (2026)
von: Zeng, Weixuan, et al.
Veröffentlicht: (2026)
LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training
von: Gwak, Minju, et al.
Veröffentlicht: (2026)
von: Gwak, Minju, et al.
Veröffentlicht: (2026)
CogGuide: Human-Like Guidance for Zero-Shot Omni-Modal Reasoning
von: Shou, Zhou-Peng, et al.
Veröffentlicht: (2025)
von: Shou, Zhou-Peng, et al.
Veröffentlicht: (2025)
Can TabPFN Compete with GNNs for Node Classification via Graph Tabularization?
von: Choi, Jeongwhan, et al.
Veröffentlicht: (2025)
von: Choi, Jeongwhan, et al.
Veröffentlicht: (2025)
OmniDPO: A Preference Optimization Framework to Address Omni-Modal Hallucination
von: Chen, Junzhe, et al.
Veröffentlicht: (2025)
von: Chen, Junzhe, et al.
Veröffentlicht: (2025)
OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments
von: Henry, Felix, et al.
Veröffentlicht: (2026)
von: Henry, Felix, et al.
Veröffentlicht: (2026)
Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models
von: Yan, Xinru, et al.
Veröffentlicht: (2026)
von: Yan, Xinru, et al.
Veröffentlicht: (2026)
OmniBench: Towards The Future of Universal Omni-Language Models
von: Li, Yizhi, et al.
Veröffentlicht: (2024)
von: Li, Yizhi, et al.
Veröffentlicht: (2024)
Zero-shot Multimodal Document Retrieval via Cross-modal Question Generation
von: Choi, Yejin, et al.
Veröffentlicht: (2025)
von: Choi, Yejin, et al.
Veröffentlicht: (2025)
OmnixR: Evaluating Omni-modality Language Models on Reasoning across Modalities
von: Chen, Lichang, et al.
Veröffentlicht: (2024)
von: Chen, Lichang, et al.
Veröffentlicht: (2024)
Cross-Domain Demo-to-Code via Neurosymbolic Counterfactual Reasoning
von: Kim, Jooyoung, et al.
Veröffentlicht: (2026)
von: Kim, Jooyoung, et al.
Veröffentlicht: (2026)
Simple Drop-in LoRA Conditioning on Attention Layers Will Improve Your Diffusion Model
von: Choi, Joo Young, et al.
Veröffentlicht: (2024)
von: Choi, Joo Young, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Two-Stage Grid Optimization for Group-wise Quantization of LLMs
von: Kim, Junhan, et al.
Veröffentlicht: (2026) -
LookaheadKV: Fast and Accurate KV Cache Eviction by Glimpsing into the Future without Generation
von: Ahn, Jinwoo, et al.
Veröffentlicht: (2026) -
Keep What Audio Cannot Say: Context-Preserving Token Pruning for Omni-LLMs
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2026) -
On the Importance of a Multi-Scale Calibration for Quantization
von: Son, Seungwoo, et al.
Veröffentlicht: (2026) -
OmniRetrieval: Unified Retrieval across Heterogeneous Knowledge Sources
von: Baek, Jinheon, et al.
Veröffentlicht: (2026)