TC-SSA: Token Compression via Semantic Slot Aggregation for Gigapixel Pathology Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Zhuo, Young, Shawn, Xu, Lijian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Discovering Pathology Rationale and Token Allocation for Efficient Multimodal Pathology Reasoning
von: Xu, Zhe, et al.
Veröffentlicht: (2025)
von: Xu, Zhe, et al.
Veröffentlicht: (2025)
WsiCaption: Multiple Instance Generation of Pathology Reports for Gigapixel Whole-Slide Images
von: Chen, Pingyi, et al.
Veröffentlicht: (2023)
von: Chen, Pingyi, et al.
Veröffentlicht: (2023)
Efficient Chest X-ray Representation Learning via Semantic-Partitioned Contrastive Learning
von: Feng, Wangyu, et al.
Veröffentlicht: (2026)
von: Feng, Wangyu, et al.
Veröffentlicht: (2026)
Neural Slot Interpreters: Grounding Object Semantics in Emergent Slot Representations
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2024)
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2024)
Fewer Tokens, Greater Scaling: Self-Adaptive Visual Bases for Efficient and Expansive Representation Learning
von: Young, Shawn, et al.
Veröffentlicht: (2025)
von: Young, Shawn, et al.
Veröffentlicht: (2025)
Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM
von: Chi, Donghwan, et al.
Veröffentlicht: (2025)
von: Chi, Donghwan, et al.
Veröffentlicht: (2025)
XrayClaw: Cooperative-Competitive Multi-Agent Alignment for Trustworthy Chest X-ray Diagnosis
von: Young, Shawn, et al.
Veröffentlicht: (2026)
von: Young, Shawn, et al.
Veröffentlicht: (2026)
Learned Image Compression and Restoration for Digital Pathology
von: Lee, SeonYeong, et al.
Veröffentlicht: (2025)
von: Lee, SeonYeong, et al.
Veröffentlicht: (2025)
Label-free Concept Based Multiple Instance Learning for Gigapixel Histopathology
von: Sun, Susu, et al.
Veröffentlicht: (2025)
von: Sun, Susu, et al.
Veröffentlicht: (2025)
Beyond Surrogate Gradients: Fully Differentiable Token Pruning for Vision-Language Models
von: He, Landi, et al.
Veröffentlicht: (2026)
von: He, Landi, et al.
Veröffentlicht: (2026)
Mixed Prototype Consistency Learning for Semi-supervised Medical Image Segmentation
von: Li, Lijian
Veröffentlicht: (2024)
von: Li, Lijian
Veröffentlicht: (2024)
Complementarity-driven Representation Learning for Multi-modal Knowledge Graph Completion
von: Li, Lijian
Veröffentlicht: (2025)
von: Li, Lijian
Veröffentlicht: (2025)
Multimodal Model for Computational Pathology:Representation Learning and Image Compression
von: Wu, Peihang, et al.
Veröffentlicht: (2026)
von: Wu, Peihang, et al.
Veröffentlicht: (2026)
Thinking in Scales: Accelerating Gigapixel Pathology Image Analysis via Adaptive Continuous Reasoning
von: Ge, Jiusong, et al.
Veröffentlicht: (2026)
von: Ge, Jiusong, et al.
Veröffentlicht: (2026)
GAS-MIL: Group-Aggregative Selection Multi-Instance Learning for Ensemble of Foundation Models in Digital Pathology Image Analysis
von: Quan, Peiran, et al.
Veröffentlicht: (2025)
von: Quan, Peiran, et al.
Veröffentlicht: (2025)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
Anatomy-Slot: Unsupervised Anatomical Factorization for Homologous Bilateral Reasoning in Retinal Diagnosis
von: Ma, Yingzhe, et al.
Veröffentlicht: (2026)
von: Ma, Yingzhe, et al.
Veröffentlicht: (2026)
Contribution-aware Token Compression for Efficient Video Understanding via Reinforcement Learning
von: Ma, Yinchao, et al.
Veröffentlicht: (2026)
von: Ma, Yinchao, et al.
Veröffentlicht: (2026)
ImgCoT: Compressing Long Chain of Thought into Compact Visual Tokens for Efficient Reasoning of Large Language Model
von: Chen, Xiaoshu, et al.
Veröffentlicht: (2026)
von: Chen, Xiaoshu, et al.
Veröffentlicht: (2026)
Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding
von: Li, Yueying, et al.
Veröffentlicht: (2026)
von: Li, Yueying, et al.
Veröffentlicht: (2026)
SemHiTok: A Unified Image Tokenizer via Semantic-Guided Hierarchical Codebook for Multimodal Understanding and Generation
von: Chen, Zisheng, et al.
Veröffentlicht: (2025)
von: Chen, Zisheng, et al.
Veröffentlicht: (2025)
AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation
von: Ma, Chaofan, et al.
Veröffentlicht: (2023)
von: Ma, Chaofan, et al.
Veröffentlicht: (2023)
InfoTok: Adaptive Discrete Video Tokenizer via Information-Theoretic Compression
von: Ye, Haotian, et al.
Veröffentlicht: (2025)
von: Ye, Haotian, et al.
Veröffentlicht: (2025)
Highly Compressed Tokenizer Can Generate Without Training
von: Beyer, L. Lao, et al.
Veröffentlicht: (2025)
von: Beyer, L. Lao, et al.
Veröffentlicht: (2025)
Towards Lossless Ultimate Vision Token Compression for VLMs
von: Zheng, Dehua, et al.
Veröffentlicht: (2025)
von: Zheng, Dehua, et al.
Veröffentlicht: (2025)
HybridToken-VLM: Hybrid Token Compression for Vision-Language Models
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
StreamingTOM: Streaming Token Compression for Efficient Video Understanding
von: Chen, Xueyi, et al.
Veröffentlicht: (2025)
von: Chen, Xueyi, et al.
Veröffentlicht: (2025)
Geometrical Properties of Text Token Embeddings for Strong Semantic Binding in Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Token-Efficient Multimodal Reasoning via Image Prompt Packaging
von: Choi, Joong Ho, et al.
Veröffentlicht: (2026)
von: Choi, Joong Ho, et al.
Veröffentlicht: (2026)
Uncertainty-aware Evidential Fusion-based Learning for Semi-supervised Medical Image Segmentation
von: He, Yuanpeng, et al.
Veröffentlicht: (2024)
von: He, Yuanpeng, et al.
Veröffentlicht: (2024)
Towards Realistic Long-tailed Semi-supervised Learning in an Open World
von: He, Yuanpeng, et al.
Veröffentlicht: (2024)
von: He, Yuanpeng, et al.
Veröffentlicht: (2024)
CIC-BART-SSA: Controllable Image Captioning with Structured Semantic Augmentation
von: Basioti, Kalliopi, et al.
Veröffentlicht: (2024)
von: Basioti, Kalliopi, et al.
Veröffentlicht: (2024)
SegMix:Shuffle-based Feedback Learning for Semantic Segmentation of Pathology Images
von: Yan, Zhiling, et al.
Veröffentlicht: (2026)
von: Yan, Zhiling, et al.
Veröffentlicht: (2026)
VISA: Group-wise Visual Token Selection and Aggregation via Graph Summarization for Efficient MLLMs Inference
von: Jiang, Pengfei, et al.
Veröffentlicht: (2025)
von: Jiang, Pengfei, et al.
Veröffentlicht: (2025)
SlotPi: Physics-informed Object-centric Reasoning Models
von: Li, Jian, et al.
Veröffentlicht: (2025)
von: Li, Jian, et al.
Veröffentlicht: (2025)
DASH: Dynamic Audio-Driven Semantic Chunking for Efficient Omnimodal Token Compression
von: Li, Bingzhou, et al.
Veröffentlicht: (2026)
von: Li, Bingzhou, et al.
Veröffentlicht: (2026)
EfficientPosterGen: Semantic-aware Efficient Poster Generation via Token Compression and Accurate Violation Detection
von: Tang, Wenxin, et al.
Veröffentlicht: (2026)
von: Tang, Wenxin, et al.
Veröffentlicht: (2026)
A Semantically Enhanced Generative Foundation Model Improves Pathological Image Synthesis
von: Guan, Xianchao, et al.
Veröffentlicht: (2025)
von: Guan, Xianchao, et al.
Veröffentlicht: (2025)
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token Merging
von: Kang, Inha, et al.
Veröffentlicht: (2025)
von: Kang, Inha, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Discovering Pathology Rationale and Token Allocation for Efficient Multimodal Pathology Reasoning
von: Xu, Zhe, et al.
Veröffentlicht: (2025) -
WsiCaption: Multiple Instance Generation of Pathology Reports for Gigapixel Whole-Slide Images
von: Chen, Pingyi, et al.
Veröffentlicht: (2023) -
Efficient Chest X-ray Representation Learning via Semantic-Partitioned Contrastive Learning
von: Feng, Wangyu, et al.
Veröffentlicht: (2026) -
Neural Slot Interpreters: Grounding Object Semantics in Emergent Slot Representations
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2024) -
Fewer Tokens, Greater Scaling: Self-Adaptive Visual Bases for Efficient and Expansive Representation Learning
von: Young, Shawn, et al.
Veröffentlicht: (2025)