Segformer++: Efficient Token-Merging Strategies for High-Resolution Semantic Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Kienzle, Daniel, Kantonis, Marco, Schön, Robin, Lienhart, Rainer |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Ball Spin and Trajectory Analysis in Table Tennis Broadcast Videos via Physically Grounded Synthetic-to-Real Transfer
by: Kienzle, Daniel, et al.
Published: (2025)
by: Kienzle, Daniel, et al.
Published: (2025)
MMMS: Multi-Modal Multi-Surface Interactive Segmentation
by: Schön, Robin, et al.
Published: (2025)
by: Schön, Robin, et al.
Published: (2025)
WSESeg: Introducing a Dataset for the Segmentation of Winter Sports Equipment with a Baseline for Interactive Segmentation
by: Schön, Robin, et al.
Published: (2024)
by: Schön, Robin, et al.
Published: (2024)
SkipClick: Combining Quick Responses and Low-Level Features for Interactive Segmentation in Winter Sports Contexts
by: Schön, Robin, et al.
Published: (2025)
by: Schön, Robin, et al.
Published: (2025)
A Review and Efficient Implementation of Scene Graph Generation Metrics
by: Lorenz, Julian, et al.
Published: (2024)
by: Lorenz, Julian, et al.
Published: (2024)
Efficient 2D to Full 3D Human Pose Uplifting including Joint Rotations
by: Ludwig, Katja, et al.
Published: (2025)
by: Ludwig, Katja, et al.
Published: (2025)
A Fair Ranking and New Model for Panoptic Scene Graph Generation
by: Lorenz, Julian, et al.
Published: (2024)
by: Lorenz, Julian, et al.
Published: (2024)
ToDo: Token Downsampling for Efficient Generation of High-Resolution Images
by: Smith, Ethan, et al.
Published: (2024)
by: Smith, Ethan, et al.
Published: (2024)
Contextual Hourglass Network for Semantic Segmentation of High Resolution Aerial Imagery
by: Li, Panfeng, et al.
Published: (2018)
by: Li, Panfeng, et al.
Published: (2018)
Adaptive Token Merging for Efficient Transformer Semantic Communication at the Edge
by: Erak, Omar, et al.
Published: (2025)
by: Erak, Omar, et al.
Published: (2025)
Uplifting Table Tennis: A Robust, Real-World Application for 3D Trajectory and Spin Estimation
by: Kienzle, Daniel, et al.
Published: (2025)
by: Kienzle, Daniel, et al.
Published: (2025)
High-Resolution Image Synthesis via Next-Token Prediction
by: Chen, Dengsheng, et al.
Published: (2024)
by: Chen, Dengsheng, et al.
Published: (2024)
ARTA: Adaptive Mixed-Resolution Token Allocation for Efficient Dense Feature Extraction
by: Hagerman, David, et al.
Published: (2026)
by: Hagerman, David, et al.
Published: (2026)
I-Segmenter: Integer-Only Vision Transformer for Efficient Semantic Segmentation
by: Sassoon, Jordan, et al.
Published: (2025)
by: Sassoon, Jordan, et al.
Published: (2025)
One-Shot Multi-Label Causal Discovery in High-Dimensional Event Sequences
by: Math, Hugo, et al.
Published: (2025)
by: Math, Hugo, et al.
Published: (2025)
Adapting the Segment Anything Model During Usage in Novel Situations
by: Schön, Robin, et al.
Published: (2024)
by: Schön, Robin, et al.
Published: (2024)
Adaptive Pareto-Optimal Token Merging for Edge Transformer Models in Semantic Communication
by: Erak, Omar, et al.
Published: (2025)
by: Erak, Omar, et al.
Published: (2025)
Negative Token Merging: Image-based Adversarial Feature Guidance
by: Singh, Jaskirat, et al.
Published: (2024)
by: Singh, Jaskirat, et al.
Published: (2024)
SkipSR: Faster Super Resolution with Token Skipping
by: Choudhury, Rohan, et al.
Published: (2025)
by: Choudhury, Rohan, et al.
Published: (2025)
Control Your View: High-Resolution Global Semantic Manipulation in Learned Image Compression
by: Liang, Jiaming, et al.
Published: (2026)
by: Liang, Jiaming, et al.
Published: (2026)
Efficient Multi-task Uncertainties for Joint Semantic Segmentation and Monocular Depth Estimation
by: Landgraf, Steven, et al.
Published: (2024)
by: Landgraf, Steven, et al.
Published: (2024)
FlashVID: Efficient Video Large Language Models via Training-free Tree-based Spatiotemporal Token Merging
by: Fan, Ziyang, et al.
Published: (2026)
by: Fan, Ziyang, et al.
Published: (2026)
CSC-Unet: A Novel Convolutional Sparse Coding Strategy Based Neural Network for Semantic Segmentation
by: Tang, Haitong, et al.
Published: (2021)
by: Tang, Haitong, et al.
Published: (2021)
Causal Unsupervised Semantic Segmentation
by: Kim, Junho, et al.
Published: (2023)
by: Kim, Junho, et al.
Published: (2023)
Lorentz Framework for Semantic Segmentation
by: Hasan, Zahid, et al.
Published: (2026)
by: Hasan, Zahid, et al.
Published: (2026)
Efficient High-Resolution Image Editing with Hallucination-Aware Loss and Adaptive Tiling
by: Kwon, Young D., et al.
Published: (2025)
by: Kwon, Young D., et al.
Published: (2025)
Multi-Scale High-Resolution Logarithmic Grapher Module for Efficient Vision GNNs
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
ATM: Improving Model Merging by Alternating Tuning and Merging
by: Zhou, Luca, et al.
Published: (2024)
by: Zhou, Luca, et al.
Published: (2024)
Occlusion-Ordered Semantic Instance Segmentation
by: Baselizadeh, Soroosh, et al.
Published: (2025)
by: Baselizadeh, Soroosh, et al.
Published: (2025)
TopoLoRA-SAM: Topology-Aware Parameter-Efficient Adaptation of Foundation Segmenters for Thin-Structure and Cross-Domain Binary Semantic Segmentation
by: Khazem, Salim
Published: (2026)
by: Khazem, Salim
Published: (2026)
ULTra: Unveiling Latent Token Interpretability in Transformer-Based Understanding and Segmentation
by: Hosseini, Hesam, et al.
Published: (2024)
by: Hosseini, Hesam, et al.
Published: (2024)
Efficient World Models with Context-Aware Tokenization
by: Micheli, Vincent, et al.
Published: (2024)
by: Micheli, Vincent, et al.
Published: (2024)
Efficient Architectures for High Resolution Vision-Language Models
by: Carvalho, Miguel, et al.
Published: (2025)
by: Carvalho, Miguel, et al.
Published: (2025)
Automatic Aorta Segmentation with Heavily Augmented, High-Resolution 3-D ResUNet: Contribution to the SEG.A Challenge
by: Wodzinski, Marek, et al.
Published: (2023)
by: Wodzinski, Marek, et al.
Published: (2023)
MagMax: Leveraging Model Merging for Seamless Continual Learning
by: Marczak, Daniel, et al.
Published: (2024)
by: Marczak, Daniel, et al.
Published: (2024)
Collaborating Foundation Models for Domain Generalized Semantic Segmentation
by: Benigmim, Yasser, et al.
Published: (2023)
by: Benigmim, Yasser, et al.
Published: (2023)
The BRAVO Semantic Segmentation Challenge Results in UNCV2024
by: Vu, Tuan-Hung, et al.
Published: (2024)
by: Vu, Tuan-Hung, et al.
Published: (2024)
Subspace-Boosted Model Merging
by: Skorobogat, Ronald, et al.
Published: (2025)
by: Skorobogat, Ronald, et al.
Published: (2025)
Unified Spatio-Temporal Token Scoring for Efficient Video VLMs
by: Zhang, Jianrui, et al.
Published: (2026)
by: Zhang, Jianrui, et al.
Published: (2026)
SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
Similar Items
-
Towards Ball Spin and Trajectory Analysis in Table Tennis Broadcast Videos via Physically Grounded Synthetic-to-Real Transfer
by: Kienzle, Daniel, et al.
Published: (2025) -
MMMS: Multi-Modal Multi-Surface Interactive Segmentation
by: Schön, Robin, et al.
Published: (2025) -
WSESeg: Introducing a Dataset for the Segmentation of Winter Sports Equipment with a Baseline for Interactive Segmentation
by: Schön, Robin, et al.
Published: (2024) -
SkipClick: Combining Quick Responses and Low-Level Features for Interactive Segmentation in Winter Sports Contexts
by: Schön, Robin, et al.
Published: (2025) -
A Review and Efficient Implementation of Scene Graph Generation Metrics
by: Lorenz, Julian, et al.
Published: (2024)