Embedding-Free Transformer with Inference Spatial Reduction for Efficient Semantic Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Hyunwoo, Cho, Yubin, Kang, Beoungwoo, Moon, Seunghun, Kong, Kyeongbo, Kang, Suk-Ju |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MetaSeg: MetaFormer-based Global Contexts-aware Network for Efficient Semantic Segmentation
by: Kang, Beoungwoo, et al.
Published: (2024)
by: Kang, Beoungwoo, et al.
Published: (2024)
VIPA: Visual Informative Part Attention for Referring Image Segmentation
by: Cho, Yubin, et al.
Published: (2026)
by: Cho, Yubin, et al.
Published: (2026)
Cross-Stage Attention Propagation for Efficient Semantic Segmentation
by: Kang, Beoungwoo
Published: (2026)
by: Kang, Beoungwoo
Published: (2026)
Cross-aware Early Fusion with Stage-divided Vision and Language Transformer Encoders for Referring Image Segmentation
by: Cho, Yubin, et al.
Published: (2024)
by: Cho, Yubin, et al.
Published: (2024)
AttentionHand: Text-driven Controllable Hand Image Generation for 3D Hand Reconstruction in the Wild
by: Park, Junho, et al.
Published: (2024)
by: Park, Junho, et al.
Published: (2024)
Programmable-Room: Interactive Textured 3D Room Meshes Generation Empowered by Large Language Models
by: Kim, Jihyun, et al.
Published: (2025)
by: Kim, Jihyun, et al.
Published: (2025)
In Defense of Lazy Visual Grounding for Open-Vocabulary Semantic Segmentation
by: Kang, Dahyun, et al.
Published: (2024)
by: Kang, Dahyun, et al.
Published: (2024)
ClickSeg3D: Few-Click Interactive Segmentation via Semantic Embeddings
by: Kang, Xueyang, et al.
Published: (2026)
by: Kang, Xueyang, et al.
Published: (2026)
HOIGS: Human-Object Interaction Gaussian Splatting
by: Kim, Taewoo, et al.
Published: (2026)
by: Kim, Taewoo, et al.
Published: (2026)
Training-Free Global Geometric Association for 4D LiDAR Panoptic Segmentation
by: Oh, Gyeongrok, et al.
Published: (2025)
by: Oh, Gyeongrok, et al.
Published: (2025)
A Hidden Semantic Bottleneck in Conditional Embeddings of Diffusion Transformers
by: Pham, Trung X., et al.
Published: (2026)
by: Pham, Trung X., et al.
Published: (2026)
LivingWorld: Interactive 4D World Generation with Environmental Dynamics
by: Mun, Hyeongju, et al.
Published: (2026)
by: Mun, Hyeongju, et al.
Published: (2026)
Self-Supervised Vision Transformers Are Efficient Segmentation Learners for Imperfect Labels
by: Lee, Seungho, et al.
Published: (2024)
by: Lee, Seungho, et al.
Published: (2024)
Efficient Redundancy Reduction for Open-Vocabulary Semantic Segmentation
by: Chen, Lin, et al.
Published: (2025)
by: Chen, Lin, et al.
Published: (2025)
ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation
by: Wang, Jingyun, et al.
Published: (2024)
by: Wang, Jingyun, et al.
Published: (2024)
Focus Matters: Phase-Aware Suppression for Hallucination in Vision-Language Models
by: Kim, Sohyeon, et al.
Published: (2026)
by: Kim, Sohyeon, et al.
Published: (2026)
MoE-GS: Mixture of Experts for Dynamic Gaussian Splatting
by: Jin, In-Hwan, et al.
Published: (2025)
by: Jin, In-Hwan, et al.
Published: (2025)
Region-Enhanced Feature Learning for Scene Semantic Segmentation
by: Kang, Xin, et al.
Published: (2023)
by: Kang, Xin, et al.
Published: (2023)
Context-Guided Spatial Feature Reconstruction for Efficient Semantic Segmentation
by: Ni, Zhenliang, et al.
Published: (2024)
by: Ni, Zhenliang, et al.
Published: (2024)
AgilePruner: An Empirical Study of Attention and Diversity for Adaptive Visual Token Pruning in Large Vision-Language Models
by: Baek, Changwoo, et al.
Published: (2026)
by: Baek, Changwoo, et al.
Published: (2026)
Stochastic Conditional Diffusion Models for Robust Semantic Image Synthesis
by: Ko, Juyeon, et al.
Published: (2024)
by: Ko, Juyeon, et al.
Published: (2024)
Diffusion Prior-Based Amortized Variational Inference for Noisy Inverse Problems
by: Lee, Sojin, et al.
Published: (2024)
by: Lee, Sojin, et al.
Published: (2024)
Revisiting Efficient Semantic Segmentation: Learning Offsets for Better Spatial and Class Feature Alignment
by: Zhang, Shi-Chen, et al.
Published: (2025)
by: Zhang, Shi-Chen, et al.
Published: (2025)
Rewrite Caption Semantics: Bridging Semantic Gaps for Language-Supervised Semantic Segmentation
by: Xing, Yun, et al.
Published: (2023)
by: Xing, Yun, et al.
Published: (2023)
Scalp Diagnostic System With Label-Free Segmentation and Training-Free Image Translation
by: Kim, Youngmin, et al.
Published: (2024)
by: Kim, Youngmin, et al.
Published: (2024)
Improving Visual Token Reduction via Rectifying Distortions for Efficient Multimodal LLM Inference
by: Cho, Hyeonwoo, et al.
Published: (2026)
by: Cho, Hyeonwoo, et al.
Published: (2026)
Attention Frequency Modulation: Training-Free Spectral Modulation of Diffusion Cross-Attention
by: Oh, Seunghun, et al.
Published: (2026)
by: Oh, Seunghun, et al.
Published: (2026)
Geometry-Constrained Monocular Scale Estimation Using Semantic Segmentation for Dynamic Scenes
by: Zhang, Hui, et al.
Published: (2025)
by: Zhang, Hui, et al.
Published: (2025)
Universal Domain Adaptation for Semantic Segmentation
by: Choe, Seun-An, et al.
Published: (2025)
by: Choe, Seun-An, et al.
Published: (2025)
CoBra: Complementary Branch Fusing Class and Semantic Knowledge for Robust Weakly Supervised Semantic Segmentation
by: Han, Woojung, et al.
Published: (2024)
by: Han, Woojung, et al.
Published: (2024)
LogoDiffuser: Training-Free Multilingual Logo Generation and Stylization via Letter-Aware Attention Control
by: Kang, Mingyu, et al.
Published: (2026)
by: Kang, Mingyu, et al.
Published: (2026)
Multi-criteria Token Fusion with One-step-ahead Attention for Efficient Vision Transformers
by: Lee, Sanghyeok, et al.
Published: (2024)
by: Lee, Sanghyeok, et al.
Published: (2024)
TRiGS: Temporal Rigid-Body Motion for Scalable 4D Gaussian Splatting
by: Yeom, Suwoong, et al.
Published: (2026)
by: Yeom, Suwoong, et al.
Published: (2026)
Video Inference for Human Mesh Recovery with Vision Transformer
by: Cho, Hanbyel, et al.
Published: (2025)
by: Cho, Hanbyel, et al.
Published: (2025)
Semi-Supervised 3D Object Detection with Channel Augmentation using Transformation Equivariance
by: Kang, Minju, et al.
Published: (2024)
by: Kang, Minju, et al.
Published: (2024)
DocParseNet: Advanced Semantic Segmentation and OCR Embeddings for Efficient Scanned Document Annotation
by: Mohammadshirazi, Ahmad, et al.
Published: (2024)
by: Mohammadshirazi, Ahmad, et al.
Published: (2024)
SIESEF-FusionNet: Spatial Inter-correlation Enhancement and Spatially-Embedded Feature Fusion Network for LiDAR Point Cloud Semantic Segmentation
by: Chen, Jiale, et al.
Published: (2024)
by: Chen, Jiale, et al.
Published: (2024)
EMA-SAM: Exponential Moving-average for SAM-based PTMC Segmentation
by: Dialameh, Maryam, et al.
Published: (2025)
by: Dialameh, Maryam, et al.
Published: (2025)
CAVIS: Context-Aware Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2024)
by: Lee, Seunghun, et al.
Published: (2024)
NUC-Net: Non-uniform Cylindrical Partition Network for Efficient LiDAR Semantic Segmentation
by: Wang, Xuzhi, et al.
Published: (2025)
by: Wang, Xuzhi, et al.
Published: (2025)
Similar Items
-
MetaSeg: MetaFormer-based Global Contexts-aware Network for Efficient Semantic Segmentation
by: Kang, Beoungwoo, et al.
Published: (2024) -
VIPA: Visual Informative Part Attention for Referring Image Segmentation
by: Cho, Yubin, et al.
Published: (2026) -
Cross-Stage Attention Propagation for Efficient Semantic Segmentation
by: Kang, Beoungwoo
Published: (2026) -
Cross-aware Early Fusion with Stage-divided Vision and Language Transformer Encoders for Referring Image Segmentation
by: Cho, Yubin, et al.
Published: (2024) -
AttentionHand: Text-driven Controllable Hand Image Generation for 3D Hand Reconstruction in the Wild
by: Park, Junho, et al.
Published: (2024)