Textual Query-Driven Mask Transformer for Domain Generalized Segmentation
Fuente:
arXiv
Guardado en:
| Autores principales: | Pak, Byeonghyun, Woo, Byeongju, Kim, Sunghwan, Kim, Dae-hwan, Kim, Hoseong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Tortoise and Hare Guidance: Accelerating Diffusion Model Inference with Multirate Integration
por: Lee, Yunghee, et al.
Publicado: (2025)
por: Lee, Yunghee, et al.
Publicado: (2025)
Pixel-level Scene Understanding in One Token: Visual States Need What-is-Where Composition
por: Lee, Seokmin, et al.
Publicado: (2026)
por: Lee, Seokmin, et al.
Publicado: (2026)
Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding
por: Woo, Byeongju, et al.
Publicado: (2026)
por: Woo, Byeongju, et al.
Publicado: (2026)
Rethinking FID Through the Geometry of the Reference Dataset
por: Lee, Yunghee, et al.
Publicado: (2026)
por: Lee, Yunghee, et al.
Publicado: (2026)
Leveraging 2D Masked Reconstruction for Domain Adaptation of 3D Pose Estimation
por: Park, Hansoo, et al.
Publicado: (2025)
por: Park, Hansoo, et al.
Publicado: (2025)
SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection
por: Kim, Junsu, et al.
Publicado: (2024)
por: Kim, Junsu, et al.
Publicado: (2024)
A Simple Baseline with Single-encoder for Referring Image Segmentation
por: Yu, Seonghoon, et al.
Publicado: (2024)
por: Yu, Seonghoon, et al.
Publicado: (2024)
Rethinking LiDAR Domain Generalization: Single Source as Multiple Density Domains
por: Kim, Jaeyeul, et al.
Publicado: (2023)
por: Kim, Jaeyeul, et al.
Publicado: (2023)
Seg4Diff: Unveiling Open-Vocabulary Segmentation in Text-to-Image Diffusion Transformers
por: Kim, Chaehyun, et al.
Publicado: (2025)
por: Kim, Chaehyun, et al.
Publicado: (2025)
BEEP3D: Box-Supervised End-to-End Pseudo-Mask Generation for 3D Instance Segmentation
por: Yoo, Youngju, et al.
Publicado: (2025)
por: Yoo, Youngju, et al.
Publicado: (2025)
Unifying Feature and Cost Aggregation with Transformers for Semantic and Visual Correspondence
por: Hong, Sunghwan, et al.
Publicado: (2024)
por: Hong, Sunghwan, et al.
Publicado: (2024)
Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
por: Kim, Jongha, et al.
Publicado: (2024)
por: Kim, Jongha, et al.
Publicado: (2024)
Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels
por: Shin, Heeseong, et al.
Publicado: (2024)
por: Shin, Heeseong, et al.
Publicado: (2024)
Mask-Free Neuron Concept Annotation for Interpreting Neural Networks in Medical Domain
por: Kim, Hyeon Bae, et al.
Publicado: (2024)
por: Kim, Hyeon Bae, et al.
Publicado: (2024)
Do We Need Perfect Data? Leveraging Noise for Domain Generalized Segmentation
por: Kim, Taeyeong, et al.
Publicado: (2025)
por: Kim, Taeyeong, et al.
Publicado: (2025)
Adaptive Residual Transformation for Enhanced Feature-Based OOD Detection in SAR Imagery
por: Lee, Kyung-hwan, et al.
Publicado: (2024)
por: Lee, Kyung-hwan, et al.
Publicado: (2024)
MMR: A Large-scale Benchmark Dataset for Multi-target and Multi-granularity Reasoning Segmentation
por: Jang, Donggon, et al.
Publicado: (2025)
por: Jang, Donggon, et al.
Publicado: (2025)
OurDB: Ouroboric Domain Bridging for Multi-Target Domain Adaptive Semantic Segmentation
por: Woo, Seungbeom, et al.
Publicado: (2024)
por: Woo, Seungbeom, et al.
Publicado: (2024)
Semantic-Aware Gaussian Process Calibration with Structured Layerwise Kernels for Deep Neural Networks
por: Lee, Kyung-hwan, et al.
Publicado: (2025)
por: Lee, Kyung-hwan, et al.
Publicado: (2025)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
por: Kim, Jeongho, et al.
Publicado: (2024)
por: Kim, Jeongho, et al.
Publicado: (2024)
BridgeTA: Bridging the Representation Gap in Knowledge Distillation via Teacher Assistant for Bird's Eye View Map Segmentation
por: Kim, Beomjun, et al.
Publicado: (2025)
por: Kim, Beomjun, et al.
Publicado: (2025)
SIDA: Synthetic Image Driven Zero-shot Domain Adaptation
por: Kim, Ye-Chan, et al.
Publicado: (2025)
por: Kim, Ye-Chan, et al.
Publicado: (2025)
Soft Segmented Randomization: Enhancing Domain Generalization in SAR-ATR for Synthetic-to-Measured
por: Kim, Minjun, et al.
Publicado: (2024)
por: Kim, Minjun, et al.
Publicado: (2024)
Cross-Domain Semantic Segmentation on Inconsistent Taxonomy using VLMs
por: Lim, Jeongkee, et al.
Publicado: (2024)
por: Lim, Jeongkee, et al.
Publicado: (2024)
MATRIX: Mask Track Alignment for Interaction-aware Video Generation
por: Jin, Siyoon, et al.
Publicado: (2025)
por: Jin, Siyoon, et al.
Publicado: (2025)
VisionTrap: Vision-Augmented Trajectory Prediction Guided by Textual Descriptions
por: Moon, Seokha, et al.
Publicado: (2024)
por: Moon, Seokha, et al.
Publicado: (2024)
CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation
por: Cho, Seokju, et al.
Publicado: (2023)
por: Cho, Seokju, et al.
Publicado: (2023)
Transformer with Leveraged Masked Autoencoder for video-based Pain Assessment
por: Nguyen, Minh-Duc, et al.
Publicado: (2024)
por: Nguyen, Minh-Duc, et al.
Publicado: (2024)
Alleviating Textual Reliance in Medical Language-guided Segmentation via Prototype-driven Semantic Approximation
por: Ye, Shuchang, et al.
Publicado: (2025)
por: Ye, Shuchang, et al.
Publicado: (2025)
Mask2Map: Vectorized HD Map Construction Using Bird's Eye View Segmentation Masks
por: Choi, Sehwan, et al.
Publicado: (2024)
por: Choi, Sehwan, et al.
Publicado: (2024)
Query-Calibrated Segmental Admission for Descriptor-Agnostic LiDAR Loop Closure in Repetitive Environments
por: Kim, Jaehyun, et al.
Publicado: (2025)
por: Kim, Jaehyun, et al.
Publicado: (2025)
Advancing Medical Image Segmentation: Morphology-Driven Learning with Diffusion Transformer
por: Kang, Sungmin, et al.
Publicado: (2024)
por: Kang, Sungmin, et al.
Publicado: (2024)
Q-Align: Alleviating Attention Leakage in Zero-Shot Appearance Transfer via Query-Query Alignment
por: Kim, Namu, et al.
Publicado: (2025)
por: Kim, Namu, et al.
Publicado: (2025)
Causal Representation-Based Domain Generalization on Gaze Estimation
por: Kim, Younghan, et al.
Publicado: (2024)
por: Kim, Younghan, et al.
Publicado: (2024)
Contour-Guided Query-Based Feature Fusion for Boundary-Aware and Generalizable Cardiac Ultrasound Segmentation
por: Ullah, Zahid, et al.
Publicado: (2026)
por: Ullah, Zahid, et al.
Publicado: (2026)
Selective Query-guided Debiasing for Video Corpus Moment Retrieval
por: Yoon, Sunjae, et al.
Publicado: (2022)
por: Yoon, Sunjae, et al.
Publicado: (2022)
FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation
por: Fang, Xueji, et al.
Publicado: (2026)
por: Fang, Xueji, et al.
Publicado: (2026)
InstantFamily: Masked Attention for Zero-shot Multi-ID Image Generation
por: Kim, Chanran, et al.
Publicado: (2024)
por: Kim, Chanran, et al.
Publicado: (2024)
AURA : Automatic Mask Generator using Randomized Input Sampling for Object Removal
por: Oh, Changsuk, et al.
Publicado: (2023)
por: Oh, Changsuk, et al.
Publicado: (2023)
GOTPR: General Outdoor Text-based Place Recognition Using Scene Graph Retrieval with OpenStreetMap
por: Jung, Donghwi, et al.
Publicado: (2025)
por: Jung, Donghwi, et al.
Publicado: (2025)
Ejemplares similares
-
Tortoise and Hare Guidance: Accelerating Diffusion Model Inference with Multirate Integration
por: Lee, Yunghee, et al.
Publicado: (2025) -
Pixel-level Scene Understanding in One Token: Visual States Need What-is-Where Composition
por: Lee, Seokmin, et al.
Publicado: (2026) -
Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding
por: Woo, Byeongju, et al.
Publicado: (2026) -
Rethinking FID Through the Geometry of the Reference Dataset
por: Lee, Yunghee, et al.
Publicado: (2026) -
Leveraging 2D Masked Reconstruction for Domain Adaptation of 3D Pose Estimation
por: Park, Hansoo, et al.
Publicado: (2025)