Salvato in:
| Autori principali: | Pak, Byeonghyun, Woo, Byeongju, Kim, Sunghwan, Kim, Dae-hwan, Kim, Hoseong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2407.09033 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Tortoise and Hare Guidance: Accelerating Diffusion Model Inference with Multirate Integration
di: Lee, Yunghee, et al.
Pubblicazione: (2025)
di: Lee, Yunghee, et al.
Pubblicazione: (2025)
Pixel-level Scene Understanding in One Token: Visual States Need What-is-Where Composition
di: Lee, Seokmin, et al.
Pubblicazione: (2026)
di: Lee, Seokmin, et al.
Pubblicazione: (2026)
Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding
di: Woo, Byeongju, et al.
Pubblicazione: (2026)
di: Woo, Byeongju, et al.
Pubblicazione: (2026)
Rethinking FID Through the Geometry of the Reference Dataset
di: Lee, Yunghee, et al.
Pubblicazione: (2026)
di: Lee, Yunghee, et al.
Pubblicazione: (2026)
Leveraging 2D Masked Reconstruction for Domain Adaptation of 3D Pose Estimation
di: Park, Hansoo, et al.
Pubblicazione: (2025)
di: Park, Hansoo, et al.
Pubblicazione: (2025)
A Simple Baseline with Single-encoder for Referring Image Segmentation
di: Yu, Seonghoon, et al.
Pubblicazione: (2024)
di: Yu, Seonghoon, et al.
Pubblicazione: (2024)
SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection
di: Kim, Junsu, et al.
Pubblicazione: (2024)
di: Kim, Junsu, et al.
Pubblicazione: (2024)
Rethinking LiDAR Domain Generalization: Single Source as Multiple Density Domains
di: Kim, Jaeyeul, et al.
Pubblicazione: (2023)
di: Kim, Jaeyeul, et al.
Pubblicazione: (2023)
Seg4Diff: Unveiling Open-Vocabulary Segmentation in Text-to-Image Diffusion Transformers
di: Kim, Chaehyun, et al.
Pubblicazione: (2025)
di: Kim, Chaehyun, et al.
Pubblicazione: (2025)
Unifying Feature and Cost Aggregation with Transformers for Semantic and Visual Correspondence
di: Hong, Sunghwan, et al.
Pubblicazione: (2024)
di: Hong, Sunghwan, et al.
Pubblicazione: (2024)
BEEP3D: Box-Supervised End-to-End Pseudo-Mask Generation for 3D Instance Segmentation
di: Yoo, Youngju, et al.
Pubblicazione: (2025)
di: Yoo, Youngju, et al.
Pubblicazione: (2025)
Adaptive Residual Transformation for Enhanced Feature-Based OOD Detection in SAR Imagery
di: Lee, Kyung-hwan, et al.
Pubblicazione: (2024)
di: Lee, Kyung-hwan, et al.
Pubblicazione: (2024)
Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels
di: Shin, Heeseong, et al.
Pubblicazione: (2024)
di: Shin, Heeseong, et al.
Pubblicazione: (2024)
Semantic-Aware Gaussian Process Calibration with Structured Layerwise Kernels for Deep Neural Networks
di: Lee, Kyung-hwan, et al.
Pubblicazione: (2025)
di: Lee, Kyung-hwan, et al.
Pubblicazione: (2025)
Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
di: Kim, Jongha, et al.
Pubblicazione: (2024)
di: Kim, Jongha, et al.
Pubblicazione: (2024)
Mask-Free Neuron Concept Annotation for Interpreting Neural Networks in Medical Domain
di: Kim, Hyeon Bae, et al.
Pubblicazione: (2024)
di: Kim, Hyeon Bae, et al.
Pubblicazione: (2024)
MMR: A Large-scale Benchmark Dataset for Multi-target and Multi-granularity Reasoning Segmentation
di: Jang, Donggon, et al.
Pubblicazione: (2025)
di: Jang, Donggon, et al.
Pubblicazione: (2025)
SIDA: Synthetic Image Driven Zero-shot Domain Adaptation
di: Kim, Ye-Chan, et al.
Pubblicazione: (2025)
di: Kim, Ye-Chan, et al.
Pubblicazione: (2025)
Do We Need Perfect Data? Leveraging Noise for Domain Generalized Segmentation
di: Kim, Taeyeong, et al.
Pubblicazione: (2025)
di: Kim, Taeyeong, et al.
Pubblicazione: (2025)
OurDB: Ouroboric Domain Bridging for Multi-Target Domain Adaptive Semantic Segmentation
di: Woo, Seungbeom, et al.
Pubblicazione: (2024)
di: Woo, Seungbeom, et al.
Pubblicazione: (2024)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation
di: Cho, Seokju, et al.
Pubblicazione: (2023)
di: Cho, Seokju, et al.
Pubblicazione: (2023)
BridgeTA: Bridging the Representation Gap in Knowledge Distillation via Teacher Assistant for Bird's Eye View Map Segmentation
di: Kim, Beomjun, et al.
Pubblicazione: (2025)
di: Kim, Beomjun, et al.
Pubblicazione: (2025)
VisionTrap: Vision-Augmented Trajectory Prediction Guided by Textual Descriptions
di: Moon, Seokha, et al.
Pubblicazione: (2024)
di: Moon, Seokha, et al.
Pubblicazione: (2024)
MATRIX: Mask Track Alignment for Interaction-aware Video Generation
di: Jin, Siyoon, et al.
Pubblicazione: (2025)
di: Jin, Siyoon, et al.
Pubblicazione: (2025)
Soft Segmented Randomization: Enhancing Domain Generalization in SAR-ATR for Synthetic-to-Measured
di: Kim, Minjun, et al.
Pubblicazione: (2024)
di: Kim, Minjun, et al.
Pubblicazione: (2024)
Alleviating Textual Reliance in Medical Language-guided Segmentation via Prototype-driven Semantic Approximation
di: Ye, Shuchang, et al.
Pubblicazione: (2025)
di: Ye, Shuchang, et al.
Pubblicazione: (2025)
Transformer with Leveraged Masked Autoencoder for video-based Pain Assessment
di: Nguyen, Minh-Duc, et al.
Pubblicazione: (2024)
di: Nguyen, Minh-Duc, et al.
Pubblicazione: (2024)
Cross-Domain Semantic Segmentation on Inconsistent Taxonomy using VLMs
di: Lim, Jeongkee, et al.
Pubblicazione: (2024)
di: Lim, Jeongkee, et al.
Pubblicazione: (2024)
Query-Calibrated Segmental Admission for Descriptor-Agnostic LiDAR Loop Closure in Repetitive Environments
di: Kim, Jaehyun, et al.
Pubblicazione: (2025)
di: Kim, Jaehyun, et al.
Pubblicazione: (2025)
Mask2Map: Vectorized HD Map Construction Using Bird's Eye View Segmentation Masks
di: Choi, Sehwan, et al.
Pubblicazione: (2024)
di: Choi, Sehwan, et al.
Pubblicazione: (2024)
Advancing Medical Image Segmentation: Morphology-Driven Learning with Diffusion Transformer
di: Kang, Sungmin, et al.
Pubblicazione: (2024)
di: Kang, Sungmin, et al.
Pubblicazione: (2024)
Selective Query-guided Debiasing for Video Corpus Moment Retrieval
di: Yoon, Sunjae, et al.
Pubblicazione: (2022)
di: Yoon, Sunjae, et al.
Pubblicazione: (2022)
UDC-VIT: A Real-World Video Dataset for Under-Display Cameras
di: Ahn, Kyusu, et al.
Pubblicazione: (2025)
di: Ahn, Kyusu, et al.
Pubblicazione: (2025)
Cross-View Completion Models are Zero-shot Correspondence Estimators
di: An, Honggyu, et al.
Pubblicazione: (2024)
di: An, Honggyu, et al.
Pubblicazione: (2024)
Q-Align: Alleviating Attention Leakage in Zero-Shot Appearance Transfer via Query-Query Alignment
di: Kim, Namu, et al.
Pubblicazione: (2025)
di: Kim, Namu, et al.
Pubblicazione: (2025)
Contour-Guided Query-Based Feature Fusion for Boundary-Aware and Generalizable Cardiac Ultrasound Segmentation
di: Ullah, Zahid, et al.
Pubblicazione: (2026)
di: Ullah, Zahid, et al.
Pubblicazione: (2026)
Emergent Outlier View Rejection in Visual Geometry Grounded Transformers
di: Han, Jisang, et al.
Pubblicazione: (2025)
di: Han, Jisang, et al.
Pubblicazione: (2025)
Causal Representation-Based Domain Generalization on Gaze Estimation
di: Kim, Younghan, et al.
Pubblicazione: (2024)
di: Kim, Younghan, et al.
Pubblicazione: (2024)
InstantFamily: Masked Attention for Zero-shot Multi-ID Image Generation
di: Kim, Chanran, et al.
Pubblicazione: (2024)
di: Kim, Chanran, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Tortoise and Hare Guidance: Accelerating Diffusion Model Inference with Multirate Integration
di: Lee, Yunghee, et al.
Pubblicazione: (2025) -
Pixel-level Scene Understanding in One Token: Visual States Need What-is-Where Composition
di: Lee, Seokmin, et al.
Pubblicazione: (2026) -
Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding
di: Woo, Byeongju, et al.
Pubblicazione: (2026) -
Rethinking FID Through the Geometry of the Reference Dataset
di: Lee, Yunghee, et al.
Pubblicazione: (2026) -
Leveraging 2D Masked Reconstruction for Domain Adaptation of 3D Pose Estimation
di: Park, Hansoo, et al.
Pubblicazione: (2025)