Activating Self-Attention for Multi-Scene Absolute Pose Regression
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Miso, Kim, Jihwan, Heo, Jae-Pil |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Long-term Pre-training for Temporal Action Detection with Transformers
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
Prediction-Feedback DETR for Temporal Action Detection
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
Masking Matters: Unlocking the Spatial Reasoning Capabilities of LLMs for 3D Scene-Language Understanding
by: Jeon, Yerim, et al.
Published: (2025)
by: Jeon, Yerim, et al.
Published: (2025)
Disambiguating 2D-3D Correspondences in Gaussian Splatting-based Feature Fields for Visual Localization
by: Lee, Miso, et al.
Published: (2026)
by: Lee, Miso, et al.
Published: (2026)
White Aggregation and Restoration for Few-shot 3D Point Cloud Semantic Segmentation
by: Im, Jiyun, et al.
Published: (2025)
by: Im, Jiyun, et al.
Published: (2025)
Mutually-Aware Feature Learning for Few-Shot Object Counting
by: Jeon, Yerim, et al.
Published: (2024)
by: Jeon, Yerim, et al.
Published: (2024)
Boundary-Recovering Network for Temporal Action Detection
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
Temporally Consistent Long-Term Memory for 3D Single Object Tracking
by: Yoo, Jaejoon, et al.
Published: (2026)
by: Yoo, Jaejoon, et al.
Published: (2026)
Noise-free Optimization in Early Training Steps for Image Super-Resolution
by: Lee, MinKyu, et al.
Published: (2023)
by: Lee, MinKyu, et al.
Published: (2023)
GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats
by: Hyun, Sangeek, et al.
Published: (2024)
by: Hyun, Sangeek, et al.
Published: (2024)
Quantifying Epistemic Uncertainty in Absolute Pose Regression
by: Zangeneh, Fereidoon, et al.
Published: (2025)
by: Zangeneh, Fereidoon, et al.
Published: (2025)
Style Injection in Diffusion: A Training-free Approach for Adapting Large-scale Diffusion Models for Style Transfer
by: Chung, Jiwoo, et al.
Published: (2023)
by: Chung, Jiwoo, et al.
Published: (2023)
Neural Refinement for Absolute Pose Regression with Feature Synthesis
by: Chen, Shuai, et al.
Published: (2023)
by: Chen, Shuai, et al.
Published: (2023)
Cross-scale Aligned Supervision for Training GANs
by: Hyun, Sangeek, et al.
Published: (2026)
by: Hyun, Sangeek, et al.
Published: (2026)
VLCounter: Text-aware Visual Representation for Zero-Shot Object Counting
by: Kang, Seunggu, et al.
Published: (2023)
by: Kang, Seunggu, et al.
Published: (2023)
KS-APR: Keyframe Selection for Robust Absolute Pose Regression
by: Liu, Changkun, et al.
Published: (2023)
by: Liu, Changkun, et al.
Published: (2023)
Correlation-Guided Query-Dependency Calibration for Video Temporal Grounding
by: Moon, WonJun, et al.
Published: (2023)
by: Moon, WonJun, et al.
Published: (2023)
Mitigating Background Shift in Class-Incremental Semantic Segmentation
by: Park, Gilhan, et al.
Published: (2024)
by: Park, Gilhan, et al.
Published: (2024)
Fine-Tuning Visual Autoregressive Models for Subject-Driven Generation
by: Chung, Jiwoo, et al.
Published: (2025)
by: Chung, Jiwoo, et al.
Published: (2025)
Analyzing the Training Dynamics of Image Restoration Transformers: A Revisit to Layer Normalization
by: Lee, MinKyu, et al.
Published: (2025)
by: Lee, MinKyu, et al.
Published: (2025)
Scalable GANs with Transformers
by: Hyun, Sangeek, et al.
Published: (2025)
by: Hyun, Sangeek, et al.
Published: (2025)
Translation of Text Embedding via Delta Vector to Suppress Strongly Entangled Content in Text-to-Image Diffusion Models
by: Koh, Eunseo, et al.
Published: (2025)
by: Koh, Eunseo, et al.
Published: (2025)
Progressive Proxy Anchor Propagation for Unsupervised Semantic Segmentation
by: Seong, Hyun Seok, et al.
Published: (2024)
by: Seong, Hyun Seok, et al.
Published: (2024)
Scene-agnostic Pose Regression for Visual Localization
by: Zheng, Junwei, et al.
Published: (2025)
by: Zheng, Junwei, et al.
Published: (2025)
Diversity-aware Channel Pruning for StyleGAN Compression
by: Chung, Jiwoo, et al.
Published: (2024)
by: Chung, Jiwoo, et al.
Published: (2024)
PDF-GS: Progressive Distractor Filtering for Robust 3D Gaussian Splatting
by: Seo, Kangmin, et al.
Published: (2026)
by: Seo, Kangmin, et al.
Published: (2026)
APR-Transformer: Initial Pose Estimation for Localization in Complex Environments through Absolute Pose Regression
by: Ravuri, Srinivas, et al.
Published: (2025)
by: Ravuri, Srinivas, et al.
Published: (2025)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
by: Park, Jihwan, et al.
Published: (2025)
by: Park, Jihwan, et al.
Published: (2025)
ConDo: Continual Domain Expansion for Absolute Pose Regression
by: Li, Zijun, et al.
Published: (2024)
by: Li, Zijun, et al.
Published: (2024)
Foreground-Covering Prototype Generation and Matching for SAM-Aided Few-Shot Segmentation
by: Park, Suho, et al.
Published: (2025)
by: Park, Suho, et al.
Published: (2025)
Auto-Encoded Supervision for Perceptual Image Super-Resolution
by: Lee, MinKyu, et al.
Published: (2024)
by: Lee, MinKyu, et al.
Published: (2024)
From Vicious to Virtuous Cycles: Synergistic Representation Learning for Unsupervised Video Object-Centric Learning
by: Seong, Hyun Seok, et al.
Published: (2026)
by: Seong, Hyun Seok, et al.
Published: (2026)
Selective Contrastive Learning for Weakly Supervised Affordance Grounding
by: Moon, WonJun, et al.
Published: (2025)
by: Moon, WonJun, et al.
Published: (2025)
Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning
by: Moon, WonJun, et al.
Published: (2026)
by: Moon, WonJun, et al.
Published: (2026)
Active Prompt Learning in Vision Language Models
by: Bang, Jihwan, et al.
Published: (2023)
by: Bang, Jihwan, et al.
Published: (2023)
Joint Learning of Pose Regression and Denoising Diffusion with Score Scaling Sampling for Category-level 6D Pose Estimation
by: Lee, Seunghyun, et al.
Published: (2025)
by: Lee, Seunghyun, et al.
Published: (2025)
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
by: Lee, SuBeen, et al.
Published: (2025)
by: Lee, SuBeen, et al.
Published: (2025)
Auxiliary Descriptive Knowledge for Few-Shot Adaptation of Vision-Language Model
by: Lee, SuBeen, et al.
Published: (2025)
by: Lee, SuBeen, et al.
Published: (2025)
NCAP: Scene Text Image Super-Resolution with Non-CAtegorical Prior
by: Park, Dongwoo, et al.
Published: (2025)
by: Park, Dongwoo, et al.
Published: (2025)
A-SCoRe: Attention-based Scene Coordinate Regression for wide-ranging scenarios
by: Bui, Huy-Hoang, et al.
Published: (2025)
by: Bui, Huy-Hoang, et al.
Published: (2025)
Similar Items
-
Long-term Pre-training for Temporal Action Detection with Transformers
by: Kim, Jihwan, et al.
Published: (2024) -
Prediction-Feedback DETR for Temporal Action Detection
by: Kim, Jihwan, et al.
Published: (2024) -
Masking Matters: Unlocking the Spatial Reasoning Capabilities of LLMs for 3D Scene-Language Understanding
by: Jeon, Yerim, et al.
Published: (2025) -
Disambiguating 2D-3D Correspondences in Gaussian Splatting-based Feature Fields for Visual Localization
by: Lee, Miso, et al.
Published: (2026) -
White Aggregation and Restoration for Few-shot 3D Point Cloud Semantic Segmentation
by: Im, Jiyun, et al.
Published: (2025)