GOAL: Global-local Object Alignment Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Choi, Hyungyu, Jang, Young Kyun, Eom, Chanho |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FAST-GOAL: Fast and Efficient Global-local Object Alignment Learning
by: Choi, Hyungyu, et al.
Published: (2026)
by: Choi, Hyungyu, et al.
Published: (2026)
CoVA: Text-Guided Composed Video Retrieval for Audio-Visual Content
by: Han, Gyuwon, et al.
Published: (2026)
by: Han, Gyuwon, et al.
Published: (2026)
DiCo: Disentangled Concept Representation for Text-to-image Person Re-identification
by: Kim, Giyeol, et al.
Published: (2026)
by: Kim, Giyeol, et al.
Published: (2026)
Visual Representation Alignment for Multimodal Large Language Models
by: Yoon, Heeji, et al.
Published: (2025)
by: Yoon, Heeji, et al.
Published: (2025)
Domain Generalization for Person Re-identification: A Survey Towards Domain-Agnostic Person Matching
by: Lee, Hyeonseo, et al.
Published: (2025)
by: Lee, Hyeonseo, et al.
Published: (2025)
MoCHA-former: Moiré-Conditioned Hybrid Adaptive Transformer for Video Demoiréing
by: Sung, Jeahun, et al.
Published: (2025)
by: Sung, Jeahun, et al.
Published: (2025)
Towards Cross-modal Backward-compatible Representation Learning for Vision-Language Models
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
Object Dynamics Modeling with Hierarchical Point Cloud-based Representations
by: Kim, Chanho, et al.
Published: (2024)
by: Kim, Chanho, et al.
Published: (2024)
R3eVision: A Survey on Robust Rendering, Restoration, and Enhancement for 3D Low-Level Vision
by: Kwon, Weeyoung, et al.
Published: (2025)
by: Kwon, Weeyoung, et al.
Published: (2025)
Leveraging Prior Knowledge of Diffusion Model for Person Search
by: Kim, Giyeol, et al.
Published: (2025)
by: Kim, Giyeol, et al.
Published: (2025)
AceVFI: A Comprehensive Survey of Advances in Video Frame Interpolation
by: Kye, Dahyeon, et al.
Published: (2025)
by: Kye, Dahyeon, et al.
Published: (2025)
SEAL: Semantic-aware Single-image Sticker Personalization with a Large-scale Sticker-tag Dataset
by: Roh, Changhyun, et al.
Published: (2026)
by: Roh, Changhyun, et al.
Published: (2026)
Disentangled Representations for Short-Term and Long-Term Person Re-Identification
by: Eom, Chanho, et al.
Published: (2024)
by: Eom, Chanho, et al.
Published: (2024)
GOAL: Geometrically Optimal Alignment for Continual Generalized Category Discovery
by: Han, Jizhou, et al.
Published: (2026)
by: Han, Jizhou, et al.
Published: (2026)
Distilling Vision-Language Pretraining for Efficient Cross-Modal Retrieval
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
Learning Adaptive Pseudo-Label Selection for Semi-Supervised 3D Object Detection
by: Kong, Taehun, et al.
Published: (2025)
by: Kong, Taehun, et al.
Published: (2025)
Cerberus: Attribute-based person re-identification using semantic IDs
by: Eom, Chanho, et al.
Published: (2024)
by: Eom, Chanho, et al.
Published: (2024)
Subnet-Aware Dynamic Supernet Training for Neural Architecture Search
by: Jeon, Jeimin, et al.
Published: (2025)
by: Jeon, Jeimin, et al.
Published: (2025)
Signal: Selective Interaction and Global-local Alignment for Multi-Modal Object Re-Identification
by: Liu, Yangyang, et al.
Published: (2025)
by: Liu, Yangyang, et al.
Published: (2025)
FLAIR: Frequency- and Locality-Aware Implicit Neural Representations
by: Ko, Sukhun, et al.
Published: (2025)
by: Ko, Sukhun, et al.
Published: (2025)
Learning a Particle Dynamics Model with Real-world Videos
by: Kim, Chanho, et al.
Published: (2026)
by: Kim, Chanho, et al.
Published: (2026)
Deep Understanding of Sign Language for Sign to Subtitle Alignment
by: Jang, Youngjoon, et al.
Published: (2025)
by: Jang, Youngjoon, et al.
Published: (2025)
Semi-Supervised 3D Object Detection with Channel Augmentation using Transformation Equivariance
by: Kang, Minju, et al.
Published: (2024)
by: Kang, Minju, et al.
Published: (2024)
MATE: Meet At The Embedding -- Connecting Images with Long Texts
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
Patch-Level Kernel Alignment for Dense Self-Supervised Learning
by: Yeo, Juan, et al.
Published: (2025)
by: Yeo, Juan, et al.
Published: (2025)
Lesion-Aware Post-Training of Latent Diffusion Models for Synthesizing Diffusion MRI from CT Perfusion
by: Lee, Junhyeok, et al.
Published: (2025)
by: Lee, Junhyeok, et al.
Published: (2025)
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
DA-Mamba: Learning Domain-Aware State Space Model for Global-Local Alignment in Domain Adaptive Object Detection
by: Li, Haochen, et al.
Published: (2026)
by: Li, Haochen, et al.
Published: (2026)
Posterior Distillation Sampling
by: Koo, Juil, et al.
Published: (2023)
by: Koo, Juil, et al.
Published: (2023)
Dynamic Full-body Motion Agent with Object Interaction via Blending Pre-trained Modular Controllers
by: Nam, Sanghyeok, et al.
Published: (2026)
by: Nam, Sanghyeok, et al.
Published: (2026)
Joint Learning of Pose Regression and Denoising Diffusion with Score Scaling Sampling for Category-level 6D Pose Estimation
by: Lee, Seunghyun, et al.
Published: (2025)
by: Lee, Seunghyun, et al.
Published: (2025)
Beyond General Prompts: Automated Prompt Refinement using Contrastive Class Alignment Scores for Disambiguating Objects in Vision-Language Models
by: Choi, Lucas, et al.
Published: (2025)
by: Choi, Lucas, et al.
Published: (2025)
PICCOLO: Point Cloud-Centric Omnidirectional Localization
by: Kim, Junho, et al.
Published: (2021)
by: Kim, Junho, et al.
Published: (2021)
CPO: Change Robust Panorama to Point Cloud Localization
by: Kim, Junho, et al.
Published: (2022)
by: Kim, Junho, et al.
Published: (2022)
FRED: Towards a Full Rotation-Equivariance in Aerial Image Object Detection
by: Lee, Chanho, et al.
Published: (2023)
by: Lee, Chanho, et al.
Published: (2023)
3Doodle: Compact Abstraction of Objects with 3D Strokes
by: Choi, Changwoon, et al.
Published: (2024)
by: Choi, Changwoon, et al.
Published: (2024)
Leveraging Image Augmentation for Object Manipulation: Towards Interpretable Controllability in Object-Centric Learning
by: Kim, Jinwoo, et al.
Published: (2023)
by: Kim, Jinwoo, et al.
Published: (2023)
Spherical Linear Interpolation and Text-Anchoring for Zero-shot Composed Image Retrieval
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
by: He, Bo, et al.
Published: (2024)
by: He, Bo, et al.
Published: (2024)
Dual Mutual Learning Network with Global-local Awareness for RGB-D Salient Object Detection
by: Yi, Kang, et al.
Published: (2025)
by: Yi, Kang, et al.
Published: (2025)
Similar Items
-
FAST-GOAL: Fast and Efficient Global-local Object Alignment Learning
by: Choi, Hyungyu, et al.
Published: (2026) -
CoVA: Text-Guided Composed Video Retrieval for Audio-Visual Content
by: Han, Gyuwon, et al.
Published: (2026) -
DiCo: Disentangled Concept Representation for Text-to-image Person Re-identification
by: Kim, Giyeol, et al.
Published: (2026) -
Visual Representation Alignment for Multimodal Large Language Models
by: Yoon, Heeji, et al.
Published: (2025) -
Domain Generalization for Person Re-identification: A Survey Towards Domain-Agnostic Person Matching
by: Lee, Hyeonseo, et al.
Published: (2025)