Improving Visual Recognition with Hyperbolical Visual Hierarchy Mapping
Fuente:
arXiv
Saved in:
| Main Authors: | Kwon, Hyeongjun, Jang, Jinhyun, Kim, Jin, Kim, Kwonyoung, Sohn, Kwanghoon |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Faster Parameter-Efficient Tuning with Token Redundancy Reduction
by: Kim, Kwonyoung, et al.
Published: (2025)
by: Kim, Kwonyoung, et al.
Published: (2025)
Descriptive Image-Text Matching with Graded Contextual Similarity
by: Jang, Jinhyun, et al.
Published: (2025)
by: Jang, Jinhyun, et al.
Published: (2025)
Enhancing Source-Free Domain Adaptive Object Detection with Low-confidence Pseudo Label Distillation
by: Yoon, Ilhoon, et al.
Published: (2024)
by: Yoon, Ilhoon, et al.
Published: (2024)
PointFix: Learning to Fix Domain Bias for Robust Online Stereo Adaptation
by: Kim, Kwonyoung, et al.
Published: (2022)
by: Kim, Kwonyoung, et al.
Published: (2022)
HypeVPR: Exploring Hyperbolic Space for Perspective to Equirectangular Visual Place Recognition
by: Woo, Suhan, et al.
Published: (2025)
by: Woo, Suhan, et al.
Published: (2025)
HyNeuralMap: Hyperbolic Mapping of Visual Semantics to Neural Hierarchies
by: Ma, Zihan, et al.
Published: (2026)
by: Ma, Zihan, et al.
Published: (2026)
Diffusion-driven GAN Inversion for Multi-Modal Face Image Generation
by: Kim, Jihyun, et al.
Published: (2024)
by: Kim, Jihyun, et al.
Published: (2024)
Learning Visual Hierarchies in Hyperbolic Space for Image Retrieval
by: Wang, Ziwei, et al.
Published: (2024)
by: Wang, Ziwei, et al.
Published: (2024)
Rethinking Open-World Semi-Supervised Learning: Distribution Mismatch and Inductive Inference
by: Park, Seongheon, et al.
Published: (2024)
by: Park, Seongheon, et al.
Published: (2024)
Learning What To Hear: Boosting Sound-Source Association For Robust Audiovisual Instance Segmentation
by: Seo, Jinbae, et al.
Published: (2025)
by: Seo, Jinbae, et al.
Published: (2025)
Bridging Vision and Language Spaces with Assignment Prediction
by: Park, Jungin, et al.
Published: (2024)
by: Park, Jungin, et al.
Published: (2024)
EBDM: Exemplar-guided Image Translation with Brownian-bridge Diffusion Models
by: Lee, Eungbean, et al.
Published: (2024)
by: Lee, Eungbean, et al.
Published: (2024)
Saliency-Aware Model Merging
by: Park, Jungin, et al.
Published: (2026)
by: Park, Jungin, et al.
Published: (2026)
Language-guided Recursive Spatiotemporal Graph Modeling for Video Summarization
by: Park, Jungin, et al.
Published: (2025)
by: Park, Jungin, et al.
Published: (2025)
Bootstrap Your Own Views: Masked Ego-Exo Modeling for Fine-grained View-invariant Video Representations
by: Park, Jungin, et al.
Published: (2025)
by: Park, Jungin, et al.
Published: (2025)
V-LynX: Token Interface Alignment for Video+X LLMs
by: Park, Jungin, et al.
Published: (2026)
by: Park, Jungin, et al.
Published: (2026)
H2G: Hierarchy-Aware Hyperbolic Grouping for 3D Scenes
by: Ko, ByungHa, et al.
Published: (2026)
by: Ko, ByungHa, et al.
Published: (2026)
OpenBox: Annotate Any Bounding Boxes in 3D
by: Lee, In-Jae, et al.
Published: (2025)
by: Lee, In-Jae, et al.
Published: (2025)
EV-CLIP: Efficient Visual Prompt Adaptation for CLIP in Few-shot Action Recognition under Visual Challenges
by: Jon, Hyo Jin, et al.
Published: (2026)
by: Jon, Hyo Jin, et al.
Published: (2026)
SEA: Evaluating Sketch Abstraction Efficiency via Element-level Commonsense Visual Question Answering
by: Park, Jiho, et al.
Published: (2026)
by: Park, Jiho, et al.
Published: (2026)
Context-Based Visual-Language Place Recognition
by: Woo, Soojin, et al.
Published: (2024)
by: Woo, Soojin, et al.
Published: (2024)
Distillation Improves Visual Place Recognition for Low Quality Images
by: Yang, Anbang, et al.
Published: (2023)
by: Yang, Anbang, et al.
Published: (2023)
SpatialBoost: Enhancing Visual Representation through Language-Guided Reasoning
by: Jeon, Byungwoo, et al.
Published: (2026)
by: Jeon, Byungwoo, et al.
Published: (2026)
First Logit Boosting: Visual Grounding Method to Mitigate Object Hallucination in Large Vision-Language Models
by: Ha, Jiwoo, et al.
Published: (2026)
by: Ha, Jiwoo, et al.
Published: (2026)
Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model
by: Shi, Yuheng, et al.
Published: (2024)
by: Shi, Yuheng, et al.
Published: (2024)
AQuA: Toward Strategic Response Generation for Ambiguous Visual Questions
by: Jang, Jihyoung, et al.
Published: (2026)
by: Jang, Jihyoung, et al.
Published: (2026)
Improving Visual Prompt Tuning by Gaussian Neighborhood Minimization for Long-Tailed Visual Recognition
by: Li, Mengke, et al.
Published: (2024)
by: Li, Mengke, et al.
Published: (2024)
VOLoc: Visual Place Recognition by Querying Compressed Lidar Map
by: Cai, Xudong, et al.
Published: (2024)
by: Cai, Xudong, et al.
Published: (2024)
UniSpector: Towards Universal Open-set Defect Recognition via Spectral-Contrastive Visual Prompting
by: Kim, Geonuk, et al.
Published: (2026)
by: Kim, Geonuk, et al.
Published: (2026)
Hyperbolic Metric Learning for Visual Outlier Detection
by: Gonzalez-Jimenez, Alvaro, et al.
Published: (2024)
by: Gonzalez-Jimenez, Alvaro, et al.
Published: (2024)
I$^2$-SLAM: Inverting Imaging Process for Robust Photorealistic Dense SLAM
by: Bae, Gwangtak, et al.
Published: (2024)
by: Bae, Gwangtak, et al.
Published: (2024)
Breaking the Visual Shortcuts in Multimodal Knowledge-Based Visual Question Answering
by: Lee, Dosung, et al.
Published: (2025)
by: Lee, Dosung, et al.
Published: (2025)
Enhancing Feature Tracking Reliability for Visual Navigation using Real-Time Safety Filter
by: Kim, Dabin, et al.
Published: (2025)
by: Kim, Dabin, et al.
Published: (2025)
Automatic Map Density Selection for Locally-Performant Visual Place Recognition
by: Hussaini, Somayeh, et al.
Published: (2026)
by: Hussaini, Somayeh, et al.
Published: (2026)
Towards Test-time Efficient Visual Place Recognition via Asymmetric Query Processing
by: Kim, Jaeyoon, et al.
Published: (2025)
by: Kim, Jaeyoon, et al.
Published: (2025)
Improving Visual Place Recognition with Sequence-Matching Receptiveness Prediction
by: Hussaini, Somayeh, et al.
Published: (2025)
by: Hussaini, Somayeh, et al.
Published: (2025)
Landmark Guided Visual Feature Extractor for Visual Speech Recognition with Limited Resource
by: Yang, Lei, et al.
Published: (2025)
by: Yang, Lei, et al.
Published: (2025)
GreenEye: Development of Real-Time Traffic Signal Recognition System for Visual Impairments
by: Kim, Danu
Published: (2024)
by: Kim, Danu
Published: (2024)
Finer: Investigating and Enhancing Fine-Grained Visual Concept Recognition in Large Vision Language Models
by: Kim, Jeonghwan, et al.
Published: (2024)
by: Kim, Jeonghwan, et al.
Published: (2024)
Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding
by: Kim, Jiwan, et al.
Published: (2026)
by: Kim, Jiwan, et al.
Published: (2026)
Similar Items
-
Faster Parameter-Efficient Tuning with Token Redundancy Reduction
by: Kim, Kwonyoung, et al.
Published: (2025) -
Descriptive Image-Text Matching with Graded Contextual Similarity
by: Jang, Jinhyun, et al.
Published: (2025) -
Enhancing Source-Free Domain Adaptive Object Detection with Low-confidence Pseudo Label Distillation
by: Yoon, Ilhoon, et al.
Published: (2024) -
PointFix: Learning to Fix Domain Bias for Robust Online Stereo Adaptation
by: Kim, Kwonyoung, et al.
Published: (2022) -
HypeVPR: Exploring Hyperbolic Space for Perspective to Equirectangular Visual Place Recognition
by: Woo, Suhan, et al.
Published: (2025)