Long-Tailed Recognition on Binary Networks by Calibrating A Pre-trained Model
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Jihun, Kim, Dahyun, Jung, Hyungrok, Oh, Taeil, Choi, Jonghyun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Training-Free Global Geometric Association for 4D LiDAR Panoptic Segmentation
di: Oh, Gyeongrok, et al.
Pubblicazione: (2025)
di: Oh, Gyeongrok, et al.
Pubblicazione: (2025)
Few-Shot Pattern Detection via Template Matching and Regression
di: Jo, Eunchan, et al.
Pubblicazione: (2025)
di: Jo, Eunchan, et al.
Pubblicazione: (2025)
SyncVSR: Data-Efficient Visual Speech Recognition with End-to-End Crossmodal Audio Token Synchronization
di: Ahn, Young Jin, et al.
Pubblicazione: (2024)
di: Ahn, Young Jin, et al.
Pubblicazione: (2024)
Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality
di: Oh, Youngtaek, et al.
Pubblicazione: (2024)
di: Oh, Youngtaek, et al.
Pubblicazione: (2024)
What Happens When: Learning Temporal Orders of Events in Videos
di: Ahn, Daechul, et al.
Pubblicazione: (2025)
di: Ahn, Daechul, et al.
Pubblicazione: (2025)
Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models
di: Woo, Sangmin, et al.
Pubblicazione: (2024)
di: Woo, Sangmin, et al.
Pubblicazione: (2024)
Adjusting Logit in Gaussian Form for Long-Tailed Visual Recognition
di: Li, Mengke, et al.
Pubblicazione: (2023)
di: Li, Mengke, et al.
Pubblicazione: (2023)
Towards Seamless Adaptation of Pre-trained Models for Visual Place Recognition
di: Lu, Feng, et al.
Pubblicazione: (2024)
di: Lu, Feng, et al.
Pubblicazione: (2024)
Contribution-based Low-Rank Adaptation with Pre-training Model for Real Image Restoration
di: Park, Donwon, et al.
Pubblicazione: (2024)
di: Park, Donwon, et al.
Pubblicazione: (2024)
MF-LPR$^2$: Multi-Frame License Plate Image Restoration and Recognition using Optical Flow
di: Na, Kihyun, et al.
Pubblicazione: (2025)
di: Na, Kihyun, et al.
Pubblicazione: (2025)
MAFA: Managing False Negatives for Vision-Language Pre-training
di: Byun, Jaeseok, et al.
Pubblicazione: (2023)
di: Byun, Jaeseok, et al.
Pubblicazione: (2023)
Advancing ALS Applications with Large-Scale Pre-training: Dataset Development and Downstream Assessment
di: Xiu, Haoyi, et al.
Pubblicazione: (2025)
di: Xiu, Haoyi, et al.
Pubblicazione: (2025)
M4CXR: Exploring Multi-task Potentials of Multi-modal Large Language Models for Chest X-ray Interpretation
di: Park, Jonggwon, et al.
Pubblicazione: (2024)
di: Park, Jonggwon, et al.
Pubblicazione: (2024)
TTA-DAME: Test-Time Adaptation with Domain Augmentation and Model Ensemble for Dynamic Driving Conditions
di: Jeon, Dongjae, et al.
Pubblicazione: (2025)
di: Jeon, Dongjae, et al.
Pubblicazione: (2025)
Online Generic Event Boundary Detection
di: Jung, Hyungrok, et al.
Pubblicazione: (2025)
di: Jung, Hyungrok, et al.
Pubblicazione: (2025)
Distribution-Aware Robust Learning from Long-Tailed Data with Noisy Labels
di: Baik, Jae Soon, et al.
Pubblicazione: (2024)
di: Baik, Jae Soon, et al.
Pubblicazione: (2024)
Revisiting Residual Connections: Orthogonal Updates for Stable and Efficient Deep Networks
di: Oh, Giyeong, et al.
Pubblicazione: (2025)
di: Oh, Giyeong, et al.
Pubblicazione: (2025)
Rethinking Classifier Re-Training in Long-Tailed Recognition: A Simple Logits Retargeting Approach
di: Lu, Han, et al.
Pubblicazione: (2024)
di: Lu, Han, et al.
Pubblicazione: (2024)
Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models
di: Jung, Mingi, et al.
Pubblicazione: (2025)
di: Jung, Mingi, et al.
Pubblicazione: (2025)
Multi-Level Knowledge Distillation and Dynamic Self-Supervised Learning for Continual Learning
di: Kim, Taeheon, et al.
Pubblicazione: (2025)
di: Kim, Taeheon, et al.
Pubblicazione: (2025)
Exploring Ordinal Bias in Action Recognition for Instructional Videos
di: Kim, Joochan, et al.
Pubblicazione: (2025)
di: Kim, Joochan, et al.
Pubblicazione: (2025)
Why Not Hyperparameter-Friendly Optimisation? A Monotonic Adaptive Norm Rescaling Approach For Long-Tailed Recognition
di: Zhang, Shuo, et al.
Pubblicazione: (2026)
di: Zhang, Shuo, et al.
Pubblicazione: (2026)
Safeguard Text-to-Image Diffusion Models with Human Feedback Inversion
di: Kim, Sanghyun, et al.
Pubblicazione: (2024)
di: Kim, Sanghyun, et al.
Pubblicazione: (2024)
RGB-Event HyperGraph Prompt for Kilometer Marker Recognition based on Pre-trained Foundation Models
di: Xian, Xiaoyu, et al.
Pubblicazione: (2026)
di: Xian, Xiaoyu, et al.
Pubblicazione: (2026)
Bridging the Missing-Modality Gap: Improving Text-Only Calibration of Vision Language Models
di: Kim, Mingyeong, et al.
Pubblicazione: (2026)
di: Kim, Mingyeong, et al.
Pubblicazione: (2026)
Towards Calibrated Robust Fine-Tuning of Vision-Language Models
di: Oh, Changdae, et al.
Pubblicazione: (2023)
di: Oh, Changdae, et al.
Pubblicazione: (2023)
VLMine: Long-Tail Data Mining with Vision Language Models
di: Ye, Mao, et al.
Pubblicazione: (2024)
di: Ye, Mao, et al.
Pubblicazione: (2024)
ReaMIL: Reasoning- and Evidence-Aware Multiple Instance Learning for Whole-Slide Histopathology
di: Jung, Hyun Do, et al.
Pubblicazione: (2026)
di: Jung, Hyun Do, et al.
Pubblicazione: (2026)
LANTERN: Accelerating Visual Autoregressive Models with Relaxed Speculative Decoding
di: Jang, Doohyuk, et al.
Pubblicazione: (2024)
di: Jang, Doohyuk, et al.
Pubblicazione: (2024)
See and Fix the Flaws: Enabling VLMs and Diffusion Models to Comprehend Visual Artifacts via Agentic Data Synthesis
di: Park, Jaehyun, et al.
Pubblicazione: (2026)
di: Park, Jaehyun, et al.
Pubblicazione: (2026)
MMeViT: Multi-Modal ensemble ViT for Post-Stroke Rehabilitation Action Recognition
di: Kim, Ye-eun, et al.
Pubblicazione: (2025)
di: Kim, Ye-eun, et al.
Pubblicazione: (2025)
Prior2Posterior: Model Prior Correction for Long-Tailed Learning
di: Bhat, S Divakar, et al.
Pubblicazione: (2024)
di: Bhat, S Divakar, et al.
Pubblicazione: (2024)
Task Prototype-Based Knowledge Retrieval for Multi-Task Learning from Partially Annotated Data
di: Oh, Youngmin, et al.
Pubblicazione: (2026)
di: Oh, Youngmin, et al.
Pubblicazione: (2026)
Pre-Deployment Robustness Stress Testing for CT Segmentation Systems Using Clinically Motivated Multi-Corruption Augmentation
di: Kang, CholMin, et al.
Pubblicazione: (2026)
di: Kang, CholMin, et al.
Pubblicazione: (2026)
Long-Tailed Learning for Generalized Category Discovery
di: Hoang, Cuong Manh
Pubblicazione: (2025)
di: Hoang, Cuong Manh
Pubblicazione: (2025)
Finetuning Pre-trained Model with Limited Data for LiDAR-based 3D Object Detection by Bridging Domain Gaps
di: Jang, Jiyun, et al.
Pubblicazione: (2024)
di: Jang, Jiyun, et al.
Pubblicazione: (2024)
Budgeted Online Continual Learning by Adaptive Layer Freezing and Frequency-based Sampling
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
Spatio-Temporal Side Tuning Pre-trained Foundation Models for Video-based Pedestrian Attribute Recognition
di: Wang, Xiao, et al.
Pubblicazione: (2024)
di: Wang, Xiao, et al.
Pubblicazione: (2024)
Heavy-Tailed Class-Conditional Priors for Long-Tailed Generative Modeling
di: Bouayed, Aymene Mohammed, et al.
Pubblicazione: (2025)
di: Bouayed, Aymene Mohammed, et al.
Pubblicazione: (2025)
Are Compact Rationales Free? Measuring Tile Selection Headroom in Frozen WSI-MIL
di: Jung, Hyun Do, et al.
Pubblicazione: (2026)
di: Jung, Hyun Do, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Training-Free Global Geometric Association for 4D LiDAR Panoptic Segmentation
di: Oh, Gyeongrok, et al.
Pubblicazione: (2025) -
Few-Shot Pattern Detection via Template Matching and Regression
di: Jo, Eunchan, et al.
Pubblicazione: (2025) -
SyncVSR: Data-Efficient Visual Speech Recognition with End-to-End Crossmodal Audio Token Synchronization
di: Ahn, Young Jin, et al.
Pubblicazione: (2024) -
Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality
di: Oh, Youngtaek, et al.
Pubblicazione: (2024) -
What Happens When: Learning Temporal Orders of Events in Videos
di: Ahn, Daechul, et al.
Pubblicazione: (2025)