Catch-Up Mix: Catch-Up Class for Struggling Filters in CNN
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kang, Minsoo, Kang, Minkoo, Kim, Suhyun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PixelBytes: Catching Unified Representation for Multimodal Generation
von: Furfaro, Fabien
Veröffentlicht: (2024)
von: Furfaro, Fabien
Veröffentlicht: (2024)
PixelBytes: Catching Unified Embedding for Multimodal Generation
von: Furfaro, Fabien
Veröffentlicht: (2024)
von: Furfaro, Fabien
Veröffentlicht: (2024)
Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight Detection
von: Um, Sung Jin, et al.
Veröffentlicht: (2025)
von: Um, Sung Jin, et al.
Veröffentlicht: (2025)
Adaptively Sampling-Reusing-Mixing Decomposed Gradients to Speed Up Sharpness Aware Minimization
von: Deng, Jiaxin, et al.
Veröffentlicht: (2025)
von: Deng, Jiaxin, et al.
Veröffentlicht: (2025)
Catch Me If You Can Describe Me: Open-Vocabulary Camouflaged Instance Segmentation with Diffusion
von: Vu, Tuan-Anh, et al.
Veröffentlicht: (2023)
von: Vu, Tuan-Anh, et al.
Veröffentlicht: (2023)
CatchBackdoor: Backdoor Detection via Critical Trojan Neural Path Fuzzing
von: Jin, Haibo, et al.
Veröffentlicht: (2021)
von: Jin, Haibo, et al.
Veröffentlicht: (2021)
Task-Specific Preconditioner for Cross-Domain Few-Shot Learning
von: Kang, Suhyun, et al.
Veröffentlicht: (2024)
von: Kang, Suhyun, et al.
Veröffentlicht: (2024)
Completely Weakly Supervised Class-Incremental Learning for Semantic Segmentation
von: Kim, David Minkwan, et al.
Veröffentlicht: (2025)
von: Kim, David Minkwan, et al.
Veröffentlicht: (2025)
Catch-Up Distillation: You Only Need to Train Once for Accelerating Sampling
von: Shao, Shitong, et al.
Veröffentlicht: (2023)
von: Shao, Shitong, et al.
Veröffentlicht: (2023)
The Effects of Mixed Sample Data Augmentation are Class Dependent
von: Lee, Haeil, et al.
Veröffentlicht: (2023)
von: Lee, Haeil, et al.
Veröffentlicht: (2023)
Generalized Class Discovery in Instance Segmentation
von: Hoang, Cuong Manh, et al.
Veröffentlicht: (2025)
von: Hoang, Cuong Manh, et al.
Veröffentlicht: (2025)
Jailbreaking on Text-to-Video Models via Scene Splitting Strategy
von: Lee, Wonjun, et al.
Veröffentlicht: (2025)
von: Lee, Wonjun, et al.
Veröffentlicht: (2025)
Upsample Guidance: Scale Up Diffusion Models without Training
von: Hwang, Juno, et al.
Veröffentlicht: (2024)
von: Hwang, Juno, et al.
Veröffentlicht: (2024)
TokenUnify: Scaling Up Autoregressive Pretraining for Neuron Segmentation
von: Chen, Yinda, et al.
Veröffentlicht: (2024)
von: Chen, Yinda, et al.
Veröffentlicht: (2024)
Face-MakeUp: Multimodal Facial Prompts for Text-to-Image Generation
von: Dai, Dawei, et al.
Veröffentlicht: (2025)
von: Dai, Dawei, et al.
Veröffentlicht: (2025)
Semi-Supervised Audio-Visual Video Action Recognition with Audio Source Localization Guided Mixup
von: Kang, Seokun, et al.
Veröffentlicht: (2025)
von: Kang, Seokun, et al.
Veröffentlicht: (2025)
CL3DOR: Contrastive Learning for 3D Large Multimodal Models via Odds Ratio on High-Resolution Point Clouds
von: Kim, Keonwoo, et al.
Veröffentlicht: (2025)
von: Kim, Keonwoo, et al.
Veröffentlicht: (2025)
No Thing, Nothing: Highlighting Safety-Critical Classes for Robust LiDAR Semantic Segmentation in Adverse Weather
von: Park, Junsung, et al.
Veröffentlicht: (2025)
von: Park, Junsung, et al.
Veröffentlicht: (2025)
LLaVA-CKD: Bottom-Up Cascaded Knowledge Distillation for Vision-Language Models
von: Gkalelis, Nikolaos, et al.
Veröffentlicht: (2026)
von: Gkalelis, Nikolaos, et al.
Veröffentlicht: (2026)
VLMs Trace Without Tracking: Diagnosing Failures in Visual Path Following
von: Hong, Hyesoo, et al.
Veröffentlicht: (2026)
von: Hong, Hyesoo, et al.
Veröffentlicht: (2026)
Two Birds, One Projection: Harmonizing Safety and Utility in LVLMs via Inference-time Feature Projection
von: Han, Yewon, et al.
Veröffentlicht: (2026)
von: Han, Yewon, et al.
Veröffentlicht: (2026)
DanQing: An Up-to-Date Large-Scale Chinese Vision-Language Pre-training Dataset
von: Shen, Hengyu, et al.
Veröffentlicht: (2026)
von: Shen, Hengyu, et al.
Veröffentlicht: (2026)
Do Vision-Language Models Measure Up? Benchmarking Visual Measurement Reading with MeasureBench
von: Lin, Fenfen, et al.
Veröffentlicht: (2025)
von: Lin, Fenfen, et al.
Veröffentlicht: (2025)
REPrune: Channel Pruning via Kernel Representative Selection
von: Park, Mincheol, et al.
Veröffentlicht: (2024)
von: Park, Mincheol, et al.
Veröffentlicht: (2024)
Tele-Catch: Adaptive Teleoperation for Dexterous Dynamic 3D Object Catching
von: Zhao, Weiguang, et al.
Veröffentlicht: (2026)
von: Zhao, Weiguang, et al.
Veröffentlicht: (2026)
Training-Free Restoration of Pruned Neural Networks
von: Lee, Keonho, et al.
Veröffentlicht: (2025)
von: Lee, Keonho, et al.
Veröffentlicht: (2025)
Advancing Medical Image Segmentation: Morphology-Driven Learning with Diffusion Transformer
von: Kang, Sungmin, et al.
Veröffentlicht: (2024)
von: Kang, Sungmin, et al.
Veröffentlicht: (2024)
CNN2GNN: How to Bridge CNN with GNN
von: Jiao, Ziheng, et al.
Veröffentlicht: (2024)
von: Jiao, Ziheng, et al.
Veröffentlicht: (2024)
Modern Deep Learning Approaches for Cricket Shot Classification: A Comprehensive Baseline Study
von: Kang, Sungwoo
Veröffentlicht: (2025)
von: Kang, Sungwoo
Veröffentlicht: (2025)
Doctoral Thesis: Geometric Deep Learning For Camera Pose Prediction, Registration, Depth Estimation, and 3D Reconstruction
von: Kang, Xueyang
Veröffentlicht: (2025)
von: Kang, Xueyang
Veröffentlicht: (2025)
Improving Interpretability and Accuracy in Neuro-Symbolic Rule Extraction Using Class-Specific Sparse Filters
von: Padalkar, Parth, et al.
Veröffentlicht: (2025)
von: Padalkar, Parth, et al.
Veröffentlicht: (2025)
Close-up-GS: Enhancing Close-Up View Synthesis in 3D Gaussian Splatting with Progressive Self-Training
von: Xia, Jiatong, et al.
Veröffentlicht: (2025)
von: Xia, Jiatong, et al.
Veröffentlicht: (2025)
Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models
von: Kim, Keuntae, et al.
Veröffentlicht: (2026)
von: Kim, Keuntae, et al.
Veröffentlicht: (2026)
A Comparative Study of Adversarial Robustness in CNN and CNN-ANFIS Architectures
von: Shankar, Kaaustaaub, et al.
Veröffentlicht: (2026)
von: Shankar, Kaaustaaub, et al.
Veröffentlicht: (2026)
See What You Are Told: Visual Attention Sink in Large Multimodal Models
von: Kang, Seil, et al.
Veröffentlicht: (2025)
von: Kang, Seil, et al.
Veröffentlicht: (2025)
Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding
von: Kang, Seil, et al.
Veröffentlicht: (2025)
von: Kang, Seil, et al.
Veröffentlicht: (2025)
Breaking Down and Building Up: Mixture of Skill-Based Vision-and-Language Navigation Agents
von: Ma, Tianyi, et al.
Veröffentlicht: (2025)
von: Ma, Tianyi, et al.
Veröffentlicht: (2025)
Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search
von: Lai, Xin, et al.
Veröffentlicht: (2025)
von: Lai, Xin, et al.
Veröffentlicht: (2025)
SearchLVLMs: A Plug-and-Play Framework for Augmenting Large Vision-Language Models by Searching Up-to-Date Internet Knowledge
von: Li, Chuanhao, et al.
Veröffentlicht: (2024)
von: Li, Chuanhao, et al.
Veröffentlicht: (2024)
Why Do Vision Language Models Struggle To Recognize Human Emotions?
von: Agarwal, Madhav, et al.
Veröffentlicht: (2026)
von: Agarwal, Madhav, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PixelBytes: Catching Unified Representation for Multimodal Generation
von: Furfaro, Fabien
Veröffentlicht: (2024) -
PixelBytes: Catching Unified Embedding for Multimodal Generation
von: Furfaro, Fabien
Veröffentlicht: (2024) -
Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight Detection
von: Um, Sung Jin, et al.
Veröffentlicht: (2025) -
Adaptively Sampling-Reusing-Mixing Decomposed Gradients to Speed Up Sharpness Aware Minimization
von: Deng, Jiaxin, et al.
Veröffentlicht: (2025) -
Catch Me If You Can Describe Me: Open-Vocabulary Camouflaged Instance Segmentation with Diffusion
von: Vu, Tuan-Anh, et al.
Veröffentlicht: (2023)