Divide, Weight, and Route: Difficulty-Aware Optimization with Dynamic Expert Fusion for Long-tailed Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wei, Xiaolei, Ouyang, Yi, Ye, Haibo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Difficulty-aware Balancing Margin Loss for Long-tailed Recognition
von: Son, Minseok, et al.
Veröffentlicht: (2024)
von: Son, Minseok, et al.
Veröffentlicht: (2024)
Latent-based Diffusion Model for Long-tailed Recognition
von: Han, Pengxiao, et al.
Veröffentlicht: (2024)
von: Han, Pengxiao, et al.
Veröffentlicht: (2024)
Semantic Data Augmentation for Long-tailed Facial Expression Recognition
von: Li, Zijian, et al.
Veröffentlicht: (2024)
von: Li, Zijian, et al.
Veröffentlicht: (2024)
Simultaneous Long-tailed Recognition and Multi-modal Fusion for Highly Imbalanced Multi-modal Data
von: Yoon, Heegeon, et al.
Veröffentlicht: (2026)
von: Yoon, Heegeon, et al.
Veröffentlicht: (2026)
Uncertainty-aware Long-tailed Weights Model the Utility of Pseudo-labels for Semi-supervised Learning
von: Wu, Jiaqi, et al.
Veröffentlicht: (2025)
von: Wu, Jiaqi, et al.
Veröffentlicht: (2025)
Divide-and-Conquer Inference for Large-Scale Visual Recognition with Multimodal Large Language Models
von: Ye, Zhipeng, et al.
Veröffentlicht: (2026)
von: Ye, Zhipeng, et al.
Veröffentlicht: (2026)
Metis-HOME: Hybrid Optimized Mixture-of-Experts for Multimodal Reasoning
von: Lan, Xiaohan, et al.
Veröffentlicht: (2025)
von: Lan, Xiaohan, et al.
Veröffentlicht: (2025)
SalientFusion: Context-Aware Compositional Zero-Shot Food Recognition
von: Song, Jiajun, et al.
Veröffentlicht: (2025)
von: Song, Jiajun, et al.
Veröffentlicht: (2025)
Ordering Matters: Rank-Aware Selective Fusion for Blended Emotion Recognition
von: Lee, Junghyun, et al.
Veröffentlicht: (2026)
von: Lee, Junghyun, et al.
Veröffentlicht: (2026)
Local and Global Feature Attention Fusion Network for Face Recognition
von: Yu, Wang, et al.
Veröffentlicht: (2024)
von: Yu, Wang, et al.
Veröffentlicht: (2024)
Dataset Awareness is not Enough: Implementing Sample-level Tail Encouragement in Long-tailed Self-supervised Learning
von: Xiao, Haowen, et al.
Veröffentlicht: (2024)
von: Xiao, Haowen, et al.
Veröffentlicht: (2024)
Soft Task-Aware Routing of Experts for Equivariant Representation Learning
von: Jeon, Jaebyeong, et al.
Veröffentlicht: (2025)
von: Jeon, Jaebyeong, et al.
Veröffentlicht: (2025)
Spatio-Semantic Expert Routing Architecture with Mixture-of-Experts for Referring Image Segmentation
von: Dalaq, Alaa, et al.
Veröffentlicht: (2026)
von: Dalaq, Alaa, et al.
Veröffentlicht: (2026)
Robust Embodied Perception in Dynamic Environments via Disentangled Weight Fusion
von: Guo, Juncen, et al.
Veröffentlicht: (2026)
von: Guo, Juncen, et al.
Veröffentlicht: (2026)
Temporal and Spatial Feature Fusion Framework for Dynamic Micro Expression Recognition
von: Liu, Feng, et al.
Veröffentlicht: (2025)
von: Liu, Feng, et al.
Veröffentlicht: (2025)
Decision Boundary-aware Generation for Long-tailed Learning
von: Yang, Jiacheng, et al.
Veröffentlicht: (2026)
von: Yang, Jiacheng, et al.
Veröffentlicht: (2026)
CoGR-MoE: Concept-Guided Expert Routing with Consistent Selection and Flexible Reasoning for Visual Question Answering
von: Zeng, Xiyin, et al.
Veröffentlicht: (2026)
von: Zeng, Xiyin, et al.
Veröffentlicht: (2026)
CAViT -- Channel-Aware Vision Transformer for Dynamic Feature Fusion
von: Safdar, Aon, et al.
Veröffentlicht: (2026)
von: Safdar, Aon, et al.
Veröffentlicht: (2026)
PAT: Pixel-wise Adaptive Training for Long-tailed Segmentation
von: Do, Khoi, et al.
Veröffentlicht: (2024)
von: Do, Khoi, et al.
Veröffentlicht: (2024)
Staircase Cascaded Fusion of Lightweight Local Pattern Recognition and Long-Range Dependencies for Structural Crack Segmentation
von: Liu, Hui, et al.
Veröffentlicht: (2024)
von: Liu, Hui, et al.
Veröffentlicht: (2024)
DREAM: Dynamic Retinal Enhancement with Adaptive Multi-modal Fusion for Expert Precision Medical Report Generation
von: Shaik, Nagur Shareef, et al.
Veröffentlicht: (2026)
von: Shaik, Nagur Shareef, et al.
Veröffentlicht: (2026)
DirMixE: Harnessing Test Agnostic Long-tail Recognition with Hierarchical Label Vartiations
von: Yang, Zhiyong, et al.
Veröffentlicht: (2024)
von: Yang, Zhiyong, et al.
Veröffentlicht: (2024)
Towards Realistic Long-tailed Semi-supervised Learning in an Open World
von: He, Yuanpeng, et al.
Veröffentlicht: (2024)
von: He, Yuanpeng, et al.
Veröffentlicht: (2024)
Divide-then-Diagnose: Weaving Clinician-Inspired Contexts for Ultra-Long Capsule Endoscopy Videos
von: Liu, Bowen, et al.
Veröffentlicht: (2026)
von: Liu, Bowen, et al.
Veröffentlicht: (2026)
Fire on Motion: Optimizing Video Pass-bands for Efficient Spiking Action Recognition
von: Ye, Shuhan, et al.
Veröffentlicht: (2026)
von: Ye, Shuhan, et al.
Veröffentlicht: (2026)
DySink: Dynamic Frame Sinks for Autoregressive Long Video Generation
von: Ye, Bo, et al.
Veröffentlicht: (2026)
von: Ye, Bo, et al.
Veröffentlicht: (2026)
Global Semantic-Guided Sub-image Feature Weight Allocation in High-Resolution Large Vision-Language Models
von: Liang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Liang, Yuxuan, et al.
Veröffentlicht: (2025)
Multi-Site Class-Incremental Learning with Weighted Experts in Echocardiography
von: Bransby, Kit M., et al.
Veröffentlicht: (2024)
von: Bransby, Kit M., et al.
Veröffentlicht: (2024)
Hierarchical Semantic-Visual Fusion of Visible and Near-infrared Images for Long-range Haze Removal
von: Li, Yi, et al.
Veröffentlicht: (2025)
von: Li, Yi, et al.
Veröffentlicht: (2025)
MDDFNet: Mamba-based Dynamic Dual Fusion Network for Traffic Sign Detection
von: Yu, TianYi
Veröffentlicht: (2025)
von: Yu, TianYi
Veröffentlicht: (2025)
Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts
von: Xu, Haolei, et al.
Veröffentlicht: (2026)
von: Xu, Haolei, et al.
Veröffentlicht: (2026)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
Divide and Conquer: Static-Dynamic Collaboration for Few-Shot Class-Incremental Learning
von: Bao, Kexin, et al.
Veröffentlicht: (2026)
von: Bao, Kexin, et al.
Veröffentlicht: (2026)
MGHFT: Multi-Granularity Hierarchical Fusion Transformer for Cross-Modal Sticker Emotion Recognition
von: Chen, Jian, et al.
Veröffentlicht: (2025)
von: Chen, Jian, et al.
Veröffentlicht: (2025)
Feature Fusion for Human Activity Recognition using Parameter-Optimized Multi-Stage Graph Convolutional Network and Transformer Models
von: Belal, Mohammad, et al.
Veröffentlicht: (2024)
von: Belal, Mohammad, et al.
Veröffentlicht: (2024)
Difficulty-guided Sampling: Bridging the Target Gap between Dataset Distillation and Downstream Tasks
von: Li, Mingzhuo, et al.
Veröffentlicht: (2026)
von: Li, Mingzhuo, et al.
Veröffentlicht: (2026)
EmoNet-Face: An Expert-Annotated Benchmark for Synthetic Emotion Recognition
von: Schuhmann, Christoph, et al.
Veröffentlicht: (2025)
von: Schuhmann, Christoph, et al.
Veröffentlicht: (2025)
DESign: Dynamic Context-Aware Convolution and Efficient Subnet Regularization for Continuous Sign Language Recognition
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception
von: Li, Xiaotong, et al.
Veröffentlicht: (2024)
von: Li, Xiaotong, et al.
Veröffentlicht: (2024)
FRAME: Forensic Routing and Adaptive Multi-path Evidence Fusion for Image Manipulation Detection
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2026)
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Difficulty-aware Balancing Margin Loss for Long-tailed Recognition
von: Son, Minseok, et al.
Veröffentlicht: (2024) -
Latent-based Diffusion Model for Long-tailed Recognition
von: Han, Pengxiao, et al.
Veröffentlicht: (2024) -
Semantic Data Augmentation for Long-tailed Facial Expression Recognition
von: Li, Zijian, et al.
Veröffentlicht: (2024) -
Simultaneous Long-tailed Recognition and Multi-modal Fusion for Highly Imbalanced Multi-modal Data
von: Yoon, Heegeon, et al.
Veröffentlicht: (2026) -
Uncertainty-aware Long-tailed Weights Model the Utility of Pseudo-labels for Semi-supervised Learning
von: Wu, Jiaqi, et al.
Veröffentlicht: (2025)