pMoE: Prompting Diverse Experts Together Wins More in Visual Adaptation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Mo, Shentong, Luo, Xufang, Li, Dongsheng |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
A Large-scale Medical Visual Task Adaptation Benchmark
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
LSPT: Long-term Spatial Prompt Tuning for Visual Representation Learning
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
Efficient 3D Shape Generation via Diffusion Mamba with Bidirectional SSMs
par: Mo, Shentong
Publié: (2024)
par: Mo, Shentong
Publié: (2024)
Improving Visual Representation Alignment Generation with GRPO
par: Mo, Shentong, et autres
Publié: (2026)
par: Mo, Shentong, et autres
Publié: (2026)
Scaling Diffusion Mamba with Bidirectional SSMs for Efficient Image and Video Generation
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
LVRPO: Language-Visual Alignment with GRPO for Multimodal Understanding and Generation
par: Mo, Shentong, et autres
Publié: (2026)
par: Mo, Shentong, et autres
Publié: (2026)
Foley-Flow: Coordinated Video-to-Audio Generation with Masked Audio-Visual Alignment and Dynamic Conditional Flows
par: Mo, Shentong, et autres
Publié: (2026)
par: Mo, Shentong, et autres
Publié: (2026)
Aligning Audio-Visual Joint Representations with an Agentic Workflow
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
The Dynamic Duo of Collaborative Masking and Target for Advanced Masked Autoencoder Learning
par: Mo, Shentong
Publié: (2024)
par: Mo, Shentong
Publié: (2024)
GMAIL: Generative Modality Alignment for generated Image Learning
par: Mo, Shentong, et autres
Publié: (2026)
par: Mo, Shentong, et autres
Publié: (2026)
DMT-JEPA: Discriminative Masked Targets for Joint-Embedding Predictive Architecture
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
MultiMed: Massively Multimodal and Multitask Medical Understanding
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
Connecting Joint-Embedding Predictive Architecture with Contrastive Self-supervised Learning
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
GMS-CAVP: Improving Audio-Video Correspondence with Multi-Scale Contrastive and Generative Pretraining
par: Mo, Shentong, et autres
Publié: (2026)
par: Mo, Shentong, et autres
Publié: (2026)
MoETTA: Test-Time Adaptation Under Mixed Distribution Shifts with MoE-LayerNorm
par: Fan, Xiao, et autres
Publié: (2025)
par: Fan, Xiao, et autres
Publié: (2025)
I2MoE: Interpretable Multimodal Interaction-aware Mixture-of-Experts
par: Xin, Jiayi, et autres
Publié: (2025)
par: Xin, Jiayi, et autres
Publié: (2025)
Text-to-Audio Generation Synchronized with Videos
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
MoPE: Mixture of Prompt Experts for Parameter-Efficient and Scalable Multimodal Fusion
par: Jiang, Ruixiang, et autres
Publié: (2024)
par: Jiang, Ruixiang, et autres
Publié: (2024)
MoLE: Enhancing Human-centric Text-to-image Diffusion via Mixture of Low-rank Experts
par: Zhu, Jie, et autres
Publié: (2024)
par: Zhu, Jie, et autres
Publié: (2024)
RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual Learning
par: Hong, Kiseong, et autres
Publié: (2025)
par: Hong, Kiseong, et autres
Publié: (2025)
Uni-Med: A Unified Medical Generalist Foundation Model For Multi-Task Learning Via Connector-MoE
par: Zhu, Xun, et autres
Publié: (2024)
par: Zhu, Xun, et autres
Publié: (2024)
MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering
par: Li, Xu, et autres
Publié: (2025)
par: Li, Xu, et autres
Publié: (2025)
DiffGAP: A Lightweight Diffusion Module in Contrastive Space for Bridging Cross-Model Gap
par: Mo, Shentong, et autres
Publié: (2025)
par: Mo, Shentong, et autres
Publié: (2025)
Unified Video-Language Pre-training with Synchronized Audio
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
MoSE: Skill-by-Skill Mixture-of-Experts Learning for Embodied Autonomous Machines
par: Xu, Lu, et autres
Publié: (2025)
par: Xu, Lu, et autres
Publié: (2025)
PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification
par: Luo, Qiuming, et autres
Publié: (2026)
par: Luo, Qiuming, et autres
Publié: (2026)
MultiIoT: Benchmarking Machine Learning for the Internet of Things
par: Mo, Shentong, et autres
Publié: (2023)
par: Mo, Shentong, et autres
Publié: (2023)
IoT-LM: Large Multisensory Language Models for the Internet of Things
par: Mo, Shentong, et autres
Publié: (2024)
par: Mo, Shentong, et autres
Publié: (2024)
BioFact-MoE: Biologically Factorized Mixture of Experts for Vision-Language Prognostic Modeling in Hepatocellular Carcinoma
par: Yang, Junlin, et autres
Publié: (2026)
par: Yang, Junlin, et autres
Publié: (2026)
Inference Optimal VLMs Need Fewer Visual Tokens and More Parameters
par: Li, Kevin Y., et autres
Publié: (2024)
par: Li, Kevin Y., et autres
Publié: (2024)
MoQE: Improve Quantization Model performance via Mixture of Quantization Experts
par: Zhang, Jinhao, et autres
Publié: (2025)
par: Zhang, Jinhao, et autres
Publié: (2025)
MoEIoU: Rethinking Bounding-Box Regression as a Mixture of Experts
par: Edula, Vinay, et autres
Publié: (2026)
par: Edula, Vinay, et autres
Publié: (2026)
Incorporating Geo-Diverse Knowledge into Prompting for Increased Geographical Robustness in Object Recognition
par: Buettner, Kyle, et autres
Publié: (2024)
par: Buettner, Kyle, et autres
Publié: (2024)
GC-MoE: Genomics-Guided Cell-Type-Specific Mixture of Experts for Histology-Based Single-Cell Spatial Transcriptomics
par: Shiku, Kaito, et autres
Publié: (2026)
par: Shiku, Kaito, et autres
Publié: (2026)
Mixture of Nested Experts: Adaptive Processing of Visual Tokens
par: Jain, Gagan, et autres
Publié: (2024)
par: Jain, Gagan, et autres
Publié: (2024)
FlyPrompt: Brain-Inspired Random-Expanded Routing with Temporal-Ensemble Experts for General Continual Learning
par: Yan, Hongwei, et autres
Publié: (2026)
par: Yan, Hongwei, et autres
Publié: (2026)
CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning
par: Liu, Yang, et autres
Publié: (2026)
par: Liu, Yang, et autres
Publié: (2026)
Is Less More? Exploring Token Condensation as Training-free Test-time Adaptation
par: Wang, Zixin, et autres
Publié: (2024)
par: Wang, Zixin, et autres
Publié: (2024)
BiPrompt: Bilateral Prompt Optimization for Visual and Textual Debiasing in Vision-Language Models
par: Gupta, Sunny, et autres
Publié: (2026)
par: Gupta, Sunny, et autres
Publié: (2026)
Continual Learning: Forget-free Winning Subnetworks for Video Representations
par: Kang, Haeyong, et autres
Publié: (2023)
par: Kang, Haeyong, et autres
Publié: (2023)
Documents similaires
-
A Large-scale Medical Visual Task Adaptation Benchmark
par: Mo, Shentong, et autres
Publié: (2024) -
LSPT: Long-term Spatial Prompt Tuning for Visual Representation Learning
par: Mo, Shentong, et autres
Publié: (2024) -
Efficient 3D Shape Generation via Diffusion Mamba with Bidirectional SSMs
par: Mo, Shentong
Publié: (2024) -
Improving Visual Representation Alignment Generation with GRPO
par: Mo, Shentong, et autres
Publié: (2026) -
Scaling Diffusion Mamba with Bidirectional SSMs for Efficient Image and Video Generation
par: Mo, Shentong, et autres
Publié: (2024)