GMGaze: MoE-Based Context-Aware Gaze Estimation with CLIP and Multiscale Transformer
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhao, Xinyuan, Wu, Yihang, Chaddad, Ahmad, Alkhodair, Sarah A., Kateb, Reem |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
GazeFormer-MoE: Context-Aware Gaze Estimation via CLIP and MoE Transformer
par: Zhao, Xinyuan, et autres
Publié: (2026)
par: Zhao, Xinyuan, et autres
Publié: (2026)
Federated Vision Transformer with Adaptive Focal Loss for Medical Image Classification
par: Zhao, Xinyuan, et autres
Publié: (2026)
par: Zhao, Xinyuan, et autres
Publié: (2026)
Deep Modeling and Optimization of Medical Image Classification
par: Wu, Yihang, et autres
Publié: (2025)
par: Wu, Yihang, et autres
Publié: (2025)
Semi-Supervised Medical Image Segmentation via Dual Networks
par: Lu, Yunyao, et autres
Publié: (2025)
par: Lu, Yunyao, et autres
Publié: (2025)
Domain Adaptation Techniques for Natural and Medical Image Classification
par: Chaddad, Ahmad, et autres
Publié: (2025)
par: Chaddad, Ahmad, et autres
Publié: (2025)
Enhancing Dual Network Based Semi-Supervised Medical Image Segmentation with Uncertainty-Guided Pseudo-Labeling
par: Lu, Yunyao, et autres
Publié: (2025)
par: Lu, Yunyao, et autres
Publié: (2025)
Federated CLIP for Resource-Efficient Heterogeneous Medical Image Classification
par: Wu, Yihang, et autres
Publié: (2025)
par: Wu, Yihang, et autres
Publié: (2025)
Generalizable and Explainable Deep Learning for Medical Image Computing: An Overview
par: Chaddad, Ahmad, et autres
Publié: (2025)
par: Chaddad, Ahmad, et autres
Publié: (2025)
Impact of domain adaptation in deep learning for medical image classifications
par: Wu, Yihang, et autres
Publié: (2026)
par: Wu, Yihang, et autres
Publié: (2026)
FACMIC: Federated Adaptative CLIP Model for Medical Image Classification
par: Wu, Yihang, et autres
Publié: (2024)
par: Wu, Yihang, et autres
Publié: (2024)
Deep Modeling and Interpretation for Bladder Cancer Classification
par: Chaddad, Ahmad, et autres
Publié: (2026)
par: Chaddad, Ahmad, et autres
Publié: (2026)
FAA-CLIP: Federated Adversarial Adaptation of CLIP
par: Wu, Yihang, et autres
Publié: (2025)
par: Wu, Yihang, et autres
Publié: (2025)
Classification based deep learning models for lung cancer and disease using medical images
par: Chaddad, Ahmad, et autres
Publié: (2025)
par: Chaddad, Ahmad, et autres
Publié: (2025)
Towards a Transparent and Interpretable AI Model for Medical Image Classifications
par: Wen, Binbin, et autres
Publié: (2025)
par: Wen, Binbin, et autres
Publié: (2025)
MiM-DiT: MoE in MoE with Diffusion Transformers for All-in-One Image Restoration
par: Kong, Lingshun, et autres
Publié: (2026)
par: Kong, Lingshun, et autres
Publié: (2026)
LongScape: Advancing Long-Horizon Embodied World Models with Context-Aware MoE
par: Shang, Yu, et autres
Publié: (2025)
par: Shang, Yu, et autres
Publié: (2025)
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
par: Zhang, Jihai, et autres
Publié: (2024)
par: Zhang, Jihai, et autres
Publié: (2024)
Dense2MoE: Restructuring Diffusion Transformer to MoE for Efficient Text-to-Image Generation
par: Zheng, Youwei, et autres
Publié: (2025)
par: Zheng, Youwei, et autres
Publié: (2025)
GazeCLIP: Enhancing Gaze Estimation Through Text-Guided Multimodal Learning
par: Wang, Jun, et autres
Publié: (2023)
par: Wang, Jun, et autres
Publié: (2023)
CLIP-Gaze: Towards General Gaze Estimation via Visual-Linguistic Model
par: Yin, Pengwei, et autres
Publié: (2024)
par: Yin, Pengwei, et autres
Publié: (2024)
GNN-MoE: Context-Aware Patch Routing using GNNs for Parameter-Efficient Domain Generalization
par: Soliman, Mahmoud, et autres
Publié: (2025)
par: Soliman, Mahmoud, et autres
Publié: (2025)
ARGaze: Autoregressive Transformers for Online Egocentric Gaze Estimation
par: Li, Jia, et autres
Publié: (2026)
par: Li, Jia, et autres
Publié: (2026)
GA3CE: Unconstrained 3D Gaze Estimation with Gaze-Aware 3D Context Encoding
par: Kawana, Yuki, et autres
Publié: (2025)
par: Kawana, Yuki, et autres
Publié: (2025)
BIG-MoE: Bypass Isolated Gating MoE for Generalized Multimodal Face Anti-Spoofing
par: Ma, Yingjie, et autres
Publié: (2024)
par: Ma, Yingjie, et autres
Publié: (2024)
GazeCLIP: Gaze-Guided CLIP with Adaptive-Enhanced Fine-Grained Language Prompt for Deepfake Attribution and Detection
par: Zhang, Yaning, et autres
Publié: (2026)
par: Zhang, Yaning, et autres
Publié: (2026)
eMoE-Tracker: Environmental MoE-based Transformer for Robust Event-guided Object Tracking
par: Chen, Yucheng, et autres
Publié: (2024)
par: Chen, Yucheng, et autres
Publié: (2024)
Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing Guidance
par: Wei, Yujie, et autres
Publié: (2025)
par: Wei, Yujie, et autres
Publié: (2025)
GazeD: Context-Aware Diffusion for Accurate 3D Gaze Estimation
par: Catalini, Riccardo, et autres
Publié: (2026)
par: Catalini, Riccardo, et autres
Publié: (2026)
MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation
par: Maldonado, Gabriel, et autres
Publié: (2025)
par: Maldonado, Gabriel, et autres
Publié: (2025)
Guiding the Experts: Semantic Priors for Efficient and Focused MoE Routing
par: Min, Chengxi, et autres
Publié: (2025)
par: Min, Chengxi, et autres
Publié: (2025)
GazeMoE: Perception of Gaze Target with Mixture-of-Experts
par: Dai, Zhuangzhuang, et autres
Publié: (2026)
par: Dai, Zhuangzhuang, et autres
Publié: (2026)
More Is Better: A MoE-Based Emotion Recognition Framework with Human Preference Alignment
par: Xie, Jun, et autres
Publié: (2025)
par: Xie, Jun, et autres
Publié: (2025)
Nucleus-Image: Sparse MoE for Image Generation
par: Akiti, Chandan, et autres
Publié: (2026)
par: Akiti, Chandan, et autres
Publié: (2026)
IMA-MoE: An Interpretable Modality-Aware Mixture-of-Experts Framework for Characterizing the Neurobiological Signatures of Binge Eating Disorder
par: Zhao, Lin, et autres
Publié: (2026)
par: Zhao, Lin, et autres
Publié: (2026)
SGAP-Gaze: Scene Grid Attention Based Point-of-Gaze Estimation Network for Driver Gaze
par: Sharma, Pavan Kumar, et autres
Publié: (2026)
par: Sharma, Pavan Kumar, et autres
Publié: (2026)
TAG-MoE: Task-Aware Gating for Unified Generative Mixture-of-Experts
par: Xu, Yu, et autres
Publié: (2026)
par: Xu, Yu, et autres
Publié: (2026)
Merging Multiple Datasets for Improved Appearance-Based Gaze Estimation
par: Wu, Liang, et autres
Publié: (2024)
par: Wu, Liang, et autres
Publié: (2024)
SHAP-Integrated Convolutional Diagnostic Networks for Feature-Selective Medical Analysis
par: Hu, Yan, et autres
Publié: (2025)
par: Hu, Yan, et autres
Publié: (2025)
MoE-GS: Mixture of Experts for Dynamic Gaussian Splatting
par: Jin, In-Hwan, et autres
Publié: (2025)
par: Jin, In-Hwan, et autres
Publié: (2025)
SMoES: Soft Modality-Guided Expert Specialization in MoE-VLMs
par: Bo, Zi-Hao, et autres
Publié: (2026)
par: Bo, Zi-Hao, et autres
Publié: (2026)
Documents similaires
-
GazeFormer-MoE: Context-Aware Gaze Estimation via CLIP and MoE Transformer
par: Zhao, Xinyuan, et autres
Publié: (2026) -
Federated Vision Transformer with Adaptive Focal Loss for Medical Image Classification
par: Zhao, Xinyuan, et autres
Publié: (2026) -
Deep Modeling and Optimization of Medical Image Classification
par: Wu, Yihang, et autres
Publié: (2025) -
Semi-Supervised Medical Image Segmentation via Dual Networks
par: Lu, Yunyao, et autres
Publié: (2025) -
Domain Adaptation Techniques for Natural and Medical Image Classification
par: Chaddad, Ahmad, et autres
Publié: (2025)