MultiMax: Sparse and Multi-Modal Attention Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Yuxuan, Fritz, Mario, Keuper, Margret |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MaxSup: Overcoming Representation Collapse in Label Smoothing
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
Balancing Diversity and Risk in LLM Sampling: How to Select Your Method and Parameter for Open-Ended Text Generation
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2024)
How Do Training Methods Influence the Utilization of Vision Models?
von: Gavrikov, Paul, et al.
Veröffentlicht: (2024)
von: Gavrikov, Paul, et al.
Veröffentlicht: (2024)
Orion-MSP: Multi-Scale Sparse Attention for Tabular In-Context Learning
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
Divide & Bind Your Attention for Improved Generative Semantic Nursing
von: Li, Yumeng, et al.
Veröffentlicht: (2023)
von: Li, Yumeng, et al.
Veröffentlicht: (2023)
A Multi-Modal CNN-LSTM Framework with Multi-Head Attention and Focal Loss for Real-Time Elderly Fall Detection
von: Zhou, Lijie, et al.
Veröffentlicht: (2026)
von: Zhou, Lijie, et al.
Veröffentlicht: (2026)
Emotion and Intention Guided Multi-Modal Learning for Sticker Response Selection
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
Multi-Modal Molecular Representation Learning via Structure Awareness
von: Yin, Rong, et al.
Veröffentlicht: (2025)
von: Yin, Rong, et al.
Veröffentlicht: (2025)
Self-Tuning Sparse Attention: Multi-Fidelity Hyperparameter Optimization for Transformer Acceleration
von: Dev, Arundhathi, et al.
Veröffentlicht: (2026)
von: Dev, Arundhathi, et al.
Veröffentlicht: (2026)
SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention
von: Xu, Hongtao, et al.
Veröffentlicht: (2026)
von: Xu, Hongtao, et al.
Veröffentlicht: (2026)
Adversarial Supervision Makes Layout-to-Image Diffusion Models Thrive
von: Li, Yumeng, et al.
Veröffentlicht: (2024)
von: Li, Yumeng, et al.
Veröffentlicht: (2024)
HISA: Efficient Hierarchical Indexing for Fine-Grained Sparse Attention
von: Xu, Yufei, et al.
Veröffentlicht: (2026)
von: Xu, Yufei, et al.
Veröffentlicht: (2026)
Learning Multi-Level Features with Matryoshka Sparse Autoencoders
von: Bussmann, Bart, et al.
Veröffentlicht: (2025)
von: Bussmann, Bart, et al.
Veröffentlicht: (2025)
A Concept-Centric Approach to Multi-Modality Learning
von: Geng, Yuchong, et al.
Veröffentlicht: (2024)
von: Geng, Yuchong, et al.
Veröffentlicht: (2024)
Generalized Neighborhood Attention: Multi-dimensional Sparse Attention at the Speed of Light
von: Hassani, Ali, et al.
Veröffentlicht: (2025)
von: Hassani, Ali, et al.
Veröffentlicht: (2025)
Detecting Scarce and Sparse Anomalous: Solving Dual Imbalance in Multi-Instance Learning
von: Jia, Lin-Han, et al.
Veröffentlicht: (2025)
von: Jia, Lin-Han, et al.
Veröffentlicht: (2025)
Multi-Objective Reinforcement Learning with Max-Min Criterion: A Game-Theoretic Approach
von: Byeon, Woohyeon, et al.
Veröffentlicht: (2025)
von: Byeon, Woohyeon, et al.
Veröffentlicht: (2025)
ProtCLIP: Function-Informed Protein Multi-Modal Learning
von: Zhou, Hanjing, et al.
Veröffentlicht: (2024)
von: Zhou, Hanjing, et al.
Veröffentlicht: (2024)
Representation Learning with Mutual Influence of Modalities for Node Classification in Multi-Modal Heterogeneous Networks
von: Li, Jiafan, et al.
Veröffentlicht: (2025)
von: Li, Jiafan, et al.
Veröffentlicht: (2025)
Multi-Modal Manipulation via Multi-Modal Policy Consensus
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
Generative Modeling of Class Probability for Multi-Modal Representation Learning
von: Shin, Jungkyoo, et al.
Veröffentlicht: (2025)
von: Shin, Jungkyoo, et al.
Veröffentlicht: (2025)
GraphT5: Unified Molecular Graph-Language Modeling via Multi-Modal Cross-Token Attention
von: Kim, Sangyeup, et al.
Veröffentlicht: (2025)
von: Kim, Sangyeup, et al.
Veröffentlicht: (2025)
vAttention: Verified Sparse Attention
von: Desai, Aditya, et al.
Veröffentlicht: (2025)
von: Desai, Aditya, et al.
Veröffentlicht: (2025)
Multi-Modal Federated Learning for Cancer Staging over Non-IID Datasets with Unbalanced Modalities
von: Borazjani, Kasra, et al.
Veröffentlicht: (2024)
von: Borazjani, Kasra, et al.
Veröffentlicht: (2024)
Multi-Granular Attention based Heterogeneous Hypergraph Neural Network
von: Jin, Hong, et al.
Veröffentlicht: (2025)
von: Jin, Hong, et al.
Veröffentlicht: (2025)
Multi-Objective Optimization for Sparse Deep Multi-Task Learning
von: Hotegni, S. S., et al.
Veröffentlicht: (2023)
von: Hotegni, S. S., et al.
Veröffentlicht: (2023)
Attention Mechanism, Max-Affine Partition, and Universal Approximation
von: Liu, Hude, et al.
Veröffentlicht: (2025)
von: Liu, Hude, et al.
Veröffentlicht: (2025)
Progressive Sparse Attention: Algorithm and System Co-design for Efficient Attention in LLM Serving
von: Zhou, Qihui, et al.
Veröffentlicht: (2025)
von: Zhou, Qihui, et al.
Veröffentlicht: (2025)
TAP: Two-Stage Adaptive Personalization of Multi-Task and Multi-Modal Foundation Models in Federated Learning
von: Lee, Seohyun, et al.
Veröffentlicht: (2025)
von: Lee, Seohyun, et al.
Veröffentlicht: (2025)
The Max-Min Formulation of Multi-Objective Reinforcement Learning: From Theory to a Model-Free Algorithm
von: Park, Giseung, et al.
Veröffentlicht: (2024)
von: Park, Giseung, et al.
Veröffentlicht: (2024)
Stem: Rethinking Causal Information Flow in Sparse Attention
von: Niu, Lin, et al.
Veröffentlicht: (2026)
von: Niu, Lin, et al.
Veröffentlicht: (2026)
Natively Trainable Sparse Attention for Hierarchical Point Cloud Datasets
von: Lapautre, Nicolas, et al.
Veröffentlicht: (2025)
von: Lapautre, Nicolas, et al.
Veröffentlicht: (2025)
Multi-Modality Collaborative Learning for Sentiment Analysis
von: Wang, Shanmin, et al.
Veröffentlicht: (2025)
von: Wang, Shanmin, et al.
Veröffentlicht: (2025)
Learning Cell-Aware Hierarchical Multi-Modal Representations for Robust Molecular Modeling
von: Li, Mengran, et al.
Veröffentlicht: (2025)
von: Li, Mengran, et al.
Veröffentlicht: (2025)
CAML: Collaborative Auxiliary Modality Learning for Multi-Agent Systems
von: Liu, Rui, et al.
Veröffentlicht: (2025)
von: Liu, Rui, et al.
Veröffentlicht: (2025)
Discrete Diffusion for Complex and Congested Multi-Agent Path Finding with Sparse Social Attention
von: Wang, Yuanzhe, et al.
Veröffentlicht: (2026)
von: Wang, Yuanzhe, et al.
Veröffentlicht: (2026)
Multi-Modal AI for Remote Patient Monitoring in Cancer Care
von: Liu, Yansong, et al.
Veröffentlicht: (2025)
von: Liu, Yansong, et al.
Veröffentlicht: (2025)
Multi-head Temporal Latent Attention
von: Deng, Keqi, et al.
Veröffentlicht: (2025)
von: Deng, Keqi, et al.
Veröffentlicht: (2025)
SeerAttention-R: Sparse Attention Adaptation for Long Reasoning
von: Gao, Yizhao, et al.
Veröffentlicht: (2025)
von: Gao, Yizhao, et al.
Veröffentlicht: (2025)
A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models
von: Lin, Zihao, et al.
Veröffentlicht: (2025)
von: Lin, Zihao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MaxSup: Overcoming Representation Collapse in Label Smoothing
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025) -
Balancing Diversity and Risk in LLM Sampling: How to Select Your Method and Parameter for Open-Ended Text Generation
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2024) -
How Do Training Methods Influence the Utilization of Vision Models?
von: Gavrikov, Paul, et al.
Veröffentlicht: (2024) -
Orion-MSP: Multi-Scale Sparse Attention for Tabular In-Context Learning
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025) -
Divide & Bind Your Attention for Improved Generative Semantic Nursing
von: Li, Yumeng, et al.
Veröffentlicht: (2023)