Enregistré dans:
| Auteurs principaux: | Harvey, Ethan, Loevlie, Dennis Johan, Hughes, Michael C. |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2605.27306 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Synthetic Data Reveals Generalization Gaps in Correlated Multiple Instance Learning
par: Harvey, Ethan, et autres
Publié: (2025)
par: Harvey, Ethan, et autres
Publié: (2025)
A Multi-Dataset Benchmark of Multiple Instance Learning for 3D Neuroimage Classification
par: Harvey, Ethan, et autres
Publié: (2026)
par: Harvey, Ethan, et autres
Publié: (2026)
Occam's Razor is Only as Sharp as Your ELBO
par: Harvey, Ethan, et autres
Publié: (2026)
par: Harvey, Ethan, et autres
Publié: (2026)
Transfer Learning with Informative Priors: Simple Baselines Better than Previously Reported
par: Harvey, Ethan, et autres
Publié: (2024)
par: Harvey, Ethan, et autres
Publié: (2024)
Learning Hyperparameters via a Data-Emphasized Variational Objective
par: Harvey, Ethan, et autres
Publié: (2025)
par: Harvey, Ethan, et autres
Publié: (2025)
Learning the Regularization Strength for Deep Fine-Tuning via a Data-Emphasized Variational Objective
par: Harvey, Ethan, et autres
Publié: (2024)
par: Harvey, Ethan, et autres
Publié: (2024)
Diverse Sampling in Diffusion Models with Marginal Preserving Particle Guidance
par: Vinograd, Gal, et autres
Publié: (2026)
par: Vinograd, Gal, et autres
Publié: (2026)
But what is your honest answer? Aiding LLM-judges with honest alternatives using steering vectors
par: Eshuijs, Leon, et autres
Publié: (2025)
par: Eshuijs, Leon, et autres
Publié: (2025)
A second order regret bound for NormalHedge
par: Freund, Yoav, et autres
Publié: (2026)
par: Freund, Yoav, et autres
Publié: (2026)
Flexible Tails for Normalizing Flows
par: Hickling, Tennessee, et autres
Publié: (2024)
par: Hickling, Tennessee, et autres
Publié: (2024)
Limitations of Normalization in Attention Mechanism
par: Mudarisov, Timur, et autres
Publié: (2025)
par: Mudarisov, Timur, et autres
Publié: (2025)
Some Attention is All You Need for Retrieval
par: Michalak, Felix, et autres
Publié: (2025)
par: Michalak, Felix, et autres
Publié: (2025)
Attention is All You Need Until You Need Retention
par: Yaslioglu, M. Murat
Publié: (2025)
par: Yaslioglu, M. Murat
Publié: (2025)
On the Normalization of Confusion Matrices: Methods and Geometric Interpretations
par: Erbani, Johan, et autres
Publié: (2025)
par: Erbani, Johan, et autres
Publié: (2025)
You Need Better Attention Priors
par: Litman, Elon, et autres
Publié: (2026)
par: Litman, Elon, et autres
Publié: (2026)
Why GRPO Needs Normalization: A Local-Curvature Perspective on Adaptive Gradients
par: Ge, Cheng, et autres
Publié: (2026)
par: Ge, Cheng, et autres
Publié: (2026)
Attention Is Not What You Need
par: Chong, Zhang
Publié: (2025)
par: Chong, Zhang
Publié: (2025)
Scalable Message Passing Neural Networks: No Need for Attention in Large Graph Representation Learning
par: Borde, Haitz Sáez de Ocáriz, et autres
Publié: (2024)
par: Borde, Haitz Sáez de Ocáriz, et autres
Publié: (2024)
Detecting Heart Disease from Multi-View Ultrasound Images via Supervised Attention Multiple Instance Learning
par: Huang, Zhe, et autres
Publié: (2023)
par: Huang, Zhe, et autres
Publié: (2023)
Element-wise Attention Is All You Need
par: Feng, Guoxin
Publié: (2025)
par: Feng, Guoxin
Publié: (2025)
Linear Memory SE(2) Invariant Attention
par: Pronovost, Ethan, et autres
Publié: (2025)
par: Pronovost, Ethan, et autres
Publié: (2025)
Attention Needs to Focus: A Unified Perspective on Attention Allocation
par: Fu, Zichuan, et autres
Publié: (2026)
par: Fu, Zichuan, et autres
Publié: (2026)
Variational Autoencoder with Normalizing flow for X-ray spectral fitting
par: Redmen, Fiona, et autres
Publié: (2026)
par: Redmen, Fiona, et autres
Publié: (2026)
Attention Is All You Need But You Don't Need All Of It For Inference of Large Language Models
par: Tyukin, Georgy, et autres
Publié: (2024)
par: Tyukin, Georgy, et autres
Publié: (2024)
Guidance is All You Need: Temperature-Guided Reasoning in Large Language Models
par: Gomaa, Eyad, et autres
Publié: (2024)
par: Gomaa, Eyad, et autres
Publié: (2024)
SINCERE: Supervised Information Noise-Contrastive Estimation REvisited
par: Feeney, Patrick, et autres
Publié: (2023)
par: Feeney, Patrick, et autres
Publié: (2023)
Geometry-Aware Attention Guidance for Diffusion Models via Modern Hopfield Dynamics
par: Kim, Kwanyoung
Publié: (2026)
par: Kim, Kwanyoung
Publié: (2026)
Attention Smoothing Is All You Need For Unlearning
par: Zade, Saleh Zare, et autres
Publié: (2026)
par: Zade, Saleh Zare, et autres
Publié: (2026)
Tensor Product Attention Is All You Need
par: Zhang, Yifan, et autres
Publié: (2025)
par: Zhang, Yifan, et autres
Publié: (2025)
What Matters in Transformers? Not All Attention is Needed
par: He, Shwai, et autres
Publié: (2024)
par: He, Shwai, et autres
Publié: (2024)
Attention Once Is All You Need: Efficient Streaming Inference with Stateful Transformers
par: Norgren, Victor
Publié: (2026)
par: Norgren, Victor
Publié: (2026)
Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance
par: Ahn, Donghoon, et autres
Publié: (2024)
par: Ahn, Donghoon, et autres
Publié: (2024)
If generative AI is the answer, what is the question?
par: Tewari, Ambuj
Publié: (2025)
par: Tewari, Ambuj
Publié: (2025)
Improving Prediction of Need for Mechanical Ventilation using Cross-Attention
par: Mohanty, Anwesh, et autres
Publié: (2024)
par: Mohanty, Anwesh, et autres
Publié: (2024)
Does Long-Term Series Forecasting Need Complex Attention and Extra Long Inputs?
par: Liang, Daojun, et autres
Publié: (2023)
par: Liang, Daojun, et autres
Publié: (2023)
Linear Log-Normal Attention with Unbiased Concentration
par: Nahshan, Yury, et autres
Publié: (2023)
par: Nahshan, Yury, et autres
Publié: (2023)
Graph Neural Networks Need Cluster-Normalize-Activate Modules
par: Skryagin, Arseny, et autres
Publié: (2024)
par: Skryagin, Arseny, et autres
Publié: (2024)
Attention-space Contrastive Guidance for Efficient Hallucination Mitigation in LVLMs
par: Jo, Yujin, et autres
Publié: (2026)
par: Jo, Yujin, et autres
Publié: (2026)
TempoControl: Temporal Attention Guidance for Text-to-Video Models
par: Schiber, Shira, et autres
Publié: (2025)
par: Schiber, Shira, et autres
Publié: (2025)
TransMLA: Multi-Head Latent Attention Is All You Need
par: Meng, Fanxu, et autres
Publié: (2025)
par: Meng, Fanxu, et autres
Publié: (2025)
Documents similaires
-
Synthetic Data Reveals Generalization Gaps in Correlated Multiple Instance Learning
par: Harvey, Ethan, et autres
Publié: (2025) -
A Multi-Dataset Benchmark of Multiple Instance Learning for 3D Neuroimage Classification
par: Harvey, Ethan, et autres
Publié: (2026) -
Occam's Razor is Only as Sharp as Your ELBO
par: Harvey, Ethan, et autres
Publié: (2026) -
Transfer Learning with Informative Priors: Simple Baselines Better than Previously Reported
par: Harvey, Ethan, et autres
Publié: (2024) -
Learning Hyperparameters via a Data-Emphasized Variational Objective
par: Harvey, Ethan, et autres
Publié: (2025)