Guardado en:
| Autores principales: | Mazaré, Pierre-Emmanuel, Szilvasy, Gergely, Lomeli, Maria, Massa, Francisco, Murray, Naila, Jégou, Hervé, Douze, Matthijs |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2502.08246 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility
por: Szilvasy, Gergely, et al.
Publicado: (2026)
por: Szilvasy, Gergely, et al.
Publicado: (2026)
Vector search with small radiuses
por: Szilvasy, Gergely, et al.
Publicado: (2024)
por: Szilvasy, Gergely, et al.
Publicado: (2024)
Short window attention enables long-term memorization
por: Cabannes, Loïc, et al.
Publicado: (2025)
por: Cabannes, Loïc, et al.
Publicado: (2025)
The Faiss library
por: Douze, Matthijs, et al.
Publicado: (2024)
por: Douze, Matthijs, et al.
Publicado: (2024)
Stochastic activations
por: Lomeli, Maria, et al.
Publicado: (2025)
por: Lomeli, Maria, et al.
Publicado: (2025)
Functional Invariants to Watermark Large Transformers
por: Fernandez, Pierre, et al.
Publicado: (2023)
por: Fernandez, Pierre, et al.
Publicado: (2023)
Evaluation data contamination in LLMs: how do we measure it and (when) does it matter?
por: Singh, Aaditya K., et al.
Publicado: (2024)
por: Singh, Aaditya K., et al.
Publicado: (2024)
Moshi: a speech-text foundation model for real-time dialogue
por: Défossez, Alexandre, et al.
Publicado: (2024)
por: Défossez, Alexandre, et al.
Publicado: (2024)
Watermarking Makes Language Models Radioactive
por: Sander, Tom, et al.
Publicado: (2024)
por: Sander, Tom, et al.
Publicado: (2024)
RA-DIT: Retrieval-Augmented Dual Instruction Tuning
por: Lin, Xi Victoria, et al.
Publicado: (2023)
por: Lin, Xi Victoria, et al.
Publicado: (2023)
Machine learning and high dimensional vector search
por: Douze, Matthijs
Publicado: (2025)
por: Douze, Matthijs
Publicado: (2025)
Neutral Residues: Revisiting Adapters for Model Extension
por: Talla, Franck Signe, et al.
Publicado: (2024)
por: Talla, Franck Signe, et al.
Publicado: (2024)
In-context Pretraining: Language Modeling Beyond Document Boundaries
por: Shi, Weijia, et al.
Publicado: (2023)
por: Shi, Weijia, et al.
Publicado: (2023)
Watermark Anything with Localized Messages
por: Sander, Tom, et al.
Publicado: (2024)
por: Sander, Tom, et al.
Publicado: (2024)
KVzap: Fast, Adaptive, and Faithful KV Cache Pruning
por: Jegou, Simon, et al.
Publicado: (2026)
por: Jegou, Simon, et al.
Publicado: (2026)
MagicPIG: LSH Sampling for Efficient LLM Generation
por: Chen, Zhuoming, et al.
Publicado: (2024)
por: Chen, Zhuoming, et al.
Publicado: (2024)
Expected Attention: KV Cache Compression by Estimating Attention from Future Queries Distribution
por: Devoto, Alessio, et al.
Publicado: (2025)
por: Devoto, Alessio, et al.
Publicado: (2025)
Verifying Chain-of-Thought Reasoning via Its Computational Graph
por: Zhao, Zheng, et al.
Publicado: (2025)
por: Zhao, Zheng, et al.
Publicado: (2025)
Aligning Spoken Dialogue Models from User Interactions
por: Wu, Anne, et al.
Publicado: (2025)
por: Wu, Anne, et al.
Publicado: (2025)
High-Fidelity Simultaneous Speech-To-Speech Translation
por: Labiausse, Tom, et al.
Publicado: (2025)
por: Labiausse, Tom, et al.
Publicado: (2025)
Syntax and Semantics of Linear Dependent Types
por: Vákár, Matthijs
Publicado: (2014)
por: Vákár, Matthijs
Publicado: (2014)
Multi-head attention debiasing and contrastive learning for mitigating Dataset Artifacts in Natural Language Inference
por: Sivakoti, Karthik
Publicado: (2024)
por: Sivakoti, Karthik
Publicado: (2024)
DeFT: Decoding with Flash Tree-attention for Efficient Tree-structured LLM Inference
por: Yao, Jinwei, et al.
Publicado: (2024)
por: Yao, Jinwei, et al.
Publicado: (2024)
CHAD: Combinatory Homomorphic Automatic Differentiation
por: Vákár, Matthijs, et al.
Publicado: (2021)
por: Vákár, Matthijs, et al.
Publicado: (2021)
DELULU: Discriminative Embedding Learning Using Latent Units for Speaker-Aware Self-Trained Speech Foundational Model
por: Baali, Massa, et al.
Publicado: (2025)
por: Baali, Massa, et al.
Publicado: (2025)
SVeritas: Benchmark for Robust Speaker Verification under Diverse Conditions
por: Baali, Massa, et al.
Publicado: (2025)
por: Baali, Massa, et al.
Publicado: (2025)
Self-attention vector output similarities reveal how machines pay attention
por: Halevi, Tal, et al.
Publicado: (2025)
por: Halevi, Tal, et al.
Publicado: (2025)
Streaming Sequence-to-Sequence Learning with Delayed Streams Modeling
por: Zeghidour, Neil, et al.
Publicado: (2025)
por: Zeghidour, Neil, et al.
Publicado: (2025)
Simulating Training Data Leakage in Multiple-Choice Benchmarks for LLM Evaluation
por: Hidayat, Naila Shafirni, et al.
Publicado: (2025)
por: Hidayat, Naila Shafirni, et al.
Publicado: (2025)
Illuminating Blind Spots of Language Models with Targeted Agent-in-the-Loop Synthetic Data
por: Lippmann, Philip, et al.
Publicado: (2024)
por: Lippmann, Philip, et al.
Publicado: (2024)
TOOLVERIFIER: Generalization to New Tools via Self-Verification
por: Mekala, Dheeraj, et al.
Publicado: (2024)
por: Mekala, Dheeraj, et al.
Publicado: (2024)
Let your LLM generate a few tokens and you will reduce the need for retrieval
por: Déjean, Hervé
Publicado: (2024)
por: Déjean, Hervé
Publicado: (2024)
What and When to Learn: CURriculum Ranking Loss for Large-Scale Speaker Verification
por: Baali, Massa, et al.
Publicado: (2026)
por: Baali, Massa, et al.
Publicado: (2026)
Higher Order Automatic Differentiation of Higher Order Functions
por: Huot, Mathieu, et al.
Publicado: (2021)
por: Huot, Mathieu, et al.
Publicado: (2021)
Winning Amazon KDD Cup'24
por: Deotte, Chris, et al.
Publicado: (2024)
por: Deotte, Chris, et al.
Publicado: (2024)
Truthful Text Sanitization Guided by Inference Attacks
por: Pilán, Ildikó, et al.
Publicado: (2024)
por: Pilán, Ildikó, et al.
Publicado: (2024)
Qinco2: Vector Compression and Search with Improved Implicit Neural Codebooks
por: Vallaeys, Théophane, et al.
Publicado: (2025)
por: Vallaeys, Théophane, et al.
Publicado: (2025)
Visualizing attention zones in machine reading comprehension models
por: Cui, Yiming, et al.
Publicado: (2024)
por: Cui, Yiming, et al.
Publicado: (2024)
Sentiment analysis with adaptive multi-head attention in Transformer
por: Meng, Fanfei, et al.
Publicado: (2023)
por: Meng, Fanfei, et al.
Publicado: (2023)
Relational inductive biases on attention mechanisms
por: Mijangos, Víctor, et al.
Publicado: (2025)
por: Mijangos, Víctor, et al.
Publicado: (2025)
Ejemplares similares
-
Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility
por: Szilvasy, Gergely, et al.
Publicado: (2026) -
Vector search with small radiuses
por: Szilvasy, Gergely, et al.
Publicado: (2024) -
Short window attention enables long-term memorization
por: Cabannes, Loïc, et al.
Publicado: (2025) -
The Faiss library
por: Douze, Matthijs, et al.
Publicado: (2024) -
Stochastic activations
por: Lomeli, Maria, et al.
Publicado: (2025)