Polynomial Mixing for Efficient Self-supervised Speech Encoders
Fuente:
arXiv
Salvato in:
| Autori principali: | Feillet, Eva, Whetten, Ryan, Picard, David, Allauzen, Alexandre |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Open Implementation and Study of BEST-RQ for Speech Processing
di: Whetten, Ryan, et al.
Pubblicazione: (2024)
di: Whetten, Ryan, et al.
Pubblicazione: (2024)
Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics
di: Colagrande, Alex, et al.
Pubblicazione: (2025)
di: Colagrande, Alex, et al.
Pubblicazione: (2025)
Limits of Resolution Equivariance in Fourier Neural Operators
di: Colagrande, Alex, et al.
Pubblicazione: (2026)
di: Colagrande, Alex, et al.
Pubblicazione: (2026)
Towards Early Prediction of Self-Supervised Speech Model Performance
di: Whetten, Ryan, et al.
Pubblicazione: (2025)
di: Whetten, Ryan, et al.
Pubblicazione: (2025)
Trading Complexity for Expressivity Through Structured Generalized Linear Token Mixing
di: Fagnou, Erwan, et al.
Pubblicazione: (2026)
di: Fagnou, Erwan, et al.
Pubblicazione: (2026)
Chain and Causal Attention for Efficient Entity Tracking
di: Fagnou, Erwan, et al.
Pubblicazione: (2024)
di: Fagnou, Erwan, et al.
Pubblicazione: (2024)
Exploring Precision and Recall to assess the quality and diversity of LLMs
di: Bronnec, Florian Le, et al.
Pubblicazione: (2024)
di: Bronnec, Florian Le, et al.
Pubblicazione: (2024)
Improving Diversity in Language Models: When Temperature Fails, Change the Loss
di: Verine, Alexandre, et al.
Pubblicazione: (2025)
di: Verine, Alexandre, et al.
Pubblicazione: (2025)
Structured-Sparse Attention for Entity Tracking with Subquadratic Sequence Complexity
di: Zhao, Hangyue, et al.
Pubblicazione: (2026)
di: Zhao, Hangyue, et al.
Pubblicazione: (2026)
An Analysis of Linear Complexity Attention Substitutes with BEST-RQ
di: Whetten, Ryan, et al.
Pubblicazione: (2024)
di: Whetten, Ryan, et al.
Pubblicazione: (2024)
TiME: Tiny Monolingual Encoders for Efficient NLP Pipelines
di: Schulmeister, David, et al.
Pubblicazione: (2025)
di: Schulmeister, David, et al.
Pubblicazione: (2025)
Efficient Sample-Specific Encoder Perturbations
di: Fathullah, Yassir, et al.
Pubblicazione: (2024)
di: Fathullah, Yassir, et al.
Pubblicazione: (2024)
Integrating Self-supervised Speech Model with Pseudo Word-level Targets from Visually-grounded Speech Model
di: Fang, Hung-Chieh, et al.
Pubblicazione: (2024)
di: Fang, Hung-Chieh, et al.
Pubblicazione: (2024)
[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic
di: Choi, Kwanghee, et al.
Pubblicazione: (2026)
di: Choi, Kwanghee, et al.
Pubblicazione: (2026)
Patent Representation Learning via Self-supervision
di: Zuo, You, et al.
Pubblicazione: (2025)
di: Zuo, You, et al.
Pubblicazione: (2025)
On Affine Homotopy between Language Encoders
di: Chan, Robin SM, et al.
Pubblicazione: (2024)
di: Chan, Robin SM, et al.
Pubblicazione: (2024)
On Importance of Code-Mixed Embeddings for Hate Speech Identification
di: Jagdale, Shruti, et al.
Pubblicazione: (2024)
di: Jagdale, Shruti, et al.
Pubblicazione: (2024)
PoM: Efficient Image and Video Generation with the Polynomial Mixer
di: Picard, David, et al.
Pubblicazione: (2024)
di: Picard, David, et al.
Pubblicazione: (2024)
Moonshine v2: Ergodic Streaming Encoder ASR for Latency-Critical Speech Applications
di: Kudlur, Manjunath, et al.
Pubblicazione: (2026)
di: Kudlur, Manjunath, et al.
Pubblicazione: (2026)
LOCOST: State-Space Models for Long Document Abstractive Summarization
di: Bronnec, Florian Le, et al.
Pubblicazione: (2024)
di: Bronnec, Florian Le, et al.
Pubblicazione: (2024)
Where Do Self-Supervised Speech Models Become Unfair?
di: Herron, Felix, et al.
Pubblicazione: (2026)
di: Herron, Felix, et al.
Pubblicazione: (2026)
Coupling Speech Encoders with Downstream Text Models
di: Chelba, Ciprian, et al.
Pubblicazione: (2024)
di: Chelba, Ciprian, et al.
Pubblicazione: (2024)
CVPD at QIAS 2025 Shared Task: An Efficient Encoder-Based Approach for Islamic Inheritance Reasoning
di: Bekhouche, Salah Eddine, et al.
Pubblicazione: (2025)
di: Bekhouche, Salah Eddine, et al.
Pubblicazione: (2025)
A Confidence-based Acquisition Model for Self-supervised Active Learning and Label Correction
di: van Niekerk, Carel, et al.
Pubblicazione: (2023)
di: van Niekerk, Carel, et al.
Pubblicazione: (2023)
Identifying and typifying demographic unfairness in phoneme-level embeddings of self-supervised speech recognition models
di: Herron, Felix, et al.
Pubblicazione: (2026)
di: Herron, Felix, et al.
Pubblicazione: (2026)
SNAP-UQ: Self-supervised Next-Activation Prediction for Single-Pass Uncertainty in TinyML
di: Lamaakal, Ismail, et al.
Pubblicazione: (2025)
di: Lamaakal, Ismail, et al.
Pubblicazione: (2025)
Recommendation of data-free class-incremental learning algorithms by simulating future data
di: Feillet, Eva, et al.
Pubblicazione: (2024)
di: Feillet, Eva, et al.
Pubblicazione: (2024)
Knowledge Graph Reasoning with Self-supervised Reinforcement Learning
di: Ma, Ying, et al.
Pubblicazione: (2024)
di: Ma, Ying, et al.
Pubblicazione: (2024)
Progressive Mixed-Precision Decoding for Efficient LLM Inference
di: Chen, Hao Mark, et al.
Pubblicazione: (2024)
di: Chen, Hao Mark, et al.
Pubblicazione: (2024)
Aligning Attention with Human Rationales for Self-Explaining Hate Speech Detection
di: Eilertsen, Brage, et al.
Pubblicazione: (2025)
di: Eilertsen, Brage, et al.
Pubblicazione: (2025)
Skipformer: A Skip-and-Recover Strategy for Efficient Speech Recognition
di: Zhu, Wenjing, et al.
Pubblicazione: (2024)
di: Zhu, Wenjing, et al.
Pubblicazione: (2024)
Dynamic Encoder Size Based on Data-Driven Layer-wise Pruning for Speech Recognition
di: Xu, Jingjing, et al.
Pubblicazione: (2024)
di: Xu, Jingjing, et al.
Pubblicazione: (2024)
Table-r1: Self-supervised and Reinforcement Learning for Program-based Table Reasoning in Small Language Models
di: Jin, Rihui, et al.
Pubblicazione: (2025)
di: Jin, Rihui, et al.
Pubblicazione: (2025)
Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models
di: De, Soham, et al.
Pubblicazione: (2024)
di: De, Soham, et al.
Pubblicazione: (2024)
Self-supervised learning of speech representations with Dutch archival data
di: Vaessen, Nik, et al.
Pubblicazione: (2025)
di: Vaessen, Nik, et al.
Pubblicazione: (2025)
ENTP: Encoder-only Next Token Prediction
di: Ewer, Ethan, et al.
Pubblicazione: (2024)
di: Ewer, Ethan, et al.
Pubblicazione: (2024)
NuNER: Entity Recognition Encoder Pre-training via LLM-Annotated Data
di: Bogdanov, Sergei, et al.
Pubblicazione: (2024)
di: Bogdanov, Sergei, et al.
Pubblicazione: (2024)
ChatGPT in Linear Algebra: Strides Forward, Steps to Go
di: Bagno, Eli, et al.
Pubblicazione: (2024)
di: Bagno, Eli, et al.
Pubblicazione: (2024)
SCOPE: A Self-supervised Framework for Improving Faithfulness in Conditional Text Generation
di: Duong, Song, et al.
Pubblicazione: (2025)
di: Duong, Song, et al.
Pubblicazione: (2025)
TaCo: Targeted Concept Erasure Prevents Non-Linear Classifiers From Detecting Protected Attributes
di: Jourdan, Fanny, et al.
Pubblicazione: (2023)
di: Jourdan, Fanny, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Open Implementation and Study of BEST-RQ for Speech Processing
di: Whetten, Ryan, et al.
Pubblicazione: (2024) -
Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics
di: Colagrande, Alex, et al.
Pubblicazione: (2025) -
Limits of Resolution Equivariance in Fourier Neural Operators
di: Colagrande, Alex, et al.
Pubblicazione: (2026) -
Towards Early Prediction of Self-Supervised Speech Model Performance
di: Whetten, Ryan, et al.
Pubblicazione: (2025) -
Trading Complexity for Expressivity Through Structured Generalized Linear Token Mixing
di: Fagnou, Erwan, et al.
Pubblicazione: (2026)