WorldSpeech: A Multilingual Speech Corpus from Around the World
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Asonitis, Antonis, Lanzendörfer, Luca A., Berdoz, Frédéric, Wattenhofer, Roger |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
High-Fidelity Speech Enhancement via Discrete Audio Tokens
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
EuroSpeech: A Multilingual Speech Corpus
von: Pfisterer, Samuel, et al.
Veröffentlicht: (2025)
von: Pfisterer, Samuel, et al.
Veröffentlicht: (2025)
Text-to-Scene with Large Reasoning Models
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
Alignment-Aware Decoding
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
N-vium: Mixture-of-Exits Transformer for Accelerated Exact Generation
von: Lorenc, Aleksander, et al.
Veröffentlicht: (2026)
von: Lorenc, Aleksander, et al.
Veröffentlicht: (2026)
Can an AI Agent Safely Run a Government? Existence of Probably Approximately Aligned Policies
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2024)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2024)
Reasoning Boosts Opinion Alignment in LLMs
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2026)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2026)
PUZZLES: A Benchmark for Neural Algorithmic Reasoning
von: Estermann, Benjamin, et al.
Veröffentlicht: (2024)
von: Estermann, Benjamin, et al.
Veröffentlicht: (2024)
Parametric Neural Amp Modeling with Active Learning
von: Grötschla, Florian, et al.
Veröffentlicht: (2025)
von: Grötschla, Florian, et al.
Veröffentlicht: (2025)
GHaLIB: A Multilingual Framework for Hope Speech Detection in Low-Resource Languages
von: Abdullah, Ahmed, et al.
Veröffentlicht: (2025)
von: Abdullah, Ahmed, et al.
Veröffentlicht: (2025)
WolBanking77: Wolof Banking Speech Intent Classification Dataset
von: Kandji, Abdou Karim, et al.
Veröffentlicht: (2025)
von: Kandji, Abdou Karim, et al.
Veröffentlicht: (2025)
Multilingual Hate Speech Detection in Social Media Using Translation-Based Approaches with Large Language Models
von: Usman, Muhammad, et al.
Veröffentlicht: (2025)
von: Usman, Muhammad, et al.
Veröffentlicht: (2025)
An Investigation of Large Language Models for Real-World Hate Speech Detection
von: Guo, Keyan, et al.
Veröffentlicht: (2024)
von: Guo, Keyan, et al.
Veröffentlicht: (2024)
Towards Generalizable Generic Harmful Speech Datasets for Implicit Hate Speech Detection
von: Almohaimeed, Saad, et al.
Veröffentlicht: (2025)
von: Almohaimeed, Saad, et al.
Veröffentlicht: (2025)
OWLS: Scaling Laws for Multilingual Speech Recognition and Translation Models
von: Chen, William, et al.
Veröffentlicht: (2025)
von: Chen, William, et al.
Veröffentlicht: (2025)
Improving Direct Persian-English Speech-to-Speech Translation with Discrete Units and Synthetic Parallel Data
von: Rashidi, Sina, et al.
Veröffentlicht: (2025)
von: Rashidi, Sina, et al.
Veröffentlicht: (2025)
EUROPA: A Legal Multilingual Keyphrase Generation Dataset
von: Salaün, Olivier, et al.
Veröffentlicht: (2024)
von: Salaün, Olivier, et al.
Veröffentlicht: (2024)
Enhancing Speech Instruction Understanding and Disambiguation in Robotics via Speech Prosody
von: Sasu, David, et al.
Veröffentlicht: (2025)
von: Sasu, David, et al.
Veröffentlicht: (2025)
Elements of World Knowledge (EWoK): A Cognition-Inspired Framework for Evaluating Basic World Knowledge in Language Models
von: Ivanova, Anna A., et al.
Veröffentlicht: (2024)
von: Ivanova, Anna A., et al.
Veröffentlicht: (2024)
WavSLM: Single-Stream Speech Language Modeling via WavLM Distillation
von: Della Libera, Luca, et al.
Veröffentlicht: (2026)
von: Della Libera, Luca, et al.
Veröffentlicht: (2026)
Understanding World or Predicting Future? A Comprehensive Survey of World Models
von: Ding, Jingtao, et al.
Veröffentlicht: (2024)
von: Ding, Jingtao, et al.
Veröffentlicht: (2024)
Analyzing and Fine-Tuning Whisper Models for Multilingual Pilot Speech Transcription in the Cockpit
von: Nareddy, Kartheek Kumar Reddy, et al.
Veröffentlicht: (2025)
von: Nareddy, Kartheek Kumar Reddy, et al.
Veröffentlicht: (2025)
HalluWorld: A Controlled Benchmark for Hallucination via Reference World Models
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
Recommender Systems for Democracy: Toward Adversarial Robustness in Voting Advice Applications
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks
von: Marius, Dumitran Adrian, et al.
Veröffentlicht: (2025)
von: Marius, Dumitran Adrian, et al.
Veröffentlicht: (2025)
DiffuSpeech: Silent Thought, Spoken Answer via Unified Speech-Text Diffusion
von: Lou, Yuxuan, et al.
Veröffentlicht: (2026)
von: Lou, Yuxuan, et al.
Veröffentlicht: (2026)
SpeechPrompt: Prompting Speech Language Models for Speech Processing Tasks
von: Chang, Kai-Wei, et al.
Veröffentlicht: (2024)
von: Chang, Kai-Wei, et al.
Veröffentlicht: (2024)
Steering Pretrained Drafters during Speculative Decoding
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
GenTranslate: Large Language Models are Generative Multilingual Speech and Machine Translators
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
Speech Emotion Recognition with Distilled Prosodic and Linguistic Affect Representations
von: Shome, Debaditya, et al.
Veröffentlicht: (2023)
von: Shome, Debaditya, et al.
Veröffentlicht: (2023)
Exploring the Plausibility of Hate and Counter Speech Detectors with Explainable AI
von: Böck, Adrian Jaques, et al.
Veröffentlicht: (2024)
von: Böck, Adrian Jaques, et al.
Veröffentlicht: (2024)
Is Child-Directed Speech Effective Training Data for Language Models?
von: Feng, Steven Y., et al.
Veröffentlicht: (2024)
von: Feng, Steven Y., et al.
Veröffentlicht: (2024)
Transformers and Ensemble methods: A solution for Hate Speech Detection in Arabic languages
von: de Paula, Angel Felipe Magnossão, et al.
Veröffentlicht: (2023)
von: de Paula, Angel Felipe Magnossão, et al.
Veröffentlicht: (2023)
Can AI Agents Agree?
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2026)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2026)
CAESAR: Enhancing Federated RL in Heterogeneous MDPs through Convergence-Aware Sampling with Screening
von: Mak, Hei Yi, et al.
Veröffentlicht: (2024)
von: Mak, Hei Yi, et al.
Veröffentlicht: (2024)
Learning to Model the World with Language
von: Lin, Jessy, et al.
Veröffentlicht: (2023)
von: Lin, Jessy, et al.
Veröffentlicht: (2023)
Symphony for Speech-to-Text: Supporting Real-Time Medical Voice Interfaces
von: Nix, Arne, et al.
Veröffentlicht: (2026)
von: Nix, Arne, et al.
Veröffentlicht: (2026)
Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps
von: Waldendorf, Jonas, et al.
Veröffentlicht: (2026)
von: Waldendorf, Jonas, et al.
Veröffentlicht: (2026)
World Properties without World Models: Recovering Spatial and Temporal Structure from Co-occurrence Statistics in Static Word Embeddings
von: Barenholtz, Elan
Veröffentlicht: (2026)
von: Barenholtz, Elan
Veröffentlicht: (2026)
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
von: Das, Amit, et al.
Veröffentlicht: (2024)
von: Das, Amit, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
High-Fidelity Speech Enhancement via Discrete Audio Tokens
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025) -
EuroSpeech: A Multilingual Speech Corpus
von: Pfisterer, Samuel, et al.
Veröffentlicht: (2025) -
Text-to-Scene with Large Reasoning Models
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025) -
Alignment-Aware Decoding
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025) -
N-vium: Mixture-of-Exits Transformer for Accelerated Exact Generation
von: Lorenc, Aleksander, et al.
Veröffentlicht: (2026)