Investigating Affect Mining Techniques for Annotation Sample Selection in the Creation of Finnish Affective Speech Corpus
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lahtinen, Kalle, Vaaras, Einari, Mustanoja, Liisa, Räsänen, Okko |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating Interactive 2D Visualization as a Sample Selection Strategy for Biomedical Time-Series Data Annotation
von: Vaaras, Einari, et al.
Veröffentlicht: (2026)
von: Vaaras, Einari, et al.
Veröffentlicht: (2026)
PFML: Self-Supervised Learning of Time-Series Data Without Representation Collapse
von: Vaaras, Einari, et al.
Veröffentlicht: (2024)
von: Vaaras, Einari, et al.
Veröffentlicht: (2024)
Evaluation of Audio-Visual Alignments in Visually Grounded Speech Models
von: Khorrami, Khazar, et al.
Veröffentlicht: (2021)
von: Khorrami, Khazar, et al.
Veröffentlicht: (2021)
Feature Space Topology Control via Hopkins Loss
von: Vaaras, Einari, et al.
Veröffentlicht: (2025)
von: Vaaras, Einari, et al.
Veröffentlicht: (2025)
Age-Dependent Analysis and Stochastic Generation of Child-Directed Speech
von: Räsänen, Okko, et al.
Veröffentlicht: (2024)
von: Räsänen, Okko, et al.
Veröffentlicht: (2024)
Can phones, syllables, and words emerge as side-products of cross-situational audiovisual learning? -- A computational investigation
von: Khorrami, Khazar, et al.
Veröffentlicht: (2021)
von: Khorrami, Khazar, et al.
Veröffentlicht: (2021)
Computational modeling of early language learning from acoustic speech and audiovisual input without linguistic priors
von: Räsänen, Okko
Veröffentlicht: (2026)
von: Räsänen, Okko
Veröffentlicht: (2026)
Enabling automatic transcription of child-centered audio recordings from real-world environments
von: Kocharov, Daniil, et al.
Veröffentlicht: (2025)
von: Kocharov, Daniil, et al.
Veröffentlicht: (2025)
FalAR: A Large-scale Speaker-Annotated European Portuguese Speech Corpus of Parliamentary Sessions
von: Teixeira, Francisco, et al.
Veröffentlicht: (2026)
von: Teixeira, Francisco, et al.
Veröffentlicht: (2026)
EuroSpeech: A Multilingual Speech Corpus
von: Pfisterer, Samuel, et al.
Veröffentlicht: (2025)
von: Pfisterer, Samuel, et al.
Veröffentlicht: (2025)
A model of early word acquisition based on realistic-scale audiovisual naming events
von: Khorrami, Khazar, et al.
Veröffentlicht: (2024)
von: Khorrami, Khazar, et al.
Veröffentlicht: (2024)
Breaking the Attention Bottleneck
von: Hilsenbek, Kalle
Veröffentlicht: (2024)
von: Hilsenbek, Kalle
Veröffentlicht: (2024)
MONOVAB : An Annotated Corpus for Bangla Multi-label Emotion Detection
von: Banshal, Sumit Kumar, et al.
Veröffentlicht: (2023)
von: Banshal, Sumit Kumar, et al.
Veröffentlicht: (2023)
BabySLM: language-acquisition-friendly benchmark of self-supervised spoken language models
von: Lavechin, Marvin, et al.
Veröffentlicht: (2023)
von: Lavechin, Marvin, et al.
Veröffentlicht: (2023)
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
von: Das, Amit, et al.
Veröffentlicht: (2024)
von: Das, Amit, et al.
Veröffentlicht: (2024)
WorldSpeech: A Multilingual Speech Corpus from Around the World
von: Asonitis, Antonis, et al.
Veröffentlicht: (2026)
von: Asonitis, Antonis, et al.
Veröffentlicht: (2026)
Annotation Sensitivity: Training Data Collection Methods Affect Model Performance
von: Kern, Christoph, et al.
Veröffentlicht: (2023)
von: Kern, Christoph, et al.
Veröffentlicht: (2023)
Large Language Model Selection with Limited Annotations
von: Durmazkeser, Yavuz, et al.
Veröffentlicht: (2026)
von: Durmazkeser, Yavuz, et al.
Veröffentlicht: (2026)
Memory Is All You Need: Testing How Model Memory Affects LLM Performance in Annotation Tasks
von: Timoneda, Joan C., et al.
Veröffentlicht: (2025)
von: Timoneda, Joan C., et al.
Veröffentlicht: (2025)
Simultaneous or Sequential Training? How Speech Representations Cooperate in a Multi-Task Self-Supervised Learning System
von: Khorrami, Khazar, et al.
Veröffentlicht: (2023)
von: Khorrami, Khazar, et al.
Veröffentlicht: (2023)
Health Text Simplification: An Annotated Corpus for Digestive Cancer Education and Novel Strategies for Reinforcement Learning
von: Rahman, Md Mushfiqur, et al.
Veröffentlicht: (2024)
von: Rahman, Md Mushfiqur, et al.
Veröffentlicht: (2024)
Efficient Knowledge Injection in LLMs via Self-Distillation
von: Kujanpää, Kalle, et al.
Veröffentlicht: (2024)
von: Kujanpää, Kalle, et al.
Veröffentlicht: (2024)
Speech Emotion Recognition with Distilled Prosodic and Linguistic Affect Representations
von: Shome, Debaditya, et al.
Veröffentlicht: (2023)
von: Shome, Debaditya, et al.
Veröffentlicht: (2023)
FeruzaSpeech: A 60 Hour Uzbek Read Speech Corpus with Punctuation, Casing, and Context
von: Povey, Anna, et al.
Veröffentlicht: (2024)
von: Povey, Anna, et al.
Veröffentlicht: (2024)
How does Multi-Task Training Affect Transformer In-Context Capabilities? Investigations with Function Classes
von: Bhasin, Harmon, et al.
Veröffentlicht: (2024)
von: Bhasin, Harmon, et al.
Veröffentlicht: (2024)
Boosting Protein Language Models with Negative Sample Mining
von: Xu, Yaoyao, et al.
Veröffentlicht: (2024)
von: Xu, Yaoyao, et al.
Veröffentlicht: (2024)
The Moral Foundations Weibo Corpus
von: Cao, Renjie, et al.
Veröffentlicht: (2024)
von: Cao, Renjie, et al.
Veröffentlicht: (2024)
Optimizing Large Language Models for Turkish: New Methodologies in Corpus Selection and Training
von: Kesgin, H. Toprak, et al.
Veröffentlicht: (2024)
von: Kesgin, H. Toprak, et al.
Veröffentlicht: (2024)
Investigating the Impact of Data Selection Strategies on Language Model Performance
von: Gu, Jiayao, et al.
Veröffentlicht: (2025)
von: Gu, Jiayao, et al.
Veröffentlicht: (2025)
The Moral Foundations Reddit Corpus
von: Trager, Jackson, et al.
Veröffentlicht: (2022)
von: Trager, Jackson, et al.
Veröffentlicht: (2022)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
Determinants of Training Corpus Size for Clinical Text Classification
von: Chaturvedi, Jaya, et al.
Veröffentlicht: (2026)
von: Chaturvedi, Jaya, et al.
Veröffentlicht: (2026)
Mining Intrinsic Rewards from LLM Hidden States for Efficient Best-of-N Sampling
von: Guo, Jizhou, et al.
Veröffentlicht: (2025)
von: Guo, Jizhou, et al.
Veröffentlicht: (2025)
Sub-SA: Strengthen In-context Learning via Submodular Selective Annotation
von: Qian, Jian, et al.
Veröffentlicht: (2024)
von: Qian, Jian, et al.
Veröffentlicht: (2024)
TaigiSpeech: A Low-Resource Real-World Speech Intent Dataset and Preliminary Results with Scalable Data Mining In-the-Wild
von: Chang, Kai-Wei, et al.
Veröffentlicht: (2026)
von: Chang, Kai-Wei, et al.
Veröffentlicht: (2026)
Analyzing Speech Unit Selection for Textless Speech-to-Speech Translation
von: Duret, Jarod, et al.
Veröffentlicht: (2024)
von: Duret, Jarod, et al.
Veröffentlicht: (2024)
Affective and Dynamic Beam Search for Story Generation
von: Huang, Tenghao, et al.
Veröffentlicht: (2023)
von: Huang, Tenghao, et al.
Veröffentlicht: (2023)
ANUBHUTI: A Comprehensive Corpus For Sentiment Analysis In Bangla Regional Languages
von: Kundu, Swastika, et al.
Veröffentlicht: (2025)
von: Kundu, Swastika, et al.
Veröffentlicht: (2025)
Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis
von: Bhowmik, Shimanto, et al.
Veröffentlicht: (2025)
von: Bhowmik, Shimanto, et al.
Veröffentlicht: (2025)
Scaling Synthetic Data Creation with 1,000,000,000 Personas
von: Ge, Tao, et al.
Veröffentlicht: (2024)
von: Ge, Tao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Evaluating Interactive 2D Visualization as a Sample Selection Strategy for Biomedical Time-Series Data Annotation
von: Vaaras, Einari, et al.
Veröffentlicht: (2026) -
PFML: Self-Supervised Learning of Time-Series Data Without Representation Collapse
von: Vaaras, Einari, et al.
Veröffentlicht: (2024) -
Evaluation of Audio-Visual Alignments in Visually Grounded Speech Models
von: Khorrami, Khazar, et al.
Veröffentlicht: (2021) -
Feature Space Topology Control via Hopkins Loss
von: Vaaras, Einari, et al.
Veröffentlicht: (2025) -
Age-Dependent Analysis and Stochastic Generation of Child-Directed Speech
von: Räsänen, Okko, et al.
Veröffentlicht: (2024)