LeWiDi-2025 at NLPerspectives: Third Edition of the Learning with Disagreements Shared Task
Fuente:
arXiv
Guardado en:
| Autores principales: | Leonardelli, Elisa, Casola, Silvia, Peng, Siyao, Rizzi, Giulia, Basile, Valerio, Fersini, Elisabetta, Frassinelli, Diego, Jang, Hyewon, Pavlovic, Maja, Plank, Barbara, Poesio, Massimo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BoN Appetit Team at LeWiDi-2025: Best-of-N Test-time Scaling Can Not Stomach Annotation Disagreements (Yet)
por: Ruiz, Tomas, et al.
Publicado: (2025)
por: Ruiz, Tomas, et al.
Publicado: (2025)
Understanding The Effect Of Temperature On Alignment With Human Opinions
por: Pavlovic, Maja, et al.
Publicado: (2024)
por: Pavlovic, Maja, et al.
Publicado: (2024)
The Effectiveness of LLMs as Annotators: A Comparative Overview and Empirical Analysis of Direct Representation
por: Pavlovic, Maja, et al.
Publicado: (2024)
por: Pavlovic, Maja, et al.
Publicado: (2024)
An Assessment of Human vs. Model Uncertainty in Soft-Label Learning and Calibration
por: Pavlovic, Maja, et al.
Publicado: (2026)
por: Pavlovic, Maja, et al.
Publicado: (2026)
Generalizable Sarcasm Detection Is Just Around The Corner, Of Course!
por: Jang, Hyewon, et al.
Publicado: (2024)
por: Jang, Hyewon, et al.
Publicado: (2024)
Mind the Uncertainty in Human Disagreement: Evaluating Discrepancies between Model Predictions and Human Responses in VQA
por: Lan, Jian, et al.
Publicado: (2024)
por: Lan, Jian, et al.
Publicado: (2024)
DeMeVa at LeWiDi-2025: Modeling Perspectives with In-Context Learning and Label Distribution Learning
por: Ignatev, Daniil, et al.
Publicado: (2025)
por: Ignatev, Daniil, et al.
Publicado: (2025)
LPI-RIT at LeWiDi-2025: Improving Distributional Predictions via Metadata and Loss Reweighting with DisCo
por: Sawkar, Mandira, et al.
Publicado: (2025)
por: Sawkar, Mandira, et al.
Publicado: (2025)
Opt-ICL at LeWiDi-2025: Maximizing In-Context Signal from Rater Examples via Meta-Learning
por: Sorensen, Taylor, et al.
Publicado: (2025)
por: Sorensen, Taylor, et al.
Publicado: (2025)
A gentle push funziona benissimo: making instructed models in Italian via contrastive activation steering
por: Scalena, Daniel, et al.
Publicado: (2024)
por: Scalena, Daniel, et al.
Publicado: (2024)
Resource-Lean Lexicon Induction for German Dialects
por: Litschko, Robert, et al.
Publicado: (2026)
por: Litschko, Robert, et al.
Publicado: (2026)
Do LLMs Give Psychometrically Plausible Responses in Educational Assessments?
por: Säuberli, Andreas, et al.
Publicado: (2025)
por: Säuberli, Andreas, et al.
Publicado: (2025)
On the Robustness of Transformer‐Based Models to Different Linguistic Perturbations: A Case of Study in Irony Detection
por: Reynier Ortega‐Bueno, et al.
Publicado: (2025)
por: Reynier Ortega‐Bueno, et al.
Publicado: (2025)
Investigating the Nature of Disagreements on Mid-Scale Ratings: A Case Study on the Abstractness-Concreteness Continuum
por: Knupleš, Urban, et al.
Publicado: (2023)
por: Knupleš, Urban, et al.
Publicado: (2023)
Understanding Model Calibration -- A gentle introduction and visual exploration of calibration and the expected calibration error (ECE)
por: Pavlovic, Maja
Publicado: (2025)
por: Pavlovic, Maja
Publicado: (2025)
Large Language Models as Minecraft Agents
por: Madge, Chris, et al.
Publicado: (2024)
por: Madge, Chris, et al.
Publicado: (2024)
Integrating knowledge bases to improve coreference and bridging resolution for the chemical domain
por: Lu, Pengcheng, et al.
Publicado: (2024)
por: Lu, Pengcheng, et al.
Publicado: (2024)
Data Augmentation for Fake Reviews Detection in Multiple Languages and Multiple Domains
por: Liu, Ming, et al.
Publicado: (2025)
por: Liu, Ming, et al.
Publicado: (2025)
A LLM Benchmark based on the Minecraft Builder Dialog Agent Task
por: Madge, Chris, et al.
Publicado: (2024)
por: Madge, Chris, et al.
Publicado: (2024)
References Matter: Investigating the Impact of Reference Set Variation on Summarization Evaluation
por: Casola, Silvia, et al.
Publicado: (2025)
por: Casola, Silvia, et al.
Publicado: (2025)
Controlling Reading Ease with Gaze-Guided Text Generation
por: Säuberli, Andreas, et al.
Publicado: (2026)
por: Säuberli, Andreas, et al.
Publicado: (2026)
I Came, I Saw, I Explained: Benchmarking Multimodal LLMs on Figurative Meaning in Memes
por: Zhou, Shijia, et al.
Publicado: (2026)
por: Zhou, Shijia, et al.
Publicado: (2026)
The Geography of Information Diffusion in Online Discourse on Europe and Migration
por: Leonardelli, Elisa, et al.
Publicado: (2024)
por: Leonardelli, Elisa, et al.
Publicado: (2024)
Referential ambiguity and clarification requests: comparing human and LLM behaviour
por: Madge, Chris, et al.
Publicado: (2025)
por: Madge, Chris, et al.
Publicado: (2025)
Grounded Misunderstandings in Asymmetric Dialogue: A Perspectivist Annotation Scheme for MapTask
por: Li, Nan, et al.
Publicado: (2025)
por: Li, Nan, et al.
Publicado: (2025)
Can LLMs Detect Ambiguous Plural Reference? An Analysis of Split-Antecedent and Mereological Reference
por: Anh, Dang, et al.
Publicado: (2025)
por: Anh, Dang, et al.
Publicado: (2025)
Hypernetworks for Perspectivist Adaptation
por: Ignatev, Daniil, et al.
Publicado: (2025)
por: Ignatev, Daniil, et al.
Publicado: (2025)
Steering Large Language Models for Machine Translation Personalization
por: Scalena, Daniel, et al.
Publicado: (2025)
por: Scalena, Daniel, et al.
Publicado: (2025)
Semantic Storyboard of Judicial Debates: A Novel Multimedia Summarization Environment
por: Fersini, E., et al.
Publicado: (2012)
por: Fersini, E., et al.
Publicado: (2012)
Make Every Letter Count: Building Dialect Variation Dictionaries from Monolingual Corpora
por: Litschko, Robert, et al.
Publicado: (2025)
por: Litschko, Robert, et al.
Publicado: (2025)
To Know or Not To Know? Analyzing Self-Consistency of Large Language Models under Ambiguity
por: Sedova, Anastasiia, et al.
Publicado: (2024)
por: Sedova, Anastasiia, et al.
Publicado: (2024)
CLIMATELI: Evaluating Entity Linking on Climate Change Data
por: Zhou, Shijia, et al.
Publicado: (2024)
por: Zhou, Shijia, et al.
Publicado: (2024)
EVADE: LLM-Based Explanation Generation and Validation for Error Detection in NLI
por: Zuo, Longfei, et al.
Publicado: (2025)
por: Zuo, Longfei, et al.
Publicado: (2025)
A Bibliometric Analysis to Study the Evolution of Artificial Intelligence in Business Ethics
por: Mario Tani, et al.
Publicado: (2025)
por: Mario Tani, et al.
Publicado: (2025)
Extending Activation Steering to Broad Skills and Multiple Behaviours
por: van der Weij, Teun, et al.
Publicado: (2024)
por: van der Weij, Teun, et al.
Publicado: (2024)
El nexo entre cambio climático y energía renovable en el Mercosur. Un análisis comparativo de las legislaciones de Argentina y Brasil
por: Laura Casola
Publicado: (2018)
por: Laura Casola
Publicado: (2018)
Formas de militancia en el Partido Comunista argentino durante la última dictadura militar (1976-1983)
por: Natalia Casola
Publicado: (2015)
por: Natalia Casola
Publicado: (2015)
La labor de la Comisión Argentina para los Refugiados (CAREF). De la emergencia humanitaria a la convergencia con el movimiento de derechos humanos y el movimiento de mujeres (1973-1992)
por: Natalia Casola
Publicado: (2022)
por: Natalia Casola
Publicado: (2022)
Inhomogeneous six-wave kinetic equation in exponentially weighted $L^\infty$ spaces
por: Pavlović, Nataša, et al.
Publicado: (2025)
por: Pavlović, Nataša, et al.
Publicado: (2025)
Spoken Records. Third Edition.
por: Roach, Helen
Publicado: (1970)
por: Roach, Helen
Publicado: (1970)
Ejemplares similares
-
BoN Appetit Team at LeWiDi-2025: Best-of-N Test-time Scaling Can Not Stomach Annotation Disagreements (Yet)
por: Ruiz, Tomas, et al.
Publicado: (2025) -
Understanding The Effect Of Temperature On Alignment With Human Opinions
por: Pavlovic, Maja, et al.
Publicado: (2024) -
The Effectiveness of LLMs as Annotators: A Comparative Overview and Empirical Analysis of Direct Representation
por: Pavlovic, Maja, et al.
Publicado: (2024) -
An Assessment of Human vs. Model Uncertainty in Soft-Label Learning and Calibration
por: Pavlovic, Maja, et al.
Publicado: (2026) -
Generalizable Sarcasm Detection Is Just Around The Corner, Of Course!
por: Jang, Hyewon, et al.
Publicado: (2024)