Exploring the Performance of Large Language Models on Subjective Span Identification Tasks
Fuente:
arXiv
Salvato in:
| Autori principali: | Dmonte, Alphaeus, Oruche, Roland, Ranasinghe, Tharindu, Zampieri, Marcos, Calyam, Prasad |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Claim Verification in the Age of Large Language Models: A Survey
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
Towards Generalized Offensive Language Identification
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
Classifying Human-Generated and AI-Generated Election Claims in Social Media
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
ALEXSIS-PT: A New Resource for Portuguese Lexical Simplification
di: North, Kai, et al.
Pubblicazione: (2022)
di: North, Kai, et al.
Pubblicazione: (2022)
Improving Training Efficiency and Reducing Maintenance Costs via Language Specific Model Merging
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2026)
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2026)
MultiLS: A Multi-task Lexical Simplification Framework
di: North, Kai, et al.
Pubblicazione: (2024)
di: North, Kai, et al.
Pubblicazione: (2024)
A Federated Learning Approach to Privacy Preserving Offensive Language Identification
di: Zampieri, Marcos, et al.
Pubblicazione: (2024)
di: Zampieri, Marcos, et al.
Pubblicazione: (2024)
A Neuro-Symbolic Multi-Agent Approach to Legal-Cybersecurity Knowledge Integration
di: Bonfanti, Chiara, et al.
Pubblicazione: (2025)
di: Bonfanti, Chiara, et al.
Pubblicazione: (2025)
SOLD: Sinhala Offensive Language Dataset
di: Ranasinghe, Tharindu, et al.
Pubblicazione: (2022)
di: Ranasinghe, Tharindu, et al.
Pubblicazione: (2022)
Do LLMs Judge Distantly Supervised Named Entity Labels Well? Constructing the JudgeWEL Dataset
di: Plum, Alistair, et al.
Pubblicazione: (2026)
di: Plum, Alistair, et al.
Pubblicazione: (2026)
AHaSIS: Shared Task on Sentiment Analysis for Arabic Dialects
di: Alharbi, Maram, et al.
Pubblicazione: (2025)
di: Alharbi, Maram, et al.
Pubblicazione: (2025)
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
di: Kumarage, Tharindu, et al.
Pubblicazione: (2024)
di: Kumarage, Tharindu, et al.
Pubblicazione: (2024)
Natural Language Satisfiability: Exploring the Problem Distribution and Evaluating Transformer-based Language Models
di: Madusanka, Tharindu, et al.
Pubblicazione: (2025)
di: Madusanka, Tharindu, et al.
Pubblicazione: (2025)
Perspective Transition of Large Language Models for Solving Subjective Tasks
di: Wang, Xiaolong, et al.
Pubblicazione: (2025)
di: Wang, Xiaolong, et al.
Pubblicazione: (2025)
Overview of the First Workshop on Language Models for Low-Resource Languages (LoResLM 2025)
di: Hettiarachchi, Hansi, et al.
Pubblicazione: (2024)
di: Hettiarachchi, Hansi, et al.
Pubblicazione: (2024)
NSINA: A News Corpus for Sinhala
di: Hettiarachchi, Hansi, et al.
Pubblicazione: (2024)
di: Hettiarachchi, Hansi, et al.
Pubblicazione: (2024)
Aggregation Artifacts in Subjective Tasks Collapse Large Language Models' Posteriors
di: Chochlakis, Georgios, et al.
Pubblicazione: (2024)
di: Chochlakis, Georgios, et al.
Pubblicazione: (2024)
Exploring Large Language Models and Hierarchical Frameworks for Classification of Large Unstructured Legal Documents
di: Prasad, Nishchal, et al.
Pubblicazione: (2024)
di: Prasad, Nishchal, et al.
Pubblicazione: (2024)
A Survey on Multilingual Mental Disorders Detection from Social Media Data
di: Bucur, Ana-Maria, et al.
Pubblicazione: (2025)
di: Bucur, Ana-Maria, et al.
Pubblicazione: (2025)
Self-Selected Attention Span for Accelerating Large Language Model Inference
di: Jin, Tian, et al.
Pubblicazione: (2024)
di: Jin, Tian, et al.
Pubblicazione: (2024)
A Survey of Multimodal Sarcasm Detection
di: Farabi, Shafkat, et al.
Pubblicazione: (2024)
di: Farabi, Shafkat, et al.
Pubblicazione: (2024)
Mathify: Evaluating Large Language Models on Mathematical Problem Solving Tasks
di: Anand, Avinash, et al.
Pubblicazione: (2024)
di: Anand, Avinash, et al.
Pubblicazione: (2024)
Language Model Council: Democratically Benchmarking Foundation Models on Highly Subjective Tasks
di: Zhao, Justin, et al.
Pubblicazione: (2024)
di: Zhao, Justin, et al.
Pubblicazione: (2024)
Social Meaning in Large Language Models: Structure, Magnitude, and Pragmatic Prompting
di: Mühlenbernd, Roland
Pubblicazione: (2026)
di: Mühlenbernd, Roland
Pubblicazione: (2026)
Is Large Language Model Performance on Reasoning Tasks Impacted by Different Ways Questions Are Asked?
di: Song, Seok Hwan, et al.
Pubblicazione: (2025)
di: Song, Seok Hwan, et al.
Pubblicazione: (2025)
TaskBench: Benchmarking Large Language Models for Task Automation
di: Shen, Yongliang, et al.
Pubblicazione: (2023)
di: Shen, Yongliang, et al.
Pubblicazione: (2023)
Dhati+: Fine-tuned Large Language Models for Arabic Subjectivity Evaluation
di: Bellaouar, Slimane, et al.
Pubblicazione: (2025)
di: Bellaouar, Slimane, et al.
Pubblicazione: (2025)
Large Language Models for Multi-Choice Question Classification of Medical Subjects
di: Ponce-López, Víctor
Pubblicazione: (2024)
di: Ponce-López, Víctor
Pubblicazione: (2024)
Same Task, More Tokens: the Impact of Input Length on the Reasoning Performance of Large Language Models
di: Levy, Mosh, et al.
Pubblicazione: (2024)
di: Levy, Mosh, et al.
Pubblicazione: (2024)
Exploring the Trade-Offs: Quantization Methods, Task Difficulty, and Model Size in Large Language Models From Edge to Giant
di: Lee, Jemin, et al.
Pubblicazione: (2024)
di: Lee, Jemin, et al.
Pubblicazione: (2024)
Large Language Model Benchmarks in Medical Tasks
di: Yan, Lawrence K. Q., et al.
Pubblicazione: (2024)
di: Yan, Lawrence K. Q., et al.
Pubblicazione: (2024)
CONTESTS: a Framework for Consistency Testing of Span Probabilities in Language Models
di: Wagner, Eitan, et al.
Pubblicazione: (2024)
di: Wagner, Eitan, et al.
Pubblicazione: (2024)
When the Majority is Wrong: Modeling Annotator Disagreement for Subjective Tasks
di: Fleisig, Eve, et al.
Pubblicazione: (2023)
di: Fleisig, Eve, et al.
Pubblicazione: (2023)
The Price of Thought: A Multilingual Analysis of Reasoning, Performance, and Cost of Negotiation in Large Language Models
di: Hakimov, Sherzod, et al.
Pubblicazione: (2025)
di: Hakimov, Sherzod, et al.
Pubblicazione: (2025)
Evaluating Ill-Defined Tasks in Large Language Models
di: Zhou, Yi, et al.
Pubblicazione: (2026)
di: Zhou, Yi, et al.
Pubblicazione: (2026)
Impact of Task Phrasing on Presumptions in Large Language Models
di: Ong, Kenneth J. K.
Pubblicazione: (2026)
di: Ong, Kenneth J. K.
Pubblicazione: (2026)
Reasoning Capabilities of Large Language Models on Dynamic Tasks
di: Wong, Annie, et al.
Pubblicazione: (2025)
di: Wong, Annie, et al.
Pubblicazione: (2025)
Task-Aligned Tool Recommendation for Large Language Models
di: Gao, Hang, et al.
Pubblicazione: (2024)
di: Gao, Hang, et al.
Pubblicazione: (2024)
Large Language Models Vote: Prompting for Rare Disease Identification
di: Oniani, David, et al.
Pubblicazione: (2023)
di: Oniani, David, et al.
Pubblicazione: (2023)
How do Scaling Laws Apply to Knowledge Graph Engineering Tasks? The Impact of Model Size on Large Language Model Performance
di: Heim, Desiree, et al.
Pubblicazione: (2025)
di: Heim, Desiree, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Claim Verification in the Age of Large Language Models: A Survey
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024) -
Towards Generalized Offensive Language Identification
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024) -
Classifying Human-Generated and AI-Generated Election Claims in Social Media
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024) -
ALEXSIS-PT: A New Resource for Portuguese Lexical Simplification
di: North, Kai, et al.
Pubblicazione: (2022) -
Improving Training Efficiency and Reducing Maintenance Costs via Language Specific Model Merging
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2026)