Can Machines Resonate with Humans? Evaluating the Emotional and Empathic Comprehension of LMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Manzoor, Muhammad Arslan, Wang, Yuxia, Wang, Minghan, Nakov, Preslav |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Rethinking STS and NLI in Large Language Models
di: Wang, Yuxia, et al.
Pubblicazione: (2023)
di: Wang, Yuxia, et al.
Pubblicazione: (2023)
Factuality of Large Language Models: A Survey
di: Wang, Yuxia, et al.
Pubblicazione: (2024)
di: Wang, Yuxia, et al.
Pubblicazione: (2024)
OpenFactCheck: Building, Benchmarking Customized Fact-Checking Systems and Evaluating the Factuality of Claims and LLMs
di: Wang, Yuxia, et al.
Pubblicazione: (2024)
di: Wang, Yuxia, et al.
Pubblicazione: (2024)
UnsafeChain: Enhancing Reasoning Model Safety via Hard Cases
di: Tomar, Raj Vardhan, et al.
Pubblicazione: (2025)
di: Tomar, Raj Vardhan, et al.
Pubblicazione: (2025)
How Does Prefix Matter in Reasoning Model Tuning?
di: Tomar, Raj Vardhan, et al.
Pubblicazione: (2026)
di: Tomar, Raj Vardhan, et al.
Pubblicazione: (2026)
A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis
di: Manzoor, Muhammad Arslan, et al.
Pubblicazione: (2026)
di: Manzoor, Muhammad Arslan, et al.
Pubblicazione: (2026)
MGM: Global Understanding of Audience Overlap Graphs for Predicting the Factuality and the Bias of News Media
di: Manzoor, Muhammad Arslan, et al.
Pubblicazione: (2024)
di: Manzoor, Muhammad Arslan, et al.
Pubblicazione: (2024)
OpenFactCheck: A Unified Framework for Factuality Evaluation of LLMs
di: Iqbal, Hasan, et al.
Pubblicazione: (2024)
di: Iqbal, Hasan, et al.
Pubblicazione: (2024)
Arabic Dataset for LLM Safeguard Evaluation
di: Ashraf, Yasser, et al.
Pubblicazione: (2024)
di: Ashraf, Yasser, et al.
Pubblicazione: (2024)
HALF: Harm-Aware LLM Fairness Evaluation Aligned with Deployment
di: Mekky, Ali, et al.
Pubblicazione: (2025)
di: Mekky, Ali, et al.
Pubblicazione: (2025)
MuDRiC: Multi-Dialect Reasoning for Arabic Commonsense Validation
di: Elozeiri, Kareem, et al.
Pubblicazione: (2025)
di: Elozeiri, Kareem, et al.
Pubblicazione: (2025)
UrduFactCheck: An Agentic Fact-Checking Framework for Urdu with Evidence Boosting and Benchmarking
di: Ahmad, Sarfraz, et al.
Pubblicazione: (2025)
di: Ahmad, Sarfraz, et al.
Pubblicazione: (2025)
A Comprehensive Survey of Machine Unlearning Techniques for Large Language Models
di: Geng, Jiahui, et al.
Pubblicazione: (2025)
di: Geng, Jiahui, et al.
Pubblicazione: (2025)
Loki: An Open-Source Tool for Fact Verification
di: Li, Haonan, et al.
Pubblicazione: (2024)
di: Li, Haonan, et al.
Pubblicazione: (2024)
CoDet-M4: Detecting Machine-Generated Code in Multi-Lingual, Multi-Generator and Multi-Domain Settings
di: Orel, Daniil, et al.
Pubblicazione: (2025)
di: Orel, Daniil, et al.
Pubblicazione: (2025)
ExaGPT: Example-Based Machine-Generated Text Detection for Human Interpretability
di: Koike, Ryuto, et al.
Pubblicazione: (2025)
di: Koike, Ryuto, et al.
Pubblicazione: (2025)
Qorgau: Evaluating LLM Safety in Kazakh-Russian Bilingual Contexts
di: Goloburda, Maiya, et al.
Pubblicazione: (2025)
di: Goloburda, Maiya, et al.
Pubblicazione: (2025)
A Survey of Confidence Estimation and Calibration in Large Language Models
di: Geng, Jiahui, et al.
Pubblicazione: (2023)
di: Geng, Jiahui, et al.
Pubblicazione: (2023)
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
di: Laiyk, Nurkhan, et al.
Pubblicazione: (2025)
di: Laiyk, Nurkhan, et al.
Pubblicazione: (2025)
A Chinese Dataset for Evaluating the Safeguards in Large Language Models
di: Wang, Yuxia, et al.
Pubblicazione: (2024)
di: Wang, Yuxia, et al.
Pubblicazione: (2024)
DUAL-Bench: Measuring Over-Refusal and Robustness in Vision-Language Models
di: Ren, Kaixuan, et al.
Pubblicazione: (2025)
di: Ren, Kaixuan, et al.
Pubblicazione: (2025)
Large Language Models are Few-Shot Training Example Generators: A Case Study in Fallacy Recognition
di: Alhindi, Tariq, et al.
Pubblicazione: (2023)
di: Alhindi, Tariq, et al.
Pubblicazione: (2023)
Profiling News Media for Factuality and Bias Using LLMs and the Fact-Checking Methodology of Human Experts
di: Mujahid, Zain Muhammad, et al.
Pubblicazione: (2025)
di: Mujahid, Zain Muhammad, et al.
Pubblicazione: (2025)
Detecting Propaganda Techniques in Code-Switched Social Media Text
di: Salman, Muhammad Umar, et al.
Pubblicazione: (2023)
di: Salman, Muhammad Umar, et al.
Pubblicazione: (2023)
Detection of Human and Machine-Authored Fake News in Urdu
di: Ali, Muhammad Zain, et al.
Pubblicazione: (2024)
di: Ali, Muhammad Zain, et al.
Pubblicazione: (2024)
From Chaos to Clarity: Claim Normalization to Empower Fact-Checking
di: Sundriyal, Megha, et al.
Pubblicazione: (2023)
di: Sundriyal, Megha, et al.
Pubblicazione: (2023)
Adapting Fake News Detection to the Era of Large Language Models
di: Su, Jinyan, et al.
Pubblicazione: (2023)
di: Su, Jinyan, et al.
Pubblicazione: (2023)
Co-FactChecker: A Framework for Human-AI Collaborative Claim Verification Using Large Reasoning Models
di: Sahnan, Dhruv, et al.
Pubblicazione: (2026)
di: Sahnan, Dhruv, et al.
Pubblicazione: (2026)
M4GT-Bench: Evaluation Benchmark for Black-Box Machine-Generated Text Detection
di: Wang, Yuxia, et al.
Pubblicazione: (2024)
di: Wang, Yuxia, et al.
Pubblicazione: (2024)
Exploring Language Model Generalization in Low-Resource Extractive QA
di: Sengupta, Saptarshi, et al.
Pubblicazione: (2024)
di: Sengupta, Saptarshi, et al.
Pubblicazione: (2024)
HumanRankEval: Automatic Evaluation of LMs as Conversational Assistants
di: Gritta, Milan, et al.
Pubblicazione: (2024)
di: Gritta, Milan, et al.
Pubblicazione: (2024)
Missci: Reconstructing Fallacies in Misrepresented Science
di: Glockner, Max, et al.
Pubblicazione: (2024)
di: Glockner, Max, et al.
Pubblicazione: (2024)
Grounding Fallacies Misrepresenting Scientific Publications in Evidence
di: Glockner, Max, et al.
Pubblicazione: (2024)
di: Glockner, Max, et al.
Pubblicazione: (2024)
ConspirED: A Dataset for Cognitive Traits of Conspiracy Theories and Large Language Model Safety
di: Bates, Luke, et al.
Pubblicazione: (2025)
di: Bates, Luke, et al.
Pubblicazione: (2025)
UNCERTAINTY-LINE: Length-Invariant Estimation of Uncertainty for Large Language Models
di: Vashurin, Roman, et al.
Pubblicazione: (2025)
di: Vashurin, Roman, et al.
Pubblicazione: (2025)
MemeMQA: Multimodal Question Answering for Memes via Rationale-Based Inferencing
di: Agarwal, Siddhant, et al.
Pubblicazione: (2024)
di: Agarwal, Siddhant, et al.
Pubblicazione: (2024)
Factcheck-Bench: Fine-Grained Evaluation Benchmark for Automatic Fact-checkers
di: Wang, Yuxia, et al.
Pubblicazione: (2023)
di: Wang, Yuxia, et al.
Pubblicazione: (2023)
Exploring the Limitations of Detecting Machine-Generated Text
di: Doughman, Jad, et al.
Pubblicazione: (2024)
di: Doughman, Jad, et al.
Pubblicazione: (2024)
Explicit and Implicit Data Augmentation for Social Event Detection
di: Ma, Congbo, et al.
Pubblicazione: (2025)
di: Ma, Congbo, et al.
Pubblicazione: (2025)
Exploring the Potential of Multimodal LLM with Knowledge-Intensive Multimodal ASR
di: Wang, Minghan, et al.
Pubblicazione: (2024)
di: Wang, Minghan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Rethinking STS and NLI in Large Language Models
di: Wang, Yuxia, et al.
Pubblicazione: (2023) -
Factuality of Large Language Models: A Survey
di: Wang, Yuxia, et al.
Pubblicazione: (2024) -
OpenFactCheck: Building, Benchmarking Customized Fact-Checking Systems and Evaluating the Factuality of Claims and LLMs
di: Wang, Yuxia, et al.
Pubblicazione: (2024) -
UnsafeChain: Enhancing Reasoning Model Safety via Hard Cases
di: Tomar, Raj Vardhan, et al.
Pubblicazione: (2025) -
How Does Prefix Matter in Reasoning Model Tuning?
di: Tomar, Raj Vardhan, et al.
Pubblicazione: (2026)