MisinfoEval: Generative AI in the Era of "Alternative Facts"
Fuente:
arXiv
Salvato in:
| Autori principali: | Gabriel, Saadia, Lyu, Liang, Siderius, James, Ghassemi, Marzyeh, Andreas, Jacob, Ozdaglar, Asu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Can AI Relate: Testing Large Language Model Response for Mental Health Support
di: Gabriel, Saadia, et al.
Pubblicazione: (2024)
di: Gabriel, Saadia, et al.
Pubblicazione: (2024)
MOSAIC: Modeling Social AI for Content Dissemination and Regulation in Multi-Agent Simulations
di: Liu, Genglin, et al.
Pubblicazione: (2025)
di: Liu, Genglin, et al.
Pubblicazione: (2025)
How AI Aggregation Affects Knowledge
di: Acemoglu, Daron, et al.
Pubblicazione: (2026)
di: Acemoglu, Daron, et al.
Pubblicazione: (2026)
Reaching Beyond the Mode: RL for Distributional Reasoning in Language Models
di: Puri, Isha, et al.
Pubblicazione: (2026)
di: Puri, Isha, et al.
Pubblicazione: (2026)
Matching of Users and Creators in Two-Sided Markets with Departures
di: Huttenlocher, Daniel, et al.
Pubblicazione: (2023)
di: Huttenlocher, Daniel, et al.
Pubblicazione: (2023)
MisinfoTeleGraph: Network-driven Misinformation Detection for German Telegram Messages
di: Kalkbrenner, Lu, et al.
Pubblicazione: (2025)
di: Kalkbrenner, Lu, et al.
Pubblicazione: (2025)
Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data
di: Xu, Xuhai, et al.
Pubblicazione: (2023)
di: Xu, Xuhai, et al.
Pubblicazione: (2023)
Wikipedia Contributions in the Wake of ChatGPT
di: Lyu, Liang, et al.
Pubblicazione: (2025)
di: Lyu, Liang, et al.
Pubblicazione: (2025)
Speak Easy: Eliciting Harmful Jailbreaks from LLMs with Simple Interactions
di: Chan, Yik Siu, et al.
Pubblicazione: (2025)
di: Chan, Yik Siu, et al.
Pubblicazione: (2025)
In the Name of Fairness: Assessing the Bias in Clinical Record De-identification
di: Xiao, Yuxin, et al.
Pubblicazione: (2023)
di: Xiao, Yuxin, et al.
Pubblicazione: (2023)
Learning the Wrong Lessons: Syntactic-Domain Spurious Correlations in Language Models
di: Shaib, Chantal, et al.
Pubblicazione: (2025)
di: Shaib, Chantal, et al.
Pubblicazione: (2025)
Disparities In Negation Understanding Across Languages In Vision-Language Models
di: Moraitaki, Charikleia, et al.
Pubblicazione: (2026)
di: Moraitaki, Charikleia, et al.
Pubblicazione: (2026)
When Style Breaks Safety: Defending LLMs Against Superficial Style Alignment
di: Xiao, Yuxin, et al.
Pubblicazione: (2025)
di: Xiao, Yuxin, et al.
Pubblicazione: (2025)
Identifying Implicit Social Biases in Vision-Language Models
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
How to Train Your Fact Verifier: Knowledge Transfer with Multimodal Open Models
di: Lee, Jaeyoung, et al.
Pubblicazione: (2024)
di: Lee, Jaeyoung, et al.
Pubblicazione: (2024)
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe
di: Xiao, Yuxin, et al.
Pubblicazione: (2024)
di: Xiao, Yuxin, et al.
Pubblicazione: (2024)
MedFactEval and MedAgentBrief: A Framework and Workflow for Generating and Evaluating Factual Clinical Summaries
di: Grolleau, François, et al.
Pubblicazione: (2025)
di: Grolleau, François, et al.
Pubblicazione: (2025)
KScope: A Framework for Characterizing the Knowledge Status of Language Models
di: Xiao, Yuxin, et al.
Pubblicazione: (2025)
di: Xiao, Yuxin, et al.
Pubblicazione: (2025)
MedPAIR: Measuring Physicians and AI Relevance Alignment in Medical Question Answering
di: Hao, Yuexing, et al.
Pubblicazione: (2025)
di: Hao, Yuexing, et al.
Pubblicazione: (2025)
Generalization in Healthcare AI: Evaluation of a Clinical Large Language Model
di: Rahman, Salman, et al.
Pubblicazione: (2024)
di: Rahman, Salman, et al.
Pubblicazione: (2024)
Vision-Language Models Do Not Understand Negation
di: Alhamoud, Kumail, et al.
Pubblicazione: (2025)
di: Alhamoud, Kumail, et al.
Pubblicazione: (2025)
Mind the Gesture: Evaluating AI Sensitivity to Culturally Offensive Non-Verbal Gestures
di: Yerukola, Akhila, et al.
Pubblicazione: (2025)
di: Yerukola, Akhila, et al.
Pubblicazione: (2025)
IYKYK (But AI Doesn't): Automated Content Moderation Does Not Capture Communities' Heterogeneous Attitudes Towards Reclaimed Language
di: Chance, Christina, et al.
Pubblicazione: (2026)
di: Chance, Christina, et al.
Pubblicazione: (2026)
OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models
di: Sun, Chongren, et al.
Pubblicazione: (2025)
di: Sun, Chongren, et al.
Pubblicazione: (2025)
UWBa at SemEval-2025 Task 7: Multilingual and Crosslingual Fact-Checked Claim Retrieval
di: Lenc, Ladislav, et al.
Pubblicazione: (2025)
di: Lenc, Ladislav, et al.
Pubblicazione: (2025)
DiagramEval: Evaluating LLM-Generated Diagrams via Graphs
di: Liang, Chumeng, et al.
Pubblicazione: (2025)
di: Liang, Chumeng, et al.
Pubblicazione: (2025)
Extractive Structures Learned in Pretraining Enable Generalization on Finetuned Facts
di: Feng, Jiahai, et al.
Pubblicazione: (2024)
di: Feng, Jiahai, et al.
Pubblicazione: (2024)
Reasoning or Rationalization? The Role of Justifications in Masked Diffusion Models for Fact Verification
di: Devasier, Jacob
Pubblicazione: (2026)
di: Devasier, Jacob
Pubblicazione: (2026)
CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval
di: Putta, Akshith Reddy, et al.
Pubblicazione: (2026)
di: Putta, Akshith Reddy, et al.
Pubblicazione: (2026)
fact check AI at SemEval-2025 Task 7: Multilingual and Crosslingual Fact-checked Claim Retrieval
di: Rastogi, Pranshu
Pubblicazione: (2025)
di: Rastogi, Pranshu
Pubblicazione: (2025)
ReviewEval: An Evaluation Framework for AI-Generated Reviews
di: Garg, Madhav Krishan, et al.
Pubblicazione: (2025)
di: Garg, Madhav Krishan, et al.
Pubblicazione: (2025)
SemEval-2025 Task 7: Multilingual and Crosslingual Fact-Checked Claim Retrieval
di: Peng, Qiwei, et al.
Pubblicazione: (2025)
di: Peng, Qiwei, et al.
Pubblicazione: (2025)
VisEval: A Benchmark for Data Visualization in the Era of Large Language Models
di: Chen, Nan, et al.
Pubblicazione: (2024)
di: Chen, Nan, et al.
Pubblicazione: (2024)
SUPERNOVA: Eliciting General Reasoning in LLMs with Reinforcement Learning on Natural Instructions
di: Suvarna, Ashima, et al.
Pubblicazione: (2026)
di: Suvarna, Ashima, et al.
Pubblicazione: (2026)
Word2winners at SemEval-2025 Task 7: Multilingual and Crosslingual Fact-Checked Claim Retrieval
di: Azadi, Amirmohammad, et al.
Pubblicazione: (2025)
di: Azadi, Amirmohammad, et al.
Pubblicazione: (2025)
What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization
di: Rezaeimanesh, Sara, et al.
Pubblicazione: (2026)
di: Rezaeimanesh, Sara, et al.
Pubblicazione: (2026)
Train Long, Think Short: Curriculum Learning for Efficient Reasoning
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2025)
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2025)
AICoderEval: Improving AI Domain Code Generation of Large Language Models
di: Xia, Yinghui, et al.
Pubblicazione: (2024)
di: Xia, Yinghui, et al.
Pubblicazione: (2024)
UFT: Unifying Supervised and Reinforcement Fine-Tuning
di: Liu, Mingyang, et al.
Pubblicazione: (2025)
di: Liu, Mingyang, et al.
Pubblicazione: (2025)
Where Fact Ends and Fairness Begins: Redefining AI Bias Evaluation through Cognitive Biases
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
di: Huang, Jen-tse, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Can AI Relate: Testing Large Language Model Response for Mental Health Support
di: Gabriel, Saadia, et al.
Pubblicazione: (2024) -
MOSAIC: Modeling Social AI for Content Dissemination and Regulation in Multi-Agent Simulations
di: Liu, Genglin, et al.
Pubblicazione: (2025) -
How AI Aggregation Affects Knowledge
di: Acemoglu, Daron, et al.
Pubblicazione: (2026) -
Reaching Beyond the Mode: RL for Distributional Reasoning in Language Models
di: Puri, Isha, et al.
Pubblicazione: (2026) -
Matching of Users and Creators in Two-Sided Markets with Departures
di: Huttenlocher, Daniel, et al.
Pubblicazione: (2023)