Enhancing LLMs for Governance with Human Oversight: Evaluating and Aligning LLMs on Expert Classification of Climate Misinformation for Detecting False or Misleading Claims about Climate Change
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Allaham, Mowafak, Lokmanoglu, Ayse D., Hart, P. Sol, Nisbet, Erik C. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating the Capabilities of LLMs for Supporting Anticipatory Impact Assessment
von: Allaham, Mowafak, et al.
Veröffentlicht: (2024)
von: Allaham, Mowafak, et al.
Veröffentlicht: (2024)
Synthetic Sources?: Auditing Generative Search Engine Citations for Evidence of AI-Generated Sources
von: Allaham, Mowafak, et al.
Veröffentlicht: (2026)
von: Allaham, Mowafak, et al.
Veröffentlicht: (2026)
Towards Leveraging News Media to Support Impact Assessment of AI Technologies
von: Allaham, Mowafak, et al.
Veröffentlicht: (2024)
von: Allaham, Mowafak, et al.
Veröffentlicht: (2024)
Global Perspectives of AI Risks and Harms: Analyzing the Negative Impacts of AI Technologies as Prioritized by News Media
von: Allaham, Mowafak, et al.
Veröffentlicht: (2025)
von: Allaham, Mowafak, et al.
Veröffentlicht: (2025)
Informing AI Risk Assessment with News Media: Analyzing National and Political Variation in the Coverage of AI Risks
von: Allaham, Mowafak, et al.
Veröffentlicht: (2025)
von: Allaham, Mowafak, et al.
Veröffentlicht: (2025)
Detecting Fallacies in Climate Misinformation: A Technocognitive Approach to Identifying Misleading Argumentation
von: Zanartu, Francisco, et al.
Veröffentlicht: (2024)
von: Zanartu, Francisco, et al.
Veröffentlicht: (2024)
Transnational Network Dynamics of Problematic Information Diffusion
von: Villa-Turek, Esteban, et al.
Veröffentlicht: (2024)
von: Villa-Turek, Esteban, et al.
Veröffentlicht: (2024)
Emergence WebVoyager: Toward Consistent and Transparent Evaluation of (Web) Agents in The Wild
von: Akkil, Deepak, et al.
Veröffentlicht: (2026)
von: Akkil, Deepak, et al.
Veröffentlicht: (2026)
VisTopics: A Visual Semantic Unsupervised Approach to Topic Modeling of Video and Image Data
von: Lokmanoglu, Ayse D, et al.
Veröffentlicht: (2025)
von: Lokmanoglu, Ayse D, et al.
Veröffentlicht: (2025)
LLMs Struggle to Reject False Presuppositions when Misinformation Stakes are High
von: Sieker, Judith, et al.
Veröffentlicht: (2025)
von: Sieker, Judith, et al.
Veröffentlicht: (2025)
Evaluating the Simulation of Human Personality-Driven Susceptibility to Misinformation with LLMs
von: Pratelli, Manuel, et al.
Veröffentlicht: (2025)
von: Pratelli, Manuel, et al.
Veröffentlicht: (2025)
How Good (Or Bad) Are LLMs at Detecting Misleading Visualizations?
von: Lo, Leo Yu-Ho, et al.
Veröffentlicht: (2024)
von: Lo, Leo Yu-Ho, et al.
Veröffentlicht: (2024)
Generative Debunking of Climate Misinformation
von: Zanartu, Francisco, et al.
Veröffentlicht: (2024)
von: Zanartu, Francisco, et al.
Veröffentlicht: (2024)
ClimateCheck 2026: Scientific Fact-Checking and Disinformation Narrative Classification of Climate-related Claims
von: Ahmad, Raia Abu, et al.
Veröffentlicht: (2026)
von: Ahmad, Raia Abu, et al.
Veröffentlicht: (2026)
ClimateChat: Designing Data and Methods for Instruction Tuning LLMs to Answer Climate Change Queries
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
von: Chen, Zhou, et al.
Veröffentlicht: (2025)
RuozhiBench: Evaluating LLMs with Logical Fallacies and Misleading Premises
von: Zhai, Zenan, et al.
Veröffentlicht: (2025)
von: Zhai, Zenan, et al.
Veröffentlicht: (2025)
Evaluating and Aligning CodeLLMs on Human Preference
von: Yang, Jian, et al.
Veröffentlicht: (2024)
von: Yang, Jian, et al.
Veröffentlicht: (2024)
Automated Fact-Checking of Climate Change Claims with Large Language Models
von: Leippold, Markus, et al.
Veröffentlicht: (2024)
von: Leippold, Markus, et al.
Veröffentlicht: (2024)
Factors Contributing to the False Diagnosis of Misleading Dermatofibromas
von: Nicholas Florin Kormos, et al.
Veröffentlicht: (2025)
von: Nicholas Florin Kormos, et al.
Veröffentlicht: (2025)
Changing Climate of Opinion about University Libraries.
von: Stambrook, F. G.
Veröffentlicht: (1983)
von: Stambrook, F. G.
Veröffentlicht: (1983)
Investigation of Misleading Techniques Used in Online Fluoride Misinformation
von: Matheus Lotto, et al.
Veröffentlicht: (2026)
von: Matheus Lotto, et al.
Veröffentlicht: (2026)
Evaluating and Aligning Human Economic Risk Preferences in LLMs
von: Liu, Jiaxin, et al.
Veröffentlicht: (2025)
von: Liu, Jiaxin, et al.
Veröffentlicht: (2025)
Unlearning Climate Misinformation in Large Language Models
von: Fore, Michael, et al.
Veröffentlicht: (2024)
von: Fore, Michael, et al.
Veröffentlicht: (2024)
Steering LLMs via Scalable Interactive Oversight
von: Zhou, Enyu, et al.
Veröffentlicht: (2026)
von: Zhou, Enyu, et al.
Veröffentlicht: (2026)
Explore the Potential of LLMs in Misinformation Detection: An Empirical Study
von: Chen, Mengyang, et al.
Veröffentlicht: (2023)
von: Chen, Mengyang, et al.
Veröffentlicht: (2023)
Artificial Intelligence and Civil Discourse: How LLMs Moderate Climate Change Conversations
von: Fan, Wenlu, et al.
Veröffentlicht: (2025)
von: Fan, Wenlu, et al.
Veröffentlicht: (2025)
Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs
von: Kim, Jinhwa, et al.
Veröffentlicht: (2025)
von: Kim, Jinhwa, et al.
Veröffentlicht: (2025)
When Claims Evolve: Evaluating and Enhancing the Robustness of Embedding Models Against Misinformation Edits
von: Magomere, Jabez, et al.
Veröffentlicht: (2025)
von: Magomere, Jabez, et al.
Veröffentlicht: (2025)
MultiClimate: Multimodal Stance Detection on Climate Change Videos
von: Wang, Jiawen, et al.
Veröffentlicht: (2024)
von: Wang, Jiawen, et al.
Veröffentlicht: (2024)
Towards Aligning Multimodal LLMs with Human Experts: A Focus on Parent-Child Interaction
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
Corporate Governance and Oversight
von: Richard Samans, et al.
Veröffentlicht: (2022)
von: Richard Samans, et al.
Veröffentlicht: (2022)
Climate Change Integration in the Multilevel Governance of Italy and Austria
Veröffentlicht: (2024)
Veröffentlicht: (2024)
Human Societies Facing Climate Change
Veröffentlicht: (2025)
Veröffentlicht: (2025)
XFacta: Contemporary, Real-World Dataset and Evaluation for Multimodal Misinformation Detection with Multimodal LLMs
von: Xiao, Yuzhuo, et al.
Veröffentlicht: (2025)
von: Xiao, Yuzhuo, et al.
Veröffentlicht: (2025)
Mapping Emerging Climate Misinformation Playbooks in the Global South
von: Locatelli, Marcelo Sartori, et al.
Veröffentlicht: (2026)
von: Locatelli, Marcelo Sartori, et al.
Veröffentlicht: (2026)
When Truth Misleads -- Phase-Aware Coherence Detection for Misinformation Correction Across Epistemic Communities
von: Müller, Heimo, et al.
Veröffentlicht: (2026)
von: Müller, Heimo, et al.
Veröffentlicht: (2026)
Modeling Human Beliefs about AI Behavior for Scalable Oversight
von: Lang, Leon, et al.
Veröffentlicht: (2025)
von: Lang, Leon, et al.
Veröffentlicht: (2025)
Are Rationales Necessary and Sufficient? Tuning LLMs for Explainable Misinformation Detection
von: Wang, Bing, et al.
Veröffentlicht: (2026)
von: Wang, Bing, et al.
Veröffentlicht: (2026)
Role of Media in moulding Public consciousness about Climate Change
von: Dhorje, Anuja
Veröffentlicht: (2025)
von: Dhorje, Anuja
Veröffentlicht: (2025)
When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2026)
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Evaluating the Capabilities of LLMs for Supporting Anticipatory Impact Assessment
von: Allaham, Mowafak, et al.
Veröffentlicht: (2024) -
Synthetic Sources?: Auditing Generative Search Engine Citations for Evidence of AI-Generated Sources
von: Allaham, Mowafak, et al.
Veröffentlicht: (2026) -
Towards Leveraging News Media to Support Impact Assessment of AI Technologies
von: Allaham, Mowafak, et al.
Veröffentlicht: (2024) -
Global Perspectives of AI Risks and Harms: Analyzing the Negative Impacts of AI Technologies as Prioritized by News Media
von: Allaham, Mowafak, et al.
Veröffentlicht: (2025) -
Informing AI Risk Assessment with News Media: Analyzing National and Political Variation in the Coverage of AI Risks
von: Allaham, Mowafak, et al.
Veröffentlicht: (2025)