Assessing Large Language Models on Climate Information
Fuente:
arXiv
Saved in:
| Main Authors: | Bulian, Jannis, Schäfer, Mike S., Amini, Afra, Lam, Heidi, Ciaramita, Massimiliano, Gaiarin, Ben, Hübscher, Michelle Chen, Buck, Christian, Mede, Niels G., Leippold, Markus, Strauß, Nadine |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CLINB: A Climate Intelligence Benchmark for Foundational Models
by: Huebscher, Michelle Chen, et al.
Published: (2025)
by: Huebscher, Michelle Chen, et al.
Published: (2025)
AI-Assisted Scientific Assessment: A Case Study on Climate Change
by: Buck, Christian, et al.
Published: (2026)
by: Buck, Christian, et al.
Published: (2026)
How Susceptible are LLMs to Influence in Prompts?
by: Anagnostidis, Sotiris, et al.
Published: (2024)
by: Anagnostidis, Sotiris, et al.
Published: (2024)
Actors, Frames and Arguments: A Multi-Decade Computational Analysis of Climate Discourse in Financial News using Large Language Models
by: Su, Ruiran, et al.
Published: (2026)
by: Su, Ruiran, et al.
Published: (2026)
Structured Voronoi Sampling
by: Amini, Afra, et al.
Published: (2023)
by: Amini, Afra, et al.
Published: (2023)
Automated Evidence Extraction and Scoring for Corporate Climate Policy Engagement: A Multilingual RAG Approach
by: Kolli, Imene, et al.
Published: (2025)
by: Kolli, Imene, et al.
Published: (2025)
Better Estimation of the Kullback--Leibler Divergence Between Language Models
by: Amini, Afra, et al.
Published: (2025)
by: Amini, Afra, et al.
Published: (2025)
Direct Preference Optimization with an Offset
by: Amini, Afra, et al.
Published: (2024)
by: Amini, Afra, et al.
Published: (2024)
AI for Climate Finance: Agentic Retrieval and Multi-Step Reasoning for Early Warning System Investments
by: Vaghefi, Saeid Ario, et al.
Published: (2025)
by: Vaghefi, Saeid Ario, et al.
Published: (2025)
Syntactic Control of Language Models by Posterior Inference
by: Xefteri, Vicky, et al.
Published: (2025)
by: Xefteri, Vicky, et al.
Published: (2025)
Variational Best-of-N Alignment
by: Amini, Afra, et al.
Published: (2024)
by: Amini, Afra, et al.
Published: (2024)
Automated Fact-Checking of Climate Change Claims with Large Language Models
by: Leippold, Markus, et al.
Published: (2024)
by: Leippold, Markus, et al.
Published: (2024)
Towards Faithful and Robust LLM Specialists for Evidence-Based Question-Answering
by: Schimanski, Tobias, et al.
Published: (2024)
by: Schimanski, Tobias, et al.
Published: (2024)
The Role of $n$-gram Smoothing in the Age of Neural Networks
by: Malagutti, Luca, et al.
Published: (2024)
by: Malagutti, Luca, et al.
Published: (2024)
Machine Translation Models are Zero-Shot Detectors of Translation Direction
by: Wastl, Michelle, et al.
Published: (2024)
by: Wastl, Michelle, et al.
Published: (2024)
SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents
by: Wastl, Michelle, et al.
Published: (2025)
by: Wastl, Michelle, et al.
Published: (2025)
AFaCTA: Assisting the Annotation of Factual Claim Detection with Reliable LLM Annotators
by: Ni, Jingwei, et al.
Published: (2024)
by: Ni, Jingwei, et al.
Published: (2024)
UsefulBench: Towards Decision-Useful Information as a Target for Information Retrieval
by: Schimanski, Tobias, et al.
Published: (2026)
by: Schimanski, Tobias, et al.
Published: (2026)
Remembering Unequally: Global and Disciplinary Bias in LLM Reconstruction of Scholarly Coauthor Lists
by: Kalhor, Ghazal, et al.
Published: (2025)
by: Kalhor, Ghazal, et al.
Published: (2025)
Reverse-Engineering the Reader
by: Kiegeland, Samuel, et al.
Published: (2024)
by: Kiegeland, Samuel, et al.
Published: (2024)
Can Reasoning Help Large Language Models Capture Human Annotator Disagreement?
by: Ni, Jingwei, et al.
Published: (2025)
by: Ni, Jingwei, et al.
Published: (2025)
Phase Transitions in the Output Distribution of Large Language Models
by: Arnold, Julian, et al.
Published: (2024)
by: Arnold, Julian, et al.
Published: (2024)
DIRAS: Efficient LLM Annotation of Document Relevance in Retrieval Augmented Generation
by: Ni, Jingwei, et al.
Published: (2024)
by: Ni, Jingwei, et al.
Published: (2024)
Exploring Nature: Datasets and Models for Analyzing Nature-Related Disclosures
by: Schimanski, Tobias, et al.
Published: (2023)
by: Schimanski, Tobias, et al.
Published: (2023)
Assessing Political Bias in Large Language Models
by: Rettenberger, Luca, et al.
Published: (2024)
by: Rettenberger, Luca, et al.
Published: (2024)
20min-XD: A Comparable Corpus of Swiss News Articles
by: Wastl, Michelle, et al.
Published: (2025)
by: Wastl, Michelle, et al.
Published: (2025)
WikiChat: Stopping the Hallucination of Large Language Model Chatbots by Few-Shot Grounding on Wikipedia
by: Semnani, Sina J., et al.
Published: (2023)
by: Semnani, Sina J., et al.
Published: (2023)
Libraries in the Cloud: Making a Case for Google and Amazon
by: Buck, Stephanie
Published: (2009)
by: Buck, Stephanie
Published: (2009)
A Linguistically Motivated Analysis of Intonational Phrasing in Text-to-Speech Systems: Revealing Gaps in Syntactic Sensitivity
by: Pouw, Charlotte, et al.
Published: (2025)
by: Pouw, Charlotte, et al.
Published: (2025)
How Language Models Prioritize Contextual Grammatical Cues?
by: Amirzadeh, Hamidreza, et al.
Published: (2024)
by: Amirzadeh, Hamidreza, et al.
Published: (2024)
pdfQA: Diverse, Challenging, and Realistic Question Answering over PDFs
by: Schimanski, Tobias, et al.
Published: (2026)
by: Schimanski, Tobias, et al.
Published: (2026)
Principled Gradient-based Markov Chain Monte Carlo for Text Generation
by: Du, Li, et al.
Published: (2023)
by: Du, Li, et al.
Published: (2023)
Source-primed Multi-turn Conversation Helps Large Language Models Translate Documents
by: Hu, Hanxu, et al.
Published: (2025)
by: Hu, Hanxu, et al.
Published: (2025)
ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models
by: Ni, Jingwei, et al.
Published: (2025)
by: Ni, Jingwei, et al.
Published: (2025)
Balancing Truthfulness and Informativeness with Uncertainty-Aware Instruction Fine-Tuning
by: Wu, Tianyi, et al.
Published: (2025)
by: Wu, Tianyi, et al.
Published: (2025)
IoTCO2: Assessing the End-To-End Carbon Footprint of Internet-of-Things-Enabled Deep Learning
by: Chen, Fan, et al.
Published: (2024)
by: Chen, Fan, et al.
Published: (2024)
ToxiTwitch: Toward Emote-Aware Hybrid Moderation for Live Streaming Platforms
by: Ansari, Baktash, et al.
Published: (2026)
by: Ansari, Baktash, et al.
Published: (2026)
Are We Paying Attention to Her? Investigating Gender Disambiguation and Attention in Machine Translation
by: Manna, Chiara, et al.
Published: (2025)
by: Manna, Chiara, et al.
Published: (2025)
Large Language Models are Zero-Shot Next Location Predictors
by: Beneduce, Ciro, et al.
Published: (2024)
by: Beneduce, Ciro, et al.
Published: (2024)
Phoenix: A Federated Generative Diffusion Model
by: Jothiraj, Fiona Victoria Stanley, et al.
Published: (2023)
by: Jothiraj, Fiona Victoria Stanley, et al.
Published: (2023)
Similar Items
-
CLINB: A Climate Intelligence Benchmark for Foundational Models
by: Huebscher, Michelle Chen, et al.
Published: (2025) -
AI-Assisted Scientific Assessment: A Case Study on Climate Change
by: Buck, Christian, et al.
Published: (2026) -
How Susceptible are LLMs to Influence in Prompts?
by: Anagnostidis, Sotiris, et al.
Published: (2024) -
Actors, Frames and Arguments: A Multi-Decade Computational Analysis of Climate Discourse in Financial News using Large Language Models
by: Su, Ruiran, et al.
Published: (2026) -
Structured Voronoi Sampling
by: Amini, Afra, et al.
Published: (2023)