CAIRNS: Balancing Readability and Scientific Accuracy in Climate Adaptation Question Answering
Fuente:
arXiv
Guardado en:
| Autores principales: | Kong, Liangji, Joshi, Aditya, Karimi, Sarvnaz |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Critical Look at Meta-evaluating Summarisation Evaluation Metrics
por: Dai, Xiang, et al.
Publicado: (2024)
por: Dai, Xiang, et al.
Publicado: (2024)
Identifying Health Risks from Family History: A Survey of Natural Language Processing Techniques
por: Dai, Xiang, et al.
Publicado: (2024)
por: Dai, Xiang, et al.
Publicado: (2024)
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language
por: Sammoudi, Mohammad, et al.
Publicado: (2024)
por: Sammoudi, Mohammad, et al.
Publicado: (2024)
SUKHSANDESH: An Avatar Therapeutic Question Answering Platform for Sexual Education in Rural India
por: Singh, Salam Michael, et al.
Publicado: (2024)
por: Singh, Salam Michael, et al.
Publicado: (2024)
MemeMQA: Multimodal Question Answering for Memes via Rationale-Based Inferencing
por: Agarwal, Siddhant, et al.
Publicado: (2024)
por: Agarwal, Siddhant, et al.
Publicado: (2024)
Social Bias in Popular Question-Answering Benchmarks
por: Kraft, Angelie, et al.
Publicado: (2025)
por: Kraft, Angelie, et al.
Publicado: (2025)
LangLingual: A Personalised, Exercise-oriented English Language Learning Tool Leveraging Large Language Models
por: Gupta, Sammriddh, et al.
Publicado: (2025)
por: Gupta, Sammriddh, et al.
Publicado: (2025)
GG-BBQ: German Gender Bias Benchmark for Question Answering
por: Satheesh, Shalaka, et al.
Publicado: (2025)
por: Satheesh, Shalaka, et al.
Publicado: (2025)
None of the Above, Less of the Right: Parallel Patterns between Humans and LLMs on Multi-Choice Questions Answering
por: Tam, Zhi Rui, et al.
Publicado: (2025)
por: Tam, Zhi Rui, et al.
Publicado: (2025)
Answering Students' Questions on Course Forums Using Multiple Chain-of-Thought Reasoning and Finetuning RAG-Enabled LLM
por: Wang, Neo, et al.
Publicado: (2025)
por: Wang, Neo, et al.
Publicado: (2025)
Towards Unsupervised Question Answering System with Multi-level Summarization for Legal Text
por: Prabhu, M Manvith, et al.
Publicado: (2024)
por: Prabhu, M Manvith, et al.
Publicado: (2024)
Reporting and Analysing the Environmental Impact of Language Models on the Example of Commonsense Question Answering with External Knowledge
por: Usmanova, Aida, et al.
Publicado: (2024)
por: Usmanova, Aida, et al.
Publicado: (2024)
SyllabusQA: A Course Logistics Question Answering Dataset
por: Fernandez, Nigel, et al.
Publicado: (2024)
por: Fernandez, Nigel, et al.
Publicado: (2024)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
por: Ma, Mingyu Derek, et al.
Publicado: (2023)
por: Ma, Mingyu Derek, et al.
Publicado: (2023)
Mental Health Equity in LLMs: Leveraging Multi-Hop Question Answering to Detect Amplified and Silenced Perspectives
por: Haider, Batool, et al.
Publicado: (2025)
por: Haider, Batool, et al.
Publicado: (2025)
PediatricsMQA: a Multi-modal Pediatrics Question Answering Benchmark
por: Bahaj, Adil, et al.
Publicado: (2025)
por: Bahaj, Adil, et al.
Publicado: (2025)
Beyond Prompting: An Efficient Embedding Framework for Open-Domain Question Answering
por: Hu, Zhanghao, et al.
Publicado: (2025)
por: Hu, Zhanghao, et al.
Publicado: (2025)
MultiADE: A Multi-domain Benchmark for Adverse Drug Event Extraction
por: Dai, Xiang, et al.
Publicado: (2024)
por: Dai, Xiang, et al.
Publicado: (2024)
DeCAP: Context-Adaptive Prompt Generation for Debiasing Zero-shot Question Answering in Large Language Models
por: Bae, Suyoung, et al.
Publicado: (2025)
por: Bae, Suyoung, et al.
Publicado: (2025)
RephQA: Evaluating Readability of Large Language Models in Public Health Question Answering
por: Qiu, Weikang, et al.
Publicado: (2025)
por: Qiu, Weikang, et al.
Publicado: (2025)
CSIRO-LT at SemEval-2025 Task 11: Adapting LLMs for Emotion Recognition for Multiple Languages
por: Chen, Jiyu, et al.
Publicado: (2025)
por: Chen, Jiyu, et al.
Publicado: (2025)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
por: Samory, Mattia, et al.
Publicado: (2025)
por: Samory, Mattia, et al.
Publicado: (2025)
MahaSQuAD: Bridging Linguistic Divides in Marathi Question-Answering
por: Ghatage, Ruturaj, et al.
Publicado: (2024)
por: Ghatage, Ruturaj, et al.
Publicado: (2024)
Computational Analysis of Climate Policy
por: Hicks, Carolyn
Publicado: (2025)
por: Hicks, Carolyn
Publicado: (2025)
Generative Debunking of Climate Misinformation
por: Zanartu, Francisco, et al.
Publicado: (2024)
por: Zanartu, Francisco, et al.
Publicado: (2024)
Multi-Hop Reasoning for Question Answering with Hyperbolic Representations
por: Welz, Simon, et al.
Publicado: (2025)
por: Welz, Simon, et al.
Publicado: (2025)
Trust, Safety, and Accuracy: Assessing LLMs for Routine Maternity Advice
por: Divya, V Sai, et al.
Publicado: (2026)
por: Divya, V Sai, et al.
Publicado: (2026)
SafeMath: Inference-time Safety improves Math Accuracy
por: Basu, Sagnik, et al.
Publicado: (2026)
por: Basu, Sagnik, et al.
Publicado: (2026)
LLMs Provide Unstable Answers to Legal Questions
por: Blair-Stanek, Andrew, et al.
Publicado: (2025)
por: Blair-Stanek, Andrew, et al.
Publicado: (2025)
Designing and Evaluating Chain-of-Hints for Scientific Question Answering
por: Jangra, Anubhav, et al.
Publicado: (2025)
por: Jangra, Anubhav, et al.
Publicado: (2025)
Can AI Extract Antecedent Factors of Human Trust in AI? An Application of Information Extraction for Scientific Literature in Behavioural and Computer Sciences
por: McGrath, Melanie, et al.
Publicado: (2024)
por: McGrath, Melanie, et al.
Publicado: (2024)
Climate Change from Large Language Models
por: Zhu, Hongyin, et al.
Publicado: (2023)
por: Zhu, Hongyin, et al.
Publicado: (2023)
IndicSQuAD: A Comprehensive Multilingual Question Answering Dataset for Indic Languages
por: Endait, Sharvi, et al.
Publicado: (2025)
por: Endait, Sharvi, et al.
Publicado: (2025)
C-QUERI: Congressional Questions, Exchanges, and Responses in Institutions Dataset
por: Rudra, Manjari, et al.
Publicado: (2025)
por: Rudra, Manjari, et al.
Publicado: (2025)
Beyond Accuracy: Diagnosing Algebraic Reasoning Failures in LLMs Across Nine Complexity Dimensions
por: Patil, Parth, et al.
Publicado: (2026)
por: Patil, Parth, et al.
Publicado: (2026)
Cancer-Myth: Evaluating Large Language Models on Patient Questions with False Presuppositions
por: Zhu, Wang Bill, et al.
Publicado: (2025)
por: Zhu, Wang Bill, et al.
Publicado: (2025)
Automated Question Generation for Science Tests in Arabic Language Using NLP Techniques
por: Tami, Mohammad, et al.
Publicado: (2024)
por: Tami, Mohammad, et al.
Publicado: (2024)
Does Scientific Writing Converge to U.S. English? Evidence from Generative AI-Assisted Publications
por: Filimonovic, Dragan, et al.
Publicado: (2025)
por: Filimonovic, Dragan, et al.
Publicado: (2025)
Artificial Intelligence and Civil Discourse: How LLMs Moderate Climate Change Conversations
por: Fan, Wenlu, et al.
Publicado: (2025)
por: Fan, Wenlu, et al.
Publicado: (2025)
When AI Takes Sides on Questions of Faith: Persistent Asymmetries in AI-Mediated Faith Guidance
por: Israelsen, Brett, et al.
Publicado: (2026)
por: Israelsen, Brett, et al.
Publicado: (2026)
Ejemplares similares
-
A Critical Look at Meta-evaluating Summarisation Evaluation Metrics
por: Dai, Xiang, et al.
Publicado: (2024) -
Identifying Health Risks from Family History: A Survey of Natural Language Processing Techniques
por: Dai, Xiang, et al.
Publicado: (2024) -
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language
por: Sammoudi, Mohammad, et al.
Publicado: (2024) -
SUKHSANDESH: An Avatar Therapeutic Question Answering Platform for Sexual Education in Rural India
por: Singh, Salam Michael, et al.
Publicado: (2024) -
MemeMQA: Multimodal Question Answering for Memes via Rationale-Based Inferencing
por: Agarwal, Siddhant, et al.
Publicado: (2024)