AI-Facilitated Analysis of Abstracts and Conclusions: Flagging Unsubstantiated Claims and Ambiguous Pronouns
Fuente:
arXiv
Salvato in:
| Autore principale: | Markhasin, Evgeny |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstracts
di: Li, Weiyue, et al.
Pubblicazione: (2026)
di: Li, Weiyue, et al.
Pubblicazione: (2026)
LLM Context Conditioning and PWP Prompting for Multimodal Validation of Chemical Formulas
di: Markhasin, Evgeny
Pubblicazione: (2025)
di: Markhasin, Evgeny
Pubblicazione: (2025)
AI-Driven Scholarly Peer Review via Persistent Workflow Prompting, Meta-Prompting, and Meta-Reasoning
di: Markhasin, Evgeny
Pubblicazione: (2025)
di: Markhasin, Evgeny
Pubblicazione: (2025)
Quality Estimation based Feedback Training for Improving Pronoun Translation
di: Dhankhar, Harshit, et al.
Pubblicazione: (2025)
di: Dhankhar, Harshit, et al.
Pubblicazione: (2025)
Persian Pronoun Resolution: Leveraging Neural Networks and Language Models
di: Mohammadi, Hassan Haji, et al.
Pubblicazione: (2024)
di: Mohammadi, Hassan Haji, et al.
Pubblicazione: (2024)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
di: Fernandes, Gustavo Lúcius, et al.
Pubblicazione: (2026)
di: Fernandes, Gustavo Lúcius, et al.
Pubblicazione: (2026)
Do They Understand Them? An Updated Evaluation on Nonbinary Pronoun Handling in Large Language Models
di: Tang, Xushuo, et al.
Pubblicazione: (2025)
di: Tang, Xushuo, et al.
Pubblicazione: (2025)
Underspecification in Language Modeling Tasks: A Causality-Informed Study of Gendered Pronoun Resolution
di: McMilin, Emily
Pubblicazione: (2022)
di: McMilin, Emily
Pubblicazione: (2022)
Silver-Tongued and Sundry: Exploring Intersectional Pronouns with ChatGPT
di: Fujii, Takao, et al.
Pubblicazione: (2024)
di: Fujii, Takao, et al.
Pubblicazione: (2024)
Reasoning about Intent for Ambiguous Requests
di: Saparina, Irina, et al.
Pubblicazione: (2025)
di: Saparina, Irina, et al.
Pubblicazione: (2025)
Accelerate Creation of Product Claims Using Generative AI
di: Liang, Po-Yu, et al.
Pubblicazione: (2025)
di: Liang, Po-Yu, et al.
Pubblicazione: (2025)
MedMeta: A Benchmark for LLMs in Synthesizing Meta-Analysis Conclusion from Medical Studies
di: Ha, Huy Hoang, et al.
Pubblicazione: (2026)
di: Ha, Huy Hoang, et al.
Pubblicazione: (2026)
Mevaker: Conclusion Extraction and Allocation Resources for the Hebrew Language
di: Shalumov, Vitaly, et al.
Pubblicazione: (2024)
di: Shalumov, Vitaly, et al.
Pubblicazione: (2024)
InteractComp: Evaluating Search Agents With Ambiguous Queries
di: Deng, Mingyi, et al.
Pubblicazione: (2025)
di: Deng, Mingyi, et al.
Pubblicazione: (2025)
Classifying Human-Generated and AI-Generated Election Claims in Social Media
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
SciClaims: An End-to-End Generative System for Biomedical Claim Analysis
di: Ortega, Raúl, et al.
Pubblicazione: (2025)
di: Ortega, Raúl, et al.
Pubblicazione: (2025)
Analyzing the Attention Heads for Pronoun Disambiguation in Context-aware Machine Translation Models
di: Mąka, Paweł, et al.
Pubblicazione: (2024)
di: Mąka, Paweł, et al.
Pubblicazione: (2024)
DEEPAMBIGQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness
di: Ji, Jiabao, et al.
Pubblicazione: (2025)
di: Ji, Jiabao, et al.
Pubblicazione: (2025)
ClaimGen-CN: A Large-scale Chinese Dataset for Legal Claim Generation
di: Zhou, Siying, et al.
Pubblicazione: (2025)
di: Zhou, Siying, et al.
Pubblicazione: (2025)
Transforming Dutch: Debiasing Dutch Coreference Resolution Systems for Non-binary Pronouns
di: van Boven, Goya, et al.
Pubblicazione: (2024)
di: van Boven, Goya, et al.
Pubblicazione: (2024)
PRACTIQ: A Practical Conversational Text-to-SQL dataset with Ambiguous and Unanswerable Queries
di: Dong, Mingwen, et al.
Pubblicazione: (2024)
di: Dong, Mingwen, et al.
Pubblicazione: (2024)
Can AI Validate Science? Benchmarking LLMs for Accurate Scientific Claim $\rightarrow$ Evidence Reasoning
di: Javaji, Shashidhar Reddy, et al.
Pubblicazione: (2025)
di: Javaji, Shashidhar Reddy, et al.
Pubblicazione: (2025)
GraphMind: Theorem Selection and Conclusion Generation Framework with Dynamic GNN for LLM Reasoning
di: Li, Yutong, et al.
Pubblicazione: (2025)
di: Li, Yutong, et al.
Pubblicazione: (2025)
ClaimIQ at CheckThat! 2025: Comparing Prompted and Fine-Tuned Language Models for Verifying Numerical Claims
di: Anik, Anirban Saha, et al.
Pubblicazione: (2025)
di: Anik, Anirban Saha, et al.
Pubblicazione: (2025)
ClaimBrush: A Novel Framework for Automated Patent Claim Refinement Based on Large Language Models
di: Kawano, Seiya, et al.
Pubblicazione: (2024)
di: Kawano, Seiya, et al.
Pubblicazione: (2024)
Can Unconfident LLM Annotations Be Used for Confident Conclusions?
di: Gligorić, Kristina, et al.
Pubblicazione: (2024)
di: Gligorić, Kristina, et al.
Pubblicazione: (2024)
Optimizing Decomposition for Optimal Claim Verification
di: Lu, Yining, et al.
Pubblicazione: (2025)
di: Lu, Yining, et al.
Pubblicazione: (2025)
MedRedFlag: Investigating how LLMs Redirect Misconceptions in Real-World Health Communication
di: Sambara, Sraavya, et al.
Pubblicazione: (2026)
di: Sambara, Sraavya, et al.
Pubblicazione: (2026)
Do AI Models Perform Human-like Abstract Reasoning Across Modalities?
di: Beger, Claas, et al.
Pubblicazione: (2025)
di: Beger, Claas, et al.
Pubblicazione: (2025)
The Alignment Bottleneck in Decomposition-Based Claim Verification
di: Akhter, Mahmud Elahi, et al.
Pubblicazione: (2026)
di: Akhter, Mahmud Elahi, et al.
Pubblicazione: (2026)
Robust Claim Verification Through Fact Detection
di: Jafari, Nazanin, et al.
Pubblicazione: (2024)
di: Jafari, Nazanin, et al.
Pubblicazione: (2024)
Knowledge AI: Fine-tuning NLP Models for Facilitating Scientific Knowledge Extraction and Understanding
di: Muralidharan, Balaji, et al.
Pubblicazione: (2024)
di: Muralidharan, Balaji, et al.
Pubblicazione: (2024)
ChatGPT Perpetuates Gender Bias in Machine Translation and Ignores Non-Gendered Pronouns: Findings across Bengali and Five other Low-Resource Languages
di: Ghosh, Sourojit, et al.
Pubblicazione: (2023)
di: Ghosh, Sourojit, et al.
Pubblicazione: (2023)
ClaimPKG: Enhancing Claim Verification via Pseudo-Subgraph Generation with Lightweight Specialized LLM
di: Pham, Hoang, et al.
Pubblicazione: (2025)
di: Pham, Hoang, et al.
Pubblicazione: (2025)
Fine-grained Claim-level RAG Benchmark for Law
di: Das, Souvick, et al.
Pubblicazione: (2026)
di: Das, Souvick, et al.
Pubblicazione: (2026)
Position: Stop Making Unscientific AGI Performance Claims
di: Altmeyer, Patrick, et al.
Pubblicazione: (2024)
di: Altmeyer, Patrick, et al.
Pubblicazione: (2024)
Examining the Metrics for Document-Level Claim Extraction in Czech and Slovak
di: Makaiova, Lucia, et al.
Pubblicazione: (2025)
di: Makaiova, Lucia, et al.
Pubblicazione: (2025)
Agent-based Automated Claim Matching with Instruction-following LLMs
di: Pisarevskaya, Dina, et al.
Pubblicazione: (2025)
di: Pisarevskaya, Dina, et al.
Pubblicazione: (2025)
Piecing It All Together: Verifying Multi-Hop Multimodal Claims
di: Wang, Haoran, et al.
Pubblicazione: (2024)
di: Wang, Haoran, et al.
Pubblicazione: (2024)
Claim Verification in the Age of Large Language Models: A Survey
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2024)
Documenti analoghi
-
MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstracts
di: Li, Weiyue, et al.
Pubblicazione: (2026) -
LLM Context Conditioning and PWP Prompting for Multimodal Validation of Chemical Formulas
di: Markhasin, Evgeny
Pubblicazione: (2025) -
AI-Driven Scholarly Peer Review via Persistent Workflow Prompting, Meta-Prompting, and Meta-Reasoning
di: Markhasin, Evgeny
Pubblicazione: (2025) -
Quality Estimation based Feedback Training for Improving Pronoun Translation
di: Dhankhar, Harshit, et al.
Pubblicazione: (2025) -
Persian Pronoun Resolution: Leveraging Neural Networks and Language Models
di: Mohammadi, Hassan Haji, et al.
Pubblicazione: (2024)