Saved in:
| Main Authors: | Vachhani, Bhavik, Shrisvastava, Kush, Nema, Pranshu, Chiranthan, Sai |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.14829 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Skin-SOAP: A Weakly Supervised Framework for Generating Structured SOAP Notes
by: Kamal, Sadia, et al.
Published: (2025)
by: Kamal, Sadia, et al.
Published: (2025)
When Reasoning Hurts: Source-Aware Evaluation of Frontier LLMs for Clinical SOAP Note Generation
by: Faisal, Faizan
Published: (2026)
by: Faisal, Faizan
Published: (2026)
Beyond Accuracy: Risk-Sensitive Evaluation of Hallucinated Medical Advice
by: Doshi, Savan
Published: (2026)
by: Doshi, Savan
Published: (2026)
Towards Scalable SOAP Note Generation: A Weakly Supervised Multimodal Framework
by: Kamal, Sadia, et al.
Published: (2025)
by: Kamal, Sadia, et al.
Published: (2025)
Are LLMs Court-Ready? Evaluating Frontier Models on Indian Legal Reasoning
by: Juvekar, Kush, et al.
Published: (2025)
by: Juvekar, Kush, et al.
Published: (2025)
fact check AI at SemEval-2025 Task 7: Multilingual and Crosslingual Fact-checked Claim Retrieval
by: Rastogi, Pranshu
Published: (2025)
by: Rastogi, Pranshu
Published: (2025)
Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM Architecture
by: Lee, Yeawon, et al.
Published: (2025)
by: Lee, Yeawon, et al.
Published: (2025)
TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization
by: Tang, Liyan, et al.
Published: (2024)
by: Tang, Liyan, et al.
Published: (2024)
Enhancing Text Classification with a Novel Multi-Agent Collaboration Framework Leveraging BERT
by: Baban, Hediyeh, et al.
Published: (2025)
by: Baban, Hediyeh, et al.
Published: (2025)
Correction with Backtracking Reduces Hallucination in Summarization
by: Liu, Zhenzhen, et al.
Published: (2023)
by: Liu, Zhenzhen, et al.
Published: (2023)
SOAP: Improving and Stabilizing Shampoo using Adam
by: Vyas, Nikhil, et al.
Published: (2024)
by: Vyas, Nikhil, et al.
Published: (2024)
Extrinsically-Focused Evaluation of Omissions in Medical Summarization
by: Schumacher, Elliot, et al.
Published: (2023)
by: Schumacher, Elliot, et al.
Published: (2023)
Low-light Pedestrian Detection in Visible and Infrared Image Feeds: Issues and Challenges
by: Akilan, Thangarajah, et al.
Published: (2023)
by: Akilan, Thangarajah, et al.
Published: (2023)
AI coach for badminton
by: Toshniwal, Dhruv, et al.
Published: (2024)
by: Toshniwal, Dhruv, et al.
Published: (2024)
Hallucination Detection-Guided Preference Optimization for Clinical Summarization
by: Seethakantha, Shamanth Kuthpadi, et al.
Published: (2026)
by: Seethakantha, Shamanth Kuthpadi, et al.
Published: (2026)
Redefining "Hallucination" in LLMs: Towards a psychology-informed framework for mitigating misinformation
by: Berberette, Elijah, et al.
Published: (2024)
by: Berberette, Elijah, et al.
Published: (2024)
Mitigating Hallucination in Abstractive Summarization with Domain-Conditional Mutual Information
by: Chae, Kyubyung, et al.
Published: (2024)
by: Chae, Kyubyung, et al.
Published: (2024)
Investigating Hallucinations in Pruned Large Language Models for Abstractive Summarization
by: Chrysostomou, George, et al.
Published: (2023)
by: Chrysostomou, George, et al.
Published: (2023)
Evaluating LLMs' Assessment of Mixed-Context Hallucination Through the Lens of Summarization
by: Qi, Siya, et al.
Published: (2025)
by: Qi, Siya, et al.
Published: (2025)
Hallucination Diversity-Aware Active Learning for Text Summarization
by: Xia, Yu, et al.
Published: (2024)
by: Xia, Yu, et al.
Published: (2024)
MDKeyChunker: Single-Call LLM Enrichment with Rolling Keys and Key-Based Restructuring for High-Accuracy RAG
by: Mangla, Bhavik
Published: (2026)
by: Mangla, Bhavik
Published: (2026)
An Algebraic Exposition of the Theory of Dyadic Morality
by: Varshney, Kush R.
Published: (2026)
by: Varshney, Kush R.
Published: (2026)
The Human-Machine Identity Blur: A Unified Framework for Cybersecurity Risk Management in 2025
by: Janani, Kush
Published: (2025)
by: Janani, Kush
Published: (2025)
Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective
by: Chandna, Bhavik, et al.
Published: (2025)
by: Chandna, Bhavik, et al.
Published: (2025)
Reducing Hallucinations in Summarization via Reinforcement Learning with Entity Hallucination Index
by: Katwe, Praveenkumar, et al.
Published: (2025)
by: Katwe, Praveenkumar, et al.
Published: (2025)
Beyond Facts: Evaluating Intent Hallucination in Large Language Models
by: Hao, Yijie, et al.
Published: (2025)
by: Hao, Yijie, et al.
Published: (2025)
Can GPT Redefine Medical Understanding? Evaluating GPT on Biomedical Machine Reading Comprehension
by: Vatsal, Shubham, et al.
Published: (2024)
by: Vatsal, Shubham, et al.
Published: (2024)
SOAP-RL: Sequential Option Advantage Propagation for Reinforcement Learning in POMDP Environments
by: Ishida, Shu, et al.
Published: (2024)
by: Ishida, Shu, et al.
Published: (2024)
Plug-in for visualizing 3D tool tracking from videos of Minimally Invasive Surgeries
by: Nema, Shubhangi, et al.
Published: (2024)
by: Nema, Shubhangi, et al.
Published: (2024)
Analyzing LLM Behavior in Dialogue Summarization: Unveiling Circumstantial Hallucination Trends
by: Ramprasad, Sanjana, et al.
Published: (2024)
by: Ramprasad, Sanjana, et al.
Published: (2024)
FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs
by: Bao, Forrest Sheng, et al.
Published: (2024)
by: Bao, Forrest Sheng, et al.
Published: (2024)
Improving Summarization with Human Edits
by: Yao, Zonghai, et al.
Published: (2023)
by: Yao, Zonghai, et al.
Published: (2023)
Code Summarization Beyond Function Level
by: Makharev, Vladimir, et al.
Published: (2025)
by: Makharev, Vladimir, et al.
Published: (2025)
MedHal: An Evaluation Dataset for Medical Hallucination Detection
by: Mehenni, Gaya, et al.
Published: (2025)
by: Mehenni, Gaya, et al.
Published: (2025)
Evaluation of Large Language Models for Summarization Tasks in the Medical Domain: A Narrative Review
by: Croxford, Emma, et al.
Published: (2024)
by: Croxford, Emma, et al.
Published: (2024)
Query-Guided Self-Supervised Summarization of Nursing Notes
by: Gao, Ya, et al.
Published: (2024)
by: Gao, Ya, et al.
Published: (2024)
Evaluating Concurrent Robustness of Language Models Across Diverse Challenge Sets
by: Gupta, Vatsal, et al.
Published: (2023)
by: Gupta, Vatsal, et al.
Published: (2023)
Think Inside the JSON: Reinforcement Strategy for Strict LLM Schema Adherence
by: Agarwal, Bhavik, et al.
Published: (2025)
by: Agarwal, Bhavik, et al.
Published: (2025)
MedSAGa: Few-shot Memory Efficient Medical Image Segmentation using Gradient Low-Rank Projection in SAM
by: Mahla, Navyansh, et al.
Published: (2024)
by: Mahla, Navyansh, et al.
Published: (2024)
Utilizing GPT to Enhance Text Summarization: A Strategy to Minimize Hallucinations
by: Shakil, Hassan, et al.
Published: (2024)
by: Shakil, Hassan, et al.
Published: (2024)
Similar Items
-
Skin-SOAP: A Weakly Supervised Framework for Generating Structured SOAP Notes
by: Kamal, Sadia, et al.
Published: (2025) -
When Reasoning Hurts: Source-Aware Evaluation of Frontier LLMs for Clinical SOAP Note Generation
by: Faisal, Faizan
Published: (2026) -
Beyond Accuracy: Risk-Sensitive Evaluation of Hallucinated Medical Advice
by: Doshi, Savan
Published: (2026) -
Towards Scalable SOAP Note Generation: A Weakly Supervised Multimodal Framework
by: Kamal, Sadia, et al.
Published: (2025) -
Are LLMs Court-Ready? Evaluating Frontier Models on Indian Legal Reasoning
by: Juvekar, Kush, et al.
Published: (2025)