Predicting LLM Correctness in Prosthodontics Using Metadata and Hallucination Signals
Fuente:
arXiv
Saved in:
| Main Authors: | Susanto, Lucky, Pranawijayana, Anasta, Sukotjo, Cortino, Prasad, Soni, Wijaya, Derry |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
by: Anugraha, David, et al.
Published: (2024)
by: Anugraha, David, et al.
Published: (2024)
MetaMetrics: Calibrating Metrics For Generation Tasks Using Human Preferences
by: Winata, Genta Indra, et al.
Published: (2024)
by: Winata, Genta Indra, et al.
Published: (2024)
NusaAksara: A Multimodal and Multilingual Benchmark for Preserving Indonesian Indigenous Scripts
by: Adilazuarda, Muhammad Farid, et al.
Published: (2025)
by: Adilazuarda, Muhammad Farid, et al.
Published: (2025)
Does Visual Rendering Bypass Tokenization? Investigating Script-Tokenizer Misalignment in Pixel-Based Language Models
by: Susanto, Lucky, et al.
Published: (2026)
by: Susanto, Lucky, et al.
Published: (2026)
Could We Have Had Better Multilingual LLMs If English Was Not the Central Language?
by: Diandaru, Ryandito, et al.
Published: (2024)
by: Diandaru, Ryandito, et al.
Published: (2024)
Is Active Persona Inference Necessary for Aligning Small Models to Personal Preferences?
by: Tang, Zilu, et al.
Published: (2025)
by: Tang, Zilu, et al.
Published: (2025)
Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation
by: Nur'aini, Khumaisa, et al.
Published: (2026)
by: Nur'aini, Khumaisa, et al.
Published: (2026)
Can We Predict the Unpredictable? Leveraging DisasterNet-LLM for Multimodal Disaster Classification
by: Kulahara, Manaswi, et al.
Published: (2025)
by: Kulahara, Manaswi, et al.
Published: (2025)
EF-LLM: Energy Forecasting LLM with AI-assisted Automation, Enhanced Sparse Prediction, Hallucination Detection
by: Qiu, Zihang, et al.
Published: (2024)
by: Qiu, Zihang, et al.
Published: (2024)
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language
by: Susanto, Lucky, et al.
Published: (2024)
by: Susanto, Lucky, et al.
Published: (2024)
Rethinking Hallucinations: Correctness, Consistency, and Prompt Multiplicity
by: Ganesh, Prakhar, et al.
Published: (2026)
by: Ganesh, Prakhar, et al.
Published: (2026)
mR3: Multilingual Rubric-Agnostic Reward Reasoning Models
by: Anugraha, David, et al.
Published: (2025)
by: Anugraha, David, et al.
Published: (2025)
Improving User Behavior Prediction: Leveraging Annotator Metadata in Supervised Machine Learning Models
by: Ng, Lynnette Hui Xian, et al.
Published: (2025)
by: Ng, Lynnette Hui Xian, et al.
Published: (2025)
Weight Decay Regimes in Grokking Transformers: Cheap Online Diagnostics
by: Verma, Lucky
Published: (2026)
by: Verma, Lucky
Published: (2026)
Squish and Release: Exposing Hidden Hallucinations by Making Them Surface as Safety Signals
by: Oh, Nathaniel, et al.
Published: (2026)
by: Oh, Nathaniel, et al.
Published: (2026)
Beyond URLs: Metadata Diversity and Position for Efficient LLM Pretraining
by: Fan, Dongyang, et al.
Published: (2025)
by: Fan, Dongyang, et al.
Published: (2025)
R3: Robust Rubric-Agnostic Reward Models
by: Anugraha, David, et al.
Published: (2025)
by: Anugraha, David, et al.
Published: (2025)
Mitigating LLM Hallucination via Behaviorally Calibrated Reinforcement Learning
by: Wu, Jiayun, et al.
Published: (2025)
by: Wu, Jiayun, et al.
Published: (2025)
LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals
by: Sun, Lihao, et al.
Published: (2026)
by: Sun, Lihao, et al.
Published: (2026)
HalluField: Detecting LLM Hallucinations via Field-Theoretic Modeling
by: Vu, Minh, et al.
Published: (2025)
by: Vu, Minh, et al.
Published: (2025)
Novel Dynamic Batch-Sensitive Adam Optimiser for Vehicular Accident Injury Severity Prediction
by: Kyei, Daniel Asare, et al.
Published: (2026)
by: Kyei, Daniel Asare, et al.
Published: (2026)
Steer LLM Latents for Hallucination Detection
by: Park, Seongheon, et al.
Published: (2025)
by: Park, Seongheon, et al.
Published: (2025)
Weakly Supervised Veracity Classification with LLM-Predicted Credibility Signals
by: Leite, João A., et al.
Published: (2023)
by: Leite, João A., et al.
Published: (2023)
Mitigating LLM Hallucinations via Conformal Abstention
by: Yadkori, Yasin Abbasi, et al.
Published: (2024)
by: Yadkori, Yasin Abbasi, et al.
Published: (2024)
Evaluating the Relevance of Uncertainty Estimators for LLM Hallucination
by: Agnimo, Yedidia, et al.
Published: (2026)
by: Agnimo, Yedidia, et al.
Published: (2026)
On Mitigating Code LLM Hallucinations with API Documentation
by: Jain, Nihal, et al.
Published: (2024)
by: Jain, Nihal, et al.
Published: (2024)
Learning to Answer from Correct Demonstrations
by: Joshi, Nirmit, et al.
Published: (2025)
by: Joshi, Nirmit, et al.
Published: (2025)
Towards Mitigation of Hallucination for LLM-empowered Agents: Progressive Generalization Bound Exploration and Watchdog Monitor
by: Liu, Siyuan, et al.
Published: (2025)
by: Liu, Siyuan, et al.
Published: (2025)
Joint Selective State Space Model and Detrending for Robust Time Series Anomaly Detection
by: Chen, Junqi, et al.
Published: (2024)
by: Chen, Junqi, et al.
Published: (2024)
AdaptStress: Online Adaptive Learning for Interpretable and Personalized Stress Prediction Using Multivariate and Sparse Physiological Signals
by: Wang, Xueyi, et al.
Published: (2026)
by: Wang, Xueyi, et al.
Published: (2026)
Graph-Grounded LLMs: Leveraging Graphical Function Calling to Minimize LLM Hallucinations
by: Gupta, Piyush, et al.
Published: (2025)
by: Gupta, Piyush, et al.
Published: (2025)
InsightBuild: LLM-Powered Causal Reasoning in Smart Building Systems
by: Neogi, Pinaki Prasad Guha, et al.
Published: (2025)
by: Neogi, Pinaki Prasad Guha, et al.
Published: (2025)
Predicting Bad Goods Risk Scores with ARIMA Time Series: A Novel Risk Assessment Approach
by: Gond, Bishwajit Prasad
Published: (2025)
by: Gond, Bishwajit Prasad
Published: (2025)
Measuring and Reducing LLM Hallucination without Gold-Standard Answers
by: Wei, Jiaheng, et al.
Published: (2024)
by: Wei, Jiaheng, et al.
Published: (2024)
Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification
by: Bouchard, Dylan, et al.
Published: (2026)
by: Bouchard, Dylan, et al.
Published: (2026)
SignalLLM: A General-Purpose LLM Agent Framework for Automated Signal Processing
by: Ke, Junlong, et al.
Published: (2025)
by: Ke, Junlong, et al.
Published: (2025)
RADAR: Mechanistic Pathways for Detecting Data Contamination in LLM Evaluation
by: Kattamuri, Ashish, et al.
Published: (2025)
by: Kattamuri, Ashish, et al.
Published: (2025)
ALCo-FM: Adaptive Long-Context Foundation Model for Accident Prediction
by: Neogi, Pinaki Prasad Guha, et al.
Published: (2025)
by: Neogi, Pinaki Prasad Guha, et al.
Published: (2025)
Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces
by: Yu, Shixing, et al.
Published: (2026)
by: Yu, Shixing, et al.
Published: (2026)
E^2-LLM: Bridging Neural Signals and Interpretable Affective Analysis
by: Ma, Fei, et al.
Published: (2026)
by: Ma, Fei, et al.
Published: (2026)
Similar Items
-
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
by: Anugraha, David, et al.
Published: (2024) -
MetaMetrics: Calibrating Metrics For Generation Tasks Using Human Preferences
by: Winata, Genta Indra, et al.
Published: (2024) -
NusaAksara: A Multimodal and Multilingual Benchmark for Preserving Indonesian Indigenous Scripts
by: Adilazuarda, Muhammad Farid, et al.
Published: (2025) -
Does Visual Rendering Bypass Tokenization? Investigating Script-Tokenizer Misalignment in Pixel-Based Language Models
by: Susanto, Lucky, et al.
Published: (2026) -
Could We Have Had Better Multilingual LLMs If English Was Not the Central Language?
by: Diandaru, Ryandito, et al.
Published: (2024)