Gespeichert in:
| Hauptverfasser: | Sakib, Abu Noman Md, Wang, Zhensen, Roby, Merjulah, Zhang, Zijie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2604.04456 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dissecting Model Failures in Abdominal Aortic Aneurysm Segmentation through Explainability-Driven Analysis
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
Explainable AI for Blind and Low-Vision Users: Navigating Trust, Modality, and Interpretability in the Agentic Era
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
Structural Rationale Distillation via Reasoning Space Compression
von: Yang, Jialin, et al.
Veröffentlicht: (2026)
von: Yang, Jialin, et al.
Veröffentlicht: (2026)
SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales
von: Xu, Tianyang, et al.
Veröffentlicht: (2024)
von: Xu, Tianyang, et al.
Veröffentlicht: (2024)
EBPO: Empirical Bayes Shrinkage for Stabilizing Group-Relative Policy Optimization
von: Han, Kevin, et al.
Veröffentlicht: (2026)
von: Han, Kevin, et al.
Veröffentlicht: (2026)
Lightning Prediction under Uncertainty: DeepLight with Hazy Loss
von: Arifin, Md Sultanul, et al.
Veröffentlicht: (2025)
von: Arifin, Md Sultanul, et al.
Veröffentlicht: (2025)
Learn-by-Wire Training Control Governance: Bounded Autonomous Training Under Stress for Stability and Efficiency
von: Radianis, Anis
Veröffentlicht: (2026)
von: Radianis, Anis
Veröffentlicht: (2026)
An Empirical Study on Preference Tuning Generalization and Diversity Under Domain Shift
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
von: Huang, Kaixuan, et al.
Veröffentlicht: (2025)
von: Huang, Kaixuan, et al.
Veröffentlicht: (2025)
Pattern Recognition or Medical Knowledge? The Problem with Multiple-Choice Questions in Medicine
von: Griot, Maxime, et al.
Veröffentlicht: (2024)
von: Griot, Maxime, et al.
Veröffentlicht: (2024)
Guided Perturbation Sensitivity (GPS): Detecting Adversarial Text via Embedding Stability and Word Importance
von: Tuck, Bryan E., et al.
Veröffentlicht: (2025)
von: Tuck, Bryan E., et al.
Veröffentlicht: (2025)
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
von: Nirmal, Ayushi, et al.
Veröffentlicht: (2024)
von: Nirmal, Ayushi, et al.
Veröffentlicht: (2024)
Self-Training Meets Consistency: Improving LLMs' Reasoning with Consistency-Driven Rationale Evaluation
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2024)
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2024)
Toward Faithful Segmentation Attribution via Benchmarking and Dual-Evidence Fusion
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Control
von: Wang, Xinyu, et al.
Veröffentlicht: (2026)
von: Wang, Xinyu, et al.
Veröffentlicht: (2026)
Exploring the Trade-off Between Model Performance and Explanation Plausibility of Text Classifiers Using Human Rationales
von: Resck, Lucas E., et al.
Veröffentlicht: (2024)
von: Resck, Lucas E., et al.
Veröffentlicht: (2024)
An Explainable Ensemble Learning Framework for Crop Classification with Optimized Feature Pyramids and Deep Networks
von: Masud, Syed Rayhan, et al.
Veröffentlicht: (2026)
von: Masud, Syed Rayhan, et al.
Veröffentlicht: (2026)
Mix-of-Granularity: Optimize the Chunking Granularity for Retrieval-Augmented Generation
von: Zhong, Zijie, et al.
Veröffentlicht: (2024)
von: Zhong, Zijie, et al.
Veröffentlicht: (2024)
Ladder: A Model-Agnostic Framework Boosting LLM-based Machine Translation to the Next Level
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
Explainable LLM Unlearning Through Reasoning
von: Liao, Junfeng, et al.
Veröffentlicht: (2026)
von: Liao, Junfeng, et al.
Veröffentlicht: (2026)
Three-Phase Transformer
von: Ayyash, Mohammad R. Abu
Veröffentlicht: (2026)
von: Ayyash, Mohammad R. Abu
Veröffentlicht: (2026)
SpaRC and SpaRP: Spatial Reasoning Characterization and Path Generation for Understanding Spatial Reasoning Capability of Large Language Models
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2024)
von: Rizvi, Md Imbesat Hassan, et al.
Veröffentlicht: (2024)
Measuring and Controlling Instruction (In)Stability in Language Model Dialogs
von: Li, Kenneth, et al.
Veröffentlicht: (2024)
von: Li, Kenneth, et al.
Veröffentlicht: (2024)
When Attention Sink Emerges in Language Models: An Empirical View
von: Gu, Xiangming, et al.
Veröffentlicht: (2024)
von: Gu, Xiangming, et al.
Veröffentlicht: (2024)
Picky LLMs and Unreliable RMs: An Empirical Study on Safety Alignment after Instruction Tuning
von: Li, Guanlin, et al.
Veröffentlicht: (2025)
von: Li, Guanlin, et al.
Veröffentlicht: (2025)
Taming Sensitive Weights : Noise Perturbation Fine-tuning for Robust LLM Quantization
von: Wang, Dongwei, et al.
Veröffentlicht: (2024)
von: Wang, Dongwei, et al.
Veröffentlicht: (2024)
Inferring from Logits: Exploring Best Practices for Decoding-Free Generative Candidate Selection
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2025)
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2025)
Grounding Language Plans in Demonstrations Through Counterfactual Perturbations
von: Wang, Yanwei, et al.
Veröffentlicht: (2024)
von: Wang, Yanwei, et al.
Veröffentlicht: (2024)
Can GRPO Help LLMs Transcend Their Pretraining Origin?
von: Ni, Kangqi, et al.
Veröffentlicht: (2025)
von: Ni, Kangqi, et al.
Veröffentlicht: (2025)
SegWithU: Uncertainty as Perturbation Energy for Single-Forward-Pass Risk-Aware Medical Image Segmentation
von: Fu, Tianhao, et al.
Veröffentlicht: (2026)
von: Fu, Tianhao, et al.
Veröffentlicht: (2026)
GTPO: Stabilizing Group Relative Policy Optimization via Gradient and Entropy Control
von: Simoni, Marco, et al.
Veröffentlicht: (2025)
von: Simoni, Marco, et al.
Veröffentlicht: (2025)
Misaligned Roles, Misplaced Images: Structural Input Perturbations Expose Multimodal Alignment Blind Spots
von: Shayegani, Erfan, et al.
Veröffentlicht: (2025)
von: Shayegani, Erfan, et al.
Veröffentlicht: (2025)
Deep Knowledge-Infusion For Explainable Depression Detection
von: Dalal, Sumit, et al.
Veröffentlicht: (2024)
von: Dalal, Sumit, et al.
Veröffentlicht: (2024)
BEExAI: Benchmark to Evaluate Explainable AI
von: Sithakoul, Samuel, et al.
Veröffentlicht: (2024)
von: Sithakoul, Samuel, et al.
Veröffentlicht: (2024)
Characterizing Pattern Matching and Its Limits on Compositional Task Structures
von: Chang, Hoyeon, et al.
Veröffentlicht: (2025)
von: Chang, Hoyeon, et al.
Veröffentlicht: (2025)
Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct
von: Ackerman, Christopher, et al.
Veröffentlicht: (2024)
von: Ackerman, Christopher, et al.
Veröffentlicht: (2024)
Toward a universal foundation model for graph-structured data
von: Mostafa, Sakib, et al.
Veröffentlicht: (2026)
von: Mostafa, Sakib, et al.
Veröffentlicht: (2026)
BANER: Boundary-Aware LLMs for Few-Shot Named Entity Recognition
von: Guo, Quanjiang, et al.
Veröffentlicht: (2024)
von: Guo, Quanjiang, et al.
Veröffentlicht: (2024)
Do Large Language Models Truly Grasp Mathematics? An Empirical Exploration From Cognitive Psychology
von: Xie, Wei, et al.
Veröffentlicht: (2024)
von: Xie, Wei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Dissecting Model Failures in Abdominal Aortic Aneurysm Segmentation through Explainability-Driven Analysis
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026) -
Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026) -
Explainable AI for Blind and Low-Vision Users: Navigating Trust, Modality, and Interpretability in the Agentic Era
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026) -
Structural Rationale Distillation via Reasoning Space Compression
von: Yang, Jialin, et al.
Veröffentlicht: (2026) -
SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales
von: Xu, Tianyang, et al.
Veröffentlicht: (2024)