Robust Uncertainty Quantification for Factual Generation of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yuhao, Yang, Zhongliang, Zhou, Linna |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VectraYX-Nano: A 42M-Parameter Spanish Cybersecurity Language Model with Curriculum Learning and Native Tool Use
von: Santillana, Juan S.
Veröffentlicht: (2026)
von: Santillana, Juan S.
Veröffentlicht: (2026)
Detecting Sleeper Agents in Large Language Models via Semantic Drift Analysis
von: Zanbaghi, Shahin, et al.
Veröffentlicht: (2025)
von: Zanbaghi, Shahin, et al.
Veröffentlicht: (2025)
Do Latent Tokens Think? A Causal and Adversarial Analysis of Chain-of-Continuous-Thought
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
Benchmarking Large Language Models for IoC Recovery under Adversarial Code Obfuscation and Encryption
von: Morales, Jaime, et al.
Veröffentlicht: (2026)
von: Morales, Jaime, et al.
Veröffentlicht: (2026)
Countermind: A Multi-Layered Security Architecture for Large Language Models
von: Schwarz, Dominik
Veröffentlicht: (2025)
von: Schwarz, Dominik
Veröffentlicht: (2025)
Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI
von: Qi, Jinhu, et al.
Veröffentlicht: (2026)
von: Qi, Jinhu, et al.
Veröffentlicht: (2026)
Hidden Reliability Risks in Large Language Models: Systematic Identification of Precision-Induced Output Disagreements
von: Wang, Yifei, et al.
Veröffentlicht: (2026)
von: Wang, Yifei, et al.
Veröffentlicht: (2026)
SALLIE: Safeguarding Against Latent Language & Image Exploits
von: Azov, Guy, et al.
Veröffentlicht: (2026)
von: Azov, Guy, et al.
Veröffentlicht: (2026)
CritBench: A Framework for Evaluating Cybersecurity Capabilities of Large Language Models in IEC 61850 Digital Substation Environments
von: Keppler, Gustav, et al.
Veröffentlicht: (2026)
von: Keppler, Gustav, et al.
Veröffentlicht: (2026)
Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks
von: Hu, Saisai
Veröffentlicht: (2026)
von: Hu, Saisai
Veröffentlicht: (2026)
Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers
von: Wang, Haochuan Kevin, et al.
Veröffentlicht: (2026)
von: Wang, Haochuan Kevin, et al.
Veröffentlicht: (2026)
Before the Last Token: Diagnosing Final-Token Safety Probe Failures
von: Doda, Shravan
Veröffentlicht: (2026)
von: Doda, Shravan
Veröffentlicht: (2026)
Jailbreak Mimicry: Automated Discovery of Narrative-Based Jailbreaks for Large Language Models
von: Ntais, Pavlos
Veröffentlicht: (2025)
von: Ntais, Pavlos
Veröffentlicht: (2025)
Formal Proofs as Structured Explanations: Proposing Several Tasks on Explainable Natural Language Inference
von: Abzianidze, Lasha
Veröffentlicht: (2023)
von: Abzianidze, Lasha
Veröffentlicht: (2023)
RMCBench: Benchmarking Large Language Models' Resistance to Malicious Code
von: Chen, Jiachi, et al.
Veröffentlicht: (2024)
von: Chen, Jiachi, et al.
Veröffentlicht: (2024)
Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection
von: Lelle, Travis
Veröffentlicht: (2026)
von: Lelle, Travis
Veröffentlicht: (2026)
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
von: Usman, Rana Muhammad
Veröffentlicht: (2026)
von: Usman, Rana Muhammad
Veröffentlicht: (2026)
MASH: Evading Black-Box AI-Generated Text Detectors via Style Humanization
von: Gu, Yongtong, et al.
Veröffentlicht: (2026)
von: Gu, Yongtong, et al.
Veröffentlicht: (2026)
Training Language Models to Use Prolog as a Tool
von: Mellgren, Niklas, et al.
Veröffentlicht: (2025)
von: Mellgren, Niklas, et al.
Veröffentlicht: (2025)
DeRAG: Black-box Adversarial Attacks on Multiple Retrieval-Augmented Generation Applications via Prompt Injection
von: Wang, Jerry, et al.
Veröffentlicht: (2025)
von: Wang, Jerry, et al.
Veröffentlicht: (2025)
NL2LOGIC: AST-Guided Translation of Natural Language into First-Order Logic with Large Language Models
von: Putra, Rizky Ramadhana, et al.
Veröffentlicht: (2026)
von: Putra, Rizky Ramadhana, et al.
Veröffentlicht: (2026)
Retrieval Augmented Classification for Confidential Documents
von: Chang, Yeseul E., et al.
Veröffentlicht: (2026)
von: Chang, Yeseul E., et al.
Veröffentlicht: (2026)
Large Language Models Are Not Strong Abstract Reasoners
von: Gendron, Gaël, et al.
Veröffentlicht: (2023)
von: Gendron, Gaël, et al.
Veröffentlicht: (2023)
Send to which account? Evaluation of an LLM-based Scambaiting System
von: Siadati, Hossein, et al.
Veröffentlicht: (2025)
von: Siadati, Hossein, et al.
Veröffentlicht: (2025)
Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps
von: Chona, Alankrit, et al.
Veröffentlicht: (2026)
von: Chona, Alankrit, et al.
Veröffentlicht: (2026)
Measuring Harmfulness of Computer-Using Agents
von: Tian, Aaron Xuxiang, et al.
Veröffentlicht: (2025)
von: Tian, Aaron Xuxiang, et al.
Veröffentlicht: (2025)
VoiceSHIELD-Small: Real-Time Malicious Speech Detection and Transcription
von: Ranjan, Sumit, et al.
Veröffentlicht: (2026)
von: Ranjan, Sumit, et al.
Veröffentlicht: (2026)
QoSGMAA: A Robust Multi-Order Graph Attention and Adversarial Framework for Sparse QoS Prediction
von: Du, Guanchen, et al.
Veröffentlicht: (2025)
von: Du, Guanchen, et al.
Veröffentlicht: (2025)
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
von: Othman, Refat
Veröffentlicht: (2026)
von: Othman, Refat
Veröffentlicht: (2026)
Evaluating the Reliability of Digital Forensic Evidence Discovered by Large Language Model: A Case Study
von: Khatiwala, Jeel Piyushkumar, et al.
Veröffentlicht: (2026)
von: Khatiwala, Jeel Piyushkumar, et al.
Veröffentlicht: (2026)
Adaptive Defense Orchestration for RAG: A Sentinel-Strategist Architecture against Multi-Vector Attacks
von: Pallerla, Pranav, et al.
Veröffentlicht: (2026)
von: Pallerla, Pranav, et al.
Veröffentlicht: (2026)
Grokking in the Wild: Data Augmentation for Real-World Multi-Hop Reasoning with Transformers
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
von: Chen, Renmiao, et al.
Veröffentlicht: (2025)
von: Chen, Renmiao, et al.
Veröffentlicht: (2025)
Enhancing Large Language Models through Neuro-Symbolic Integration and Ontological Reasoning
von: Vsevolodovna, Ruslan Idelfonso Magana, et al.
Veröffentlicht: (2025)
von: Vsevolodovna, Ruslan Idelfonso Magana, et al.
Veröffentlicht: (2025)
Same Payload, Different Channel: Measuring Trust Asymmetry in Tool-Using Language Models
von: Syed, Mohammed Sameer, et al.
Veröffentlicht: (2026)
von: Syed, Mohammed Sameer, et al.
Veröffentlicht: (2026)
Natural Language Interface for Firewall Configuration
von: Taghiyev, F., et al.
Veröffentlicht: (2025)
von: Taghiyev, F., et al.
Veröffentlicht: (2025)
Can Large Language Models Learn Independent Causal Mechanisms?
von: Gendron, Gaël, et al.
Veröffentlicht: (2024)
von: Gendron, Gaël, et al.
Veröffentlicht: (2024)
AIRTBench: Measuring Autonomous AI Red Teaming Capabilities in Language Models
von: Dawson, Ads, et al.
Veröffentlicht: (2025)
von: Dawson, Ads, et al.
Veröffentlicht: (2025)
AegisShield: Democratizing Cyber Threat Modeling with Generative AI
von: Grofsky, Matthew
Veröffentlicht: (2025)
von: Grofsky, Matthew
Veröffentlicht: (2025)
Ähnliche Einträge
-
VectraYX-Nano: A 42M-Parameter Spanish Cybersecurity Language Model with Curriculum Learning and Native Tool Use
von: Santillana, Juan S.
Veröffentlicht: (2026) -
Detecting Sleeper Agents in Large Language Models via Semantic Drift Analysis
von: Zanbaghi, Shahin, et al.
Veröffentlicht: (2025) -
Do Latent Tokens Think? A Causal and Adversarial Analysis of Chain-of-Continuous-Thought
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025) -
Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk
von: Wu, Shuai, et al.
Veröffentlicht: (2026) -
Benchmarking Large Language Models for IoC Recovery under Adversarial Code Obfuscation and Encryption
von: Morales, Jaime, et al.
Veröffentlicht: (2026)