Gespeichert in:
| 1. Verfasser: | Advani, Laksh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.00513 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Trajectory Guard -- A Lightweight, Sequence-Aware Model for Real-Time Anomaly Detection in Agentic AI
von: Advani, Laksh
Veröffentlicht: (2026)
von: Advani, Laksh
Veröffentlicht: (2026)
Clever Materials: When Models Identify Good Materials for the Wrong Reasons
von: Jablonka, Kevin Maik
Veröffentlicht: (2026)
von: Jablonka, Kevin Maik
Veröffentlicht: (2026)
Data Cartography for Detecting Memorization Hotspots and Guiding Data Interventions in Generative Models
von: Patel, Laksh, et al.
Veröffentlicht: (2025)
von: Patel, Laksh, et al.
Veröffentlicht: (2025)
Wrong Model, Right Uncertainty: Spatial Associations for Discrete Data with Misspecification
von: Burt, David R., et al.
Veröffentlicht: (2025)
von: Burt, David R., et al.
Veröffentlicht: (2025)
Perplexity Cannot Always Tell Right from Wrong
von: Veličković, Petar, et al.
Veröffentlicht: (2026)
von: Veličković, Petar, et al.
Veröffentlicht: (2026)
When World Models Dream Wrong: Physical-Conditioned Adversarial Attacks against World Models
von: Guo, Zhixiang, et al.
Veröffentlicht: (2026)
von: Guo, Zhixiang, et al.
Veröffentlicht: (2026)
Right Now, Wrong Then: Non-Stationary Direct Preference Optimization under Preference Drift
von: Son, Seongho, et al.
Veröffentlicht: (2024)
von: Son, Seongho, et al.
Veröffentlicht: (2024)
Stable but Wrong: When More Data Degrades Scientific Conclusions
von: Zhang, Zhipeng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhipeng, et al.
Veröffentlicht: (2026)
When to Trust the Cheap Check: Weak and Strong Verification for Reasoning
von: Kiyani, Shayan, et al.
Veröffentlicht: (2026)
von: Kiyani, Shayan, et al.
Veröffentlicht: (2026)
When PINNs Go Wrong: Pseudo-Time Stepping Against Spurious Solutions
von: Wang, Sifan, et al.
Veröffentlicht: (2026)
von: Wang, Sifan, et al.
Veröffentlicht: (2026)
[Experiments & Analysis] Evaluating the Feasibility of Sampling-Based Techniques for Training Multilayer Perceptrons
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2023)
von: Ebrahimi, Sana, et al.
Veröffentlicht: (2023)
When LLMs Learn to Be Consistently Wrong: A Multi-Model Study of Linear Representations of Synthetic Deception
von: Zolfaghari, Vahideh
Veröffentlicht: (2026)
von: Zolfaghari, Vahideh
Veröffentlicht: (2026)
Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents
von: Long, Carol Xuan
Veröffentlicht: (2026)
von: Long, Carol Xuan
Veröffentlicht: (2026)
When To Solve, When To Verify: Compute-Optimal Problem Solving and Generative Verification for LLM Reasoning
von: Singhi, Nishad, et al.
Veröffentlicht: (2025)
von: Singhi, Nishad, et al.
Veröffentlicht: (2025)
Thinking Wrong in Silence: Backdoor Attacks on Continuous Latent Reasoning
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
Know When You're Wrong: Aligning Confidence with Correctness for LLM Error Detection
von: Xiaohu, Xie, et al.
Veröffentlicht: (2026)
von: Xiaohu, Xie, et al.
Veröffentlicht: (2026)
When Learning Rates Go Wrong: Early Structural Signals in PPO Actor-Critic
von: Fernández-Hernández, Alberto, et al.
Veröffentlicht: (2026)
von: Fernández-Hernández, Alberto, et al.
Veröffentlicht: (2026)
Acting for the Right Reasons: Creating Reason-Sensitive Artificial Moral Agents
von: Baum, Kevin, et al.
Veröffentlicht: (2024)
von: Baum, Kevin, et al.
Veröffentlicht: (2024)
Right for the Right Reasons: Avoiding Reasoning Shortcuts via Prototypical Neurosymbolic AI
von: Andolfi, Luca, et al.
Veröffentlicht: (2025)
von: Andolfi, Luca, et al.
Veröffentlicht: (2025)
Towards Trustworthy GUI Agents: A Survey
von: Shi, Yucheng, et al.
Veröffentlicht: (2025)
von: Shi, Yucheng, et al.
Veröffentlicht: (2025)
ThermEval: A Structured Benchmark for Evaluation of Vision-Language Models on Thermal Imagery
von: Shrivastava, Ayush, et al.
Veröffentlicht: (2026)
von: Shrivastava, Ayush, et al.
Veröffentlicht: (2026)
Wrong Code, Right Structure: Learning Netlist Representations from Imperfect LLM-Generated RTL
von: Cai, Siyang, et al.
Veröffentlicht: (2026)
von: Cai, Siyang, et al.
Veröffentlicht: (2026)
Decidable By Construction: Design-Time Verification for Trustworthy AI
von: Haynes, Houston
Veröffentlicht: (2026)
von: Haynes, Houston
Veröffentlicht: (2026)
The Right Answer, the Wrong Direction: Why Transformers Fail at Counting and How to Fix It
von: Garcia, Gabriel
Veröffentlicht: (2026)
von: Garcia, Gabriel
Veröffentlicht: (2026)
All AI Models are Wrong, but Some are Optimal
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
LLMs as Assessors: Right for the Right Reason?
von: Saha, Sourav, et al.
Veröffentlicht: (2026)
von: Saha, Sourav, et al.
Veröffentlicht: (2026)
VerificAgent: Domain-Specific Memory Verification for Scalable Oversight of Aligned Computer-Use Agents
von: Nguyen, Thong Q., et al.
Veröffentlicht: (2025)
von: Nguyen, Thong Q., et al.
Veröffentlicht: (2025)
What is Wrong with Perplexity for Long-context Language Modeling?
von: Fang, Lizhe, et al.
Veröffentlicht: (2024)
von: Fang, Lizhe, et al.
Veröffentlicht: (2024)
Beyond Benchmarks: Dynamic, Automatic And Systematic Red-Teaming Agents For Trustworthy Medical Language Models
von: Pan, Jiazhen, et al.
Veröffentlicht: (2025)
von: Pan, Jiazhen, et al.
Veröffentlicht: (2025)
Trustworthy Prediction with Gaussian Process Knowledge Scores
von: Butler, Kurt, et al.
Veröffentlicht: (2025)
von: Butler, Kurt, et al.
Veröffentlicht: (2025)
Verification and Validation for Trustworthy Scientific Machine Learning
von: Jakeman, John D., et al.
Veröffentlicht: (2025)
von: Jakeman, John D., et al.
Veröffentlicht: (2025)
Towards Interpretable and Trustworthy Time Series Reasoning: A BlueSky Vision
von: Ning, Kanghui, et al.
Veröffentlicht: (2025)
von: Ning, Kanghui, et al.
Veröffentlicht: (2025)
Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA
von: Martinez, John Ray B.
Veröffentlicht: (2026)
von: Martinez, John Ray B.
Veröffentlicht: (2026)
An Accurate and Interpretable Framework for Trustworthy Process Monitoring
von: Wang, Hao, et al.
Veröffentlicht: (2023)
von: Wang, Hao, et al.
Veröffentlicht: (2023)
ManifoldMind: Dynamic Hyperbolic Reasoning for Trustworthy Recommendations
von: Harit, Anoushka, et al.
Veröffentlicht: (2025)
von: Harit, Anoushka, et al.
Veröffentlicht: (2025)
The Geometry of Self-Verification in a Task-Specific Reasoning Model
von: Lee, Andrew, et al.
Veröffentlicht: (2025)
von: Lee, Andrew, et al.
Veröffentlicht: (2025)
To Backtrack or Not to Backtrack: When Sequential Search Limits Model Reasoning
von: Qin, Tian, et al.
Veröffentlicht: (2025)
von: Qin, Tian, et al.
Veröffentlicht: (2025)
Step-by-Step Diffusion: An Elementary Tutorial
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2024)
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2024)
Out-of-Distribution Detection Methods Answer the Wrong Questions
von: Li, Yucen Lily, et al.
Veröffentlicht: (2025)
von: Li, Yucen Lily, et al.
Veröffentlicht: (2025)
Your Assumed DAG is Wrong and Here's How To Deal With It
von: Padh, Kirtan, et al.
Veröffentlicht: (2025)
von: Padh, Kirtan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Trajectory Guard -- A Lightweight, Sequence-Aware Model for Real-Time Anomaly Detection in Agentic AI
von: Advani, Laksh
Veröffentlicht: (2026) -
Clever Materials: When Models Identify Good Materials for the Wrong Reasons
von: Jablonka, Kevin Maik
Veröffentlicht: (2026) -
Data Cartography for Detecting Memorization Hotspots and Guiding Data Interventions in Generative Models
von: Patel, Laksh, et al.
Veröffentlicht: (2025) -
Wrong Model, Right Uncertainty: Spatial Associations for Discrete Data with Misspecification
von: Burt, David R., et al.
Veröffentlicht: (2025) -
Perplexity Cannot Always Tell Right from Wrong
von: Veličković, Petar, et al.
Veröffentlicht: (2026)