Guideline-Grounded Evidence Accumulation for High-Stakes Agent Verification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yichi, Seedat, Nabeel, Dong, Yinpeng, Cui, Peng, Zhu, Jun, van de Schaar, Mihaela |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DC-Check: A Data-Centric AI checklist to guide the development of reliable machine learning systems
von: Seedat, Nabeel, et al.
Veröffentlicht: (2022)
von: Seedat, Nabeel, et al.
Veröffentlicht: (2022)
You can't handle the (dirty) truth: Data-centric insights improve pseudo-labeling
von: Seedat, Nabeel, et al.
Veröffentlicht: (2024)
von: Seedat, Nabeel, et al.
Veröffentlicht: (2024)
Large Language Models to Enhance Bayesian Optimization
von: Liu, Tennison, et al.
Veröffentlicht: (2024)
von: Liu, Tennison, et al.
Veröffentlicht: (2024)
Self-Healing Machine Learning: A Framework for Autonomous Adaptation in Real-World Environments
von: Rauba, Paulius, et al.
Veröffentlicht: (2024)
von: Rauba, Paulius, et al.
Veröffentlicht: (2024)
Curated LLM: Synergy of LLMs and Data Curation for tabular augmentation in low-data regimes
von: Seedat, Nabeel, et al.
Veröffentlicht: (2023)
von: Seedat, Nabeel, et al.
Veröffentlicht: (2023)
Cascaded Language Models for Cost-effective Human-AI Decision-Making
von: Fanconi, Claudio, et al.
Veröffentlicht: (2025)
von: Fanconi, Claudio, et al.
Veröffentlicht: (2025)
When is Off-Policy Evaluation (Reward Modeling) Useful in Contextual Bandits? A Data-Centric Perspective
von: Sun, Hao, et al.
Veröffentlicht: (2023)
von: Sun, Hao, et al.
Veröffentlicht: (2023)
DAGnosis: Localized Identification of Data Inconsistencies using Structures
von: Huynh, Nicolas, et al.
Veröffentlicht: (2024)
von: Huynh, Nicolas, et al.
Veröffentlicht: (2024)
Language Bottleneck Models for Qualitative Knowledge State Modeling
von: Berthon, Antonin, et al.
Veröffentlicht: (2025)
von: Berthon, Antonin, et al.
Veröffentlicht: (2025)
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities
von: Sun, Hao, et al.
Veröffentlicht: (2025)
von: Sun, Hao, et al.
Veröffentlicht: (2025)
To Whom Do Language Models Align? Measuring Principal Hierarchies Under High-Stakes Competing Demands
von: Yu, Fangyi, et al.
Veröffentlicht: (2026)
von: Yu, Fangyi, et al.
Veröffentlicht: (2026)
Matchmaker: Self-Improving Large Language Model Programs for Schema Matching
von: Seedat, Nabeel, et al.
Veröffentlicht: (2024)
von: Seedat, Nabeel, et al.
Veröffentlicht: (2024)
Influence-Guided Symbolic Regression: Scientific Discovery via LLM-Driven Equation Search with Granular Feedback
von: Saveliev, Evgeny S., et al.
Veröffentlicht: (2026)
von: Saveliev, Evgeny S., et al.
Veröffentlicht: (2026)
What's the next frontier for Data-centric AI? Data Savvy Agents
von: Seedat, Nabeel, et al.
Veröffentlicht: (2025)
von: Seedat, Nabeel, et al.
Veröffentlicht: (2025)
Query-Dependent Prompt Evaluation and Optimization with Offline Inverse RL
von: Sun, Hao, et al.
Veröffentlicht: (2023)
von: Sun, Hao, et al.
Veröffentlicht: (2023)
RealSafe-R1: Safety-Aligned DeepSeek-R1 without Compromising Reasoning Capability
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
The Synergy of LLMs & RL Unlocks Offline Learning of Generalizable Language-Conditioned Policies with Low-fidelity Data
von: Pouplin, Thomas, et al.
Veröffentlicht: (2024)
von: Pouplin, Thomas, et al.
Veröffentlicht: (2024)
Beyond Pointwise Scores: Decomposed Criteria-Based Evaluation of LLM Responses
von: Yu, Fangyi, et al.
Veröffentlicht: (2025)
von: Yu, Fangyi, et al.
Veröffentlicht: (2025)
Truly Self-Improving Agents Require Intrinsic Metacognitive Learning
von: Liu, Tennison, et al.
Veröffentlicht: (2025)
von: Liu, Tennison, et al.
Veröffentlicht: (2025)
TTPrint: Evidence-Grounded TTP Extraction via Diverge-then-Converge Verification
von: Cheng, Yutong, et al.
Veröffentlicht: (2026)
von: Cheng, Yutong, et al.
Veröffentlicht: (2026)
DeceptionBench: A Comprehensive Benchmark for AI Deception Behaviors in Real-world Scenarios
von: Huang, Yao, et al.
Veröffentlicht: (2025)
von: Huang, Yao, et al.
Veröffentlicht: (2025)
RepoMirage: Probing Repository Context Reasoning in Code Agents with Perturbations
von: Li, Hanyu, et al.
Veröffentlicht: (2026)
von: Li, Hanyu, et al.
Veröffentlicht: (2026)
Breaking the Ceiling: Exploring the Potential of Jailbreak Attacks through Expanding Strategy Space
von: Huang, Yao, et al.
Veröffentlicht: (2025)
von: Huang, Yao, et al.
Veröffentlicht: (2025)
Mitigating Overthinking in Large Reasoning Models via Manifold Steering
von: Huang, Yao, et al.
Veröffentlicht: (2025)
von: Huang, Yao, et al.
Veröffentlicht: (2025)
Towards Safe Reasoning in Large Reasoning Models via Corrective Intervention
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
Active Task Disambiguation with LLMs
von: Kobalczyk, Katarzyna, et al.
Veröffentlicht: (2025)
von: Kobalczyk, Katarzyna, et al.
Veröffentlicht: (2025)
A Survey on Autonomy-Induced Security Risks in Large Model-Based Agents
von: Su, Hang, et al.
Veröffentlicht: (2025)
von: Su, Hang, et al.
Veröffentlicht: (2025)
Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets
von: Zhang, Simpson, et al.
Veröffentlicht: (2025)
von: Zhang, Simpson, et al.
Veröffentlicht: (2025)
Reusing Embeddings: Reproducible Reward Model Research in Large Language Model Alignment without GPUs
von: Sun, Hao, et al.
Veröffentlicht: (2025)
von: Sun, Hao, et al.
Veröffentlicht: (2025)
Dissecting Sample Hardness: A Fine-Grained Analysis of Hardness Characterization Methods for Data-Centric AI
von: Seedat, Nabeel, et al.
Veröffentlicht: (2024)
von: Seedat, Nabeel, et al.
Veröffentlicht: (2024)
Strategic Self-Improvement for Competitive Agents in AI Labour Markets
von: Chiu, Christopher, et al.
Veröffentlicht: (2025)
von: Chiu, Christopher, et al.
Veröffentlicht: (2025)
Implicit Behavioral Alignment of Language Agents in High-Stakes Crowd Simulations
von: Wang, Yunzhe, et al.
Veröffentlicht: (2025)
von: Wang, Yunzhe, et al.
Veröffentlicht: (2025)
Unveiling Trust in Multimodal Large Language Models: Evaluation, Analysis, and Mitigation
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
von: Zhang, Yichi, et al.
Veröffentlicht: (2025)
StakeBench: Evaluating Language Understanding Grounded in Market Commitment
von: Pei, Yunhua, et al.
Veröffentlicht: (2026)
von: Pei, Yunhua, et al.
Veröffentlicht: (2026)
STAR-PólyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision
von: Wu, Jiaao, et al.
Veröffentlicht: (2026)
von: Wu, Jiaao, et al.
Veröffentlicht: (2026)
Not All Explanations for Deep Learning Phenomena Are Equally Valuable
von: Jeffares, Alan, et al.
Veröffentlicht: (2025)
von: Jeffares, Alan, et al.
Veröffentlicht: (2025)
Preference Learning for AI Alignment: a Causal Perspective
von: Kobalczyk, Katarzyna, et al.
Veröffentlicht: (2025)
von: Kobalczyk, Katarzyna, et al.
Veröffentlicht: (2025)
Active Timepoint Selection for Learning Measure-Valued Trajectories
von: Huynh, Nicolas, et al.
Veröffentlicht: (2026)
von: Huynh, Nicolas, et al.
Veröffentlicht: (2026)
Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning
von: Sun, Hao, et al.
Veröffentlicht: (2024)
von: Sun, Hao, et al.
Veröffentlicht: (2024)
Hyperparameter Trajectory Inference with Conditional Lagrangian Optimal Transport
von: Amad, Harry, et al.
Veröffentlicht: (2026)
von: Amad, Harry, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DC-Check: A Data-Centric AI checklist to guide the development of reliable machine learning systems
von: Seedat, Nabeel, et al.
Veröffentlicht: (2022) -
You can't handle the (dirty) truth: Data-centric insights improve pseudo-labeling
von: Seedat, Nabeel, et al.
Veröffentlicht: (2024) -
Large Language Models to Enhance Bayesian Optimization
von: Liu, Tennison, et al.
Veröffentlicht: (2024) -
Self-Healing Machine Learning: A Framework for Autonomous Adaptation in Real-World Environments
von: Rauba, Paulius, et al.
Veröffentlicht: (2024) -
Curated LLM: Synergy of LLMs and Data Curation for tabular augmentation in low-data regimes
von: Seedat, Nabeel, et al.
Veröffentlicht: (2023)