Verifying Memoryless Sequential Decision-making of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gross, Dennis, Spieker, Helge, Gotlieb, Arnaud |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bounded PCTL Model Checking of Large Language Model Outputs
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
Enhancing Manufacturing Quality Prediction Models through the Integration of Explainability Methods
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
Co-Activation Graph Analysis of Safety-Verified and Explainable Deep Reinforcement Learning Policies
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
Translating the Rashomon Effect to Sequential Decision-Making Tasks
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
Probabilistic Model Checking of Stochastic Reinforcement Learning Policies
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
Towards Trustworthy Automated Driving through Qualitative Scene Understanding and Explanations
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2024)
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2024)
Semi-supervised CAPP Transformer Learning via Pseudo-labeling
von: Gross, Dennis, et al.
Veröffentlicht: (2026)
von: Gross, Dennis, et al.
Veröffentlicht: (2026)
Evaluating Human Trajectory Prediction with Metamorphic Testing
von: Spieker, Helge, et al.
Veröffentlicht: (2024)
von: Spieker, Helge, et al.
Veröffentlicht: (2024)
Trustworthy Automated Driving through Qualitative Scene Understanding and Explanations
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2024)
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2024)
Explainable Scene Understanding with Qualitative Representations and Graph Neural Networks
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2025)
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2025)
Rashomon in the Streets: Explanation Ambiguity in Scene Understanding
von: Spieker, Helge, et al.
Veröffentlicht: (2025)
von: Spieker, Helge, et al.
Veröffentlicht: (2025)
Efficient Milling Quality Prediction with Explainable Machine Learning
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
Prompting for Performance: Exploring LLMs for Configuring Software
von: Spieker, Helge, et al.
Veröffentlicht: (2025)
von: Spieker, Helge, et al.
Veröffentlicht: (2025)
Constraint-Guided Test Execution Scheduling: An Experience Report at ABB Robotics
von: Gotlieb, Arnaud, et al.
Veröffentlicht: (2023)
von: Gotlieb, Arnaud, et al.
Veröffentlicht: (2023)
Formally Verifying and Explaining Sepsis Treatment Policies with COOL-MC
von: Gross, Dennis
Veröffentlicht: (2026)
von: Gross, Dennis
Veröffentlicht: (2026)
COOL-MC: Verifying and Explaining RL Policies for Platelet Inventory Management
von: Gross, Dennis
Veröffentlicht: (2026)
von: Gross, Dennis
Veröffentlicht: (2026)
COOL-MC: Verifying and Explaining RL Policies for Multi-bridge Network Maintenance
von: Gross, Dennis
Veröffentlicht: (2026)
von: Gross, Dennis
Veröffentlicht: (2026)
On the Modeling Capabilities of Large Language Models for Sequential Decision Making
von: Klissarov, Martin, et al.
Veröffentlicht: (2024)
von: Klissarov, Martin, et al.
Veröffentlicht: (2024)
Policy Testing with MDPFuzz (Replicability Study)
von: Mazouni, Quentin, et al.
Veröffentlicht: (2025)
von: Mazouni, Quentin, et al.
Veröffentlicht: (2025)
Efficiently Ranking Software Variants with Minimal Benchmarks
von: Matricon, Théo, et al.
Veröffentlicht: (2025)
von: Matricon, Théo, et al.
Veröffentlicht: (2025)
Metamorphic Testing of Multimodal Human Trajectory Prediction
von: Spieker, Helge, et al.
Veröffentlicht: (2025)
von: Spieker, Helge, et al.
Veröffentlicht: (2025)
Testing for Fault Diversity in Reinforcement Learning
von: Mazouni, Quentin, et al.
Veröffentlicht: (2024)
von: Mazouni, Quentin, et al.
Veröffentlicht: (2024)
Enhancing RL Safety with Counterfactual LLM Reasoning
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
Safety-Oriented Pruning and Interpretation of Reinforcement Learning Policies
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and Perception
von: Yang, Yunhao, et al.
Veröffentlicht: (2023)
von: Yang, Yunhao, et al.
Veröffentlicht: (2023)
Optimizing Ethical Risk Reduction for Medical Intelligent Systems with Constraint Programming
von: Brayé, Clotilde, et al.
Veröffentlicht: (2025)
von: Brayé, Clotilde, et al.
Veröffentlicht: (2025)
Position Paper: Rethinking Privacy in RL for Sequential Decision-making in the Age of LLMs
von: Fan, Flint Xiaofeng, et al.
Veröffentlicht: (2025)
von: Fan, Flint Xiaofeng, et al.
Veröffentlicht: (2025)
A Unified Assessment of the Poverty of the Stimulus Argument for Neural Language Models
von: Yang, Xiulin, et al.
Veröffentlicht: (2026)
von: Yang, Xiulin, et al.
Veröffentlicht: (2026)
MazeEval: A Benchmark for Testing Sequential Decision-Making in Language Models
von: Einarsson, Hafsteinn
Veröffentlicht: (2025)
von: Einarsson, Hafsteinn
Veröffentlicht: (2025)
Function Words as Statistical Cues for Language Learning
von: Yang, Xiulin, et al.
Veröffentlicht: (2026)
von: Yang, Xiulin, et al.
Veröffentlicht: (2026)
DLM: Unified Decision Language Models for Offline Multi-Agent Sequential Decision Making
von: Zhang, Zhuohui, et al.
Veröffentlicht: (2026)
von: Zhang, Zhuohui, et al.
Veröffentlicht: (2026)
Mutation‐Guided Metamorphic Testing of Optimality in AI Planning
von: Quentin Mazouni, et al.
Veröffentlicht: (2024)
von: Quentin Mazouni, et al.
Veröffentlicht: (2024)
Large Language Models for Sequential Decision-Making: Improving In-Context Learning via Supervised Fine-Tuning
von: Zhang, Minmin, et al.
Veröffentlicht: (2026)
von: Zhang, Minmin, et al.
Veröffentlicht: (2026)
Sequential Enumeration in Large Language Models
von: Hou, Kuinan, et al.
Veröffentlicht: (2025)
von: Hou, Kuinan, et al.
Veröffentlicht: (2025)
Cognitive LLMs: Towards Integrating Cognitive Architectures and Large Language Models for Manufacturing Decision-making
von: Wu, Siyu, et al.
Veröffentlicht: (2024)
von: Wu, Siyu, et al.
Veröffentlicht: (2024)
Turn-based Multi-Agent Reinforcement Learning Model Checking
von: Gross, Dennis
Veröffentlicht: (2025)
von: Gross, Dennis
Veröffentlicht: (2025)
O3D: Offline Data-driven Discovery and Distillation for Sequential Decision-Making with Large Language Models
von: Xiao, Yuchen, et al.
Veröffentlicht: (2023)
von: Xiao, Yuchen, et al.
Veröffentlicht: (2023)
A Sequential Decision-Making Model for Perimeter Identification
von: Taitler, Ayal
Veröffentlicht: (2024)
von: Taitler, Ayal
Veröffentlicht: (2024)
DecisionLLM: Large Language Models for Long Sequence Decision Exploration
von: Lv, Xiaowei, et al.
Veröffentlicht: (2026)
von: Lv, Xiaowei, et al.
Veröffentlicht: (2026)
Large Language Models for Supply Chain Decisions
von: Simchi-Levi, David, et al.
Veröffentlicht: (2025)
von: Simchi-Levi, David, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Bounded PCTL Model Checking of Large Language Model Outputs
von: Gross, Dennis, et al.
Veröffentlicht: (2025) -
Enhancing Manufacturing Quality Prediction Models through the Integration of Explainability Methods
von: Gross, Dennis, et al.
Veröffentlicht: (2024) -
Co-Activation Graph Analysis of Safety-Verified and Explainable Deep Reinforcement Learning Policies
von: Gross, Dennis, et al.
Veröffentlicht: (2025) -
Translating the Rashomon Effect to Sequential Decision-Making Tasks
von: Gross, Dennis, et al.
Veröffentlicht: (2025) -
Probabilistic Model Checking of Stochastic Reinforcement Learning Policies
von: Gross, Dennis, et al.
Veröffentlicht: (2024)