Trust at Your Own Peril: A Mixed Methods Exploration of the Ability of Large Language Models to Generate Expert-Like Systems Engineering Artifacts and a Characterization of Failure Modes
Fuente:
arXiv
Guardado en:
| Autores principales: | Topcu, Taylan G., Husain, Mohammed, Ofsa, Max, Wach, Paul |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
An Empirical Exploration of ChatGPT's Ability to Support Problem Formulation Tasks for Mission Engineering and a Documentation of its Performance Variability
por: Ofsa, Max, et al.
Publicado: (2025)
por: Ofsa, Max, et al.
Publicado: (2025)
What is the Return on Investment of Digital Engineering for Complex Systems Development? Findings from a Mixed-Methods Study on the Post-production Design Change Process of Navy Assets
por: Shefa, Jannatul, et al.
Publicado: (2025)
por: Shefa, Jannatul, et al.
Publicado: (2025)
Digital Engineering Testbed for Test and Evaluation: Operation Safe Passage Status and Lessons Learned
por: Brandt Sandman, et al.
Publicado: (2025)
por: Brandt Sandman, et al.
Publicado: (2025)
An Analysis of Early-Stage Functional Safety Analysis Methods and Their Integration into Model-Based Systems Engineering
por: Shefa, Jannatul, et al.
Publicado: (2025)
por: Shefa, Jannatul, et al.
Publicado: (2025)
A Rapid Review of How Model‐based Systems Engineering is Used in Healthcare Systems
por: Md Doulotuzzaman Xames, et al.
Publicado: (2024)
por: Md Doulotuzzaman Xames, et al.
Publicado: (2024)
A systematic literature review on the mathematical underpinning of model‐based systems engineering
por: Paul Wach, et al.
Publicado: (2024)
por: Paul Wach, et al.
Publicado: (2024)
Digital Engineering Transformation as a Sociotechnical Challenge: Categorization of Barriers and Their Mapping to DoD's Policy Goals
por: Xames, Md Doulotuzzaman, et al.
Publicado: (2025)
por: Xames, Md Doulotuzzaman, et al.
Publicado: (2025)
Digital Engineering Transformation as a Sociotechnical Challenge: Categorization of Barriers and Their Mapping to DoD's Policy Goals
por: Md Doulotuzzaman Xames, et al.
Publicado: (2026)
por: Md Doulotuzzaman Xames, et al.
Publicado: (2026)
Building Your Own Trusted Execution Environments Using FPGA
por: Armanuzzaman, Md, et al.
Publicado: (2022)
por: Armanuzzaman, Md, et al.
Publicado: (2022)
Leveraging High-Fidelity Digital Models and Reinforcement Learning for Mission Engineering: A Case Study of Aerial Firefighting Under Perfect Information
por: Çetinkaya, İbrahim Oğuz, et al.
Publicado: (2025)
por: Çetinkaya, İbrahim Oğuz, et al.
Publicado: (2025)
Stakeholder Perspectives on Digital Twin Implementation Challenges in Healthcare: Insights from a Provider Digital Twin Case Study
por: Xames, Md Doulotuzzaman, et al.
Publicado: (2025)
por: Xames, Md Doulotuzzaman, et al.
Publicado: (2025)
Predictive Failure Mode Analysis: The Potential Role of Artificial Intelligence in Predicting Failure Modes During Root Canal Retreatment
por: Mohammed Turky, et al.
Publicado: (2025)
por: Mohammed Turky, et al.
Publicado: (2025)
Diagnosing FP4 inference: a layer-wise and block-wise sensitivity analysis of NVFP4 and MXFP4
por: Cim, Musa, et al.
Publicado: (2026)
por: Cim, Musa, et al.
Publicado: (2026)
BYOL: Bring Your Own Language Into LLMs
por: Zamir, Syed Waqas, et al.
Publicado: (2026)
por: Zamir, Syed Waqas, et al.
Publicado: (2026)
YOYO (You're On Your Own!).
por: Young, Laura
Publicado: (1998)
por: Young, Laura
Publicado: (1998)
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models
por: Yu, Longhui, et al.
Publicado: (2023)
por: Yu, Longhui, et al.
Publicado: (2023)
Competing Bandits: The Perils of Exploration Under Competition
por: Aridor, Guy, et al.
Publicado: (2020)
por: Aridor, Guy, et al.
Publicado: (2020)
Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition
por: Goldstein, Ariel, et al.
Publicado: (2024)
por: Goldstein, Ariel, et al.
Publicado: (2024)
Graph Your Own Prompt
por: Ding, Xi, et al.
Publicado: (2025)
por: Ding, Xi, et al.
Publicado: (2025)
Create Your Own Holiday
Publicado: (2024)
Publicado: (2024)
Burning Your Own CDs.
por: Ekhaml, Leticia
Publicado: (2001)
por: Ekhaml, Leticia
Publicado: (2001)
Mind Your Own Business
por: Nixon, Judith M., et al.
Publicado: (2004)
por: Nixon, Judith M., et al.
Publicado: (2004)
BYON: Bring Your Own Networks for Digital Agriculture Applications
por: Sie, Emerson, et al.
Publicado: (2025)
por: Sie, Emerson, et al.
Publicado: (2025)
AI Mental Models & Trust: The Promises and Perils of Interaction Design
por: SOOJIN JEONG, et al.
Publicado: (2024)
por: SOOJIN JEONG, et al.
Publicado: (2024)
Reconstructing Trust Embeddings from Siamese Trust Scores: A Direct-Sum Approach with Fixed-Point Semantics
por: Alpay, Faruk, et al.
Publicado: (2025)
por: Alpay, Faruk, et al.
Publicado: (2025)
The Perils & Promises of Fact-checking with Large Language Models
por: Quelle, Dorian, et al.
Publicado: (2023)
por: Quelle, Dorian, et al.
Publicado: (2023)
Following Experts at Work in Their Own Information Spaces: Using Observational Methods To Develop Tools for the Digital Library.
por: Gorman, Paul, et al.
Publicado: (2002)
por: Gorman, Paul, et al.
Publicado: (2002)
Parallel Context Compaction for Long-Horizon LLM Agent Serving
por: Cim, Musa, et al.
Publicado: (2026)
por: Cim, Musa, et al.
Publicado: (2026)
Bootstrap Your Own Context Length
por: Wang, Liang, et al.
Publicado: (2024)
por: Wang, Liang, et al.
Publicado: (2024)
Making Your Own Gene Library.
por: Perez-Ortin, Jose E., et al.
Publicado: (1997)
por: Perez-Ortin, Jose E., et al.
Publicado: (1997)
Taking the Law into Your Own Hands.
por: Pedzich, Joan
Publicado: (1993)
por: Pedzich, Joan
Publicado: (1993)
To Trust Or Not To Trust Your Vision-Language Model's Prediction
por: Dong, Hao, et al.
Publicado: (2025)
por: Dong, Hao, et al.
Publicado: (2025)
Bring Your Own Knowledge: A Survey of Methods for LLM Knowledge Expansion
por: Wang, Mingyang, et al.
Publicado: (2025)
por: Wang, Mingyang, et al.
Publicado: (2025)
Theoretical Underpinnings to Establish Fidelity Conditions for Defining Verification Models
por: Paul Wach, et al.
Publicado: (2024)
por: Paul Wach, et al.
Publicado: (2024)
BYOS: Knowledge-driven Large Language Models Bring Your Own Operating System More Excellent
por: Lin, Hongyu, et al.
Publicado: (2025)
por: Lin, Hongyu, et al.
Publicado: (2025)
Metamorphic Malware Evolution: The Potential and Peril of Large Language Models
por: Madani, Pooria
Publicado: (2024)
por: Madani, Pooria
Publicado: (2024)
In Transformer We Trust? A Perspective on Transformer Architecture Failure Modes
por: Mondal, Trishit, et al.
Publicado: (2026)
por: Mondal, Trishit, et al.
Publicado: (2026)
Theatrical Compliance: A Failure Mode in Large Language Models
por: Nowickij (Navitski), Kirill Vladimirovich
Publicado: (2026)
por: Nowickij (Navitski), Kirill Vladimirovich
Publicado: (2026)
Books in Peril: Cooperative Approaches to Conservation
por: Banks, Paul N.
Publicado: (1976)
por: Banks, Paul N.
Publicado: (1976)
If Left to Your Own Devices, Consider Colophony
por: E. Dimitra Bednar, et al.
Publicado: (2025)
por: E. Dimitra Bednar, et al.
Publicado: (2025)
Ejemplares similares
-
An Empirical Exploration of ChatGPT's Ability to Support Problem Formulation Tasks for Mission Engineering and a Documentation of its Performance Variability
por: Ofsa, Max, et al.
Publicado: (2025) -
What is the Return on Investment of Digital Engineering for Complex Systems Development? Findings from a Mixed-Methods Study on the Post-production Design Change Process of Navy Assets
por: Shefa, Jannatul, et al.
Publicado: (2025) -
Digital Engineering Testbed for Test and Evaluation: Operation Safe Passage Status and Lessons Learned
por: Brandt Sandman, et al.
Publicado: (2025) -
An Analysis of Early-Stage Functional Safety Analysis Methods and Their Integration into Model-Based Systems Engineering
por: Shefa, Jannatul, et al.
Publicado: (2025) -
A Rapid Review of How Model‐based Systems Engineering is Used in Healthcare Systems
por: Md Doulotuzzaman Xames, et al.
Publicado: (2024)