Automating the Enterprise with Foundation Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Wornow, Michael, Narayan, Avanika, Opsahl-Ong, Krista, McIntyre, Quinn, Shah, Nigam H., Re, Christopher |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
WONDERBREAD: A Benchmark for Evaluating Multimodal Foundation Models on Business Process Management Tasks
por: Wornow, Michael, et al.
Publicado: (2024)
por: Wornow, Michael, et al.
Publicado: (2024)
Verifying LLM-Generated Code in the Context of Software Verification with Ada/SPARK
por: Cramer, Marcos, et al.
Publicado: (2025)
por: Cramer, Marcos, et al.
Publicado: (2025)
Redundancy and Concept Analysis for Code-trained Language Models
por: Sharma, Arushi, et al.
Publicado: (2023)
por: Sharma, Arushi, et al.
Publicado: (2023)
Generating Minimalist Adversarial Perturbations to Test Object-Detection Models: An Adaptive Multi-Metric Evolutionary Search Approach
por: McIntyre-Garcia, Cristopher, et al.
Publicado: (2024)
por: McIntyre-Garcia, Cristopher, et al.
Publicado: (2024)
Analyzing Latent Concepts in Code Language Models
por: Sharma, Arushi, et al.
Publicado: (2025)
por: Sharma, Arushi, et al.
Publicado: (2025)
Context Clues: Evaluating Long Context Models for Clinical Prediction Tasks on EHRs
por: Wornow, Michael, et al.
Publicado: (2024)
por: Wornow, Michael, et al.
Publicado: (2024)
Foundation Model Engineering: Engineering Foundation Models Just as Engineering Software
por: Ran, Dezhi, et al.
Publicado: (2024)
por: Ran, Dezhi, et al.
Publicado: (2024)
Automated Customization of LLMs for Enterprise Code Repositories Using Semantic Scopes
por: Finkler, Ulrich, et al.
Publicado: (2026)
por: Finkler, Ulrich, et al.
Publicado: (2026)
Automated Creation and Enrichment Framework for Improved Invocation of Enterprise APIs as Tools
por: Agarwal, Prerna, et al.
Publicado: (2025)
por: Agarwal, Prerna, et al.
Publicado: (2025)
Evaluating Robustness of Large Language Models in Enterprise Applications: Benchmarks for Perturbation Consistency Across Formats and Languages
por: Bogavelli, Tara, et al.
Publicado: (2026)
por: Bogavelli, Tara, et al.
Publicado: (2026)
Terminal Agents Suffice for Enterprise Automation
por: Bechard, Patrice, et al.
Publicado: (2026)
por: Bechard, Patrice, et al.
Publicado: (2026)
Z-Space: A Multi-Agent Tool Orchestration Framework for Enterprise-Grade LLM Automation
por: He, Qingsong, et al.
Publicado: (2025)
por: He, Qingsong, et al.
Publicado: (2025)
Strategic Decision Framework for Enterprise LLM Adoption
por: Trusov, Michael, et al.
Publicado: (2025)
por: Trusov, Michael, et al.
Publicado: (2025)
KernelBench: Can LLMs Write Efficient GPU Kernels?
por: Ouyang, Anne, et al.
Publicado: (2025)
por: Ouyang, Anne, et al.
Publicado: (2025)
Deploying Geospatial Foundation Models in the Real World: Lessons from WorldCereal
por: Butsko, Christina, et al.
Publicado: (2025)
por: Butsko, Christina, et al.
Publicado: (2025)
The Hitchhikers Guide to Production-ready Trustworthy Foundation Model powered Software (FMware)
por: Vasilevski, Kirill, et al.
Publicado: (2025)
por: Vasilevski, Kirill, et al.
Publicado: (2025)
R-LAM: Reproducibility-Constrained Large Action Models for Scientific Workflow Automation
por: Sureshkumar, Suriya
Publicado: (2026)
por: Sureshkumar, Suriya
Publicado: (2026)
CONSTRUCTA: Automating Commercial Construction Schedules in Fabrication Facilities with Large Language Models
por: Zhang, Yifan, et al.
Publicado: (2025)
por: Zhang, Yifan, et al.
Publicado: (2025)
AIPC: Agent-Based Automation for AI Model Deployment with Qualcomm AI Runtime
por: Su, Jianhao, et al.
Publicado: (2026)
por: Su, Jianhao, et al.
Publicado: (2026)
Assessing Large Language Models for Automated Feedback Generation in Learning Programming Problem Solving
por: Silva, Priscylla, et al.
Publicado: (2025)
por: Silva, Priscylla, et al.
Publicado: (2025)
Planning-Driven Programming: A Large Language Model Programming Workflow
por: Lei, Chao, et al.
Publicado: (2024)
por: Lei, Chao, et al.
Publicado: (2024)
RAG Does Not Work for Enterprises
por: Bruckhaus, Tilmann
Publicado: (2024)
por: Bruckhaus, Tilmann
Publicado: (2024)
World of Workflows: A Benchmark for Bringing World Models to Enterprise Systems
por: Gupta, Lakshya, et al.
Publicado: (2026)
por: Gupta, Lakshya, et al.
Publicado: (2026)
Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation
por: Jacopin, Éric
Publicado: (2026)
por: Jacopin, Éric
Publicado: (2026)
Automated Cloud Infrastructure-as-Code Reconciliation with AI Agents
por: Yang, Zhenning, et al.
Publicado: (2025)
por: Yang, Zhenning, et al.
Publicado: (2025)
Monitizer: Automating Design and Evaluation of Neural Network Monitors
por: Azeem, Muqsit, et al.
Publicado: (2024)
por: Azeem, Muqsit, et al.
Publicado: (2024)
Automating Code Adaptation for MLOps -- A Benchmarking Study on LLMs
por: Patel, Harsh, et al.
Publicado: (2024)
por: Patel, Harsh, et al.
Publicado: (2024)
On the Impact of Code Comments for Automated Bug-Fixing: An Empirical Study
por: Vitale, Antonio, et al.
Publicado: (2026)
por: Vitale, Antonio, et al.
Publicado: (2026)
Imitation Game: Reproducing Deep Learning Bugs Leveraging an Intelligent Agent
por: Shah, Mehil B, et al.
Publicado: (2025)
por: Shah, Mehil B, et al.
Publicado: (2025)
Data Wrangling Task Automation Using Code-Generating Language Models
por: Akella, Ashlesha, et al.
Publicado: (2025)
por: Akella, Ashlesha, et al.
Publicado: (2025)
FuzzTheREST: An Intelligent Automated Black-box RESTful API Fuzzer
por: Dias, Tiago, et al.
Publicado: (2024)
por: Dias, Tiago, et al.
Publicado: (2024)
PostTrainBench: Can LLM Agents Automate LLM Post-Training?
por: Rank, Ben, et al.
Publicado: (2026)
por: Rank, Ben, et al.
Publicado: (2026)
ENCO: Life-Cycle Management of Enterprise-Grade Copilots
por: Zhu, Yiwen, et al.
Publicado: (2024)
por: Zhu, Yiwen, et al.
Publicado: (2024)
FLOW-BENCH: Towards Conversational Generation of Enterprise Workflows
por: Duesterwald, Evelyn, et al.
Publicado: (2025)
por: Duesterwald, Evelyn, et al.
Publicado: (2025)
The Foundations of Computational Management: A Systematic Approach to Task Automation for the Integration of Artificial Intelligence into Existing Workflows
por: Jadad-Garcia, Tamen, et al.
Publicado: (2024)
por: Jadad-Garcia, Tamen, et al.
Publicado: (2024)
BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice
por: Li, Yuanpeng, et al.
Publicado: (2025)
por: Li, Yuanpeng, et al.
Publicado: (2025)
Automated Machine Learning: A Case Study on Non-Intrusive Appliance Load Monitoring
por: Moin, Armin, et al.
Publicado: (2022)
por: Moin, Armin, et al.
Publicado: (2022)
Software Engineering and Foundation Models: Insights from Industry Blogs Using a Jury of Foundation Models
por: Li, Hao, et al.
Publicado: (2024)
por: Li, Hao, et al.
Publicado: (2024)
High-Dimensional Fault Tolerance Testing of Highly Automated Vehicles Based on Low-Rank Models
por: Mei, Yuewen, et al.
Publicado: (2024)
por: Mei, Yuewen, et al.
Publicado: (2024)
Proving the Coding Interview: A Benchmark for Formally Verified Code Generation
por: Dougherty, Quinn, et al.
Publicado: (2025)
por: Dougherty, Quinn, et al.
Publicado: (2025)
Ejemplares similares
-
WONDERBREAD: A Benchmark for Evaluating Multimodal Foundation Models on Business Process Management Tasks
por: Wornow, Michael, et al.
Publicado: (2024) -
Verifying LLM-Generated Code in the Context of Software Verification with Ada/SPARK
por: Cramer, Marcos, et al.
Publicado: (2025) -
Redundancy and Concept Analysis for Code-trained Language Models
por: Sharma, Arushi, et al.
Publicado: (2023) -
Generating Minimalist Adversarial Perturbations to Test Object-Detection Models: An Adaptive Multi-Metric Evolutionary Search Approach
por: McIntyre-Garcia, Cristopher, et al.
Publicado: (2024) -
Analyzing Latent Concepts in Code Language Models
por: Sharma, Arushi, et al.
Publicado: (2025)