Computer Use at the Edge of the Statistical Precipice
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | D'Oro, Pierluca, Silwal, Sneha, Wong, William, Sun, Yuxuan, Xiao, Fanyi, Wang, Manchen, Gan, Eric, Bolourchi, Allen, Tighe, Joseph |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OpenComputer: Verifiable Software Worlds for Computer-Use Agents
von: Wei, Jinbiao, et al.
Veröffentlicht: (2026)
von: Wei, Jinbiao, et al.
Veröffentlicht: (2026)
ADEPTS: A Capability Framework for Human-Centered Agent Design
von: D'Oro, Pierluca, et al.
Veröffentlicht: (2025)
von: D'Oro, Pierluca, et al.
Veröffentlicht: (2025)
DigiData: Training and Evaluating General-Purpose Mobile Control Agents
von: Sun, Yuxuan, et al.
Veröffentlicht: (2025)
von: Sun, Yuxuan, et al.
Veröffentlicht: (2025)
UCRBench: Benchmarking LLMs on Use Case Recovery
von: Xiao, Shuyuan, et al.
Veröffentlicht: (2025)
von: Xiao, Shuyuan, et al.
Veröffentlicht: (2025)
Lessons from a Big-Bang Integration: Challenges in Edge Computing and Machine Learning
von: Aneggi, Alessandro, et al.
Veröffentlicht: (2025)
von: Aneggi, Alessandro, et al.
Veröffentlicht: (2025)
Automated Statistical Testing and Certification of a Reliable Model-Coupling Server for Scientific Computing
von: Wolfgang, Seth, et al.
Veröffentlicht: (2025)
von: Wolfgang, Seth, et al.
Veröffentlicht: (2025)
Real-Time Agile Software Management for Edge and Fog Computing Based Smart City Infrastructure
von: Jana, Debasish, et al.
Veröffentlicht: (2025)
von: Jana, Debasish, et al.
Veröffentlicht: (2025)
Many-Objective Search-Based Coverage-Guided Automatic Test Generation for Deep Neural Networks
von: Li, Dongcheng, et al.
Veröffentlicht: (2024)
von: Li, Dongcheng, et al.
Veröffentlicht: (2024)
Software Fault Localization Based on Multi-objective Feature Fusion and Deep Learning
von: Hu, Xiaolei, et al.
Veröffentlicht: (2024)
von: Hu, Xiaolei, et al.
Veröffentlicht: (2024)
A Hybrid Sampling and Multi-Objective Optimization Approach for Enhanced Software Defect Prediction
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
RAG-MCP: Mitigating Prompt Bloat in LLM Tool Selection via Retrieval-Augmented Generation
von: Gan, Tiantian, et al.
Veröffentlicht: (2025)
von: Gan, Tiantian, et al.
Veröffentlicht: (2025)
Empirical Computation
von: Tang, Eric, et al.
Veröffentlicht: (2025)
von: Tang, Eric, et al.
Veröffentlicht: (2025)
Exploring the Impact of Integrating UI Testing in CI/CD Workflows on GitHub
von: Gan, Xiaoxiao, et al.
Veröffentlicht: (2025)
von: Gan, Xiaoxiao, et al.
Veröffentlicht: (2025)
UIBenchKit: A unified toolkit for design-to-code model evaluation
von: Le, Chinh T., et al.
Veröffentlicht: (2026)
von: Le, Chinh T., et al.
Veröffentlicht: (2026)
On the Rationale and Use of Assertion Messages in Test Code: Insights from Software Practitioners
von: Peruma, Anthony, et al.
Veröffentlicht: (2024)
von: Peruma, Anthony, et al.
Veröffentlicht: (2024)
A Systematic Literature Review of the Use of GenAI Assistants for Code Comprehension: Implications for Computing Education Research and Practice
von: Qiao, Yunhan, et al.
Veröffentlicht: (2025)
von: Qiao, Yunhan, et al.
Veröffentlicht: (2025)
Smart Contract Vulnerability Detection based on Static Analysis and Multi-Objective Search
von: Li, Dongcheng, et al.
Veröffentlicht: (2024)
von: Li, Dongcheng, et al.
Veröffentlicht: (2024)
EdgeMLBalancer: A Self-Adaptive Approach for Dynamic Model Switching on Resource-Constrained Edge Devices
von: Matathammal, Akhila, et al.
Veröffentlicht: (2025)
von: Matathammal, Akhila, et al.
Veröffentlicht: (2025)
From Runnable to Shippable: Multi-Agent Test-Driven Development for Generating Full-Stack Web Applications from Requirements
von: Wan, Yuxuan, et al.
Veröffentlicht: (2026)
von: Wan, Yuxuan, et al.
Veröffentlicht: (2026)
Statistical Software Engineering with Tuned Variables
von: Busany, Nimrod
Veröffentlicht: (2026)
von: Busany, Nimrod
Veröffentlicht: (2026)
Verification and Validation of Autonomous Systems
von: Shetiya, Sneha Sudhir, et al.
Veröffentlicht: (2024)
von: Shetiya, Sneha Sudhir, et al.
Veröffentlicht: (2024)
HardRace: A Dynamic Data Race Monitor for Production Use
von: Sun, Xudong, et al.
Veröffentlicht: (2024)
von: Sun, Xudong, et al.
Veröffentlicht: (2024)
Teaching Code LLMs to Use Autocompletion Tools in Repository-Level Code Generation
von: Wang, Chong, et al.
Veröffentlicht: (2024)
von: Wang, Chong, et al.
Veröffentlicht: (2024)
Statistical Independence Aware Caching for LLM Workflows
von: Dai, Yihan, et al.
Veröffentlicht: (2025)
von: Dai, Yihan, et al.
Veröffentlicht: (2025)
Understanding on the Edge: LLM-generated Boundary Test Explanations
von: Akbarova, Sabinakhon, et al.
Veröffentlicht: (2026)
von: Akbarova, Sabinakhon, et al.
Veröffentlicht: (2026)
Quality Engineering for Agile and DevOps on the Cloud and Edge
von: Farchi, Eitan, et al.
Veröffentlicht: (2023)
von: Farchi, Eitan, et al.
Veröffentlicht: (2023)
A Study to Evaluate the Impact of LoRA Fine-tuning on the Performance of Non-functional Requirements Classification
von: Li, Xia, et al.
Veröffentlicht: (2025)
von: Li, Xia, et al.
Veröffentlicht: (2025)
Implicit Test Oracles for Quantum Computing
von: Langdon, William B.
Veröffentlicht: (2024)
von: Langdon, William B.
Veröffentlicht: (2024)
OS-Harm: A Benchmark for Measuring Safety of Computer Use Agents
von: Kuntz, Thomas, et al.
Veröffentlicht: (2025)
von: Kuntz, Thomas, et al.
Veröffentlicht: (2025)
Programming with Pixels: Can Computer-Use Agents do Software Engineering?
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
A Framework for QoS of Integration Testing in Satellite Edge Clouds
von: Zeng, Guogen, et al.
Veröffentlicht: (2024)
von: Zeng, Guogen, et al.
Veröffentlicht: (2024)
Tutorial Debriefing: Applied Statistical Causal Inference in Requirements Engineering
von: Frattini, Julian, et al.
Veröffentlicht: (2025)
von: Frattini, Julian, et al.
Veröffentlicht: (2025)
Statistical complexity of software systems represented as multi-layer networks
von: Žižka, Jan
Veröffentlicht: (2025)
von: Žižka, Jan
Veröffentlicht: (2025)
StatsClaw: An AI-Collaborative Workflow for Statistical Software Development
von: Qin, Tianzhu, et al.
Veröffentlicht: (2026)
von: Qin, Tianzhu, et al.
Veröffentlicht: (2026)
Defining a Reference Architecture for Edge Systems in Highly-Uncertain Environments
von: Pitstick, Kevin, et al.
Veröffentlicht: (2024)
von: Pitstick, Kevin, et al.
Veröffentlicht: (2024)
Use of Agile Practices in Start-ups
von: Klotins, Eriks, et al.
Veröffentlicht: (2024)
von: Klotins, Eriks, et al.
Veröffentlicht: (2024)
Towards Richer Challenge Problems for Scientific Computing Correctness
von: Sottile, Matthew, et al.
Veröffentlicht: (2025)
von: Sottile, Matthew, et al.
Veröffentlicht: (2025)
A Serverless Edge-Native Data Processing Architecture for Autonomous Driving Training
von: Bally, Fabian, et al.
Veröffentlicht: (2026)
von: Bally, Fabian, et al.
Veröffentlicht: (2026)
Understanding the Characteristics of LLM-Generated Property-Based Tests in Exploring Edge Cases
von: Tanaka, Hidetake, et al.
Veröffentlicht: (2025)
von: Tanaka, Hidetake, et al.
Veröffentlicht: (2025)
Which Syntactic Capabilities Are Statistically Learned by Masked Language Models for Code?
von: Velasco, Alejandro, et al.
Veröffentlicht: (2024)
von: Velasco, Alejandro, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
OpenComputer: Verifiable Software Worlds for Computer-Use Agents
von: Wei, Jinbiao, et al.
Veröffentlicht: (2026) -
ADEPTS: A Capability Framework for Human-Centered Agent Design
von: D'Oro, Pierluca, et al.
Veröffentlicht: (2025) -
DigiData: Training and Evaluating General-Purpose Mobile Control Agents
von: Sun, Yuxuan, et al.
Veröffentlicht: (2025) -
UCRBench: Benchmarking LLMs on Use Case Recovery
von: Xiao, Shuyuan, et al.
Veröffentlicht: (2025) -
Lessons from a Big-Bang Integration: Challenges in Edge Computing and Machine Learning
von: Aneggi, Alessandro, et al.
Veröffentlicht: (2025)