Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction
Fuente:
arXiv
Guardado en:
| Autores principales: | Dennis, Simon, Shabahang, Kevin, Guo, Hao, Patil, Rivaan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost
por: Dennis, Simon, et al.
Publicado: (2026)
por: Dennis, Simon, et al.
Publicado: (2026)
When Mean CE Fails: Median CE Can Better Track Language Model Quality
por: Guo, Hao, et al.
Publicado: (2026)
por: Guo, Hao, et al.
Publicado: (2026)
In-Context Prompting Obsoletes Agent Orchestration for Procedural Tasks
por: Dennis, Simon, et al.
Publicado: (2026)
por: Dennis, Simon, et al.
Publicado: (2026)
Beyond Local Code Optimization: Multi-Agent Reasoning for Software System Optimization
por: Peng, Huiyun, et al.
Publicado: (2026)
por: Peng, Huiyun, et al.
Publicado: (2026)
Towards Specification-Driven LLM-Based Generation of Embedded Automotive Software
por: Patil, Minal Suresh, et al.
Publicado: (2024)
por: Patil, Minal Suresh, et al.
Publicado: (2024)
Deploy-Master: Automating the Deployment of 50,000+ Agent-Ready Scientific Tools in One Day
por: Wang, Yi, et al.
Publicado: (2026)
por: Wang, Yi, et al.
Publicado: (2026)
Towards Better Code Understanding in Decoder-Only Models with Contrastive Learning
por: Lin, Jiayi, et al.
Publicado: (2024)
por: Lin, Jiayi, et al.
Publicado: (2024)
Test Before You Deploy: Governing Updates in the LLM Supply Chain
por: Chishti, Mohd Sameen, et al.
Publicado: (2026)
por: Chishti, Mohd Sameen, et al.
Publicado: (2026)
Enhancing Deployment-Time Predictive Model Robustness for Code Analysis and Optimization
por: Wang, Huanting, et al.
Publicado: (2024)
por: Wang, Huanting, et al.
Publicado: (2024)
DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs
por: Liu, Zhihan, et al.
Publicado: (2024)
por: Liu, Zhihan, et al.
Publicado: (2024)
WhatsCode: Large-Scale GenAI Deployment for Developer Efficiency at WhatsApp
por: Mao, Ke, et al.
Publicado: (2025)
por: Mao, Ke, et al.
Publicado: (2025)
Beyond Blind Spots: Analytic Hints for Mitigating LLM-Based Evaluation Pitfalls
por: Fandina, Ora Nova, et al.
Publicado: (2025)
por: Fandina, Ora Nova, et al.
Publicado: (2025)
MEMRES: A Memory-Augmented Resolver with Confidence Cascade for Agentic Python Dependency Resolution
por: Minh, Dao Sy Duy, et al.
Publicado: (2026)
por: Minh, Dao Sy Duy, et al.
Publicado: (2026)
What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook
por: Huo, Junyu, et al.
Publicado: (2026)
por: Huo, Junyu, et al.
Publicado: (2026)
Tu(r)ning AI Green: Exploring Energy Efficiency Cascading with Orthogonal Optimizations
por: Rajput, Saurabhsingh, et al.
Publicado: (2025)
por: Rajput, Saurabhsingh, et al.
Publicado: (2025)
A Blueprint for AI-Driven Software Quality: Integrating LLMs with Established Standards
por: Patil, Avinash
Publicado: (2025)
por: Patil, Avinash
Publicado: (2025)
Bootstrapping Code Translation with Weighted Multilanguage Exploration
por: Wu, Yuhan, et al.
Publicado: (2026)
por: Wu, Yuhan, et al.
Publicado: (2026)
Help Without Being Asked: A Deployed Proactive Agent System for On-Call Support with Continuous Self-Improvement
por: Liu, Fengrui, et al.
Publicado: (2026)
por: Liu, Fengrui, et al.
Publicado: (2026)
The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment
por: Yu, Shasha, et al.
Publicado: (2026)
por: Yu, Shasha, et al.
Publicado: (2026)
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation
por: Arrieta, Aitor, et al.
Publicado: (2025)
por: Arrieta, Aitor, et al.
Publicado: (2025)
Enhancing Medical Learning and Reasoning Systems: A Boxology-Based Comparative Analysis of Design Patterns
por: Ng, Chi Him
Publicado: (2024)
por: Ng, Chi Him
Publicado: (2024)
ChatGPT vs. DeepSeek: A Comparative Study on AI-Based Code Generation
por: Manik, Md Motaleb Hossen
Publicado: (2025)
por: Manik, Md Motaleb Hossen
Publicado: (2025)
Consolidating TinyML Lifecycle with Large Language Models: Reality, Illusion, or Opportunity?
por: Wu, Guanghan, et al.
Publicado: (2025)
por: Wu, Guanghan, et al.
Publicado: (2025)
AIPC: Agent-Based Automation for AI Model Deployment with Qualcomm AI Runtime
por: Su, Jianhao, et al.
Publicado: (2026)
por: Su, Jianhao, et al.
Publicado: (2026)
PBT-Bench: Benchmarking AI Agents on Property-Based Testing
por: Jing, Lucas, et al.
Publicado: (2026)
por: Jing, Lucas, et al.
Publicado: (2026)
Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries
por: Nirujan, Hinduja, et al.
Publicado: (2026)
por: Nirujan, Hinduja, et al.
Publicado: (2026)
Meta-Engineering Harnesses for AI-Native Software Production: A Contract-Driven Adversarial Verification Architecture with Early Deployment Report
por: Sengupta, Satadru, et al.
Publicado: (2026)
por: Sengupta, Satadru, et al.
Publicado: (2026)
Uncovering Systematic Failures of LLMs in Verifying Code Against Natural Language Specifications
por: Jin, Haolin, et al.
Publicado: (2025)
por: Jin, Haolin, et al.
Publicado: (2025)
Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?
por: Orvalho, Pedro, et al.
Publicado: (2025)
por: Orvalho, Pedro, et al.
Publicado: (2025)
SWE-PRBench: Benchmarking AI Code Review Quality Against Pull Request Feedback
por: Kumar, Deepak
Publicado: (2026)
por: Kumar, Deepak
Publicado: (2026)
Instruction-Tuning Open-Weight Language Models for BPMN Model Generation
por: Çelikmasat, Gökberk, et al.
Publicado: (2025)
por: Çelikmasat, Gökberk, et al.
Publicado: (2025)
AST-PAC: AST-guided Membership Inference for Code
por: Koohestani, Roham, et al.
Publicado: (2026)
por: Koohestani, Roham, et al.
Publicado: (2026)
TENET: Leveraging Tests Beyond Validation for Code Generation
por: Hu, Yiran, et al.
Publicado: (2025)
por: Hu, Yiran, et al.
Publicado: (2025)
Efficient Story Point Estimation With Comparative Learning
por: Khan, Monoshiz Mahbub, et al.
Publicado: (2025)
por: Khan, Monoshiz Mahbub, et al.
Publicado: (2025)
Fuzzy Inference System for Test Case Prioritization in Software Testing
por: Karatayev, Aron, et al.
Publicado: (2024)
por: Karatayev, Aron, et al.
Publicado: (2024)
Deploying Privacy Guardrails for LLMs: A Comparative Analysis of Real-World Applications
por: Asthana, Shubhi, et al.
Publicado: (2025)
por: Asthana, Shubhi, et al.
Publicado: (2025)
Beyond Retrieval: A Multitask Benchmark and Model for Code Search
por: Xue, Siqiao, et al.
Publicado: (2026)
por: Xue, Siqiao, et al.
Publicado: (2026)
Beyond the 'Diff': Addressing Agentic Entropy in Agentic Software Development
por: Casserini, Matteo, et al.
Publicado: (2026)
por: Casserini, Matteo, et al.
Publicado: (2026)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
por: Liu, Fang, et al.
Publicado: (2024)
por: Liu, Fang, et al.
Publicado: (2024)
Stop Comparing LLM Agents Without Disclosing the Harness
por: Zhang, Yunbei, et al.
Publicado: (2026)
por: Zhang, Yunbei, et al.
Publicado: (2026)
Ejemplares similares
-
Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost
por: Dennis, Simon, et al.
Publicado: (2026) -
When Mean CE Fails: Median CE Can Better Track Language Model Quality
por: Guo, Hao, et al.
Publicado: (2026) -
In-Context Prompting Obsoletes Agent Orchestration for Procedural Tasks
por: Dennis, Simon, et al.
Publicado: (2026) -
Beyond Local Code Optimization: Multi-Agent Reasoning for Software System Optimization
por: Peng, Huiyun, et al.
Publicado: (2026) -
Towards Specification-Driven LLM-Based Generation of Embedded Automotive Software
por: Patil, Minal Suresh, et al.
Publicado: (2024)