NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents
Fuente:
arXiv
Salvato in:
| Autori principali: | Song, Yang, Vyas, Anoushka, Wei, Zirui, Pakazad, Sina Khoshfetrat, Ohlsson, Henrik, Neubig, Graham |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Time to Embed: Unlocking Foundation Models for Time Series with Channel Descriptions
di: Dutta, Utsav, et al.
Pubblicazione: (2025)
di: Dutta, Utsav, et al.
Pubblicazione: (2025)
Giving Sensors a Voice: Multimodal JEPA for Semantic Time-Series Embeddings
di: Dutta, Utsav, et al.
Pubblicazione: (2026)
di: Dutta, Utsav, et al.
Pubblicazione: (2026)
Better Instruction-Following Through Minimum Bayes Risk
di: Wu, Ian, et al.
Pubblicazione: (2024)
di: Wu, Ian, et al.
Pubblicazione: (2024)
Synthetic Multimodal Question Generation
di: Wu, Ian, et al.
Pubblicazione: (2024)
di: Wu, Ian, et al.
Pubblicazione: (2024)
Training Versatile Coding Agents in Synthetic Environments
di: Zhu, Yiqi, et al.
Pubblicazione: (2025)
di: Zhu, Yiqi, et al.
Pubblicazione: (2025)
Effective Strategies for Asynchronous Software Engineering Agents
di: Geng, Jiayi, et al.
Pubblicazione: (2026)
di: Geng, Jiayi, et al.
Pubblicazione: (2026)
Can Large Language Models be Trusted for Evaluation? Scalable Meta-Evaluation of LLMs as Evaluators via Agent Debate
di: Chern, Steffi, et al.
Pubblicazione: (2024)
di: Chern, Steffi, et al.
Pubblicazione: (2024)
Gym-Anything: Turn any Software into an Agent Environment
di: Aggarwal, Pranjal, et al.
Pubblicazione: (2026)
di: Aggarwal, Pranjal, et al.
Pubblicazione: (2026)
Recursive Agent Optimization
di: Gandhi, Apurva, et al.
Pubblicazione: (2026)
di: Gandhi, Apurva, et al.
Pubblicazione: (2026)
Investigating Execution-Aware Language Models for Code Optimization
di: Di Menna, Federico, et al.
Pubblicazione: (2025)
di: Di Menna, Federico, et al.
Pubblicazione: (2025)
PaperScout: An Autonomous Agent for Academic Paper Search with Process-Aware Sequence-Level Policy Optimization
di: Pan, Tingyue, et al.
Pubblicazione: (2026)
di: Pan, Tingyue, et al.
Pubblicazione: (2026)
PARC: An Autonomous Self-Reflective Coding Agent for Robust Execution of Long-Horizon Tasks
di: Orimo, Yuki, et al.
Pubblicazione: (2025)
di: Orimo, Yuki, et al.
Pubblicazione: (2025)
Fault-Tolerant Sandboxing for AI Coding Agents: A Transactional Approach to Safe Autonomous Execution
di: Yan, Boyang
Pubblicazione: (2025)
di: Yan, Boyang
Pubblicazione: (2025)
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
di: Guo, Chengquan, et al.
Pubblicazione: (2024)
di: Guo, Chengquan, et al.
Pubblicazione: (2024)
Ambig-SWE: Interactive Agents to Overcome Underspecificity in Software Engineering
di: Vijayvargiya, Sanidhya, et al.
Pubblicazione: (2025)
di: Vijayvargiya, Sanidhya, et al.
Pubblicazione: (2025)
Causal Spherical Hypergraph Networks for Modelling Social Uncertainty
di: Harit, Anoushka, et al.
Pubblicazione: (2025)
di: Harit, Anoushka, et al.
Pubblicazione: (2025)
NEMO-4-PAYPAL: Leveraging NVIDIA's Nemo Framework for empowering PayPal's Commerce Agent
di: Garg, Sudhanshu, et al.
Pubblicazione: (2025)
di: Garg, Sudhanshu, et al.
Pubblicazione: (2025)
TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks
di: Wang, Zhiruo, et al.
Pubblicazione: (2024)
di: Wang, Zhiruo, et al.
Pubblicazione: (2024)
Autonomous Agents on Blockchains: Standards, Execution Models, and Trust Boundaries
di: Alqithami, Saad
Pubblicazione: (2026)
di: Alqithami, Saad
Pubblicazione: (2026)
Executable World Models for ARC-AGI-3 in the Era of Coding Agents
di: Rodionov, Sergey
Pubblicazione: (2026)
di: Rodionov, Sergey
Pubblicazione: (2026)
TOM-SWE: User Mental Modeling For Software Engineering Agents
di: Zhou, Xuhui, et al.
Pubblicazione: (2025)
di: Zhou, Xuhui, et al.
Pubblicazione: (2025)
CowPilot: A Framework for Autonomous and Human-Agent Collaborative Web Navigation
di: Huq, Faria, et al.
Pubblicazione: (2025)
di: Huq, Faria, et al.
Pubblicazione: (2025)
Asking What Matters: Reward-Driven Clarification for Software Engineering Tasks
di: Vijayvargiya, Sanidhya, et al.
Pubblicazione: (2026)
di: Vijayvargiya, Sanidhya, et al.
Pubblicazione: (2026)
APEX: Agent Payment Execution with Policy for Autonomous Agent API Access
di: Uddin, Mohd Safwan, et al.
Pubblicazione: (2026)
di: Uddin, Mohd Safwan, et al.
Pubblicazione: (2026)
Actionable Interpretability via Causal Hypergraphs: Unravelling Batch Size Effects in Deep Learning
di: Sun, Zhongtian, et al.
Pubblicazione: (2025)
di: Sun, Zhongtian, et al.
Pubblicazione: (2025)
Key Safety Design Overview in AI-driven Autonomous Vehicles
di: Vyas, Vikas, et al.
Pubblicazione: (2024)
di: Vyas, Vikas, et al.
Pubblicazione: (2024)
OpenAgentSafety: A Comprehensive Framework for Evaluating Real-World AI Agent Safety
di: Vijayvargiya, Sanidhya, et al.
Pubblicazione: (2025)
di: Vijayvargiya, Sanidhya, et al.
Pubblicazione: (2025)
Transferable Expertise for Autonomous Agents via Real-World Case-Based Learning
di: Ma, Zhenyu, et al.
Pubblicazione: (2026)
di: Ma, Zhenyu, et al.
Pubblicazione: (2026)
Executable Code Actions Elicit Better LLM Agents
di: Wang, Xingyao, et al.
Pubblicazione: (2024)
di: Wang, Xingyao, et al.
Pubblicazione: (2024)
CodeScout: An Effective Recipe for Reinforcement Learning of Code Search Agents
di: Sutawika, Lintang, et al.
Pubblicazione: (2026)
di: Sutawika, Lintang, et al.
Pubblicazione: (2026)
Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
di: Mai, Xinji, et al.
Pubblicazione: (2025)
di: Mai, Xinji, et al.
Pubblicazione: (2025)
B-PASTE: Beam-Aware Pattern-Guided Speculative Execution for Resource-Constrained LLM Agents
di: Song, Yanfei
Pubblicazione: (2026)
di: Song, Yanfei
Pubblicazione: (2026)
The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution
di: Li, Junlong, et al.
Pubblicazione: (2025)
di: Li, Junlong, et al.
Pubblicazione: (2025)
OmegaUse: Building a General-Purpose GUI Agent for Autonomous Task Execution
di: Zhang, Le, et al.
Pubblicazione: (2026)
di: Zhang, Le, et al.
Pubblicazione: (2026)
LEMON: Learning Executable Multi-Agent Orchestration via Counterfactual Reinforcement Learning
di: Chen, Xudong, et al.
Pubblicazione: (2026)
di: Chen, Xudong, et al.
Pubblicazione: (2026)
AgentForge: Execution-Grounded Multi-Agent LLM Framework for Autonomous Software Engineering
di: Kumar, Rajesh, et al.
Pubblicazione: (2026)
di: Kumar, Rajesh, et al.
Pubblicazione: (2026)
SELF-GUIDE: Better Task-Specific Instruction Following via Self-Synthetic Finetuning
di: Zhao, Chenyang, et al.
Pubblicazione: (2024)
di: Zhao, Chenyang, et al.
Pubblicazione: (2024)
Verification and Validation of Autonomous Systems
di: Shetiya, Sneha Sudhir, et al.
Pubblicazione: (2024)
di: Shetiya, Sneha Sudhir, et al.
Pubblicazione: (2024)
RicciFlowRec: A Geometric Root Cause Recommender Using Ricci Curvature on Financial Graphs
di: Sun, Zhongtian, et al.
Pubblicazione: (2025)
di: Sun, Zhongtian, et al.
Pubblicazione: (2025)
WebArena: A Realistic Web Environment for Building Autonomous Agents
di: Zhou, Shuyan, et al.
Pubblicazione: (2023)
di: Zhou, Shuyan, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Time to Embed: Unlocking Foundation Models for Time Series with Channel Descriptions
di: Dutta, Utsav, et al.
Pubblicazione: (2025) -
Giving Sensors a Voice: Multimodal JEPA for Semantic Time-Series Embeddings
di: Dutta, Utsav, et al.
Pubblicazione: (2026) -
Better Instruction-Following Through Minimum Bayes Risk
di: Wu, Ian, et al.
Pubblicazione: (2024) -
Synthetic Multimodal Question Generation
di: Wu, Ian, et al.
Pubblicazione: (2024) -
Training Versatile Coding Agents in Synthetic Environments
di: Zhu, Yiqi, et al.
Pubblicazione: (2025)