NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Yang, Vyas, Anoushka, Wei, Zirui, Pakazad, Sina Khoshfetrat, Ohlsson, Henrik, Neubig, Graham |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Time to Embed: Unlocking Foundation Models for Time Series with Channel Descriptions
by: Dutta, Utsav, et al.
Published: (2025)
by: Dutta, Utsav, et al.
Published: (2025)
Giving Sensors a Voice: Multimodal JEPA for Semantic Time-Series Embeddings
by: Dutta, Utsav, et al.
Published: (2026)
by: Dutta, Utsav, et al.
Published: (2026)
Better Instruction-Following Through Minimum Bayes Risk
by: Wu, Ian, et al.
Published: (2024)
by: Wu, Ian, et al.
Published: (2024)
Synthetic Multimodal Question Generation
by: Wu, Ian, et al.
Published: (2024)
by: Wu, Ian, et al.
Published: (2024)
Training Versatile Coding Agents in Synthetic Environments
by: Zhu, Yiqi, et al.
Published: (2025)
by: Zhu, Yiqi, et al.
Published: (2025)
Effective Strategies for Asynchronous Software Engineering Agents
by: Geng, Jiayi, et al.
Published: (2026)
by: Geng, Jiayi, et al.
Published: (2026)
Can Large Language Models be Trusted for Evaluation? Scalable Meta-Evaluation of LLMs as Evaluators via Agent Debate
by: Chern, Steffi, et al.
Published: (2024)
by: Chern, Steffi, et al.
Published: (2024)
Gym-Anything: Turn any Software into an Agent Environment
by: Aggarwal, Pranjal, et al.
Published: (2026)
by: Aggarwal, Pranjal, et al.
Published: (2026)
Recursive Agent Optimization
by: Gandhi, Apurva, et al.
Published: (2026)
by: Gandhi, Apurva, et al.
Published: (2026)
Investigating Execution-Aware Language Models for Code Optimization
by: Di Menna, Federico, et al.
Published: (2025)
by: Di Menna, Federico, et al.
Published: (2025)
PaperScout: An Autonomous Agent for Academic Paper Search with Process-Aware Sequence-Level Policy Optimization
by: Pan, Tingyue, et al.
Published: (2026)
by: Pan, Tingyue, et al.
Published: (2026)
PARC: An Autonomous Self-Reflective Coding Agent for Robust Execution of Long-Horizon Tasks
by: Orimo, Yuki, et al.
Published: (2025)
by: Orimo, Yuki, et al.
Published: (2025)
Fault-Tolerant Sandboxing for AI Coding Agents: A Transactional Approach to Safe Autonomous Execution
by: Yan, Boyang
Published: (2025)
by: Yan, Boyang
Published: (2025)
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
by: Guo, Chengquan, et al.
Published: (2024)
by: Guo, Chengquan, et al.
Published: (2024)
Ambig-SWE: Interactive Agents to Overcome Underspecificity in Software Engineering
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
Causal Spherical Hypergraph Networks for Modelling Social Uncertainty
by: Harit, Anoushka, et al.
Published: (2025)
by: Harit, Anoushka, et al.
Published: (2025)
NEMO-4-PAYPAL: Leveraging NVIDIA's Nemo Framework for empowering PayPal's Commerce Agent
by: Garg, Sudhanshu, et al.
Published: (2025)
by: Garg, Sudhanshu, et al.
Published: (2025)
TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks
by: Wang, Zhiruo, et al.
Published: (2024)
by: Wang, Zhiruo, et al.
Published: (2024)
Autonomous Agents on Blockchains: Standards, Execution Models, and Trust Boundaries
by: Alqithami, Saad
Published: (2026)
by: Alqithami, Saad
Published: (2026)
Executable World Models for ARC-AGI-3 in the Era of Coding Agents
by: Rodionov, Sergey
Published: (2026)
by: Rodionov, Sergey
Published: (2026)
TOM-SWE: User Mental Modeling For Software Engineering Agents
by: Zhou, Xuhui, et al.
Published: (2025)
by: Zhou, Xuhui, et al.
Published: (2025)
CowPilot: A Framework for Autonomous and Human-Agent Collaborative Web Navigation
by: Huq, Faria, et al.
Published: (2025)
by: Huq, Faria, et al.
Published: (2025)
Asking What Matters: Reward-Driven Clarification for Software Engineering Tasks
by: Vijayvargiya, Sanidhya, et al.
Published: (2026)
by: Vijayvargiya, Sanidhya, et al.
Published: (2026)
APEX: Agent Payment Execution with Policy for Autonomous Agent API Access
by: Uddin, Mohd Safwan, et al.
Published: (2026)
by: Uddin, Mohd Safwan, et al.
Published: (2026)
Actionable Interpretability via Causal Hypergraphs: Unravelling Batch Size Effects in Deep Learning
by: Sun, Zhongtian, et al.
Published: (2025)
by: Sun, Zhongtian, et al.
Published: (2025)
Key Safety Design Overview in AI-driven Autonomous Vehicles
by: Vyas, Vikas, et al.
Published: (2024)
by: Vyas, Vikas, et al.
Published: (2024)
OpenAgentSafety: A Comprehensive Framework for Evaluating Real-World AI Agent Safety
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
by: Vijayvargiya, Sanidhya, et al.
Published: (2025)
Transferable Expertise for Autonomous Agents via Real-World Case-Based Learning
by: Ma, Zhenyu, et al.
Published: (2026)
by: Ma, Zhenyu, et al.
Published: (2026)
Executable Code Actions Elicit Better LLM Agents
by: Wang, Xingyao, et al.
Published: (2024)
by: Wang, Xingyao, et al.
Published: (2024)
CodeScout: An Effective Recipe for Reinforcement Learning of Code Search Agents
by: Sutawika, Lintang, et al.
Published: (2026)
by: Sutawika, Lintang, et al.
Published: (2026)
B-PASTE: Beam-Aware Pattern-Guided Speculative Execution for Resource-Constrained LLM Agents
by: Song, Yanfei
Published: (2026)
by: Song, Yanfei
Published: (2026)
Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
by: Mai, Xinji, et al.
Published: (2025)
by: Mai, Xinji, et al.
Published: (2025)
The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution
by: Li, Junlong, et al.
Published: (2025)
by: Li, Junlong, et al.
Published: (2025)
OmegaUse: Building a General-Purpose GUI Agent for Autonomous Task Execution
by: Zhang, Le, et al.
Published: (2026)
by: Zhang, Le, et al.
Published: (2026)
LEMON: Learning Executable Multi-Agent Orchestration via Counterfactual Reinforcement Learning
by: Chen, Xudong, et al.
Published: (2026)
by: Chen, Xudong, et al.
Published: (2026)
AgentForge: Execution-Grounded Multi-Agent LLM Framework for Autonomous Software Engineering
by: Kumar, Rajesh, et al.
Published: (2026)
by: Kumar, Rajesh, et al.
Published: (2026)
SELF-GUIDE: Better Task-Specific Instruction Following via Self-Synthetic Finetuning
by: Zhao, Chenyang, et al.
Published: (2024)
by: Zhao, Chenyang, et al.
Published: (2024)
Verification and Validation of Autonomous Systems
by: Shetiya, Sneha Sudhir, et al.
Published: (2024)
by: Shetiya, Sneha Sudhir, et al.
Published: (2024)
RicciFlowRec: A Geometric Root Cause Recommender Using Ricci Curvature on Financial Graphs
by: Sun, Zhongtian, et al.
Published: (2025)
by: Sun, Zhongtian, et al.
Published: (2025)
WebArena: A Realistic Web Environment for Building Autonomous Agents
by: Zhou, Shuyan, et al.
Published: (2023)
by: Zhou, Shuyan, et al.
Published: (2023)
Similar Items
-
Time to Embed: Unlocking Foundation Models for Time Series with Channel Descriptions
by: Dutta, Utsav, et al.
Published: (2025) -
Giving Sensors a Voice: Multimodal JEPA for Semantic Time-Series Embeddings
by: Dutta, Utsav, et al.
Published: (2026) -
Better Instruction-Following Through Minimum Bayes Risk
by: Wu, Ian, et al.
Published: (2024) -
Synthetic Multimodal Question Generation
by: Wu, Ian, et al.
Published: (2024) -
Training Versatile Coding Agents in Synthetic Environments
by: Zhu, Yiqi, et al.
Published: (2025)