Toward a Trustworthy Optimization Modeling Agent via Verifiable Synthetic Data Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Lima, Vinicius, Phan, Dzung T., Kalagnanam, Jayant, Patel, Dhaval, Zhou, Nianjun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Embedding-based Synthetic Data Generation for Complex Reasoning Tasks
by: Jayaraman, Srideepika, et al.
Published: (2026)
by: Jayaraman, Srideepika, et al.
Published: (2026)
Towards Automated Solution Recipe Generation for Industrial Asset Management with LLM
by: Zhou, Nianjun, et al.
Published: (2024)
by: Zhou, Nianjun, et al.
Published: (2024)
AssetOpsBench: Benchmarking AI Agents for Task Automation in Industrial Asset Operations and Maintenance
by: Patel, Dhaval, et al.
Published: (2025)
by: Patel, Dhaval, et al.
Published: (2025)
Learning to Shuffle: Block Reshuffling and Reversal Schemes for Stochastic Optimization
by: Nguyen, Lam M., et al.
Published: (2026)
by: Nguyen, Lam M., et al.
Published: (2026)
DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules
by: De Silva, Devin Yasith, et al.
Published: (2026)
by: De Silva, Devin Yasith, et al.
Published: (2026)
MCP-Cosmos: World Model-Augmented Agents for Complex Task Execution in MCP Environments
by: Ganapavarapu, Giridhar, et al.
Published: (2026)
by: Ganapavarapu, Giridhar, et al.
Published: (2026)
From Static Templates to Dynamic Runtime Graphs: A Survey of Workflow Optimization for LLM Agents
by: Yue, Ling, et al.
Published: (2026)
by: Yue, Ling, et al.
Published: (2026)
Results and Retrospective Analysis of the CODS 2025 AssetOpsBench Challenge
by: Patel, Dhaval, et al.
Published: (2026)
by: Patel, Dhaval, et al.
Published: (2026)
SPIN: Structural LLM Planning via Iterative Navigation for Industrial Tasks
by: Ozaki, Yusuke, et al.
Published: (2026)
by: Ozaki, Yusuke, et al.
Published: (2026)
Adaptive Conformal Anomaly Detection with Time Series Foundation Models for Signal Monitoring
by: Gil, Natalia Martinez, et al.
Published: (2026)
by: Gil, Natalia Martinez, et al.
Published: (2026)
Chat-of-Thought: Collaborative Multi-Agent System for Generating Domain Specific Information
by: Constantinides, Christodoulos, et al.
Published: (2025)
by: Constantinides, Christodoulos, et al.
Published: (2025)
Forging Time Series with Language: A Large Language Model Approach to Synthetic Data Generation
by: Rousseau, Cécile, et al.
Published: (2025)
by: Rousseau, Cécile, et al.
Published: (2025)
Cardinality-Regularized Hawkes-Granger Model
by: Idé, Tsuyoshi, et al.
Published: (2022)
by: Idé, Tsuyoshi, et al.
Published: (2022)
IndustryAssetEQA: A Neurosymbolic Operational Intelligence System for Embodied Question Answering in Industrial Asset Maintenance
by: Shyalika, Chathurangi, et al.
Published: (2026)
by: Shyalika, Chathurangi, et al.
Published: (2026)
Towards Trustworthy Multi-Turn LLM Agents via Behavioral Guidance
by: Gürsun, Gonca
Published: (2025)
by: Gürsun, Gonca
Published: (2025)
TSPulse: Tiny Pre-Trained Models with Disentangled Representations for Rapid Time-Series Analysis
by: Ekambaram, Vijay, et al.
Published: (2025)
by: Ekambaram, Vijay, et al.
Published: (2025)
Optimizing AI Agent Attacks With Synthetic Data
by: Loughridge, Chloe, et al.
Published: (2025)
by: Loughridge, Chloe, et al.
Published: (2025)
Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents
by: Ye, Bowen, et al.
Published: (2026)
by: Ye, Bowen, et al.
Published: (2026)
Fine-Tuned Thoughts: Leveraging Chain-of-Thought Reasoning for Industrial Asset Health Monitoring
by: Lin, Shuxin, et al.
Published: (2025)
by: Lin, Shuxin, et al.
Published: (2025)
Tiny Time Mixers (TTMs): Fast Pre-trained Models for Enhanced Zero/Few-Shot Forecasting of Multivariate Time Series
by: Ekambaram, Vijay, et al.
Published: (2024)
by: Ekambaram, Vijay, et al.
Published: (2024)
Enhancing Computer Programming Education with LLMs: A Study on Effective Prompt Engineering for Python Code Generation
by: Wang, Tianyu, et al.
Published: (2024)
by: Wang, Tianyu, et al.
Published: (2024)
NP-Engine: Empowering Optimization Reasoning in Large Language Models with Verifiable Synthetic NP Problems
by: Li, Xiaozhe, et al.
Published: (2025)
by: Li, Xiaozhe, et al.
Published: (2025)
AEMA: Verifiable Evaluation Framework for Trustworthy and Controlled Agentic LLM Systems
by: Lee, YenTing, et al.
Published: (2026)
by: Lee, YenTing, et al.
Published: (2026)
TrustGeoGen: Formal-Verified Data Engine for Trustworthy Multi-modal Geometric Problem Solving
by: Fu, Daocheng, et al.
Published: (2025)
by: Fu, Daocheng, et al.
Published: (2025)
Towards Trustworthy Legal AI through LLM Agents and Formal Reasoning
by: Chen, Linze, et al.
Published: (2025)
by: Chen, Linze, et al.
Published: (2025)
Grounding Generative Planners in Verifiable Logic: A Hybrid Architecture for Trustworthy Embodied AI
by: Wu, Feiyu, et al.
Published: (2026)
by: Wu, Feiyu, et al.
Published: (2026)
Scaling Synthetic Task Generation for Agents via Exploration
by: Ramrakhya, Ram, et al.
Published: (2025)
by: Ramrakhya, Ram, et al.
Published: (2025)
AutoOpt: A Dataset and a Unified Framework for Automating Optimization Problem Solving
by: Sinha, Ankur, et al.
Published: (2025)
by: Sinha, Ankur, et al.
Published: (2025)
CALICO: Conversational Agent Localization via Synthetic Data Generation
by: Rosenbaum, Andy, et al.
Published: (2024)
by: Rosenbaum, Andy, et al.
Published: (2024)
Deep Policy Iteration with Integer Programming for Inventory Management
by: Harsha, Pavithra, et al.
Published: (2021)
by: Harsha, Pavithra, et al.
Published: (2021)
SWE-Synth: Synthesizing Verifiable Bug-Fix Data to Enable Large Language Models in Resolving Real-World Bugs
by: Pham, Minh V. T., et al.
Published: (2025)
by: Pham, Minh V. T., et al.
Published: (2025)
Towards Trustworthy Report Generation: A Deep Research Agent with Progressive Confidence Estimation and Calibration
by: Yuan, Yi, et al.
Published: (2026)
by: Yuan, Yi, et al.
Published: (2026)
FailureSensorIQ: A Multi-Choice QA Dataset for Understanding Sensor Relationships and Failure Modes
by: Constantinides, Christodoulos, et al.
Published: (2025)
by: Constantinides, Christodoulos, et al.
Published: (2025)
The LLM Data Auditor: A Metric-oriented Survey on Quality and Trustworthiness in Evaluating Synthetic Data
by: Zhang, Kaituo, et al.
Published: (2026)
by: Zhang, Kaituo, et al.
Published: (2026)
Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering
by: Guan, Xinyan, et al.
Published: (2024)
by: Guan, Xinyan, et al.
Published: (2024)
Multi-Agent Legal Verifier Systems for Data Transfer Planning
by: Nguyen, Ha-Thanh, et al.
Published: (2025)
by: Nguyen, Ha-Thanh, et al.
Published: (2025)
VET Your Agent: Towards Host-Independent Autonomy via Verifiable Execution Traces
by: Grigor, Artem, et al.
Published: (2025)
by: Grigor, Artem, et al.
Published: (2025)
Toward Safe and Responsible AI Agents: A Three-Pillar Model for Transparency, Accountability, and Trustworthiness
by: Cheng, Edward C., et al.
Published: (2026)
by: Cheng, Edward C., et al.
Published: (2026)
Towards the Development of Balanced Synthetic Data for Correcting Grammatical Errors in Arabic: An Approach Based on Error Tagging Model and Synthetic Data Generating Model
by: Alrehili, Ahlam, et al.
Published: (2025)
by: Alrehili, Ahlam, et al.
Published: (2025)
Towards Trustworthy Retrieval Augmented Generation for Large Language Models: A Survey
by: Ni, Bo, et al.
Published: (2025)
by: Ni, Bo, et al.
Published: (2025)
Similar Items
-
Efficient Embedding-based Synthetic Data Generation for Complex Reasoning Tasks
by: Jayaraman, Srideepika, et al.
Published: (2026) -
Towards Automated Solution Recipe Generation for Industrial Asset Management with LLM
by: Zhou, Nianjun, et al.
Published: (2024) -
AssetOpsBench: Benchmarking AI Agents for Task Automation in Industrial Asset Operations and Maintenance
by: Patel, Dhaval, et al.
Published: (2025) -
Learning to Shuffle: Block Reshuffling and Reversal Schemes for Stochastic Optimization
by: Nguyen, Lam M., et al.
Published: (2026) -
DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules
by: De Silva, Devin Yasith, et al.
Published: (2026)