MM-Agent: LLM as Agents for Real-world Mathematical Modeling Problem
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Liu, Fan, Yang, Zherui, Liu, Cancheng, Song, Tianrui, Gao, Xiaofeng, Liu, Hao |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
MCPAgentBench: A Real-world Task Benchmark for Evaluating LLM Agent MCP Tool Use
par: Liu, Wenrui, et autres
Publié: (2025)
par: Liu, Wenrui, et autres
Publié: (2025)
ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis
par: Wang, Hao, et autres
Publié: (2026)
par: Wang, Hao, et autres
Publié: (2026)
Mobile GUI Agents under Real-world Threats: Are We There Yet?
par: Liu, Guohong, et autres
Publié: (2025)
par: Liu, Guohong, et autres
Publié: (2025)
OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories
par: Liu, Yibing, et autres
Publié: (2026)
par: Liu, Yibing, et autres
Publié: (2026)
Foundation Models for Scientific Discovery: From Paradigm Enhancement to Paradigm Transition
par: Liu, Fan, et autres
Publié: (2025)
par: Liu, Fan, et autres
Publié: (2025)
VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications
par: He, Wei, et autres
Publié: (2025)
par: He, Wei, et autres
Publié: (2025)
Large Language Model Enhanced Hard Sample Identification for Denoising Recommendation
par: Song, Tianrui, et autres
Publié: (2024)
par: Song, Tianrui, et autres
Publié: (2024)
Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving
par: Gao, Songyang, et autres
Publié: (2025)
par: Gao, Songyang, et autres
Publié: (2025)
CORBA: Contagious Recursive Blocking Attacks on Multi-Agent Systems Based on Large Language Models
par: Zhou, Zhenhong, et autres
Publié: (2025)
par: Zhou, Zhenhong, et autres
Publié: (2025)
Hard vs. Noise: Resolving Hard-Noisy Sample Confusion in Recommender Systems via Large Language Models
par: Song, Tianrui, et autres
Publié: (2025)
par: Song, Tianrui, et autres
Publié: (2025)
Agent-Omit: Adaptive Context Omission for Efficient LLM Agents
par: Ning, Yansong, et autres
Publié: (2026)
par: Ning, Yansong, et autres
Publié: (2026)
Mozi: Governed Autonomy for Drug Discovery LLM Agents
par: Cao, He, et autres
Publié: (2026)
par: Cao, He, et autres
Publié: (2026)
DataGovBench: Benchmarking LLM Agents for Real-World Data Governance Workflows
par: Liu, Zhou, et autres
Publié: (2025)
par: Liu, Zhou, et autres
Publié: (2025)
Sibyl: Simple yet Effective Agent Framework for Complex Real-world Reasoning
par: Wang, Yulong, et autres
Publié: (2024)
par: Wang, Yulong, et autres
Publié: (2024)
Information Fidelity in Tool-Using LLM Agents: A Martingale Analysis of the Model Context Protocol
par: Fan, Flint Xiaofeng, et autres
Publié: (2026)
par: Fan, Flint Xiaofeng, et autres
Publié: (2026)
Prover Agent: An Agent-Based Framework for Formal Mathematical Proofs
par: Baba, Kaito, et autres
Publié: (2025)
par: Baba, Kaito, et autres
Publié: (2025)
A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
par: Liu, Zijun, et autres
Publié: (2023)
par: Liu, Zijun, et autres
Publié: (2023)
Exploring Communication Strategies for Collaborative LLM Agents in Mathematical Problem-Solving
par: Zhang, Liang, et autres
Publié: (2025)
par: Zhang, Liang, et autres
Publié: (2025)
Can We Trust a Black-box LLM? LLM Untrustworthy Boundary Detection via Bias-Diffusion and Multi-Agent Reinforcement Learning
par: Zhou, Xiaotian, et autres
Publié: (2026)
par: Zhou, Xiaotian, et autres
Publié: (2026)
CoMaTrack: Competitive Multi-Agent Game-Theoretic Tracking with Vision-Language-Action Models
par: Liu, Youzhi, et autres
Publié: (2026)
par: Liu, Youzhi, et autres
Publié: (2026)
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
par: Liu, Youwei, et autres
Publié: (2026)
par: Liu, Youwei, et autres
Publié: (2026)
Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
par: Mai, Xinji, et autres
Publié: (2025)
par: Mai, Xinji, et autres
Publié: (2025)
RuleSmith: Multi-Agent LLMs for Automated Game Balancing
par: Zeng, Ziyao, et autres
Publié: (2026)
par: Zeng, Ziyao, et autres
Publié: (2026)
Regulator-Manufacturer AI Agents Modeling: Mathematical Feedback-Driven Multi-Agent LLM Framework
par: Han, Yu, et autres
Publié: (2024)
par: Han, Yu, et autres
Publié: (2024)
LLM-Powered Hierarchical Language Agent for Real-time Human-AI Coordination
par: Liu, Jijia, et autres
Publié: (2023)
par: Liu, Jijia, et autres
Publié: (2023)
SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking
par: Liu, Guohong, et autres
Publié: (2026)
par: Liu, Guohong, et autres
Publié: (2026)
Toward LLM-Agent-Based Modeling of Transportation Systems: A Conceptual Framework
par: Liu, Tianming, et autres
Publié: (2024)
par: Liu, Tianming, et autres
Publié: (2024)
AgentRec: Next-Generation LLM-Powered Multi-Agent Collaborative Recommendation with Adaptive Intelligence
par: Ma, Bo, et autres
Publié: (2025)
par: Ma, Bo, et autres
Publié: (2025)
Experience Transfer for Multimodal LLM Agents in Minecraft Game
par: Li, Chenghao, et autres
Publié: (2026)
par: Li, Chenghao, et autres
Publié: (2026)
Harnessing LLM Agents with Skill Programs
par: Liu, Hongjun, et autres
Publié: (2026)
par: Liu, Hongjun, et autres
Publié: (2026)
MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments
par: Gao, Yicheng, et autres
Publié: (2026)
par: Gao, Yicheng, et autres
Publié: (2026)
UIS-Digger: Towards Comprehensive Research Agent Systems for Real-world Unindexed Information Seeking
par: Liu, Chang, et autres
Publié: (2026)
par: Liu, Chang, et autres
Publié: (2026)
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL
par: Li, Weizhen, et autres
Publié: (2025)
par: Li, Weizhen, et autres
Publié: (2025)
ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges
par: Qian, Cheng, et autres
Publié: (2025)
par: Qian, Cheng, et autres
Publié: (2025)
Patient-Zero: Scaling Synthetic Patient Agents to Real-World Distributions without Real Patient Data
par: Lai, Yunghwei, et autres
Publié: (2025)
par: Lai, Yunghwei, et autres
Publié: (2025)
OR-LLM-Agent: Automating Modeling and Solving of Operations Research Optimization Problems with Reasoning LLM
par: Zhang, Bowen, et autres
Publié: (2025)
par: Zhang, Bowen, et autres
Publié: (2025)
CheatAgent: Attacking LLM-Empowered Recommender Systems via LLM Agent
par: Ning, Liang-bo, et autres
Publié: (2025)
par: Ning, Liang-bo, et autres
Publié: (2025)
MM-WebAgent: A Hierarchical Multimodal Web Agent for Webpage Generation
par: Li, Yan, et autres
Publié: (2026)
par: Li, Yan, et autres
Publié: (2026)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
par: Trae Research Team, et autres
Publié: (2025)
par: Trae Research Team, et autres
Publié: (2025)
Simulating Rumor Spreading in Social Networks using LLM Agents
par: Hu, Tianrui, et autres
Publié: (2025)
par: Hu, Tianrui, et autres
Publié: (2025)
Documents similaires
-
MCPAgentBench: A Real-world Task Benchmark for Evaluating LLM Agent MCP Tool Use
par: Liu, Wenrui, et autres
Publié: (2025) -
ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis
par: Wang, Hao, et autres
Publié: (2026) -
Mobile GUI Agents under Real-world Threats: Are We There Yet?
par: Liu, Guohong, et autres
Publié: (2025) -
OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories
par: Liu, Yibing, et autres
Publié: (2026) -
Foundation Models for Scientific Discovery: From Paradigm Enhancement to Paradigm Transition
par: Liu, Fan, et autres
Publié: (2025)