MM-Agent: LLM as Agents for Real-world Mathematical Modeling Problem
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Fan, Yang, Zherui, Liu, Cancheng, Song, Tianrui, Gao, Xiaofeng, Liu, Hao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MCPAgentBench: A Real-world Task Benchmark for Evaluating LLM Agent MCP Tool Use
por: Liu, Wenrui, et al.
Publicado: (2025)
por: Liu, Wenrui, et al.
Publicado: (2025)
ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis
por: Wang, Hao, et al.
Publicado: (2026)
por: Wang, Hao, et al.
Publicado: (2026)
Mobile GUI Agents under Real-world Threats: Are We There Yet?
por: Liu, Guohong, et al.
Publicado: (2025)
por: Liu, Guohong, et al.
Publicado: (2025)
OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories
por: Liu, Yibing, et al.
Publicado: (2026)
por: Liu, Yibing, et al.
Publicado: (2026)
Foundation Models for Scientific Discovery: From Paradigm Enhancement to Paradigm Transition
por: Liu, Fan, et al.
Publicado: (2025)
por: Liu, Fan, et al.
Publicado: (2025)
VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications
por: He, Wei, et al.
Publicado: (2025)
por: He, Wei, et al.
Publicado: (2025)
Large Language Model Enhanced Hard Sample Identification for Denoising Recommendation
por: Song, Tianrui, et al.
Publicado: (2024)
por: Song, Tianrui, et al.
Publicado: (2024)
Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving
por: Gao, Songyang, et al.
Publicado: (2025)
por: Gao, Songyang, et al.
Publicado: (2025)
CORBA: Contagious Recursive Blocking Attacks on Multi-Agent Systems Based on Large Language Models
por: Zhou, Zhenhong, et al.
Publicado: (2025)
por: Zhou, Zhenhong, et al.
Publicado: (2025)
Hard vs. Noise: Resolving Hard-Noisy Sample Confusion in Recommender Systems via Large Language Models
por: Song, Tianrui, et al.
Publicado: (2025)
por: Song, Tianrui, et al.
Publicado: (2025)
Agent-Omit: Adaptive Context Omission for Efficient LLM Agents
por: Ning, Yansong, et al.
Publicado: (2026)
por: Ning, Yansong, et al.
Publicado: (2026)
Mozi: Governed Autonomy for Drug Discovery LLM Agents
por: Cao, He, et al.
Publicado: (2026)
por: Cao, He, et al.
Publicado: (2026)
DataGovBench: Benchmarking LLM Agents for Real-World Data Governance Workflows
por: Liu, Zhou, et al.
Publicado: (2025)
por: Liu, Zhou, et al.
Publicado: (2025)
Sibyl: Simple yet Effective Agent Framework for Complex Real-world Reasoning
por: Wang, Yulong, et al.
Publicado: (2024)
por: Wang, Yulong, et al.
Publicado: (2024)
Information Fidelity in Tool-Using LLM Agents: A Martingale Analysis of the Model Context Protocol
por: Fan, Flint Xiaofeng, et al.
Publicado: (2026)
por: Fan, Flint Xiaofeng, et al.
Publicado: (2026)
Prover Agent: An Agent-Based Framework for Formal Mathematical Proofs
por: Baba, Kaito, et al.
Publicado: (2025)
por: Baba, Kaito, et al.
Publicado: (2025)
A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
por: Liu, Zijun, et al.
Publicado: (2023)
por: Liu, Zijun, et al.
Publicado: (2023)
Exploring Communication Strategies for Collaborative LLM Agents in Mathematical Problem-Solving
por: Zhang, Liang, et al.
Publicado: (2025)
por: Zhang, Liang, et al.
Publicado: (2025)
Can We Trust a Black-box LLM? LLM Untrustworthy Boundary Detection via Bias-Diffusion and Multi-Agent Reinforcement Learning
por: Zhou, Xiaotian, et al.
Publicado: (2026)
por: Zhou, Xiaotian, et al.
Publicado: (2026)
CoMaTrack: Competitive Multi-Agent Game-Theoretic Tracking with Vision-Language-Action Models
por: Liu, Youzhi, et al.
Publicado: (2026)
por: Liu, Youzhi, et al.
Publicado: (2026)
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
por: Liu, Youwei, et al.
Publicado: (2026)
por: Liu, Youwei, et al.
Publicado: (2026)
Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
por: Mai, Xinji, et al.
Publicado: (2025)
por: Mai, Xinji, et al.
Publicado: (2025)
RuleSmith: Multi-Agent LLMs for Automated Game Balancing
por: Zeng, Ziyao, et al.
Publicado: (2026)
por: Zeng, Ziyao, et al.
Publicado: (2026)
Regulator-Manufacturer AI Agents Modeling: Mathematical Feedback-Driven Multi-Agent LLM Framework
por: Han, Yu, et al.
Publicado: (2024)
por: Han, Yu, et al.
Publicado: (2024)
LLM-Powered Hierarchical Language Agent for Real-time Human-AI Coordination
por: Liu, Jijia, et al.
Publicado: (2023)
por: Liu, Jijia, et al.
Publicado: (2023)
SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking
por: Liu, Guohong, et al.
Publicado: (2026)
por: Liu, Guohong, et al.
Publicado: (2026)
Toward LLM-Agent-Based Modeling of Transportation Systems: A Conceptual Framework
por: Liu, Tianming, et al.
Publicado: (2024)
por: Liu, Tianming, et al.
Publicado: (2024)
AgentRec: Next-Generation LLM-Powered Multi-Agent Collaborative Recommendation with Adaptive Intelligence
por: Ma, Bo, et al.
Publicado: (2025)
por: Ma, Bo, et al.
Publicado: (2025)
Experience Transfer for Multimodal LLM Agents in Minecraft Game
por: Li, Chenghao, et al.
Publicado: (2026)
por: Li, Chenghao, et al.
Publicado: (2026)
Harnessing LLM Agents with Skill Programs
por: Liu, Hongjun, et al.
Publicado: (2026)
por: Liu, Hongjun, et al.
Publicado: (2026)
MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments
por: Gao, Yicheng, et al.
Publicado: (2026)
por: Gao, Yicheng, et al.
Publicado: (2026)
UIS-Digger: Towards Comprehensive Research Agent Systems for Real-world Unindexed Information Seeking
por: Liu, Chang, et al.
Publicado: (2026)
por: Liu, Chang, et al.
Publicado: (2026)
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL
por: Li, Weizhen, et al.
Publicado: (2025)
por: Li, Weizhen, et al.
Publicado: (2025)
ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges
por: Qian, Cheng, et al.
Publicado: (2025)
por: Qian, Cheng, et al.
Publicado: (2025)
Patient-Zero: Scaling Synthetic Patient Agents to Real-World Distributions without Real Patient Data
por: Lai, Yunghwei, et al.
Publicado: (2025)
por: Lai, Yunghwei, et al.
Publicado: (2025)
OR-LLM-Agent: Automating Modeling and Solving of Operations Research Optimization Problems with Reasoning LLM
por: Zhang, Bowen, et al.
Publicado: (2025)
por: Zhang, Bowen, et al.
Publicado: (2025)
CheatAgent: Attacking LLM-Empowered Recommender Systems via LLM Agent
por: Ning, Liang-bo, et al.
Publicado: (2025)
por: Ning, Liang-bo, et al.
Publicado: (2025)
MM-WebAgent: A Hierarchical Multimodal Web Agent for Webpage Generation
por: Li, Yan, et al.
Publicado: (2026)
por: Li, Yan, et al.
Publicado: (2026)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
por: Trae Research Team, et al.
Publicado: (2025)
por: Trae Research Team, et al.
Publicado: (2025)
Simulating Rumor Spreading in Social Networks using LLM Agents
por: Hu, Tianrui, et al.
Publicado: (2025)
por: Hu, Tianrui, et al.
Publicado: (2025)
Ejemplares similares
-
MCPAgentBench: A Real-world Task Benchmark for Evaluating LLM Agent MCP Tool Use
por: Liu, Wenrui, et al.
Publicado: (2025) -
ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis
por: Wang, Hao, et al.
Publicado: (2026) -
Mobile GUI Agents under Real-world Threats: Are We There Yet?
por: Liu, Guohong, et al.
Publicado: (2025) -
OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories
por: Liu, Yibing, et al.
Publicado: (2026) -
Foundation Models for Scientific Discovery: From Paradigm Enhancement to Paradigm Transition
por: Liu, Fan, et al.
Publicado: (2025)