Can Vibe Coding Beat Graduate CS Students? An LLM vs. Human Coding Tournament on Market-driven Strategic Planning
Fuente:
arXiv
Guardado en:
| Autores principales: | Danassis, Panayiotis, Goel, Naman |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Data-Driven Discretized CS:GO Simulation Environment to Facilitate Strategic Multi-Agent Planning Research
por: Wang, Yunzhe, et al.
Publicado: (2025)
por: Wang, Yunzhe, et al.
Publicado: (2025)
Soft Tournament Equilibrium
por: Alqithami, Saad
Publicado: (2026)
por: Alqithami, Saad
Publicado: (2026)
Improving Health Information Access in the World's Largest Maternal Mobile Health Program via Bandit Algorithms
por: Lalan, Arshika, et al.
Publicado: (2024)
por: Lalan, Arshika, et al.
Publicado: (2024)
Deep Reinforcement Learning Agents for Strategic Production Policies in Microeconomic Market Simulations
por: Garrido-Merchán, Eduardo C., et al.
Publicado: (2024)
por: Garrido-Merchán, Eduardo C., et al.
Publicado: (2024)
Using Analytics on Student Created Data to Content Validate Pedagogical Tools
por: Kos, John, et al.
Publicado: (2023)
por: Kos, John, et al.
Publicado: (2023)
Frontier Coding Agents Can Now Implement an AlphaZero Self-Play Machine Learning Pipeline For Connect Four That Performs Comparably to an External Solver
por: Sherwood, Joshua, et al.
Publicado: (2026)
por: Sherwood, Joshua, et al.
Publicado: (2026)
AKIBoards: A Structure-Following Multiagent System for Predicting Acute Kidney Injury
por: Gordon, David, et al.
Publicado: (2025)
por: Gordon, David, et al.
Publicado: (2025)
A Subgoal-driven Framework for Improving Long-Horizon LLM Agents
por: Wang, Taiyi, et al.
Publicado: (2026)
por: Wang, Taiyi, et al.
Publicado: (2026)
SPIRAL: Symbolic LLM Planning via Grounded and Reflective Search
por: Zhang, Yifan, et al.
Publicado: (2025)
por: Zhang, Yifan, et al.
Publicado: (2025)
PIVOT: Bridging Planning and Execution in LLM Agents via Trajectory Refinement
por: Zhang, Tuo, et al.
Publicado: (2026)
por: Zhang, Tuo, et al.
Publicado: (2026)
The Influence of Human-inspired Agentic Sophistication in LLM-driven Strategic Reasoners
por: Trencsenyi, Vince, et al.
Publicado: (2025)
por: Trencsenyi, Vince, et al.
Publicado: (2025)
LERO: LLM-driven Evolutionary framework with Hybrid Rewards and Enhanced Observation for Multi-Agent Reinforcement Learning
por: Wei, Yuan, et al.
Publicado: (2025)
por: Wei, Yuan, et al.
Publicado: (2025)
From Glue-Code to Protocols: A Critical Analysis of A2A and MCP Integration for Scalable Agent Systems
por: Li, Qiaomu, et al.
Publicado: (2025)
por: Li, Qiaomu, et al.
Publicado: (2025)
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
por: Xu, Zelai, et al.
Publicado: (2023)
por: Xu, Zelai, et al.
Publicado: (2023)
Training Generalizable Collaborative Agents via Strategic Risk Aversion
por: Qu, Chengrui, et al.
Publicado: (2026)
por: Qu, Chengrui, et al.
Publicado: (2026)
Explaining Strategic Decisions in Multi-Agent Reinforcement Learning for Aerial Combat Tactics
por: Selmonaj, Ardian, et al.
Publicado: (2025)
por: Selmonaj, Ardian, et al.
Publicado: (2025)
Can Agents Judge Systematic Reviews Like Humans? Evaluating SLRs with LLM-based Multi-Agent System
por: Mushtaq, Abdullah, et al.
Publicado: (2025)
por: Mushtaq, Abdullah, et al.
Publicado: (2025)
Carbon Market Simulation with Adaptive Mechanism Design
por: Wang, Han, et al.
Publicado: (2024)
por: Wang, Han, et al.
Publicado: (2024)
Dynamic Speculative Agent Planning
por: Guan, Yilin, et al.
Publicado: (2025)
por: Guan, Yilin, et al.
Publicado: (2025)
STACKFEED: Structured Textual Actor-Critic Knowledge Base Editing with FeedBack
por: Kirtania, Shashank, et al.
Publicado: (2024)
por: Kirtania, Shashank, et al.
Publicado: (2024)
GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis
por: Liu, Haoyang, et al.
Publicado: (2025)
por: Liu, Haoyang, et al.
Publicado: (2025)
Agent-Oriented Planning in Multi-Agent Systems
por: Li, Ao, et al.
Publicado: (2024)
por: Li, Ao, et al.
Publicado: (2024)
Efficient Multi-agent Reinforcement Learning by Planning
por: Liu, Qihan, et al.
Publicado: (2024)
por: Liu, Qihan, et al.
Publicado: (2024)
Agentic Agile-V: From Vibe Coding to Verified Engineering in Software and Hardware Development
por: Koch, Christopher
Publicado: (2026)
por: Koch, Christopher
Publicado: (2026)
Incorporating Human Flexibility through Reward Preferences in Human-AI Teaming
por: Bhambri, Siddhant, et al.
Publicado: (2023)
por: Bhambri, Siddhant, et al.
Publicado: (2023)
Combining Planning and Reinforcement Learning for Solving Relational Multiagent Domains
por: Prabhakar, Nikhilesh, et al.
Publicado: (2025)
por: Prabhakar, Nikhilesh, et al.
Publicado: (2025)
Code Like Humans: A Multi-Agent Solution for Medical Coding
por: Motzfeldt, Andreas, et al.
Publicado: (2025)
por: Motzfeldt, Andreas, et al.
Publicado: (2025)
EvoMem: Improving Multi-Agent Planning with Dual-Evolving Memory
por: Fan, Wenzhe, et al.
Publicado: (2025)
por: Fan, Wenzhe, et al.
Publicado: (2025)
Lessons Learned: A Multi-Agent Framework for Code LLMs to Learn and Improve
por: Liu, Yuanzhe, et al.
Publicado: (2025)
por: Liu, Yuanzhe, et al.
Publicado: (2025)
Learning to Cooperate with Humans using Generative Agents
por: Liang, Yancheng, et al.
Publicado: (2024)
por: Liang, Yancheng, et al.
Publicado: (2024)
SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems
por: Shindo, Hikaru, et al.
Publicado: (2026)
por: Shindo, Hikaru, et al.
Publicado: (2026)
Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning
por: Wu, Wenyi, et al.
Publicado: (2026)
por: Wu, Wenyi, et al.
Publicado: (2026)
Domain-driven Metrics for Reinforcement Learning: A Case Study on Epidemic Control using Agent-based Simulation
por: Gaur, Rishabh, et al.
Publicado: (2025)
por: Gaur, Rishabh, et al.
Publicado: (2025)
Moving Out: Physically-grounded Human-AI Collaboration
por: Kang, Xuhui, et al.
Publicado: (2025)
por: Kang, Xuhui, et al.
Publicado: (2025)
A Hierarchical Deep Reinforcement Learning Framework for Traffic Signal Control with Predictable Cycle Planning
por: Gu, Hankang, et al.
Publicado: (2025)
por: Gu, Hankang, et al.
Publicado: (2025)
CleanAgent: Automating Data Standardization with LLM-based Agents
por: Qi, Danrui, et al.
Publicado: (2024)
por: Qi, Danrui, et al.
Publicado: (2024)
LLM-Guided Communication for Cooperative Multi-Agent Reinforcement Learning
por: Bae, Sangjun, et al.
Publicado: (2026)
por: Bae, Sangjun, et al.
Publicado: (2026)
Healthcare Mechanisms from Policy-as-Code Search under Strategic Provider Response
por: Wang, Zihan, et al.
Publicado: (2026)
por: Wang, Zihan, et al.
Publicado: (2026)
pAI/MSc: ML Theory Research with Humans on the Loop
por: Abdelmoneum, Mahmoud, et al.
Publicado: (2026)
por: Abdelmoneum, Mahmoud, et al.
Publicado: (2026)
Multi-Agent Path Finding via Offline RL and LLM Collaboration
por: Atasever, Merve, et al.
Publicado: (2025)
por: Atasever, Merve, et al.
Publicado: (2025)
Ejemplares similares
-
A Data-Driven Discretized CS:GO Simulation Environment to Facilitate Strategic Multi-Agent Planning Research
por: Wang, Yunzhe, et al.
Publicado: (2025) -
Soft Tournament Equilibrium
por: Alqithami, Saad
Publicado: (2026) -
Improving Health Information Access in the World's Largest Maternal Mobile Health Program via Bandit Algorithms
por: Lalan, Arshika, et al.
Publicado: (2024) -
Deep Reinforcement Learning Agents for Strategic Production Policies in Microeconomic Market Simulations
por: Garrido-Merchán, Eduardo C., et al.
Publicado: (2024) -
Using Analytics on Student Created Data to Content Validate Pedagogical Tools
por: Kos, John, et al.
Publicado: (2023)