Credit-Budgeted ICPC-Style Coding: When Agents Must Pay for Every Decision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Lingfeng, Shi, Junhao, Gao, Jin, Wang, Dequan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FlockVote: LLM-Empowered Agent-Based Modeling for Simulating U.S. Presidential Elections
von: Zhou, Lingfeng, et al.
Veröffentlicht: (2025)
von: Zhou, Lingfeng, et al.
Veröffentlicht: (2025)
Aligned Agents, Biased Swarm: Measuring Bias Amplification in Multi-Agent Systems
von: Li, Keyu, et al.
Veröffentlicht: (2026)
von: Li, Keyu, et al.
Veröffentlicht: (2026)
Style-Preserving Policy Optimization for Game Agents
von: Li, Lingfeng, et al.
Veröffentlicht: (2025)
von: Li, Lingfeng, et al.
Veröffentlicht: (2025)
Intelligent Computing Social Modeling and Methodological Innovations in Political Science in the Era of Large Language Models
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
ICPC: In-context Prompt Compression with Faster Inference
von: Yu, Ziyang, et al.
Veröffentlicht: (2025)
von: Yu, Ziyang, et al.
Veröffentlicht: (2025)
Data-Centric Foundation Models in Computational Healthcare: A Survey
von: Zhang, Yunkun, et al.
Veröffentlicht: (2024)
von: Zhang, Yunkun, et al.
Veröffentlicht: (2024)
Trustworthy Agent Network: Trust in Agent Networks Must Be Baked In, Not Bolted On
von: Yao, Yixiang, et al.
Veröffentlicht: (2026)
von: Yao, Yixiang, et al.
Veröffentlicht: (2026)
LLMs Should Not Yet Be Credited with Decision Explanation
von: Wang, Wenshuo
Veröffentlicht: (2026)
von: Wang, Wenshuo
Veröffentlicht: (2026)
MAC: A Live Benchmark for Multimodal Large Language Models in Scientific Understanding
von: Jiang, Mohan, et al.
Veröffentlicht: (2025)
von: Jiang, Mohan, et al.
Veröffentlicht: (2025)
Dissecting Dissonance: Benchmarking Large Multimodal Models Against Self-Contradictory Instructions
von: Gao, Jin, et al.
Veröffentlicht: (2024)
von: Gao, Jin, et al.
Veröffentlicht: (2024)
Position: AI Safety Must Embrace an Antifragile Perspective
von: Jin, Ming, et al.
Veröffentlicht: (2025)
von: Jin, Ming, et al.
Veröffentlicht: (2025)
What Capable Agents Must Know: Selection Theorems for Robust Decision-Making under Uncertainty
von: Nayebi, Aran
Veröffentlicht: (2026)
von: Nayebi, Aran
Veröffentlicht: (2026)
Search-Based Credit Assignment for Offline Preference-Based Reinforcement Learning
von: Gao, Xiancheng, et al.
Veröffentlicht: (2025)
von: Gao, Xiancheng, et al.
Veröffentlicht: (2025)
Some[Body] Must Receive That Pain for Agent Accountability
von: Hu, Botao Amber, et al.
Veröffentlicht: (2026)
von: Hu, Botao Amber, et al.
Veröffentlicht: (2026)
Detection of LLM-Paraphrased Code and Identification of the Responsible LLM Using Coding Style Features
von: Park, Shinwoo, et al.
Veröffentlicht: (2025)
von: Park, Shinwoo, et al.
Veröffentlicht: (2025)
QLLM: Do We Really Need a Mixing Network for Credit Assignment in Multi-Agent Reinforcement Learning?
von: Li, Yuanjun, et al.
Veröffentlicht: (2025)
von: Li, Yuanjun, et al.
Veröffentlicht: (2025)
Budget-Aware Tool-Use Enables Effective Agent Scaling
von: Liu, Tengxiao, et al.
Veröffentlicht: (2025)
von: Liu, Tengxiao, et al.
Veröffentlicht: (2025)
AgencyBench: Benchmarking the Frontiers of Autonomous Agents in 1M-Token Real-World Contexts
von: Li, Keyu, et al.
Veröffentlicht: (2026)
von: Li, Keyu, et al.
Veröffentlicht: (2026)
Parallax: Why AI Agents That Think Must Never Act
von: Fokou, Joel
Veröffentlicht: (2026)
von: Fokou, Joel
Veröffentlicht: (2026)
Every Software as an Agent: Blueprint and Case Study
von: Xu, Mengwei
Veröffentlicht: (2025)
von: Xu, Mengwei
Veröffentlicht: (2025)
BAGEN: Are LLM Agents Budget-Aware?
von: Lin, Yuxiang, et al.
Veröffentlicht: (2026)
von: Lin, Yuxiang, et al.
Veröffentlicht: (2026)
ContextBudget: Budget-Aware Context Management for Long-Horizon Search Agents
von: Wu, Yong, et al.
Veröffentlicht: (2026)
von: Wu, Yong, et al.
Veröffentlicht: (2026)
Reasoning-Style Poisoning of LLM Agents via Stealthy Style Transfer: Process-Level Attacks and Runtime Monitoring in RSV Space
von: Zhou, Xingfu, et al.
Veröffentlicht: (2025)
von: Zhou, Xingfu, et al.
Veröffentlicht: (2025)
Simulating Multi-Stakeholder Decision-Making with Generative Agents in Urban Planning
von: Gao, Jin, et al.
Veröffentlicht: (2024)
von: Gao, Jin, et al.
Veröffentlicht: (2024)
AI Must not be Fully Autonomous
von: Adewumi, Tosin, et al.
Veröffentlicht: (2025)
von: Adewumi, Tosin, et al.
Veröffentlicht: (2025)
Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts
von: Zheng, Haizhong, et al.
Veröffentlicht: (2025)
von: Zheng, Haizhong, et al.
Veröffentlicht: (2025)
Proximity-Based Multi-Turn Optimization: Practical Credit Assignment for LLM Agent Training
von: Fang, Yangyi, et al.
Veröffentlicht: (2026)
von: Fang, Yangyi, et al.
Veröffentlicht: (2026)
High Noise Scheduling is a Must
von: Gokmen, Mahmut S., et al.
Veröffentlicht: (2024)
von: Gokmen, Mahmut S., et al.
Veröffentlicht: (2024)
When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search
von: Robertson, John T., et al.
Veröffentlicht: (2026)
von: Robertson, John T., et al.
Veröffentlicht: (2026)
Counterfactual Credit Policy Optimization for Multi-Agent Collaboration
von: Li, Zhongyi, et al.
Veröffentlicht: (2026)
von: Li, Zhongyi, et al.
Veröffentlicht: (2026)
Exact Is Easier: Credit Assignment for Cooperative LLM Agents
von: Chen, Yanjun, et al.
Veröffentlicht: (2026)
von: Chen, Yanjun, et al.
Veröffentlicht: (2026)
Phase Transition for Budgeted Multi-Agent Synergy
von: Liu, Bang, et al.
Veröffentlicht: (2026)
von: Liu, Bang, et al.
Veröffentlicht: (2026)
Agentic AI for Autonomous, Explainable, and Real-Time Credit Risk Decision-Making
von: Kubam, Chandra Sekhar
Veröffentlicht: (2025)
von: Kubam, Chandra Sekhar
Veröffentlicht: (2025)
Knowledge Distillation Must Account for What It Loses
von: Wang, Wenshuo
Veröffentlicht: (2026)
von: Wang, Wenshuo
Veröffentlicht: (2026)
Budget-Efficient Automatic Algorithm Design via Code Graph
von: Bouscary, Maxime, et al.
Veröffentlicht: (2026)
von: Bouscary, Maxime, et al.
Veröffentlicht: (2026)
When Agents Overtrust Environmental Evidence: An Extensible Agentic Framework for Benchmarking Evidence-Grounding Defects in LLM Agents
von: Sheng, Strick, et al.
Veröffentlicht: (2026)
von: Sheng, Strick, et al.
Veröffentlicht: (2026)
Quality-Aware Exploration Budget Allocation for Cooperative Multi-Agent Reinforcement Learning
von: Oh, Dahyun, et al.
Veröffentlicht: (2026)
von: Oh, Dahyun, et al.
Veröffentlicht: (2026)
Function-to-Style Guidance of LLMs for Code Translation
von: Zhang, Longhui, et al.
Veröffentlicht: (2025)
von: Zhang, Longhui, et al.
Veröffentlicht: (2025)
Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
von: Li, Lingfeng, et al.
Veröffentlicht: (2025)
von: Li, Lingfeng, et al.
Veröffentlicht: (2025)
ARCA: Adapter-Residual Credit Assignment When Token Signals Degenerate
von: Lafuente-Mercado, Rodney
Veröffentlicht: (2026)
von: Lafuente-Mercado, Rodney
Veröffentlicht: (2026)
Ähnliche Einträge
-
FlockVote: LLM-Empowered Agent-Based Modeling for Simulating U.S. Presidential Elections
von: Zhou, Lingfeng, et al.
Veröffentlicht: (2025) -
Aligned Agents, Biased Swarm: Measuring Bias Amplification in Multi-Agent Systems
von: Li, Keyu, et al.
Veröffentlicht: (2026) -
Style-Preserving Policy Optimization for Game Agents
von: Li, Lingfeng, et al.
Veröffentlicht: (2025) -
Intelligent Computing Social Modeling and Methodological Innovations in Political Science in the Era of Large Language Models
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024) -
ICPC: In-context Prompt Compression with Faster Inference
von: Yu, Ziyang, et al.
Veröffentlicht: (2025)