Greedy Is a Strong Default: Agents as Iterative Optimizers
Fuente:
arXiv
Guardado en:
| Autor principal: | Li, Yitao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Lightweight Baselines for Medical Abstract Classification: DistilBERT with Cross-Entropy as a Strong Default
por: Liu, Jiaqi, et al.
Publicado: (2025)
por: Liu, Jiaqi, et al.
Publicado: (2025)
MACPO: Weak-to-Strong Alignment via Multi-Agent Contrastive Preference Optimization
por: Lyu, Yougang, et al.
Publicado: (2024)
por: Lyu, Yougang, et al.
Publicado: (2024)
Greedy-DiM: Greedy Algorithms for Unreasonably Effective Face Morphs
por: Blasingame, Zander W., et al.
Publicado: (2024)
por: Blasingame, Zander W., et al.
Publicado: (2024)
Iterative Reasoning Preference Optimization
por: Pang, Richard Yuanzhe, et al.
Publicado: (2024)
por: Pang, Richard Yuanzhe, et al.
Publicado: (2024)
Generics and Default Reasoning in Large Language Models
por: Kirkpatrick, James Ravi, et al.
Publicado: (2025)
por: Kirkpatrick, James Ravi, et al.
Publicado: (2025)
ABD: Default Exception Abduction in Finite First Order Worlds
por: Batzoglou, Serafim
Publicado: (2026)
por: Batzoglou, Serafim
Publicado: (2026)
Strongly Polynomial Time Complexity of Policy Iteration for $L_\infty$ Robust MDPs
por: Asadi, Ali, et al.
Publicado: (2026)
por: Asadi, Ali, et al.
Publicado: (2026)
The Good, The Bad, and The Greedy: Evaluation of LLMs Should Not Ignore Non-Determinism
por: Song, Yifan, et al.
Publicado: (2024)
por: Song, Yifan, et al.
Publicado: (2024)
OpenWebVoyager: Building Multimodal Web Agents via Iterative Real-World Exploration, Feedback and Optimization
por: He, Hongliang, et al.
Publicado: (2024)
por: He, Hongliang, et al.
Publicado: (2024)
Refined Iterated Pareto Greedy for Energy-aware Hybrid Flowshop Scheduling with Blocking Constraints
por: Missaoui, Ahmed, et al.
Publicado: (2025)
por: Missaoui, Ahmed, et al.
Publicado: (2025)
Evolutionary Rule Extraction from Corporate Default Prediction Models
por: Fabbretti, Desirè, et al.
Publicado: (2026)
por: Fabbretti, Desirè, et al.
Publicado: (2026)
A Training-Free Length Extrapolation Approach for LLMs: Greedy Attention Logit Interpolation (GALI)
por: Li, Yan, et al.
Publicado: (2025)
por: Li, Yan, et al.
Publicado: (2025)
EPOCH: An Agentic Protocol for Multi-Round System Optimization
por: Liu, Zhanlin, et al.
Publicado: (2026)
por: Liu, Zhanlin, et al.
Publicado: (2026)
Explainable Benchmarking for Iterative Optimization Heuristics
por: van Stein, Niki, et al.
Publicado: (2024)
por: van Stein, Niki, et al.
Publicado: (2024)
Contextual Experience Replay for Self-Improvement of Language Agents
por: Liu, Yitao, et al.
Publicado: (2025)
por: Liu, Yitao, et al.
Publicado: (2025)
More Than a Quick Glance: Overcoming the Greedy Bias in KV-Cache Compression
por: Sood, Aryan, et al.
Publicado: (2026)
por: Sood, Aryan, et al.
Publicado: (2026)
CATArena: Evaluating Evolutionary Capabilities of Code Agents via Iterative Tournaments
por: Fu, Lingyue, et al.
Publicado: (2025)
por: Fu, Lingyue, et al.
Publicado: (2025)
AgentOccam: A Simple Yet Strong Baseline for LLM-Based Web Agents
por: Yang, Ke, et al.
Publicado: (2024)
por: Yang, Ke, et al.
Publicado: (2024)
Explainable Iterative Data Visualisation Refinement via an LLM Agent
por: Susam, Burak, et al.
Publicado: (2026)
por: Susam, Burak, et al.
Publicado: (2026)
FLAIRR-TS -- Forecasting LLM-Agents with Iterative Refinement and Retrieval for Time Series
por: Jalori, Gunjan, et al.
Publicado: (2025)
por: Jalori, Gunjan, et al.
Publicado: (2025)
Adaptive Greedy Frame Selection for Long Video Understanding
por: Huang, Yuning, et al.
Publicado: (2026)
por: Huang, Yuning, et al.
Publicado: (2026)
CDQuant: Greedy Coordinate Descent for Accurate LLM Quantization
por: Nair, Pranav Ajit, et al.
Publicado: (2024)
por: Nair, Pranav Ajit, et al.
Publicado: (2024)
Aligning Frozen LLMs by Reinforcement Learning: An Iterative Reweight-then-Optimize Approach
por: Zhang, Xinnan, et al.
Publicado: (2025)
por: Zhang, Xinnan, et al.
Publicado: (2025)
MobileIPL: Enhancing Mobile Agents Thinking Process via Iterative Preference Learning
por: Huang, Kun, et al.
Publicado: (2025)
por: Huang, Kun, et al.
Publicado: (2025)
Iterative Experience Refinement of Software-Developing Agents
por: Qian, Chen, et al.
Publicado: (2024)
por: Qian, Chen, et al.
Publicado: (2024)
Dual-IPO: Dual-Iterative Preference Optimization for Text-to-Video Generation
por: Yang, Xiaomeng, et al.
Publicado: (2025)
por: Yang, Xiaomeng, et al.
Publicado: (2025)
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
por: Usman, Rana Muhammad
Publicado: (2026)
por: Usman, Rana Muhammad
Publicado: (2026)
Teola: Towards End-to-End Optimization of LLM-based Applications
por: Tan, Xin, et al.
Publicado: (2024)
por: Tan, Xin, et al.
Publicado: (2024)
Dynamic Sampling that Adapts: Self-Aware Iterative Data Persistent Optimization for Mathematical Reasoning
por: Rao, Jun, et al.
Publicado: (2025)
por: Rao, Jun, et al.
Publicado: (2025)
MCU: An Evaluation Framework for Open-Ended Game Agents
por: Zheng, Xinyue, et al.
Publicado: (2023)
por: Zheng, Xinyue, et al.
Publicado: (2023)
Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
por: Xiong, Weimin, et al.
Publicado: (2024)
por: Xiong, Weimin, et al.
Publicado: (2024)
ICR: Iterative Clarification and Rewriting for Conversational Search
por: Cao, Zhiyu, et al.
Publicado: (2025)
por: Cao, Zhiyu, et al.
Publicado: (2025)
Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss
por: Xu, Jing, et al.
Publicado: (2023)
por: Xu, Jing, et al.
Publicado: (2023)
Selective Weak-to-Strong Generalization
por: Lang, Hao, et al.
Publicado: (2025)
por: Lang, Hao, et al.
Publicado: (2025)
ALSO: Adversarial Online Strategy Optimization for Social Agents
por: Li, Xiang, et al.
Publicado: (2026)
por: Li, Xiang, et al.
Publicado: (2026)
Iterative Formalization and Planning in Partially Observable Environments
por: Gong, Liancheng, et al.
Publicado: (2025)
por: Gong, Liancheng, et al.
Publicado: (2025)
RAT: Retrieval Augmented Thoughts Elicit Context-Aware Reasoning in Long-Horizon Generation
por: Wang, Zihao, et al.
Publicado: (2024)
por: Wang, Zihao, et al.
Publicado: (2024)
SAC-Opt: Semantic Anchors for Iterative Correction in Optimization Modeling
por: Zhang, Yansen, et al.
Publicado: (2025)
por: Zhang, Yansen, et al.
Publicado: (2025)
OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces
por: Li, Xiaozhe, et al.
Publicado: (2026)
por: Li, Xiaozhe, et al.
Publicado: (2026)
Iterative Pretraining Framework for Interatomic Potentials
por: Cui, Taoyong, et al.
Publicado: (2025)
por: Cui, Taoyong, et al.
Publicado: (2025)
Ejemplares similares
-
Lightweight Baselines for Medical Abstract Classification: DistilBERT with Cross-Entropy as a Strong Default
por: Liu, Jiaqi, et al.
Publicado: (2025) -
MACPO: Weak-to-Strong Alignment via Multi-Agent Contrastive Preference Optimization
por: Lyu, Yougang, et al.
Publicado: (2024) -
Greedy-DiM: Greedy Algorithms for Unreasonably Effective Face Morphs
por: Blasingame, Zander W., et al.
Publicado: (2024) -
Iterative Reasoning Preference Optimization
por: Pang, Richard Yuanzhe, et al.
Publicado: (2024) -
Generics and Default Reasoning in Large Language Models
por: Kirkpatrick, James Ravi, et al.
Publicado: (2025)