AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations
Fuente:
arXiv
Saved in:
| Main Authors: | Verma, Gaurav, Kaur, Rachneet, Srishankar, Nishan, Zeng, Zhen, Balch, Tucker, Veloso, Manuela |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering
by: Kaur, Rachneet, et al.
Published: (2025)
by: Kaur, Rachneet, et al.
Published: (2025)
LAW: Legal Agentic Workflows for Custody and Fund Services Contracts
by: Watson, William, et al.
Published: (2024)
by: Watson, William, et al.
Published: (2024)
LETS-C: Leveraging Text Embedding for Time Series Classification
by: Kaur, Rachneet, et al.
Published: (2024)
by: Kaur, Rachneet, et al.
Published: (2024)
From Pixels to Predictions: Spectrogram and Vision Transformer for Better Time Series Forecasting
by: Zeng, Zhen, et al.
Published: (2024)
by: Zeng, Zhen, et al.
Published: (2024)
Is There No Such Thing as a Bad Question? H4R: HalluciBot For Ratiocination, Rewriting, Ranking, and Routing
by: Watson, William, et al.
Published: (2024)
by: Watson, William, et al.
Published: (2024)
SynthAgent: Adapting Web Agents with Synthetic Supervision
by: Wang, Zhaoyang, et al.
Published: (2025)
by: Wang, Zhaoyang, et al.
Published: (2025)
FISHNET: Financial Intelligence from Sub-querying, Harmonizing, Neural-Conditioning, Expert Swarms, and Task Planning
by: Cho, Nicole, et al.
Published: (2024)
by: Cho, Nicole, et al.
Published: (2024)
Adapting LLM Agents with Universal Feedback in Communication
by: Wang, Kuan, et al.
Published: (2023)
by: Wang, Kuan, et al.
Published: (2023)
SpidR-Adapt: A Universal Speech Representation Model for Few-Shot Adaptation
by: Luthra, Mahi, et al.
Published: (2025)
by: Luthra, Mahi, et al.
Published: (2025)
Evaluating Large Language Models on Time Series Feature Understanding: A Comprehensive Taxonomy and Benchmark
by: Fons, Elizabeth, et al.
Published: (2024)
by: Fons, Elizabeth, et al.
Published: (2024)
FlowMind: Automatic Workflow Generation with LLMs
by: Zeng, Zhen, et al.
Published: (2024)
by: Zeng, Zhen, et al.
Published: (2024)
TADACap: Time-series Adaptive Domain-Aware Captioning
by: Fons, Elizabeth, et al.
Published: (2025)
by: Fons, Elizabeth, et al.
Published: (2025)
HiddenTables & PyQTax: A Cooperative Game and Dataset For TableQA to Ensure Scale and Data Privacy Across a Myriad of Taxonomies
by: Watson, William, et al.
Published: (2024)
by: Watson, William, et al.
Published: (2024)
LLM-based Few-Shot Early Rumor Detection with Imitation Agent
by: Zeng, Fengzhu, et al.
Published: (2025)
by: Zeng, Fengzhu, et al.
Published: (2025)
Avenir-Web: Human-Experience-Imitating Multimodal Web Agents with Mixture of Grounding Experts
by: Li, Aiden Yiliu, et al.
Published: (2026)
by: Li, Aiden Yiliu, et al.
Published: (2026)
BrowserAgent: Building Web Agents with Human-Inspired Web Browsing Actions
by: Yu, Tao, et al.
Published: (2025)
by: Yu, Tao, et al.
Published: (2025)
AdaptEvolve: Improving Efficiency of Evolutionary AI Agents through Adaptive Model Selection
by: Ray, Pretam, et al.
Published: (2026)
by: Ray, Pretam, et al.
Published: (2026)
Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration
by: Chen, Weile, et al.
Published: (2026)
by: Chen, Weile, et al.
Published: (2026)
EnvGen: Generating and Adapting Environments via LLMs for Training Embodied Agents
by: Zala, Abhay, et al.
Published: (2024)
by: Zala, Abhay, et al.
Published: (2024)
MM-WebAgent: A Hierarchical Multimodal Web Agent for Webpage Generation
by: Li, Yan, et al.
Published: (2026)
by: Li, Yan, et al.
Published: (2026)
Instruction Agent: Enhancing Agent with Expert Demonstration
by: Li, Yinheng, et al.
Published: (2025)
by: Li, Yinheng, et al.
Published: (2025)
SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories
by: Yu, Zhuoyun, et al.
Published: (2026)
by: Yu, Zhuoyun, et al.
Published: (2026)
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
by: He, Hongliang, et al.
Published: (2024)
by: He, Hongliang, et al.
Published: (2024)
LLM-based MOFs Synthesis Condition Extraction using Few-Shot Demonstrations
by: Shi, Lei, et al.
Published: (2024)
by: Shi, Lei, et al.
Published: (2024)
WebXSkill: Skill Learning for Autonomous Web Agents
by: Wang, Zhaoyang, et al.
Published: (2026)
by: Wang, Zhaoyang, et al.
Published: (2026)
Grounded Relational Inference: Domain Knowledge Driven Explainable Autonomous Driving
by: Tang, Chen, et al.
Published: (2021)
by: Tang, Chen, et al.
Published: (2021)
PD$^3$: A Project Duplication Detection Framework via Adapted Multi-Agent Debate
by: Bao, Dezheng, et al.
Published: (2025)
by: Bao, Dezheng, et al.
Published: (2025)
SynAdapt: Learning Adaptive Reasoning in Large Language Models via Synthetic Continuous Chain-of-Thought
by: Wang, Jianwei, et al.
Published: (2025)
by: Wang, Jianwei, et al.
Published: (2025)
DynaWeb: Model-Based Reinforcement Learning of Web Agents
by: Ding, Hang, et al.
Published: (2026)
by: Ding, Hang, et al.
Published: (2026)
Demonstrating ViviDoc: Generating Interactive Documents through Human-Agent Collaboration
by: Tang, Yinghao, et al.
Published: (2026)
by: Tang, Yinghao, et al.
Published: (2026)
Do LLMs Really Adapt to Domains? An Ontology Learning Perspective
by: Mai, Huu Tan, et al.
Published: (2024)
by: Mai, Huu Tan, et al.
Published: (2024)
PARM: Pipeline-Adapted Reward Model
by: Fan, Xingyu, et al.
Published: (2026)
by: Fan, Xingyu, et al.
Published: (2026)
PLHF: Prompt Optimization with Few-Shot Human Feedback
by: Yang, Chun-Pai, et al.
Published: (2025)
by: Yang, Chun-Pai, et al.
Published: (2025)
SlideAgent: Hierarchical Agentic Framework for Multi-Page Visual Document Understanding
by: Jin, Yiqiao, et al.
Published: (2025)
by: Jin, Yiqiao, et al.
Published: (2025)
ResAdapt: Adaptive Resolution for Efficient Multimodal Reasoning
by: Liao, Huanxuan, et al.
Published: (2026)
by: Liao, Huanxuan, et al.
Published: (2026)
Tracking the Behavioral Trajectories of Adapting Agents
by: Leshin, Jonah, et al.
Published: (2026)
by: Leshin, Jonah, et al.
Published: (2026)
From Random to Informed Data Selection: A Diversity-Based Approach to Optimize Human Annotation and Few-Shot Learning
by: Alcoforado, Alexandre, et al.
Published: (2024)
by: Alcoforado, Alexandre, et al.
Published: (2024)
TASER: Table Agents for Schema-guided Extraction and Recommendation
by: Cho, Nicole, et al.
Published: (2025)
by: Cho, Nicole, et al.
Published: (2025)
What Makes a Good Query? Measuring the Impact of Human-Confusing Linguistic Features on LLM Performance
by: Watson, William, et al.
Published: (2026)
by: Watson, William, et al.
Published: (2026)
GRAD: Generative Retrieval-Aligned Demonstration Sampler for Efficient Few-Shot Reasoning
by: Gabouj, Oussama, et al.
Published: (2025)
by: Gabouj, Oussama, et al.
Published: (2025)
Similar Items
-
ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering
by: Kaur, Rachneet, et al.
Published: (2025) -
LAW: Legal Agentic Workflows for Custody and Fund Services Contracts
by: Watson, William, et al.
Published: (2024) -
LETS-C: Leveraging Text Embedding for Time Series Classification
by: Kaur, Rachneet, et al.
Published: (2024) -
From Pixels to Predictions: Spectrogram and Vision Transformer for Better Time Series Forecasting
by: Zeng, Zhen, et al.
Published: (2024) -
Is There No Such Thing as a Bad Question? H4R: HalluciBot For Ratiocination, Rewriting, Ranking, and Routing
by: Watson, William, et al.
Published: (2024)