Navigating WebAI: Training Agents to Complete Web Tasks with Large Language Models and Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Thil, Lucas-Andreï, Popa, Mirela, Spanakis, Gerasimos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TRIZ Agents: A Multi-Agent LLM Approach for TRIZ-Based Innovation
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
Collaborative LLM Agents for C4 Software Architecture Design Automation
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
GraphWalk: Enabling Reasoning in Large Language Models through Tool-Based Graph Navigation
von: Ghandi, Taraneh, et al.
Veröffentlicht: (2026)
von: Ghandi, Taraneh, et al.
Veröffentlicht: (2026)
Open-TI: Open Traffic Intelligence with Augmented Language Model
von: Da, Longchao, et al.
Veröffentlicht: (2023)
von: Da, Longchao, et al.
Veröffentlicht: (2023)
FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation
von: Hildebrand, Samuel, et al.
Veröffentlicht: (2025)
von: Hildebrand, Samuel, et al.
Veröffentlicht: (2025)
Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning
von: Pather, Kaviraj, et al.
Veröffentlicht: (2025)
von: Pather, Kaviraj, et al.
Veröffentlicht: (2025)
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
Cooperative Patrol Routing: Optimizing Urban Crime Surveillance through Multi-Agent Reinforcement Learning
von: Palma-Borda, Juan, et al.
Veröffentlicht: (2025)
von: Palma-Borda, Juan, et al.
Veröffentlicht: (2025)
BudgetMLAgent: A Cost-Effective LLM Multi-Agent system for Automating Machine Learning Tasks
von: Gandhi, Shubham, et al.
Veröffentlicht: (2024)
von: Gandhi, Shubham, et al.
Veröffentlicht: (2024)
Streaming Continual Learning for Unified Adaptive Intelligence in Dynamic Environments
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
From Search to Reasoning: A Five-Level RAG Capability Framework for Enterprise Data
von: Gill, Gurbinder, et al.
Veröffentlicht: (2025)
von: Gill, Gurbinder, et al.
Veröffentlicht: (2025)
XAutoLM: Efficient Fine-Tuning of Language Models via Meta-Learning and AutoML
von: Estevanell-Valladares, Ernesto L., et al.
Veröffentlicht: (2025)
von: Estevanell-Valladares, Ernesto L., et al.
Veröffentlicht: (2025)
AI Agents-as-Judge: Automated Assessment of Accuracy, Consistency, Completeness and Clarity for Enterprise Documents
von: Dasgupta, Sudip, et al.
Veröffentlicht: (2025)
von: Dasgupta, Sudip, et al.
Veröffentlicht: (2025)
Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
Can AI Assist in Olympiad Coding
von: Ren, Samuel
Veröffentlicht: (2025)
von: Ren, Samuel
Veröffentlicht: (2025)
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
von: Kaiser, Daniel, et al.
Veröffentlicht: (2025)
von: Kaiser, Daniel, et al.
Veröffentlicht: (2025)
Improving LLM Agent Planning with In-Context Learning via Atomic Fact Augmentation and Lookahead Search
von: Holt, Samuel, et al.
Veröffentlicht: (2025)
von: Holt, Samuel, et al.
Veröffentlicht: (2025)
LLMs taking shortcuts in test generation: A study with SAP HANA and LevelDB
von: Bekmyradov, Vekil, et al.
Veröffentlicht: (2026)
von: Bekmyradov, Vekil, et al.
Veröffentlicht: (2026)
Recent Advances in Data-Driven Business Process Management
von: Ackermann, Lars, et al.
Veröffentlicht: (2024)
von: Ackermann, Lars, et al.
Veröffentlicht: (2024)
AgentDS Technical Report: Benchmarking the Future of Human-AI Collaboration in Domain-Specific Data Science
von: Luo, An, et al.
Veröffentlicht: (2026)
von: Luo, An, et al.
Veröffentlicht: (2026)
Intersymbolic AI: Interlinking Symbolic AI and Subsymbolic AI
von: Platzer, André
Veröffentlicht: (2024)
von: Platzer, André
Veröffentlicht: (2024)
Pattern Recognition Tasks with Personalized Federated Learning
von: Rahman, Md. Arifur, et al.
Veröffentlicht: (2026)
von: Rahman, Md. Arifur, et al.
Veröffentlicht: (2026)
TrafficRAG: A Multimodal RAG Framework for Traffic Accident Liability Determination
von: Li, Xu, et al.
Veröffentlicht: (2026)
von: Li, Xu, et al.
Veröffentlicht: (2026)
ChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level Generation
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
Learning Natural Language Constraints for Safe Reinforcement Learning of Language Agents
von: Chua, Jaymari, et al.
Veröffentlicht: (2025)
von: Chua, Jaymari, et al.
Veröffentlicht: (2025)
The Curious Case of In-Training Compression of State Space Models
von: Chahine, Makram, et al.
Veröffentlicht: (2025)
von: Chahine, Makram, et al.
Veröffentlicht: (2025)
Approaches to Semantic Textual Similarity in Slovak Language: From Algorithms to Transformers
von: Radosky, Lukas, et al.
Veröffentlicht: (2026)
von: Radosky, Lukas, et al.
Veröffentlicht: (2026)
Can Agentic AI Match the Performance of Human Data Scientists?
von: Luo, An, et al.
Veröffentlicht: (2025)
von: Luo, An, et al.
Veröffentlicht: (2025)
AI Model for Predicting Binding Affinity of Antidiabetic Compounds Targeting PPAR
von: Aman, La Ode, et al.
Veröffentlicht: (2024)
von: Aman, La Ode, et al.
Veröffentlicht: (2024)
Critical Insights into Leading Conversational AI Models
von: Kohli, Urja, et al.
Veröffentlicht: (2025)
von: Kohli, Urja, et al.
Veröffentlicht: (2025)
Murphys Laws of AI Alignment: Why the Gap Always Wins
von: Gaikwad, Madhava
Veröffentlicht: (2025)
von: Gaikwad, Madhava
Veröffentlicht: (2025)
NOTAI.AI: Explainable Detection of Machine-Generated Text via Curvature and Feature Attribution
von: Breneur, Oleksandr Marchenko, et al.
Veröffentlicht: (2026)
von: Breneur, Oleksandr Marchenko, et al.
Veröffentlicht: (2026)
CAPE: Corrective Actions from Precondition Errors using Large Language Models
von: Raman, Shreyas Sundara, et al.
Veröffentlicht: (2022)
von: Raman, Shreyas Sundara, et al.
Veröffentlicht: (2022)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
Constitution or Collapse? Exploring Constitutional AI with Llama 3-8B
von: Zhang, Xue
Veröffentlicht: (2025)
von: Zhang, Xue
Veröffentlicht: (2025)
Federated Distributional Reinforcement Learning with Distributional Critic Regularization
von: Millard, David, et al.
Veröffentlicht: (2026)
von: Millard, David, et al.
Veröffentlicht: (2026)
Pioneer Agent: Continual Improvement of Small Language Models in Production
von: Atreja, Dhruv, et al.
Veröffentlicht: (2026)
von: Atreja, Dhruv, et al.
Veröffentlicht: (2026)
An Improved Adaptive PID Optimizer with Enhanced Convergence and Stability for Deep Learning
von: Saini, Saurabh, et al.
Veröffentlicht: (2026)
von: Saini, Saurabh, et al.
Veröffentlicht: (2026)
NeuronSpark: A Spiking Neural Network Language Model with Selective State Space Dynamics
von: Tang, Zhengzheng
Veröffentlicht: (2026)
von: Tang, Zhengzheng
Veröffentlicht: (2026)
Emotion-Inspired Learning Signals (EILS): A Homeostatic Framework for Adaptive Autonomous Agents
von: Tiwari, Dhruv
Veröffentlicht: (2025)
von: Tiwari, Dhruv
Veröffentlicht: (2025)
Ähnliche Einträge
-
TRIZ Agents: A Multi-Agent LLM Approach for TRIZ-Based Innovation
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025) -
Collaborative LLM Agents for C4 Software Architecture Design Automation
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025) -
GraphWalk: Enabling Reasoning in Large Language Models through Tool-Based Graph Navigation
von: Ghandi, Taraneh, et al.
Veröffentlicht: (2026) -
Open-TI: Open Traffic Intelligence with Augmented Language Model
von: Da, Longchao, et al.
Veröffentlicht: (2023) -
FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation
von: Hildebrand, Samuel, et al.
Veröffentlicht: (2025)