Web Agents with World Models: Learning and Leveraging Environment Dynamics in Web Navigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chae, Hyungjoo, Kim, Namyoung, Ong, Kai Tzu-iunn, Gwak, Minju, Song, Gwanwoo, Kim, Jihoon, Kim, Sunghwan, Lee, Dongha, Yeo, Jinyoung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Lifelong Dialogue Agents via Timeline-based Memory Management
von: Ong, Kai Tzu-iunn, et al.
Veröffentlicht: (2024)
von: Ong, Kai Tzu-iunn, et al.
Veröffentlicht: (2024)
PRINCIPLES: Synthetic Strategy Memory for Proactive Dialogue Agents
von: Kim, Namyoung, et al.
Veröffentlicht: (2025)
von: Kim, Namyoung, et al.
Veröffentlicht: (2025)
Commonsense-augmented Memory Construction and Management in Long-term Conversations via Context-aware Persona Refinement
von: Kim, Hana, et al.
Veröffentlicht: (2024)
von: Kim, Hana, et al.
Veröffentlicht: (2024)
Web-Shepherd: Advancing PRMs for Reinforcing Web Agents
von: Chae, Hyungjoo, et al.
Veröffentlicht: (2025)
von: Chae, Hyungjoo, et al.
Veröffentlicht: (2025)
Rethinking Reward Model Evaluation Through the Lens of Reward Overoptimization
von: Kim, Sunghwan, et al.
Veröffentlicht: (2025)
von: Kim, Sunghwan, et al.
Veröffentlicht: (2025)
ToolHaystack: Stress-Testing Tool-Augmented Language Models in Realistic Long-Term Interactions
von: Kwak, Beong-woo, et al.
Veröffentlicht: (2025)
von: Kwak, Beong-woo, et al.
Veröffentlicht: (2025)
AgenticShop: Benchmarking Agentic Product Curation for Personalized Web Shopping
von: Kim, Sunghwan, et al.
Veröffentlicht: (2026)
von: Kim, Sunghwan, et al.
Veröffentlicht: (2026)
VerifiNER: Verification-augmented NER via Knowledge-grounded Reasoning with Large Language Models
von: Kim, Seoyeon, et al.
Veröffentlicht: (2024)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2024)
Evaluating Robustness of Reward Models for Mathematical Reasoning
von: Kim, Sunghwan, et al.
Veröffentlicht: (2024)
von: Kim, Sunghwan, et al.
Veröffentlicht: (2024)
Large Language Models Are Self-Taught Reasoners: Enhancing LLM Applications via Tailored Problem-Solving Demonstrations
von: Ong, Kai Tzu-iunn, et al.
Veröffentlicht: (2024)
von: Ong, Kai Tzu-iunn, et al.
Veröffentlicht: (2024)
Stop Playing the Guessing Game! Target-free User Simulation for Evaluating Conversational Recommender Systems
von: Kim, Sunghwan, et al.
Veröffentlicht: (2024)
von: Kim, Sunghwan, et al.
Veröffentlicht: (2024)
Language Models as Compilers: Simulating Pseudocode Execution Improves Algorithmic Reasoning in Language Models
von: Chae, Hyungjoo, et al.
Veröffentlicht: (2024)
von: Chae, Hyungjoo, et al.
Veröffentlicht: (2024)
Coffee: Boost Your Code LLMs by Fixing Bugs with Feedback
von: Moon, Seungjun, et al.
Veröffentlicht: (2023)
von: Moon, Seungjun, et al.
Veröffentlicht: (2023)
Evidence-Focused Fact Summarization for Knowledge-Augmented Zero-Shot Question Answering
von: Ko, Sungho, et al.
Veröffentlicht: (2024)
von: Ko, Sungho, et al.
Veröffentlicht: (2024)
Can You Share Your Story? Modeling Clients' Metacognition and Openness for LLM Therapist Evaluation
von: Kim, Minju, et al.
Veröffentlicht: (2025)
von: Kim, Minju, et al.
Veröffentlicht: (2025)
Coffee-Gym: An Environment for Evaluating and Improving Natural Language Feedback on Erroneous Code
von: Chae, Hyungjoo, et al.
Veröffentlicht: (2024)
von: Chae, Hyungjoo, et al.
Veröffentlicht: (2024)
CONDESION-BENCH: Conditional Decision-Making of Large Language Models in Compositional Action Space
von: Hwang, Yeonjun, et al.
Veröffentlicht: (2026)
von: Hwang, Yeonjun, et al.
Veröffentlicht: (2026)
YA-TA: Towards Personalized Question-Answering Teaching Assistants using Instructor-Student Dual Retrieval-augmented Knowledge Fusion
von: Yang, Dongil, et al.
Veröffentlicht: (2024)
von: Yang, Dongil, et al.
Veröffentlicht: (2024)
Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History
von: Kim, Serin, et al.
Veröffentlicht: (2026)
von: Kim, Serin, et al.
Veröffentlicht: (2026)
Safe and Scalable Web Agent Learning via Recreated Websites
von: Chae, Hyungjoo, et al.
Veröffentlicht: (2026)
von: Chae, Hyungjoo, et al.
Veröffentlicht: (2026)
Fast and Fluent Diffusion Language Models via Convolutional Decoding and Rejective Fine-tuning
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
Do LLMs Have Distinct and Consistent Personality? TRAIT: Personality Testset designed for LLMs with Psychometrics
von: Lee, Seungbeen, et al.
Veröffentlicht: (2024)
von: Lee, Seungbeen, et al.
Veröffentlicht: (2024)
Towards Direct Evaluation of Harness Optimizers via Priority Ranking
von: Ong, Kai Tzu-iunn, et al.
Veröffentlicht: (2026)
von: Ong, Kai Tzu-iunn, et al.
Veröffentlicht: (2026)
SAGEO Arena: A Realistic Environment for Evaluating Search-Augmented Generative Engine Optimization
von: Kim, Sunghwan, et al.
Veröffentlicht: (2026)
von: Kim, Sunghwan, et al.
Veröffentlicht: (2026)
Can Code-Switched Texts Activate a Knowledge Switch in LLMs? A Case Study on English-Korean Code-Switching
von: Kim, Seoyeon, et al.
Veröffentlicht: (2024)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2024)
COCOA: CBT-based Conversational Counseling Agent using Memory Specialized in Cognitive Distortions and Dynamic Prompt
von: Lee, Suyeon, et al.
Veröffentlicht: (2024)
von: Lee, Suyeon, et al.
Veröffentlicht: (2024)
Towards Personalized Conversational Sales Agents: Contextual User Profiling for Strategic Action
von: Kim, Tongyoung, et al.
Veröffentlicht: (2025)
von: Kim, Tongyoung, et al.
Veröffentlicht: (2025)
Large Language Models are Clinical Reasoners: Reasoning-Aware Diagnosis Framework with Prompt-Generated Rationales
von: Kwon, Taeyoon, et al.
Veröffentlicht: (2023)
von: Kwon, Taeyoon, et al.
Veröffentlicht: (2023)
Region4Web: Rethinking Observation Space Granularity for Web Agents
von: Kwon, Donguk, et al.
Veröffentlicht: (2026)
von: Kwon, Donguk, et al.
Veröffentlicht: (2026)
LEGO-Eval: Towards Fine-Grained Evaluation on Synthesizing 3D Embodied Environments with Tool Augmentation
von: Hwangbo, Gyeom, et al.
Veröffentlicht: (2025)
von: Hwangbo, Gyeom, et al.
Veröffentlicht: (2025)
Polymerization‐Induced Direct Photolithography of Quantum Dots
von: Taehyung Kim, et al.
Veröffentlicht: (2025)
von: Taehyung Kim, et al.
Veröffentlicht: (2025)
Can Large Language Models be Good Emotional Supporter? Mitigating Preference Bias on Emotional Support Conversation
von: Kang, Dongjin, et al.
Veröffentlicht: (2024)
von: Kang, Dongjin, et al.
Veröffentlicht: (2024)
Cactus: Towards Psychological Counseling Conversations using Cognitive Behavioral Theory
von: Lee, Suyeon, et al.
Veröffentlicht: (2024)
von: Lee, Suyeon, et al.
Veröffentlicht: (2024)
Train-Attention: Meta-Learning Where to Focus in Continual Knowledge Learning
von: Seo, Yeongbin, et al.
Veröffentlicht: (2024)
von: Seo, Yeongbin, et al.
Veröffentlicht: (2024)
Quantifying Genuine Awareness in Hallucination Prediction Beyond Question-Side Shortcuts
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
Unveiling Implicit Table Knowledge with Question-Then-Pinpoint Reasoner for Insightful Table Summarization
von: Seo, Kwangwook, et al.
Veröffentlicht: (2024)
von: Seo, Kwangwook, et al.
Veröffentlicht: (2024)
Pearl: A Review-driven Persona-Knowledge Grounded Conversational Recommendation Dataset
von: Kim, Minjin, et al.
Veröffentlicht: (2024)
von: Kim, Minjin, et al.
Veröffentlicht: (2024)
Revisiting the Uniform Information Density Hypothesis in LLM Reasoning
von: Gwak, Minju, et al.
Veröffentlicht: (2025)
von: Gwak, Minju, et al.
Veröffentlicht: (2025)
Revisiting the UID Hypothesis in LLM Reasoning Traces
von: Gwak, Minju, et al.
Veröffentlicht: (2025)
von: Gwak, Minju, et al.
Veröffentlicht: (2025)
MVIGER: Multi-View Variational Integration of Complementary Knowledge for Generative Recommender
von: Kim, Tongyoung, et al.
Veröffentlicht: (2024)
von: Kim, Tongyoung, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Lifelong Dialogue Agents via Timeline-based Memory Management
von: Ong, Kai Tzu-iunn, et al.
Veröffentlicht: (2024) -
PRINCIPLES: Synthetic Strategy Memory for Proactive Dialogue Agents
von: Kim, Namyoung, et al.
Veröffentlicht: (2025) -
Commonsense-augmented Memory Construction and Management in Long-term Conversations via Context-aware Persona Refinement
von: Kim, Hana, et al.
Veröffentlicht: (2024) -
Web-Shepherd: Advancing PRMs for Reinforcing Web Agents
von: Chae, Hyungjoo, et al.
Veröffentlicht: (2025) -
Rethinking Reward Model Evaluation Through the Lens of Reward Overoptimization
von: Kim, Sunghwan, et al.
Veröffentlicht: (2025)