Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Serin, Lee, Sangam, Lee, Dongha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Region4Web: Rethinking Observation Space Granularity for Web Agents
von: Kwon, Donguk, et al.
Veröffentlicht: (2026)
von: Kwon, Donguk, et al.
Veröffentlicht: (2026)
MoCoRP: Modeling Consistent Relations between Persona and Response for Persona-based Dialogue
von: Lee, Kyungro, et al.
Veröffentlicht: (2025)
von: Lee, Kyungro, et al.
Veröffentlicht: (2025)
OPSD Compresses What RLVR Teaches: A Post-RL Compaction Stage for Reasoning Models
von: Kim, Jaehoon, et al.
Veröffentlicht: (2026)
von: Kim, Jaehoon, et al.
Veröffentlicht: (2026)
SAGEO Arena: A Realistic Environment for Evaluating Search-Augmented Generative Engine Optimization
von: Kim, Sunghwan, et al.
Veröffentlicht: (2026)
von: Kim, Sunghwan, et al.
Veröffentlicht: (2026)
Beyond the Final Answer: Evaluating the Reasoning Trajectories of Tool-Augmented Agents
von: Kim, Wonjoong, et al.
Veröffentlicht: (2025)
von: Kim, Wonjoong, et al.
Veröffentlicht: (2025)
BESPOKE: Benchmark for Search-Augmented Large Language Model Personalization via Diagnostic Feedback
von: Kim, Hyunseo, et al.
Veröffentlicht: (2025)
von: Kim, Hyunseo, et al.
Veröffentlicht: (2025)
WebSailor: Navigating Super-human Reasoning for Web Agent
von: Li, Kuan, et al.
Veröffentlicht: (2025)
von: Li, Kuan, et al.
Veröffentlicht: (2025)
Web-CogReasoner: Towards Knowledge-Induced Cognitive Reasoning for Web Agents
von: Guo, Yuhan, et al.
Veröffentlicht: (2025)
von: Guo, Yuhan, et al.
Veröffentlicht: (2025)
Towards Personalized Conversational Sales Agents: Contextual User Profiling for Strategic Action
von: Kim, Tongyoung, et al.
Veröffentlicht: (2025)
von: Kim, Tongyoung, et al.
Veröffentlicht: (2025)
Commonsense-augmented Memory Construction and Management in Long-term Conversations via Context-aware Persona Refinement
von: Kim, Hana, et al.
Veröffentlicht: (2024)
von: Kim, Hana, et al.
Veröffentlicht: (2024)
Personalizing Large Language Models using Retrieval Augmented Generation and Knowledge Graph
von: Prahlad, Deeksha, et al.
Veröffentlicht: (2025)
von: Prahlad, Deeksha, et al.
Veröffentlicht: (2025)
WebWalker: Benchmarking LLMs in Web Traversal
von: Wu, Jialong, et al.
Veröffentlicht: (2025)
von: Wu, Jialong, et al.
Veröffentlicht: (2025)
Persona Dynamics: Unveiling the Impact of Personality Traits on Agents in Text-Based Games
von: Lim, Seungwon, et al.
Veröffentlicht: (2025)
von: Lim, Seungwon, et al.
Veröffentlicht: (2025)
Ego2Web: A Web Agent Benchmark Grounded in Egocentric Videos
von: Yu, Shoubin, et al.
Veröffentlicht: (2026)
von: Yu, Shoubin, et al.
Veröffentlicht: (2026)
AgenticShop: Benchmarking Agentic Product Curation for Personalized Web Shopping
von: Kim, Sunghwan, et al.
Veröffentlicht: (2026)
von: Kim, Sunghwan, et al.
Veröffentlicht: (2026)
Beyond Ontology in Dialogue State Tracking for Goal-Oriented Chatbot
von: Lee, Sejin, et al.
Veröffentlicht: (2024)
von: Lee, Sejin, et al.
Veröffentlicht: (2024)
Agentic Test-Time Scaling for WebAgents
von: Lee, Nicholas, et al.
Veröffentlicht: (2026)
von: Lee, Nicholas, et al.
Veröffentlicht: (2026)
WebXSkill: Skill Learning for Autonomous Web Agents
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2026)
COCOA: CBT-based Conversational Counseling Agent using Memory Specialized in Cognitive Distortions and Dynamic Prompt
von: Lee, Suyeon, et al.
Veröffentlicht: (2024)
von: Lee, Suyeon, et al.
Veröffentlicht: (2024)
BrowserAgent: Building Web Agents with Human-Inspired Web Browsing Actions
von: Yu, Tao, et al.
Veröffentlicht: (2025)
von: Yu, Tao, et al.
Veröffentlicht: (2025)
Large Language Models Empowered Personalized Web Agents
von: Cai, Hongru, et al.
Veröffentlicht: (2024)
von: Cai, Hongru, et al.
Veröffentlicht: (2024)
DynaWeb: Model-Based Reinforcement Learning of Web Agents
von: Ding, Hang, et al.
Veröffentlicht: (2026)
von: Ding, Hang, et al.
Veröffentlicht: (2026)
Needle in the Web: A Benchmark for Retrieving Targeted Web Pages in the Wild
von: Wang, Yumeng, et al.
Veröffentlicht: (2025)
von: Wang, Yumeng, et al.
Veröffentlicht: (2025)
WebDS: An End-to-End Benchmark for Web-based Data Science
von: Hsu, Ethan, et al.
Veröffentlicht: (2025)
von: Hsu, Ethan, et al.
Veröffentlicht: (2025)
WebRollback: Enhancing Web Agents with Explicit Rollback Mechanisms
von: Zhang, Zhisong, et al.
Veröffentlicht: (2025)
von: Zhang, Zhisong, et al.
Veröffentlicht: (2025)
MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments
von: Wang, Han, et al.
Veröffentlicht: (2026)
von: Wang, Han, et al.
Veröffentlicht: (2026)
Auto-Intent: Automated Intent Discovery and Self-Exploration for Large Language Model Web Agents
von: Kim, Jaekyeom, et al.
Veröffentlicht: (2024)
von: Kim, Jaekyeom, et al.
Veröffentlicht: (2024)
Do LLMs Have Distinct and Consistent Personality? TRAIT: Personality Testset designed for LLMs with Psychometrics
von: Lee, Seungbeen, et al.
Veröffentlicht: (2024)
von: Lee, Seungbeen, et al.
Veröffentlicht: (2024)
Can Agent Conquer Web? Exploring the Frontiers of ChatGPT Atlas Agent in Web Games
von: Zhang, Jingran, et al.
Veröffentlicht: (2025)
von: Zhang, Jingran, et al.
Veröffentlicht: (2025)
A Functionality-Grounded Benchmark for Evaluating Web Agents in E-commerce Domains
von: Zhang, Xianren, et al.
Veröffentlicht: (2025)
von: Zhang, Xianren, et al.
Veröffentlicht: (2025)
Infogent: An Agent-Based Framework for Web Information Aggregation
von: Reddy, Revanth Gangi, et al.
Veröffentlicht: (2024)
von: Reddy, Revanth Gangi, et al.
Veröffentlicht: (2024)
WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance
von: Liu, Genglin, et al.
Veröffentlicht: (2025)
von: Liu, Genglin, et al.
Veröffentlicht: (2025)
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
von: He, Hongliang, et al.
Veröffentlicht: (2024)
von: He, Hongliang, et al.
Veröffentlicht: (2024)
AutoScraper: A Progressive Understanding Web Agent for Web Scraper Generation
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
CUPID: Evaluating Personalized and Contextualized Alignment of LLMs from Interactions
von: Kim, Tae Soo, et al.
Veröffentlicht: (2025)
von: Kim, Tae Soo, et al.
Veröffentlicht: (2025)
Lost in the Noise: How Reasoning Models Fail with Contextual Distractors
von: Lee, Seongyun, et al.
Veröffentlicht: (2026)
von: Lee, Seongyun, et al.
Veröffentlicht: (2026)
RealWebAssist: A Benchmark for Long-Horizon Web Assistance with Real-World Users
von: Ye, Suyu, et al.
Veröffentlicht: (2025)
von: Ye, Suyu, et al.
Veröffentlicht: (2025)
WebCanvas: Benchmarking Web Agents in Online Environments
von: Pan, Yichen, et al.
Veröffentlicht: (2024)
von: Pan, Yichen, et al.
Veröffentlicht: (2024)
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
Avenir-Web: Human-Experience-Imitating Multimodal Web Agents with Mixture of Grounding Experts
von: Li, Aiden Yiliu, et al.
Veröffentlicht: (2026)
von: Li, Aiden Yiliu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Region4Web: Rethinking Observation Space Granularity for Web Agents
von: Kwon, Donguk, et al.
Veröffentlicht: (2026) -
MoCoRP: Modeling Consistent Relations between Persona and Response for Persona-based Dialogue
von: Lee, Kyungro, et al.
Veröffentlicht: (2025) -
OPSD Compresses What RLVR Teaches: A Post-RL Compaction Stage for Reasoning Models
von: Kim, Jaehoon, et al.
Veröffentlicht: (2026) -
SAGEO Arena: A Realistic Environment for Evaluating Search-Augmented Generative Engine Optimization
von: Kim, Sunghwan, et al.
Veröffentlicht: (2026) -
Beyond the Final Answer: Evaluating the Reasoning Trajectories of Tool-Augmented Agents
von: Kim, Wonjoong, et al.
Veröffentlicht: (2025)