Weasel: Out-of-Domain Generalization for Web Agents via Importance-Diversity Data Selection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zadeh, Fatemeh Pesaran, Choi, Seyeon, Lù, Xing Han, Reddy, Siva, Kim, Gunhee |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Text2Chart31: Instruction Tuning for Chart Generation with Automatic Feedback
von: Zadeh, Fatemeh Pesaran, et al.
Veröffentlicht: (2024)
von: Zadeh, Fatemeh Pesaran, et al.
Veröffentlicht: (2024)
Structured Distillation of Web Agent Capabilities Enables Generalization
von: Lù, Xing Han, et al.
Veröffentlicht: (2026)
von: Lù, Xing Han, et al.
Veröffentlicht: (2026)
LPOI: Listwise Preference Optimization for Vision Language Models
von: Zadeh, Fatemeh Pesaran, et al.
Veröffentlicht: (2025)
von: Zadeh, Fatemeh Pesaran, et al.
Veröffentlicht: (2025)
WebLINX: Real-World Website Navigation with Multi-Turn Dialogue
von: Lù, Xing Han, et al.
Veröffentlicht: (2024)
von: Lù, Xing Han, et al.
Veröffentlicht: (2024)
SafeArena: Evaluating the Safety of Autonomous Web Agents
von: Tur, Ada Defne, et al.
Veröffentlicht: (2025)
von: Tur, Ada Defne, et al.
Veröffentlicht: (2025)
AgentRewardBench: Evaluating Automatic Evaluations of Web Agent Trajectories
von: Lù, Xing Han, et al.
Veröffentlicht: (2025)
von: Lù, Xing Han, et al.
Veröffentlicht: (2025)
Build the web for agents, not agents for the web
von: Lù, Xing Han, et al.
Veröffentlicht: (2025)
von: Lù, Xing Han, et al.
Veröffentlicht: (2025)
Out-Of-Domain Unlabeled Data Improves Generalization
von: Saberi, Amir Hossein, et al.
Veröffentlicht: (2023)
von: Saberi, Amir Hossein, et al.
Veröffentlicht: (2023)
Generalized Gaussian Temporal Difference Error for Uncertainty-aware Reinforcement Learning
von: Kim, Seyeon, et al.
Veröffentlicht: (2024)
von: Kim, Seyeon, et al.
Veröffentlicht: (2024)
Towards Optimization and Model Selection for Domain Generalization: A Mixup-guided Solution
von: Lu, Wang, et al.
Veröffentlicht: (2022)
von: Lu, Wang, et al.
Veröffentlicht: (2022)
Sample Selection via Contrastive Fragmentation for Noisy Label Regression
von: Kim, Chris Dongjoo, et al.
Veröffentlicht: (2025)
von: Kim, Chris Dongjoo, et al.
Veröffentlicht: (2025)
Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning
von: Kim, Minung, et al.
Veröffentlicht: (2026)
von: Kim, Minung, et al.
Veröffentlicht: (2026)
The BrowserGym Ecosystem for Web Agent Research
von: De Chezelles, Thibault Le Sellier, et al.
Veröffentlicht: (2024)
von: De Chezelles, Thibault Le Sellier, et al.
Veröffentlicht: (2024)
Estimating Subgraph Importance with Structural Prior Domain Knowledge
von: Kim, Changhyun, et al.
Veröffentlicht: (2026)
von: Kim, Changhyun, et al.
Veröffentlicht: (2026)
When Meta-Learning Meets Online and Continual Learning: A Survey
von: Son, Jaehyeon, et al.
Veröffentlicht: (2023)
von: Son, Jaehyeon, et al.
Veröffentlicht: (2023)
Unveiling AI's Blind Spots: An Oracle for In-Domain, Out-of-Domain, and Adversarial Errors
von: Han, Shuangpeng, et al.
Veröffentlicht: (2024)
von: Han, Shuangpeng, et al.
Veröffentlicht: (2024)
Out-of-Context Misinformation Detection via Variational Domain-Invariant Learning with Test-Time Training
von: Yang, Xi, et al.
Veröffentlicht: (2025)
von: Yang, Xi, et al.
Veröffentlicht: (2025)
Towards Diverse Perspective Learning with Selection over Multiple Temporal Poolings
von: Seong, Jihyeon, et al.
Veröffentlicht: (2024)
von: Seong, Jihyeon, et al.
Veröffentlicht: (2024)
Dynamic Importance Learning using Fisher Information Matrix (FIM) for Nonlinear Dynamic Mapping
von: Eivaghi, Vahid MohammadZadeh, et al.
Veröffentlicht: (2024)
von: Eivaghi, Vahid MohammadZadeh, et al.
Veröffentlicht: (2024)
Consistency-Guided Temperature Scaling Using Style and Content Information for Out-of-Domain Calibration
von: Choi, Wonjeong, et al.
Veröffentlicht: (2024)
von: Choi, Wonjeong, et al.
Veröffentlicht: (2024)
Towards a Better Evaluation of Out-of-Domain Generalization
von: Hwang, Duhun, et al.
Veröffentlicht: (2024)
von: Hwang, Duhun, et al.
Veröffentlicht: (2024)
Recasting Continual Learning as Sequence Modeling
von: Lee, Soochan, et al.
Veröffentlicht: (2023)
von: Lee, Soochan, et al.
Veröffentlicht: (2023)
Distilling Reinforcement Learning Algorithms for In-Context Model-Based Planning
von: Son, Jaehyeon, et al.
Veröffentlicht: (2025)
von: Son, Jaehyeon, et al.
Veröffentlicht: (2025)
Faithfulness Measurable Masked Language Models
von: Madsen, Andreas, et al.
Veröffentlicht: (2023)
von: Madsen, Andreas, et al.
Veröffentlicht: (2023)
ACE and Diverse Generalization via Selective Disagreement
von: Daniels, Oliver, et al.
Veröffentlicht: (2025)
von: Daniels, Oliver, et al.
Veröffentlicht: (2025)
Improving Out-of-Domain Audio Deepfake Detection via Layer Selection and Fusion of SSL-Based Countermeasures
von: Serrano, Pierre, et al.
Veröffentlicht: (2025)
von: Serrano, Pierre, et al.
Veröffentlicht: (2025)
Mixture Data for Training Cannot Ensure Out-of-distribution Generalization
von: Zhang, Songming, et al.
Veröffentlicht: (2023)
von: Zhang, Songming, et al.
Veröffentlicht: (2023)
Compositional Conservatism: A Transductive Approach in Offline Reinforcement Learning
von: Song, Yeda, et al.
Veröffentlicht: (2024)
von: Song, Yeda, et al.
Veröffentlicht: (2024)
Entropy Meets Importance: A Unified Head Importance-Entropy Score for Stable and Efficient Transformer Pruning
von: Choi, Minsik, et al.
Veröffentlicht: (2025)
von: Choi, Minsik, et al.
Veröffentlicht: (2025)
WebGames: Challenging General-Purpose Web-Browsing AI Agents
von: Thomas, George, et al.
Veröffentlicht: (2025)
von: Thomas, George, et al.
Veröffentlicht: (2025)
Temp-SCONE: A Novel Out-of-Distribution Detection and Domain Generalization Framework for Wild Data with Temporal Shift
von: Naiknaware, Aditi, et al.
Veröffentlicht: (2025)
von: Naiknaware, Aditi, et al.
Veröffentlicht: (2025)
How Many Domains Suffice for Domain Generalization? A Tight Characterization via the Domain Shattering Dimension
von: Dwork, Cynthia, et al.
Veröffentlicht: (2025)
von: Dwork, Cynthia, et al.
Veröffentlicht: (2025)
Noise-Aware Generalization: Robustness to In-Domain Noise and Out-of-Domain Generalization
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents
von: Zambrano, Alejandra, et al.
Veröffentlicht: (2026)
von: Zambrano, Alejandra, et al.
Veröffentlicht: (2026)
Bridging Reasoning to Learning: Unmasking Illusions using Complexity Out of Distribution Generalization
von: Paqaleh, Mohammad Mahdi Samiei, et al.
Veröffentlicht: (2025)
von: Paqaleh, Mohammad Mahdi Samiei, et al.
Veröffentlicht: (2025)
Coverage-Aware Web Crawling for Domain-Specific Supplier Discovery via a Web--Knowledge--Web Pipeline
von: Qi, Yijiashun, et al.
Veröffentlicht: (2026)
von: Qi, Yijiashun, et al.
Veröffentlicht: (2026)
Operationalising the Superficial Alignment Hypothesis via Task Complexity
von: Vergara-Browne, Tomás, et al.
Veröffentlicht: (2026)
von: Vergara-Browne, Tomás, et al.
Veröffentlicht: (2026)
Information-Preserving Domain Transfer with Unlabeled Data in Misspecified Simulation-Based Inference
von: Jang, Joon, et al.
Veröffentlicht: (2026)
von: Jang, Joon, et al.
Veröffentlicht: (2026)
Are self-explanations from Large Language Models faithful?
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
Out-of-Core Dimensionality Reduction for Large Data via Out-of-Sample Extensions
von: Reichmann, Luca, et al.
Veröffentlicht: (2024)
von: Reichmann, Luca, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Text2Chart31: Instruction Tuning for Chart Generation with Automatic Feedback
von: Zadeh, Fatemeh Pesaran, et al.
Veröffentlicht: (2024) -
Structured Distillation of Web Agent Capabilities Enables Generalization
von: Lù, Xing Han, et al.
Veröffentlicht: (2026) -
LPOI: Listwise Preference Optimization for Vision Language Models
von: Zadeh, Fatemeh Pesaran, et al.
Veröffentlicht: (2025) -
WebLINX: Real-World Website Navigation with Multi-Turn Dialogue
von: Lù, Xing Han, et al.
Veröffentlicht: (2024) -
SafeArena: Evaluating the Safety of Autonomous Web Agents
von: Tur, Ada Defne, et al.
Veröffentlicht: (2025)