VeriWeb: Verifiable Long-Chain Web Benchmark for Agentic Information-Seeking
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Liu, Shunyu, Liu, Minghao, Zhou, Huichi, Cui, Zhenyu, Zhou, Yang, Zhou, Yuhao, Gao, Jialiang, Zhou, Heng, Yang, Yunhao, Fan, Wendong, zhang, puzhen, Zhang, Ge, Shi, Jiajun, Xuan, Weihao, Huang, Jiaxing, Luo, Shuang, Wu, Fang, Qi, Heli, Zeng, Qingcheng, Wang, Junjie, Feng, Aosong, Lv, Jindi, Jiang, Sicong, Ren, Ziqi, Zhou, Wangchunshu, Yin, Zhenfei, Zhang, Wenlong, Li, Guohao, Yu, Wenhao, Ma, Lei, Bai, Lei, Lin, Qunshu, Song, Mingli, Tao, Dacheng |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Temporal Evolution and Prevention Efficacy of Ritually Induced Fire Ignition Probability in Urban‐Forest Interface Zones: An Empirical Model Based on Forest Fire Risk Data From Kaifu District
par: Sicong Zhou
Publié: (2026)
par: Sicong Zhou
Publié: (2026)
WRAP++: Web discoveRy Amplified Pretraining
par: Zhou, Jiang, et autres
Publié: (2026)
par: Zhou, Jiang, et autres
Publié: (2026)
A New World Framework: The Three Realms and Six Layers Model
par: Zhou, Dacheng
Publié: (2024)
par: Zhou, Dacheng
Publié: (2024)
Examining Search Functions of EAD Finding Aids Web Sites
par: Zhou, Xiaomu
Publié: (2006)
par: Zhou, Xiaomu
Publié: (2006)
Bi-level Mean Field: Dynamic Grouping for Large-Scale MARL
par: Zheng, Yuxuan, et autres
Publié: (2025)
par: Zheng, Yuxuan, et autres
Publié: (2025)
SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Data
par: Fang, Wenkai, et autres
Publié: (2025)
par: Fang, Wenkai, et autres
Publié: (2025)
Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models
par: Xuan, Weihao, et autres
Publié: (2025)
par: Xuan, Weihao, et autres
Publié: (2025)
REAL-IoT: Characterizing GNN Intrusion Detection Robustness under Practical Adversarial Attack
par: Zhan, Zhonghao, et autres
Publié: (2025)
par: Zhan, Zhonghao, et autres
Publié: (2025)
Poster: Enhancing GNN Robustness for Network Intrusion Detection via Agent-based Analysis
par: Zhan, Zhonghao, et autres
Publié: (2025)
par: Zhan, Zhonghao, et autres
Publié: (2025)
Negative contact surgery on Legendrian non-simple knots
par: Wan, Shunyu, et autres
Publié: (2024)
par: Wan, Shunyu, et autres
Publié: (2024)
InfoSeeker: A Scalable Hierarchical Parallel Agent Framework for Web Information Seeking
par: Lee, Ka Yiu, et autres
Publié: (2026)
par: Lee, Ka Yiu, et autres
Publié: (2026)
Code-Switching Information Retrieval: Benchmarks, Analysis, and the Limits of Current Retrievers
par: Zeng, Qingcheng, et autres
Publié: (2026)
par: Zeng, Qingcheng, et autres
Publié: (2026)
The Devil Behind the Mirror: Tracking the Campaigns of Cryptocurrency Abuses on the Dark Web
par: Xia, Pengcheng, et autres
Publié: (2024)
par: Xia, Pengcheng, et autres
Publié: (2024)
WebCanvas: Benchmarking Web Agents in Online Environments
par: Pan, Yichen, et autres
Publié: (2024)
par: Pan, Yichen, et autres
Publié: (2024)
Leveraging Web-Crawled Data for High-Quality Fine-Tuning
par: Zhou, Jing, et autres
Publié: (2024)
par: Zhou, Jing, et autres
Publié: (2024)
Kill Webs by Collaborative & Self-organizing Agents (CSOAs)
par: Zhao, Ying, et autres
Publié: (2026)
par: Zhao, Ying, et autres
Publié: (2026)
Sharp Gain Laws for Extraction Words in Recursive Interval Geometry
par: Zhou, Lei
Publié: (2026)
par: Zhou, Lei
Publié: (2026)
Rules and Algorithms for Objective Construction of Fuzzy Sets
par: Zhou, Lei
Publié: (2024)
par: Zhou, Lei
Publié: (2024)
Temporal Prototype-Aware Learning for Active Voltage Control on Power Distribution Networks
par: Xu, Feiyang, et autres
Publié: (2024)
par: Xu, Feiyang, et autres
Publié: (2024)
A2PO: Towards Effective Offline Reinforcement Learning from an Advantage-aware Perspective
par: Qing, Yunpeng, et autres
Publié: (2024)
par: Qing, Yunpeng, et autres
Publié: (2024)
Reasoning with Reinforced Functional Token Tuning
par: Zhang, Kongcheng, et autres
Publié: (2025)
par: Zhang, Kongcheng, et autres
Publié: (2025)
The Confidence Dichotomy: Analyzing and Mitigating Miscalibration in Tool-Use Agents
par: Xuan, Weihao, et autres
Publié: (2026)
par: Xuan, Weihao, et autres
Publié: (2026)
WebInject: Prompt Injection Attack to Web Agents
par: Wang, Xilong, et autres
Publié: (2025)
par: Wang, Xilong, et autres
Publié: (2025)
WebArena: A Realistic Web Environment for Building Autonomous Agents
par: Zhou, Shuyan, et autres
Publié: (2023)
par: Zhou, Shuyan, et autres
Publié: (2023)
BetaWeb: Towards a Blockchain-enabled Trustworthy Agentic Web
par: Guo, Zihan, et autres
Publié: (2025)
par: Guo, Zihan, et autres
Publié: (2025)
Prune4Web: DOM Tree Pruning Programming for Web Agent
par: Zhang, Jiayuan, et autres
Publié: (2025)
par: Zhang, Jiayuan, et autres
Publié: (2025)
SemanticShield: LLM-Powered Audits Expose Shilling Attacks in Recommender Systems
par: Li, Kaihong, et autres
Publié: (2025)
par: Li, Kaihong, et autres
Publié: (2025)
MPAT: Building Robust Deep Neural Networks against Textual Adversarial Attacks
par: Zhang, Fangyuan, et autres
Publié: (2024)
par: Zhang, Fangyuan, et autres
Publié: (2024)
DiffuseDef: Improved Robustness to Adversarial Attacks via Iterative Denoising
par: Li, Zhenhao, et autres
Publié: (2024)
par: Li, Zhenhao, et autres
Publié: (2024)
Beyond the Hype: A dispassionate look at vision-language models in medical scenario
par: Nan, Yang, et autres
Publié: (2024)
par: Nan, Yang, et autres
Publié: (2024)
Web2BigTable: A Bi-Level Multi-Agent LLM System for Internet-Scale Information Search and Extraction
par: Huang, Yuxuan, et autres
Publié: (2026)
par: Huang, Yuxuan, et autres
Publié: (2026)
SupReMix: Supervised Contrastive Learning for Medical Imaging Regression with Mixup
par: Wu, Yilei, et autres
Publié: (2023)
par: Wu, Yilei, et autres
Publié: (2023)
Beyond Browsing: API-Based Web Agents
par: Song, Yueqi, et autres
Publié: (2024)
par: Song, Yueqi, et autres
Publié: (2024)
High-throughput Biomedical Relation Extraction for Semi-Structured Web Articles Empowered by Large Language Models
par: Zhou, Songchi, et autres
Publié: (2023)
par: Zhou, Songchi, et autres
Publié: (2023)
Surgeries on knots and tight contact structures
par: Li, Zhenkun, et autres
Publié: (2025)
par: Li, Zhenkun, et autres
Publié: (2025)
SkeFi: Cross-Modal Knowledge Transfer for Wireless Skeleton-Based Action Recognition
par: Huang, Shunyu, et autres
Publié: (2026)
par: Huang, Shunyu, et autres
Publié: (2026)
AI-Governed Agent Architecture for Web-Trustworthy Tokenization of Alternative Assets
par: Borjigin, Ailiya, et autres
Publié: (2025)
par: Borjigin, Ailiya, et autres
Publié: (2025)
WebWeaver: Structuring Web-Scale Evidence with Dynamic Outlines for Open-Ended Deep Research
par: Li, Zijian, et autres
Publié: (2025)
par: Li, Zijian, et autres
Publié: (2025)
Using an anime module as a fun way to teach microbial pathogenesis in medical microbiology
par: Dan Lei Zhou
Publié: (2023)
par: Dan Lei Zhou
Publié: (2023)
See and Remember: A Multimodal Agent for Web Traversal
par: Wang, Xinjun, et autres
Publié: (2026)
par: Wang, Xinjun, et autres
Publié: (2026)
Documents similaires
-
Temporal Evolution and Prevention Efficacy of Ritually Induced Fire Ignition Probability in Urban‐Forest Interface Zones: An Empirical Model Based on Forest Fire Risk Data From Kaifu District
par: Sicong Zhou
Publié: (2026) -
WRAP++: Web discoveRy Amplified Pretraining
par: Zhou, Jiang, et autres
Publié: (2026) -
A New World Framework: The Three Realms and Six Layers Model
par: Zhou, Dacheng
Publié: (2024) -
Examining Search Functions of EAD Finding Aids Web Sites
par: Zhou, Xiaomu
Publié: (2006) -
Bi-level Mean Field: Dynamic Grouping for Large-Scale MARL
par: Zheng, Yuxuan, et autres
Publié: (2025)