LiveWeb-IE: A Benchmark For Online Web Information Extraction
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Seungbin, Kim, Jihwan, Choi, Jaemin, Kim, Dongjin, Yang, Soyoung, Park, ChaeHun, Choo, Jaegul |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can Tool-augmented Large Language Models be Aware of Incomplete Conditions?
by: Yang, Seungbin, et al.
Published: (2024)
by: Yang, Seungbin, et al.
Published: (2024)
Breaking Chains: Unraveling the Links in Multi-Hop Knowledge Unlearning
by: Choi, Minseok, et al.
Published: (2024)
by: Choi, Minseok, et al.
Published: (2024)
PairEval: Open-domain Dialogue Evaluation with Pairwise Comparison
by: Park, ChaeHun, et al.
Published: (2024)
by: Park, ChaeHun, et al.
Published: (2024)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
by: Park, ChaeHun, et al.
Published: (2024)
by: Park, ChaeHun, et al.
Published: (2024)
Retrieve Only Relevant Tables Whether Few or Many: Adaptive Table Retrieval Method
by: Kim, Taehee, et al.
Published: (2026)
by: Kim, Taehee, et al.
Published: (2026)
The Comparative Trap: Pairwise Comparisons Amplifies Biased Preferences of LLM Evaluators
by: Jeong, Hawon, et al.
Published: (2024)
by: Jeong, Hawon, et al.
Published: (2024)
Evaluating Visual and Cultural Interpretation: The K-Viscuit Benchmark with Human-VLM Collaboration
by: Park, ChaeHun, et al.
Published: (2024)
by: Park, ChaeHun, et al.
Published: (2024)
ExpGuard: LLM Content Moderation in Specialized Domains
by: Choi, Minseok, et al.
Published: (2026)
by: Choi, Minseok, et al.
Published: (2026)
Not the Example, but the Process: How Self-Generated Examples Enhance LLM Reasoning
by: Gwak, Daehoon, et al.
Published: (2026)
by: Gwak, Daehoon, et al.
Published: (2026)
Reward-Weighted Sampling: Enhancing Non-Autoregressive Characteristics in Masked Diffusion LLMs
by: Gwak, Daehoon, et al.
Published: (2025)
by: Gwak, Daehoon, et al.
Published: (2025)
Translation Deserves Better: Analyzing Translation Artifacts in Cross-lingual Visual Question Answering
by: Park, ChaeHun, et al.
Published: (2024)
by: Park, ChaeHun, et al.
Published: (2024)
Building Resource-Constrained Language Agents: A Korean Case Study on Chemical Toxicity Information
by: Cho, Hojun, et al.
Published: (2025)
by: Cho, Hojun, et al.
Published: (2025)
BankMathBench: A Benchmark for Numerical Reasoning in Banking Scenarios
by: Lee, Yunseung, et al.
Published: (2026)
by: Lee, Yunseung, et al.
Published: (2026)
Web-Shepherd: Advancing PRMs for Reinforcing Web Agents
by: Chae, Hyungjoo, et al.
Published: (2025)
by: Chae, Hyungjoo, et al.
Published: (2025)
Talk to Your Slides: High-Efficiency Slide Editing via Language-Driven Structured Data Manipulation
by: Jung, Kyudan, et al.
Published: (2025)
by: Jung, Kyudan, et al.
Published: (2025)
Cross-Lingual Unlearning of Selective Knowledge in Multilingual Language Models
by: Choi, Minseok, et al.
Published: (2024)
by: Choi, Minseok, et al.
Published: (2024)
Single Ground Truth Is Not Enough: Adding Flexibility to Aspect-Based Sentiment Analysis Evaluation
by: Yang, Soyoung, et al.
Published: (2024)
by: Yang, Soyoung, et al.
Published: (2024)
Web Agents with World Models: Learning and Leveraging Environment Dynamics in Web Navigation
by: Chae, Hyungjoo, et al.
Published: (2024)
by: Chae, Hyungjoo, et al.
Published: (2024)
Safe and Scalable Web Agent Learning via Recreated Websites
by: Chae, Hyungjoo, et al.
Published: (2026)
by: Chae, Hyungjoo, et al.
Published: (2026)
Opt-Out: Investigating Entity-Level Unlearning for Large Language Models via Optimal Transport
by: Choi, Minseok, et al.
Published: (2024)
by: Choi, Minseok, et al.
Published: (2024)
Protecting Privacy Through Approximating Optimal Parameters for Sequence Unlearning in Language Models
by: Lee, Dohyun, et al.
Published: (2024)
by: Lee, Dohyun, et al.
Published: (2024)
TV-LiVE: Training-Free, Text-Guided Video Editing via Layer Informed Vitality Exploitation
by: Kim, Min-Jung, et al.
Published: (2025)
by: Kim, Min-Jung, et al.
Published: (2025)
K-BrowseComp: A Web Browsing Agent Benchmark Grounded in Korean Contexts
by: Lee, Nahyun, et al.
Published: (2026)
by: Lee, Nahyun, et al.
Published: (2026)
Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History
by: Kim, Serin, et al.
Published: (2026)
by: Kim, Serin, et al.
Published: (2026)
IE as Cache: Information Extraction Enhanced Agentic Reasoning
by: Lv, Hang, et al.
Published: (2026)
by: Lv, Hang, et al.
Published: (2026)
WCXB: A Multi-Type Web Content Extraction Benchmark
by: Foley, Murrough
Published: (2026)
by: Foley, Murrough
Published: (2026)
Evaluating Span Extraction in Generative Paradigm: A Reflection on Aspect-Based Sentiment Analysis
by: Yang, Soyoung, et al.
Published: (2024)
by: Yang, Soyoung, et al.
Published: (2024)
Forecasting Future International Events: A Reliable Dataset for Text-Based Event Modeling
by: Gwak, Daehoon, et al.
Published: (2024)
by: Gwak, Daehoon, et al.
Published: (2024)
Cross-lingual Collapse: How Language-Centric Foundation Models Shape Reasoning in Large Language Models
by: Park, Cheonbok, et al.
Published: (2025)
by: Park, Cheonbok, et al.
Published: (2025)
WebCanvas: Benchmarking Web Agents in Online Environments
by: Pan, Yichen, et al.
Published: (2024)
by: Pan, Yichen, et al.
Published: (2024)
$\textit{BenchIE}^{FL}$ : A Manually Re-Annotated Fact-Based Open Information Extraction Benchmark
by: Lamarche, Fabrice, et al.
Published: (2024)
by: Lamarche, Fabrice, et al.
Published: (2024)
WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
by: Qi, Zehan, et al.
Published: (2024)
by: Qi, Zehan, et al.
Published: (2024)
VotIE: Information Extraction from Meeting Minutes
by: Evans, José Pedro, et al.
Published: (2026)
by: Evans, José Pedro, et al.
Published: (2026)
Rethinking KenLM: Good and Bad Model Ensembles for Efficient Text Quality Filtering in Large Web Corpora
by: Kim, Yungi, et al.
Published: (2024)
by: Kim, Yungi, et al.
Published: (2024)
Wikidata as a seed for Web Extraction
by: Guo, Kunpeng, et al.
Published: (2024)
by: Guo, Kunpeng, et al.
Published: (2024)
Cross-Domain Web Information Extraction at Pinterest
by: Farag, Michael, et al.
Published: (2025)
by: Farag, Michael, et al.
Published: (2025)
PyTorch-IE: Fast and Reproducible Prototyping for Information Extraction
by: Binder, Arne, et al.
Published: (2024)
by: Binder, Arne, et al.
Published: (2024)
WebWalker: Benchmarking LLMs in Web Traversal
by: Wu, Jialong, et al.
Published: (2025)
by: Wu, Jialong, et al.
Published: (2025)
Exploring In-context Example Generation for Machine Translation
by: Lee, Dohyun, et al.
Published: (2025)
by: Lee, Dohyun, et al.
Published: (2025)
AXE: Low-Cost Cross-Domain Web Structured Information Extraction
by: Mansour, Abdelrahman, et al.
Published: (2026)
by: Mansour, Abdelrahman, et al.
Published: (2026)
Similar Items
-
Can Tool-augmented Large Language Models be Aware of Incomplete Conditions?
by: Yang, Seungbin, et al.
Published: (2024) -
Breaking Chains: Unraveling the Links in Multi-Hop Knowledge Unlearning
by: Choi, Minseok, et al.
Published: (2024) -
PairEval: Open-domain Dialogue Evaluation with Pairwise Comparison
by: Park, ChaeHun, et al.
Published: (2024) -
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
by: Park, ChaeHun, et al.
Published: (2024) -
Retrieve Only Relevant Tables Whether Few or Many: Adaptive Table Retrieval Method
by: Kim, Taehee, et al.
Published: (2026)