AXE: Low-Cost Cross-Domain Web Structured Information Extraction
Fuente:
arXiv
Saved in:
| Main Authors: | Mansour, Abdelrahman, Alshaer, Khaled W., Elsaban, Moataz |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-Domain Web Information Extraction at Pinterest
by: Farag, Michael, et al.
Published: (2025)
by: Farag, Michael, et al.
Published: (2025)
LiveWeb-IE: A Benchmark For Online Web Information Extraction
by: Yang, Seungbin, et al.
Published: (2026)
by: Yang, Seungbin, et al.
Published: (2026)
Generator-Retriever-Generator Approach for Open-Domain Question Answering
by: Abdallah, Abdelrahman, et al.
Published: (2023)
by: Abdallah, Abdelrahman, et al.
Published: (2023)
AMuRD: Annotated Arabic-English Receipt Dataset for Key Information Extraction and Classification
by: Abdallah, Abdelrahman, et al.
Published: (2023)
by: Abdallah, Abdelrahman, et al.
Published: (2023)
R1-RE: Cross-Domain Relation Extraction with RLVR
by: Dai, Runpeng, et al.
Published: (2025)
by: Dai, Runpeng, et al.
Published: (2025)
The Arabic AI Fingerprint: Stylometric Analysis and Detection of Large Language Models Text
by: Al-Shaibani, Maged S., et al.
Published: (2025)
by: Al-Shaibani, Maged S., et al.
Published: (2025)
Prototype-Regularized Federated Learning for Cross-Domain Aspect Sentiment Triplet Extraction
by: Cai, Zongming, et al.
Published: (2026)
by: Cai, Zongming, et al.
Published: (2026)
WebRouter: Query-specific Router via Variational Information Bottleneck for Cost-sensitive Web Agent
by: Li, Tao, et al.
Published: (2025)
by: Li, Tao, et al.
Published: (2025)
Peering Inside the Black Box: Uncovering LLM Errors in Optimization Modelling through Component-Level Evaluation
by: Refai, Dania, et al.
Published: (2025)
by: Refai, Dania, et al.
Published: (2025)
Crossing Domains without Labels: Distant Supervision for Term Extraction
by: Senger, Elena, et al.
Published: (2025)
by: Senger, Elena, et al.
Published: (2025)
SCRIBES: Web-Scale Script-Based Semi-Structured Data Extraction with Reinforcement Learning
by: Liu, Shicheng, et al.
Published: (2025)
by: Liu, Shicheng, et al.
Published: (2025)
Wikidata as a seed for Web Extraction
by: Guo, Kunpeng, et al.
Published: (2024)
by: Guo, Kunpeng, et al.
Published: (2024)
Understanding Cross-Domain Adaptation in Low-Resource Topic Modeling
by: Akash, Pritom Saha, et al.
Published: (2025)
by: Akash, Pritom Saha, et al.
Published: (2025)
Low-resource Information Extraction with the European Clinical Case Corpus
by: Ghosh, Soumitra, et al.
Published: (2025)
by: Ghosh, Soumitra, et al.
Published: (2025)
Combining Language and Graph Models for Semi-structured Information Extraction on the Web
by: Hong, Zhi, et al.
Published: (2024)
by: Hong, Zhi, et al.
Published: (2024)
Multi-Record Web Page Information Extraction From News Websites
by: Kustenkov, Alexander, et al.
Published: (2025)
by: Kustenkov, Alexander, et al.
Published: (2025)
Pyramid MoA: A Probabilistic Framework for Cost-Optimized Anytime Inference
by: Khaled, Arindam
Published: (2026)
by: Khaled, Arindam
Published: (2026)
Translation and Fusion Improves Zero-shot Cross-lingual Information Extraction
by: Chen, Yang, et al.
Published: (2023)
by: Chen, Yang, et al.
Published: (2023)
Cross-Domain Keyword Extraction with Keyness Patterns
by: Zhou, Dongmei, et al.
Published: (2024)
by: Zhou, Dongmei, et al.
Published: (2024)
Dialogue Benchmark Generation from Knowledge Graphs with Cost-Effective Retrieval-Augmented LLMs
by: Omar, Reham, et al.
Published: (2025)
by: Omar, Reham, et al.
Published: (2025)
STAR: Boosting Low-Resource Information Extraction by Structure-to-Text Data Generation with Large Language Models
by: Ma, Mingyu Derek, et al.
Published: (2023)
by: Ma, Mingyu Derek, et al.
Published: (2023)
Domain-Adapted Pre-trained Language Models for Implicit Information Extraction in Crash Narratives
by: Wang, Xixi, et al.
Published: (2025)
by: Wang, Xixi, et al.
Published: (2025)
Exploring Hint Generation Approaches in Open-Domain Question Answering
by: Mozafari, Jamshid, et al.
Published: (2024)
by: Mozafari, Jamshid, et al.
Published: (2024)
Cross-Lingual LLM-Judge Transfer via Evaluation Decomposition
by: Sheth, Ivaxi, et al.
Published: (2026)
by: Sheth, Ivaxi, et al.
Published: (2026)
WCXB: A Multi-Type Web Content Extraction Benchmark
by: Foley, Murrough
Published: (2026)
by: Foley, Murrough
Published: (2026)
PromptPort: A Reliability Layer for Cross-Model Structured Extraction
by: Kotte, Varun
Published: (2026)
by: Kotte, Varun
Published: (2026)
Lexicalized Constituency Parsing for Middle Dutch: Low-resource Training and Cross-Domain Generalization
by: Liang, Yiming, et al.
Published: (2026)
by: Liang, Yiming, et al.
Published: (2026)
Low-Resource Fine-Tuning for Multi-Task Structured Information Extraction with a Billion-Parameter Instruction-Tuned Model
by: Chih, Yu Cheng, et al.
Published: (2025)
by: Chih, Yu Cheng, et al.
Published: (2025)
Leveraging Prompt-Learning for Structured Information Extraction from Crohn's Disease Radiology Reports in a Low-Resource Language
by: Hazan, Liam, et al.
Published: (2024)
by: Hazan, Liam, et al.
Published: (2024)
A Domain-Adapted Pipeline for Structured Information Extraction from Police Incident Announcements on Social Media
by: Shen, Mengfan, et al.
Published: (2025)
by: Shen, Mengfan, et al.
Published: (2025)
High-throughput Biomedical Relation Extraction for Semi-Structured Web Articles Empowered by Large Language Models
by: Zhou, Songchi, et al.
Published: (2023)
by: Zhou, Songchi, et al.
Published: (2023)
Zero-Shot Cross-Domain Dialogue State Tracking via Dual Low-Rank Adaptation
by: Luo, Xiang, et al.
Published: (2024)
by: Luo, Xiang, et al.
Published: (2024)
Multilingual Attribute Extraction from News Web Pages
by: Bedrin, Pavel, et al.
Published: (2025)
by: Bedrin, Pavel, et al.
Published: (2025)
Cross-Domain Transfer and Few-Shot Learning for Personal Identifiable Information Recognition
by: Ye, Junhong, et al.
Published: (2025)
by: Ye, Junhong, et al.
Published: (2025)
Leveraging the Cross-Domain & Cross-Linguistic Corpus for Low Resource NMT: A Case Study On Bhili-Hindi-English Parallel Corpus
by: Singh, Pooja, et al.
Published: (2025)
by: Singh, Pooja, et al.
Published: (2025)
Leveraging Open Information Extraction for More Robust Domain Transfer of Event Trigger Detection
by: Dukić, David, et al.
Published: (2023)
by: Dukić, David, et al.
Published: (2023)
A Few Good Clauses: Comparing LLMs vs Domain-Trained Small Language Models on Structured Contract Extraction
by: Lincoln, Nicole, et al.
Published: (2026)
by: Lincoln, Nicole, et al.
Published: (2026)
Fine Tuning Named Entity Extraction Models for the Fantasy Domain
by: Sivaganeshan, Aravinth, et al.
Published: (2024)
by: Sivaganeshan, Aravinth, et al.
Published: (2024)
Defining Boundaries: The Impact of Domain Specification on Cross-Language and Cross-Domain Transfer in Machine Translation
by: Shahnazaryan, Lia, et al.
Published: (2024)
by: Shahnazaryan, Lia, et al.
Published: (2024)
MSA at ImageCLEF 2025 Multimodal Reasoning: Multilingual Multimodal Reasoning With Ensemble Vision Language Models
by: Ahmed, Seif, et al.
Published: (2025)
by: Ahmed, Seif, et al.
Published: (2025)
Similar Items
-
Cross-Domain Web Information Extraction at Pinterest
by: Farag, Michael, et al.
Published: (2025) -
LiveWeb-IE: A Benchmark For Online Web Information Extraction
by: Yang, Seungbin, et al.
Published: (2026) -
Generator-Retriever-Generator Approach for Open-Domain Question Answering
by: Abdallah, Abdelrahman, et al.
Published: (2023) -
AMuRD: Annotated Arabic-English Receipt Dataset for Key Information Extraction and Classification
by: Abdallah, Abdelrahman, et al.
Published: (2023) -
R1-RE: Cross-Domain Relation Extraction with RLVR
by: Dai, Runpeng, et al.
Published: (2025)