AXE: Low-Cost Cross-Domain Web Structured Information Extraction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mansour, Abdelrahman, Alshaer, Khaled W., Elsaban, Moataz |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Cross-Domain Web Information Extraction at Pinterest
von: Farag, Michael, et al.
Veröffentlicht: (2025)
von: Farag, Michael, et al.
Veröffentlicht: (2025)
LiveWeb-IE: A Benchmark For Online Web Information Extraction
von: Yang, Seungbin, et al.
Veröffentlicht: (2026)
von: Yang, Seungbin, et al.
Veröffentlicht: (2026)
Generator-Retriever-Generator Approach for Open-Domain Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
AMuRD: Annotated Arabic-English Receipt Dataset for Key Information Extraction and Classification
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
R1-RE: Cross-Domain Relation Extraction with RLVR
von: Dai, Runpeng, et al.
Veröffentlicht: (2025)
von: Dai, Runpeng, et al.
Veröffentlicht: (2025)
The Arabic AI Fingerprint: Stylometric Analysis and Detection of Large Language Models Text
von: Al-Shaibani, Maged S., et al.
Veröffentlicht: (2025)
von: Al-Shaibani, Maged S., et al.
Veröffentlicht: (2025)
Prototype-Regularized Federated Learning for Cross-Domain Aspect Sentiment Triplet Extraction
von: Cai, Zongming, et al.
Veröffentlicht: (2026)
von: Cai, Zongming, et al.
Veröffentlicht: (2026)
WebRouter: Query-specific Router via Variational Information Bottleneck for Cost-sensitive Web Agent
von: Li, Tao, et al.
Veröffentlicht: (2025)
von: Li, Tao, et al.
Veröffentlicht: (2025)
Peering Inside the Black Box: Uncovering LLM Errors in Optimization Modelling through Component-Level Evaluation
von: Refai, Dania, et al.
Veröffentlicht: (2025)
von: Refai, Dania, et al.
Veröffentlicht: (2025)
Crossing Domains without Labels: Distant Supervision for Term Extraction
von: Senger, Elena, et al.
Veröffentlicht: (2025)
von: Senger, Elena, et al.
Veröffentlicht: (2025)
SCRIBES: Web-Scale Script-Based Semi-Structured Data Extraction with Reinforcement Learning
von: Liu, Shicheng, et al.
Veröffentlicht: (2025)
von: Liu, Shicheng, et al.
Veröffentlicht: (2025)
Wikidata as a seed for Web Extraction
von: Guo, Kunpeng, et al.
Veröffentlicht: (2024)
von: Guo, Kunpeng, et al.
Veröffentlicht: (2024)
Understanding Cross-Domain Adaptation in Low-Resource Topic Modeling
von: Akash, Pritom Saha, et al.
Veröffentlicht: (2025)
von: Akash, Pritom Saha, et al.
Veröffentlicht: (2025)
Low-resource Information Extraction with the European Clinical Case Corpus
von: Ghosh, Soumitra, et al.
Veröffentlicht: (2025)
von: Ghosh, Soumitra, et al.
Veröffentlicht: (2025)
Combining Language and Graph Models for Semi-structured Information Extraction on the Web
von: Hong, Zhi, et al.
Veröffentlicht: (2024)
von: Hong, Zhi, et al.
Veröffentlicht: (2024)
Multi-Record Web Page Information Extraction From News Websites
von: Kustenkov, Alexander, et al.
Veröffentlicht: (2025)
von: Kustenkov, Alexander, et al.
Veröffentlicht: (2025)
Pyramid MoA: A Probabilistic Framework for Cost-Optimized Anytime Inference
von: Khaled, Arindam
Veröffentlicht: (2026)
von: Khaled, Arindam
Veröffentlicht: (2026)
Translation and Fusion Improves Zero-shot Cross-lingual Information Extraction
von: Chen, Yang, et al.
Veröffentlicht: (2023)
von: Chen, Yang, et al.
Veröffentlicht: (2023)
Cross-Domain Keyword Extraction with Keyness Patterns
von: Zhou, Dongmei, et al.
Veröffentlicht: (2024)
von: Zhou, Dongmei, et al.
Veröffentlicht: (2024)
Dialogue Benchmark Generation from Knowledge Graphs with Cost-Effective Retrieval-Augmented LLMs
von: Omar, Reham, et al.
Veröffentlicht: (2025)
von: Omar, Reham, et al.
Veröffentlicht: (2025)
STAR: Boosting Low-Resource Information Extraction by Structure-to-Text Data Generation with Large Language Models
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
Domain-Adapted Pre-trained Language Models for Implicit Information Extraction in Crash Narratives
von: Wang, Xixi, et al.
Veröffentlicht: (2025)
von: Wang, Xixi, et al.
Veröffentlicht: (2025)
Exploring Hint Generation Approaches in Open-Domain Question Answering
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
Cross-Lingual LLM-Judge Transfer via Evaluation Decomposition
von: Sheth, Ivaxi, et al.
Veröffentlicht: (2026)
von: Sheth, Ivaxi, et al.
Veröffentlicht: (2026)
WCXB: A Multi-Type Web Content Extraction Benchmark
von: Foley, Murrough
Veröffentlicht: (2026)
von: Foley, Murrough
Veröffentlicht: (2026)
PromptPort: A Reliability Layer for Cross-Model Structured Extraction
von: Kotte, Varun
Veröffentlicht: (2026)
von: Kotte, Varun
Veröffentlicht: (2026)
Lexicalized Constituency Parsing for Middle Dutch: Low-resource Training and Cross-Domain Generalization
von: Liang, Yiming, et al.
Veröffentlicht: (2026)
von: Liang, Yiming, et al.
Veröffentlicht: (2026)
Low-Resource Fine-Tuning for Multi-Task Structured Information Extraction with a Billion-Parameter Instruction-Tuned Model
von: Chih, Yu Cheng, et al.
Veröffentlicht: (2025)
von: Chih, Yu Cheng, et al.
Veröffentlicht: (2025)
Leveraging Prompt-Learning for Structured Information Extraction from Crohn's Disease Radiology Reports in a Low-Resource Language
von: Hazan, Liam, et al.
Veröffentlicht: (2024)
von: Hazan, Liam, et al.
Veröffentlicht: (2024)
A Domain-Adapted Pipeline for Structured Information Extraction from Police Incident Announcements on Social Media
von: Shen, Mengfan, et al.
Veröffentlicht: (2025)
von: Shen, Mengfan, et al.
Veröffentlicht: (2025)
High-throughput Biomedical Relation Extraction for Semi-Structured Web Articles Empowered by Large Language Models
von: Zhou, Songchi, et al.
Veröffentlicht: (2023)
von: Zhou, Songchi, et al.
Veröffentlicht: (2023)
Zero-Shot Cross-Domain Dialogue State Tracking via Dual Low-Rank Adaptation
von: Luo, Xiang, et al.
Veröffentlicht: (2024)
von: Luo, Xiang, et al.
Veröffentlicht: (2024)
Multilingual Attribute Extraction from News Web Pages
von: Bedrin, Pavel, et al.
Veröffentlicht: (2025)
von: Bedrin, Pavel, et al.
Veröffentlicht: (2025)
Cross-Domain Transfer and Few-Shot Learning for Personal Identifiable Information Recognition
von: Ye, Junhong, et al.
Veröffentlicht: (2025)
von: Ye, Junhong, et al.
Veröffentlicht: (2025)
Leveraging the Cross-Domain & Cross-Linguistic Corpus for Low Resource NMT: A Case Study On Bhili-Hindi-English Parallel Corpus
von: Singh, Pooja, et al.
Veröffentlicht: (2025)
von: Singh, Pooja, et al.
Veröffentlicht: (2025)
Leveraging Open Information Extraction for More Robust Domain Transfer of Event Trigger Detection
von: Dukić, David, et al.
Veröffentlicht: (2023)
von: Dukić, David, et al.
Veröffentlicht: (2023)
A Few Good Clauses: Comparing LLMs vs Domain-Trained Small Language Models on Structured Contract Extraction
von: Lincoln, Nicole, et al.
Veröffentlicht: (2026)
von: Lincoln, Nicole, et al.
Veröffentlicht: (2026)
Fine Tuning Named Entity Extraction Models for the Fantasy Domain
von: Sivaganeshan, Aravinth, et al.
Veröffentlicht: (2024)
von: Sivaganeshan, Aravinth, et al.
Veröffentlicht: (2024)
Defining Boundaries: The Impact of Domain Specification on Cross-Language and Cross-Domain Transfer in Machine Translation
von: Shahnazaryan, Lia, et al.
Veröffentlicht: (2024)
von: Shahnazaryan, Lia, et al.
Veröffentlicht: (2024)
MSA at ImageCLEF 2025 Multimodal Reasoning: Multilingual Multimodal Reasoning With Ensemble Vision Language Models
von: Ahmed, Seif, et al.
Veröffentlicht: (2025)
von: Ahmed, Seif, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Cross-Domain Web Information Extraction at Pinterest
von: Farag, Michael, et al.
Veröffentlicht: (2025) -
LiveWeb-IE: A Benchmark For Online Web Information Extraction
von: Yang, Seungbin, et al.
Veröffentlicht: (2026) -
Generator-Retriever-Generator Approach for Open-Domain Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023) -
AMuRD: Annotated Arabic-English Receipt Dataset for Key Information Extraction and Classification
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023) -
R1-RE: Cross-Domain Relation Extraction with RLVR
von: Dai, Runpeng, et al.
Veröffentlicht: (2025)