MeXtract: Light-Weight Metadata Extraction from Scientific Papers
Fuente:
arXiv
Saved in:
| Main Authors: | Alyafeai, Zaid, Al-Shaibani, Maged S., Ghanem, Bernard |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MOLE: Metadata Extraction and Validation in Scientific Papers Using LLMs
by: Alyafeai, Zaid, et al.
Published: (2025)
by: Alyafeai, Zaid, et al.
Published: (2025)
Poem Meter Classification of Recited Arabic Poetry: Integrating High-Resource Systems for a Low-Resource Task
by: Al-Shaibani, Maged S., et al.
Published: (2025)
by: Al-Shaibani, Maged S., et al.
Published: (2025)
The Arabic AI Fingerprint: Stylometric Analysis and Detection of Large Language Models Text
by: Al-Shaibani, Maged S., et al.
Published: (2025)
by: Al-Shaibani, Maged S., et al.
Published: (2025)
CIDAR: Culturally Relevant Instruction Dataset For Arabic
by: Alyafeai, Zaid, et al.
Published: (2024)
by: Alyafeai, Zaid, et al.
Published: (2024)
Automated Extraction of Acronym-Expansion Pairs from Scientific Papers
by: Ali, Izhar, et al.
Published: (2024)
by: Ali, Izhar, et al.
Published: (2024)
CounterCount: A Diagnostic Framework for Counting Bias in Vision Language Models
by: Alzahrani, Reem, et al.
Published: (2026)
by: Alzahrani, Reem, et al.
Published: (2026)
Reasoning Vectors: Transferring Chain-of-Thought Capabilities via Task Arithmetic
by: Zbeeb, Mohammad, et al.
Published: (2025)
by: Zbeeb, Mohammad, et al.
Published: (2025)
Hala Technical Report: Building Arabic-Centric Instruction & Translation Models at Scale
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
Beyond the Last Answer: Your Reasoning Trace Uncovers More than You Think
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
MiNER: A Two-Stage Pipeline for Metadata Extraction from Municipal Meeting Minutes
by: Batista, Rodrigo, et al.
Published: (2026)
by: Batista, Rodrigo, et al.
Published: (2026)
Paper2Code: Automating Code Generation from Scientific Papers in Machine Learning
by: Seo, Minju, et al.
Published: (2025)
by: Seo, Minju, et al.
Published: (2025)
Scientific Discourse Tagging for Evidence Extraction
by: Li, Xiangci, et al.
Published: (2019)
by: Li, Xiangci, et al.
Published: (2019)
Clinical Document Metadata Extraction: A Scoping Review
by: Miller, Kurt, et al.
Published: (2025)
by: Miller, Kurt, et al.
Published: (2025)
Papilusion at DAGPap24: Paper or Illusion? Detecting AI-generated Scientific Papers
by: Andreev, Nikita, et al.
Published: (2024)
by: Andreev, Nikita, et al.
Published: (2024)
Arabic Stable LM: Adapting Stable LM 2 1.6B to Arabic
by: Alyafeai, Zaid, et al.
Published: (2024)
by: Alyafeai, Zaid, et al.
Published: (2024)
Automatic Extraction of Metadata from Scientific Publications for CRIS Systems
by: Kovacevic, Aleksandar, et al.
Published: (2011)
by: Kovacevic, Aleksandar, et al.
Published: (2011)
CLAIMCHECK: How Grounded are LLM Critiques of Scientific Papers?
by: Ou, Jiefu, et al.
Published: (2025)
by: Ou, Jiefu, et al.
Published: (2025)
MMCR: Benchmarking Cross-Source Reasoning in Scientific Papers
by: Tian, Yang, et al.
Published: (2025)
by: Tian, Yang, et al.
Published: (2025)
MARG: Multi-Agent Review Generation for Scientific Papers
by: D'Arcy, Mike, et al.
Published: (2024)
by: D'Arcy, Mike, et al.
Published: (2024)
SAVeS: Steering Safety Judgments in Vision-Language Models via Semantic Cues
by: Hinojosa, Carlos, et al.
Published: (2026)
by: Hinojosa, Carlos, et al.
Published: (2026)
TAPS: Task Aware Proposal Distributions for Speculative Sampling
by: Zbib, Mohamad, et al.
Published: (2026)
by: Zbib, Mohamad, et al.
Published: (2026)
Rephrasing natural text data with different languages and quality levels for Large Language Model pre-training
by: Pieler, Michael, et al.
Published: (2024)
by: Pieler, Michael, et al.
Published: (2024)
Russian-Language Multimodal Dataset for Automatic Summarization of Scientific Papers
by: Tsanda, Alena, et al.
Published: (2024)
by: Tsanda, Alena, et al.
Published: (2024)
An Embarrassingly Simple Defense Against LLM Abliteration Attacks
by: Shairah, Harethah Abu, et al.
Published: (2025)
by: Shairah, Harethah Abu, et al.
Published: (2025)
Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection
by: Shairah, Harethah Abu, et al.
Published: (2025)
by: Shairah, Harethah Abu, et al.
Published: (2025)
SciEvent: Benchmarking Multi-domain Scientific Event Extraction
by: Dong, Bofu, et al.
Published: (2025)
by: Dong, Bofu, et al.
Published: (2025)
Automatic Construction of Multiple Classification Dimensions for Managing Approaches in Scientific Papers
by: Ma, Bing, et al.
Published: (2025)
by: Ma, Bing, et al.
Published: (2025)
SciClaimEval: Cross-modal Claim Verification in Scientific Papers
by: Ho, Xanh, et al.
Published: (2026)
by: Ho, Xanh, et al.
Published: (2026)
HyperPIE: Hyperparameter Information Extraction from Scientific Publications
by: Saier, Tarek, et al.
Published: (2023)
by: Saier, Tarek, et al.
Published: (2023)
Data Augmentation Techniques for Process Extraction from Scientific Publications
by: Susanti, Yuni
Published: (2024)
by: Susanti, Yuni
Published: (2024)
Comparison of Feature Learning Methods for Metadata Extraction from PDF Scholarly Documents
by: Boukhers, Zeyd, et al.
Published: (2025)
by: Boukhers, Zeyd, et al.
Published: (2025)
Paper2Video: Automatic Video Generation from Scientific Papers
by: Zhu, Zeyu, et al.
Published: (2025)
by: Zhu, Zeyu, et al.
Published: (2025)
Evaluating Arabic Large Language Models: A Survey of Benchmarks, Methods, and Gaps
by: Alzubaidi, Ahmed, et al.
Published: (2025)
by: Alzubaidi, Ahmed, et al.
Published: (2025)
3LM: Bridging Arabic, STEM, and Code through Benchmarking
by: Boussaha, Basma El Amel, et al.
Published: (2025)
by: Boussaha, Basma El Amel, et al.
Published: (2025)
Paper2Poster: Towards Multimodal Poster Automation from Scientific Papers
by: Pang, Wei, et al.
Published: (2025)
by: Pang, Wei, et al.
Published: (2025)
Unforgotten Safety: Preserving Safety Alignment of Large Language Models with Continual Learning
by: Alssum, Lama, et al.
Published: (2025)
by: Alssum, Lama, et al.
Published: (2025)
DAGverse: Building Document-Grounded Semantic DAGs from Scientific Papers
by: Wan, Shu, et al.
Published: (2026)
by: Wan, Shu, et al.
Published: (2026)
ArabicMMLU: Assessing Massive Multitask Language Understanding in Arabic
by: Koto, Fajri, et al.
Published: (2024)
by: Koto, Fajri, et al.
Published: (2024)
Collage: Decomposable Rapid Prototyping for Information Extraction on Scientific PDFs
by: Gururaja, Sireesh, et al.
Published: (2024)
by: Gururaja, Sireesh, et al.
Published: (2024)
Behind the Magic, MERLIM: Multi-modal Evaluation Benchmark for Large Image-Language Models
by: Villa, Andrés, et al.
Published: (2023)
by: Villa, Andrés, et al.
Published: (2023)
Similar Items
-
MOLE: Metadata Extraction and Validation in Scientific Papers Using LLMs
by: Alyafeai, Zaid, et al.
Published: (2025) -
Poem Meter Classification of Recited Arabic Poetry: Integrating High-Resource Systems for a Low-Resource Task
by: Al-Shaibani, Maged S., et al.
Published: (2025) -
The Arabic AI Fingerprint: Stylometric Analysis and Detection of Large Language Models Text
by: Al-Shaibani, Maged S., et al.
Published: (2025) -
CIDAR: Culturally Relevant Instruction Dataset For Arabic
by: Alyafeai, Zaid, et al.
Published: (2024) -
Automated Extraction of Acronym-Expansion Pairs from Scientific Papers
by: Ali, Izhar, et al.
Published: (2024)