MOLE: Metadata Extraction and Validation in Scientific Papers Using LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Alyafeai, Zaid, Al-Shaibani, Maged S., Ghanem, Bernard |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MeXtract: Light-Weight Metadata Extraction from Scientific Papers
by: Alyafeai, Zaid, et al.
Published: (2025)
by: Alyafeai, Zaid, et al.
Published: (2025)
Poem Meter Classification of Recited Arabic Poetry: Integrating High-Resource Systems for a Low-Resource Task
by: Al-Shaibani, Maged S., et al.
Published: (2025)
by: Al-Shaibani, Maged S., et al.
Published: (2025)
The Arabic AI Fingerprint: Stylometric Analysis and Detection of Large Language Models Text
by: Al-Shaibani, Maged S., et al.
Published: (2025)
by: Al-Shaibani, Maged S., et al.
Published: (2025)
CIDAR: Culturally Relevant Instruction Dataset For Arabic
by: Alyafeai, Zaid, et al.
Published: (2024)
by: Alyafeai, Zaid, et al.
Published: (2024)
Exploring LLMs for Scientific Information Extraction Using The SciEx Framework
by: Li, Sha, et al.
Published: (2025)
by: Li, Sha, et al.
Published: (2025)
CounterCount: A Diagnostic Framework for Counting Bias in Vision Language Models
by: Alzahrani, Reem, et al.
Published: (2026)
by: Alzahrani, Reem, et al.
Published: (2026)
Automated Extraction of Acronym-Expansion Pairs from Scientific Papers
by: Ali, Izhar, et al.
Published: (2024)
by: Ali, Izhar, et al.
Published: (2024)
Reasoning Vectors: Transferring Chain-of-Thought Capabilities via Task Arithmetic
by: Zbeeb, Mohammad, et al.
Published: (2025)
by: Zbeeb, Mohammad, et al.
Published: (2025)
Reviewing Scientific Papers for Critical Problems With Reasoning LLMs: Baseline Approaches and Automatic Evaluation
by: Zhang, Tianmai M., et al.
Published: (2025)
by: Zhang, Tianmai M., et al.
Published: (2025)
Unveiling the Merits and Defects of LLMs in Automatic Review Generation for Scientific Papers
by: Li, Ruochi, et al.
Published: (2025)
by: Li, Ruochi, et al.
Published: (2025)
Hala Technical Report: Building Arabic-Centric Instruction & Translation Models at Scale
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
Beyond the Last Answer: Your Reasoning Trace Uncovers More than You Think
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
Mapping the Increasing Use of LLMs in Scientific Papers
by: Liang, Weixin, et al.
Published: (2024)
by: Liang, Weixin, et al.
Published: (2024)
Scientific Discourse Tagging for Evidence Extraction
by: Li, Xiangci, et al.
Published: (2019)
by: Li, Xiangci, et al.
Published: (2019)
Using LLMs for the Extraction and Normalization of Product Attribute Values
by: Brinkmann, Alexander, et al.
Published: (2024)
by: Brinkmann, Alexander, et al.
Published: (2024)
Clinical Document Metadata Extraction: A Scoping Review
by: Miller, Kurt, et al.
Published: (2025)
by: Miller, Kurt, et al.
Published: (2025)
Can LLMs Identify Critical Limitations within Scientific Research? A Systematic Evaluation on AI Research Papers
by: Xu, Zhijian, et al.
Published: (2025)
by: Xu, Zhijian, et al.
Published: (2025)
Papilusion at DAGPap24: Paper or Illusion? Detecting AI-generated Scientific Papers
by: Andreev, Nikita, et al.
Published: (2024)
by: Andreev, Nikita, et al.
Published: (2024)
Usefulness of LLMs as an Author Checklist Assistant for Scientific Papers: NeurIPS'24 Experiment
by: Goldberg, Alexander, et al.
Published: (2024)
by: Goldberg, Alexander, et al.
Published: (2024)
Arabic Stable LM: Adapting Stable LM 2 1.6B to Arabic
by: Alyafeai, Zaid, et al.
Published: (2024)
by: Alyafeai, Zaid, et al.
Published: (2024)
MiNER: A Two-Stage Pipeline for Metadata Extraction from Municipal Meeting Minutes
by: Batista, Rodrigo, et al.
Published: (2026)
by: Batista, Rodrigo, et al.
Published: (2026)
Paper2Code: Automating Code Generation from Scientific Papers in Machine Learning
by: Seo, Minju, et al.
Published: (2025)
by: Seo, Minju, et al.
Published: (2025)
Enhancing Abstractive Summarization of Scientific Papers Using Structure Information
by: Bao, Tong, et al.
Published: (2025)
by: Bao, Tong, et al.
Published: (2025)
CLAIMCHECK: How Grounded are LLM Critiques of Scientific Papers?
by: Ou, Jiefu, et al.
Published: (2025)
by: Ou, Jiefu, et al.
Published: (2025)
MMCR: Benchmarking Cross-Source Reasoning in Scientific Papers
by: Tian, Yang, et al.
Published: (2025)
by: Tian, Yang, et al.
Published: (2025)
MARG: Multi-Agent Review Generation for Scientific Papers
by: D'Arcy, Mike, et al.
Published: (2024)
by: D'Arcy, Mike, et al.
Published: (2024)
SAVeS: Steering Safety Judgments in Vision-Language Models via Semantic Cues
by: Hinojosa, Carlos, et al.
Published: (2026)
by: Hinojosa, Carlos, et al.
Published: (2026)
TAPS: Task Aware Proposal Distributions for Speculative Sampling
by: Zbib, Mohamad, et al.
Published: (2026)
by: Zbib, Mohamad, et al.
Published: (2026)
Dataset Mention Extraction in Scientific Articles Using Bi-LSTM-CRF Model
by: Zeng, Tong, et al.
Published: (2024)
by: Zeng, Tong, et al.
Published: (2024)
Information Extraction From Fiscal Documents Using LLMs
by: Aggarwal, Vikram, et al.
Published: (2025)
by: Aggarwal, Vikram, et al.
Published: (2025)
Rephrasing natural text data with different languages and quality levels for Large Language Model pre-training
by: Pieler, Michael, et al.
Published: (2024)
by: Pieler, Michael, et al.
Published: (2024)
Question Suggestion for Conversational Shopping Assistants Using Product Metadata
by: Vedula, Nikhita, et al.
Published: (2024)
by: Vedula, Nikhita, et al.
Published: (2024)
Can AI Validate Science? Benchmarking LLMs for Accurate Scientific Claim $\rightarrow$ Evidence Reasoning
by: Javaji, Shashidhar Reddy, et al.
Published: (2025)
by: Javaji, Shashidhar Reddy, et al.
Published: (2025)
Enhancing Legal LLMs through Metadata-Enriched RAG Pipelines and Direct Preference Optimization
by: Maniyar, Suyash, et al.
Published: (2026)
by: Maniyar, Suyash, et al.
Published: (2026)
Generating Hierarchical JSON Representations of Scientific Sentences Using LLMs
by: Nimmagadda, Satya Sri Rajiteswari, et al.
Published: (2026)
by: Nimmagadda, Satya Sri Rajiteswari, et al.
Published: (2026)
Russian-Language Multimodal Dataset for Automatic Summarization of Scientific Papers
by: Tsanda, Alena, et al.
Published: (2024)
by: Tsanda, Alena, et al.
Published: (2024)
An Embarrassingly Simple Defense Against LLM Abliteration Attacks
by: Shairah, Harethah Abu, et al.
Published: (2025)
by: Shairah, Harethah Abu, et al.
Published: (2025)
Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection
by: Shairah, Harethah Abu, et al.
Published: (2025)
by: Shairah, Harethah Abu, et al.
Published: (2025)
Iterative Prompt Refinement for Radiation Oncology Symptom Extraction Using Teacher-Student Large Language Models
by: Khanmohammadi, Reza, et al.
Published: (2024)
by: Khanmohammadi, Reza, et al.
Published: (2024)
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers
by: Lin, Liang, et al.
Published: (2025)
by: Lin, Liang, et al.
Published: (2025)
Similar Items
-
MeXtract: Light-Weight Metadata Extraction from Scientific Papers
by: Alyafeai, Zaid, et al.
Published: (2025) -
Poem Meter Classification of Recited Arabic Poetry: Integrating High-Resource Systems for a Low-Resource Task
by: Al-Shaibani, Maged S., et al.
Published: (2025) -
The Arabic AI Fingerprint: Stylometric Analysis and Detection of Large Language Models Text
by: Al-Shaibani, Maged S., et al.
Published: (2025) -
CIDAR: Culturally Relevant Instruction Dataset For Arabic
by: Alyafeai, Zaid, et al.
Published: (2024) -
Exploring LLMs for Scientific Information Extraction Using The SciEx Framework
by: Li, Sha, et al.
Published: (2025)