Encoder Fine-tuning with Stochastic Sampling Outperforms Open-weight GPT in Astronomy Knowledge Extraction
Fuente:
arXiv
Saved in:
| Main Authors: | Rawat, Shivam, Flek, Lucie, Karimi, Akbar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Robustness of LLMs to Paraphrasing Based on Sociodemographic Factors
by: Arora, Pulkit, et al.
Published: (2025)
by: Arora, Pulkit, et al.
Published: (2025)
Exploring Robustness of Multilingual LLMs on Real-World Noisy Data
by: Aliakbarzadeh, Amirhossein, et al.
Published: (2025)
by: Aliakbarzadeh, Amirhossein, et al.
Published: (2025)
Multi-Hop Reasoning for Question Answering with Hyperbolic Representations
by: Welz, Simon, et al.
Published: (2025)
by: Welz, Simon, et al.
Published: (2025)
Label-Consistent Data Generation for Aspect-Based Sentiment Analysis Using LLM Agents
by: Monfared, Mohammad H. A., et al.
Published: (2026)
by: Monfared, Mohammad H. A., et al.
Published: (2026)
More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persists
by: Alavi, Khashayar, et al.
Published: (2025)
by: Alavi, Khashayar, et al.
Published: (2025)
ArithmAttack: Evaluating Robustness of LLMs to Noisy Context in Math Problem Solving
by: Abedin, Zain Ul, et al.
Published: (2025)
by: Abedin, Zain Ul, et al.
Published: (2025)
Reasoning Primitives in Hybrid and Non-Hybrid LLMs: Do Architectural Differences Yield Advantages in State-Tracking and Recall?
by: Rawat, Shivam, et al.
Published: (2026)
by: Rawat, Shivam, et al.
Published: (2026)
Plausible but Wrong: A case study on Agentic Failures in Astrophysical Workflows
by: Rawat, Shivam, et al.
Published: (2026)
by: Rawat, Shivam, et al.
Published: (2026)
Can LLM Agents Identify Spoken Dialects like a Linguist?
by: Bystrich, Tobias, et al.
Published: (2026)
by: Bystrich, Tobias, et al.
Published: (2026)
Improving Low-Resource Dialect Classification Using Retrieval-based Voice Conversion
by: Fischbach, Lea, et al.
Published: (2025)
by: Fischbach, Lea, et al.
Published: (2025)
IKnow: Instruction-Knowledge-Aware Continual Pretraining for Effective Domain Adaptation
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
Probing the Robustness of Theory of Mind in Large Language Models
by: Nickel, Christian, et al.
Published: (2024)
by: Nickel, Christian, et al.
Published: (2024)
Knowledge AI: Fine-tuning NLP Models for Facilitating Scientific Knowledge Extraction and Understanding
by: Muralidharan, Balaji, et al.
Published: (2024)
by: Muralidharan, Balaji, et al.
Published: (2024)
The Muddy Waters of Modeling Empathy in Language: The Practical Impacts of Theoretical Constructs
by: Lahnala, Allison, et al.
Published: (2025)
by: Lahnala, Allison, et al.
Published: (2025)
A Critical Reflection and Forward Perspective on Empathy and Natural Language Processing
by: Lahnala, Allison, et al.
Published: (2022)
by: Lahnala, Allison, et al.
Published: (2022)
Do Multilingual Large Language Models Mitigate Stereotype Bias?
by: Nie, Shangrui, et al.
Published: (2024)
by: Nie, Shangrui, et al.
Published: (2024)
Fine-tuning the SwissBERT Encoder Model for Embedding Sentences and Documents
by: Grosjean, Juri, et al.
Published: (2024)
by: Grosjean, Juri, et al.
Published: (2024)
Funzac at CoMeDi Shared Task: Modeling Annotator Disagreement from Word-In-Context Perspectives
by: Sarumi, Olufunke O., et al.
Published: (2025)
by: Sarumi, Olufunke O., et al.
Published: (2025)
Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models
by: Nickel, Christian, et al.
Published: (2026)
by: Nickel, Christian, et al.
Published: (2026)
UltraLink: An Open-Source Knowledge-Enhanced Multilingual Supervised Fine-tuning Dataset
by: Wang, Haoyu, et al.
Published: (2024)
by: Wang, Haoyu, et al.
Published: (2024)
Automated Information Extraction from Thyroid Operation Narrative: A Comparative Study of GPT-4 and Fine-tuned KoELECTRA
by: Jang, Dongsuk, et al.
Published: (2024)
by: Jang, Dongsuk, et al.
Published: (2024)
PERSPECTRA: A Scalable and Configurable Pluralist Benchmark of Perspectives from Arguments
by: Nie, Shangrui, et al.
Published: (2026)
by: Nie, Shangrui, et al.
Published: (2026)
Fine-tuning Transformer-based Encoder for Turkish Language Understanding Tasks
by: Yildirim, Savas
Published: (2024)
by: Yildirim, Savas
Published: (2024)
Tucano 2 Cool: Better Open Source LLMs for Portuguese
by: Corrêa, Nicholas Kluge, et al.
Published: (2026)
by: Corrêa, Nicholas Kluge, et al.
Published: (2026)
Pitfalls of Conversational LLMs on News Debiasing
by: Schlicht, Ipek Baris, et al.
Published: (2024)
by: Schlicht, Ipek Baris, et al.
Published: (2024)
Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!
by: Zhang, Zhexin, et al.
Published: (2025)
by: Zhang, Zhexin, et al.
Published: (2025)
On the Limitations of Language Targeted Pruning: Investigating the Calibration Language Impact in Multilingual LLM Pruning
by: Kurz, Simon, et al.
Published: (2024)
by: Kurz, Simon, et al.
Published: (2024)
Survey-to-Behavior: Downstream Alignment of Human Values in LLMs via Survey Questions
by: Nie, Shangrui, et al.
Published: (2025)
by: Nie, Shangrui, et al.
Published: (2025)
Reinforcement Learning Amplifies Emergent Misalignment from Harmless Rewards
by: Jørgenvåg, Magnus, et al.
Published: (2026)
by: Jørgenvåg, Magnus, et al.
Published: (2026)
Fine-tuning ChatGPT for Automatic Scoring of Written Scientific Explanations in Chinese
by: Yang, Jie, et al.
Published: (2025)
by: Yang, Jie, et al.
Published: (2025)
Semantic-preserved Augmentation with Confidence-weighted Fine-tuning for Aspect Category Sentiment Analysis
by: Chai, Yaping, et al.
Published: (2025)
by: Chai, Yaping, et al.
Published: (2025)
Aloe: A Family of Fine-tuned Open Healthcare LLMs
by: Gururajan, Ashwin Kumar, et al.
Published: (2024)
by: Gururajan, Ashwin Kumar, et al.
Published: (2024)
ChatGPT v.s. Media Bias: A Comparative Study of GPT-3.5 and Fine-tuned Language Models
by: Wen, Zehao, et al.
Published: (2024)
by: Wen, Zehao, et al.
Published: (2024)
Delving into the Utilisation of ChatGPT in Scientific Publications in Astronomy
by: Astarita, Simone, et al.
Published: (2024)
by: Astarita, Simone, et al.
Published: (2024)
Relational Knowledge Distillation Using Fine-tuned Function Vectors
by: Kang, Andrea, et al.
Published: (2026)
by: Kang, Andrea, et al.
Published: (2026)
Benign Samples Matter! Fine-tuning On Outlier Benign Samples Severely Breaks Safety
by: Guan, Zihan, et al.
Published: (2025)
by: Guan, Zihan, et al.
Published: (2025)
Outlier-weighed Layerwise Sampling for LLM Fine-tuning
by: Li, Pengxiang, et al.
Published: (2024)
by: Li, Pengxiang, et al.
Published: (2024)
UNH at CheckThat! 2025: Fine-tuning Vs Prompting in Claim Extraction
by: Wilder, Joe, et al.
Published: (2025)
by: Wilder, Joe, et al.
Published: (2025)
Unifying the Extremes: Developing a Unified Model for Detecting and Predicting Extremist Traits and Radicalization
by: Lahnala, Allison, et al.
Published: (2025)
by: Lahnala, Allison, et al.
Published: (2025)
Can Stories Help LLMs Reason? Curating Information Space Through Narrative
by: Javadi, Vahid Sadiri, et al.
Published: (2024)
by: Javadi, Vahid Sadiri, et al.
Published: (2024)
Similar Items
-
Exploring Robustness of LLMs to Paraphrasing Based on Sociodemographic Factors
by: Arora, Pulkit, et al.
Published: (2025) -
Exploring Robustness of Multilingual LLMs on Real-World Noisy Data
by: Aliakbarzadeh, Amirhossein, et al.
Published: (2025) -
Multi-Hop Reasoning for Question Answering with Hyperbolic Representations
by: Welz, Simon, et al.
Published: (2025) -
Label-Consistent Data Generation for Aspect-Based Sentiment Analysis Using LLM Agents
by: Monfared, Mohammad H. A., et al.
Published: (2026) -
More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persists
by: Alavi, Khashayar, et al.
Published: (2025)