Fine-Tuning Language Models on Multiple Datasets for Citation Intention Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Shui, Zeren, Karypis, Petros, Karls, Daniel S., Wen, Mingjian, Manchanda, Saurav, Tadmor, Ellad B., Karypis, George |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pack of LLMs: Model Fusion at Test-Time via Perplexity Optimization
by: Mavromatis, Costas, et al.
Published: (2024)
by: Mavromatis, Costas, et al.
Published: (2024)
Extending Input Contexts of Language Models through Training on Segmented Sequences
by: Karypis, Petros, et al.
Published: (2023)
by: Karypis, Petros, et al.
Published: (2023)
SemPool: Simple, robust, and interpretable KG pooling for enhancing language models
by: Mavromatis, Costas, et al.
Published: (2024)
by: Mavromatis, Costas, et al.
Published: (2024)
Learning to Generate Answers with Citations via Factual Consistency Models
by: Aly, Rami, et al.
Published: (2024)
by: Aly, Rami, et al.
Published: (2024)
GNN-RAG: Graph Neural Retrieval for Large Language Model Reasoning
by: Mavromatis, Costas, et al.
Published: (2024)
by: Mavromatis, Costas, et al.
Published: (2024)
Differentially Private Bias-Term Fine-tuning of Foundation Models
by: Bu, Zhiqi, et al.
Published: (2022)
by: Bu, Zhiqi, et al.
Published: (2022)
Extending OpenKIM with an Uncertainty Quantification Toolkit for Molecular Modeling
by: Kurniawan, Yonatan, et al.
Published: (2022)
by: Kurniawan, Yonatan, et al.
Published: (2022)
Parameter-Efficient Tuning Large Language Models for Graph Representation Learning
by: Zhu, Qi, et al.
Published: (2024)
by: Zhu, Qi, et al.
Published: (2024)
Multimodal Chain-of-Thought Reasoning in Language Models
by: Zhang, Zhuosheng, et al.
Published: (2023)
by: Zhang, Zhuosheng, et al.
Published: (2023)
Revisiting SMoE Language Models by Evaluating Inefficiencies with Task Specific Expert Pruning
by: Sarkar, Soumajyoti, et al.
Published: (2024)
by: Sarkar, Soumajyoti, et al.
Published: (2024)
Beyond instruction-conditioning, MoTE: Mixture of Task Experts for Multi-task Embedding Models
by: Romero, Miguel, et al.
Published: (2025)
by: Romero, Miguel, et al.
Published: (2025)
Extreme Miscalibration and the Illusion of Adversarial Robustness
by: Raina, Vyas, et al.
Published: (2024)
by: Raina, Vyas, et al.
Published: (2024)
Scalable Prompt Routing via Fine-Grained Latent Task Discovery
by: Zhang, Yunyi, et al.
Published: (2026)
by: Zhang, Yunyi, et al.
Published: (2026)
KLIFF: A framework to develop physics-based and machine learning interatomic potentials
by: Wen, Mingjian, et al.
Published: (2021)
by: Wen, Mingjian, et al.
Published: (2021)
MaxCode: A Max-Reward Reinforcement Learning Framework for Automated Code Optimization
by: Ou, Jiefu, et al.
Published: (2026)
by: Ou, Jiefu, et al.
Published: (2026)
AgentOccam: A Simple Yet Strong Baseline for LLM-Based Web Agents
by: Yang, Ke, et al.
Published: (2024)
by: Yang, Ke, et al.
Published: (2024)
Comparative study of ensemble-based uncertainty quantification methods for neural network interatomic potentials
by: Kurniawan, Yonatan, et al.
Published: (2025)
by: Kurniawan, Yonatan, et al.
Published: (2025)
Small Language Models Fine-tuned to Coordinate Larger Language Models improve Complex Reasoning
by: Juneja, Gurusha, et al.
Published: (2023)
by: Juneja, Gurusha, et al.
Published: (2023)
Safeguard Fine-Tuned LLMs Through Pre- and Post-Tuning Model Merging
by: Farn, Hua, et al.
Published: (2024)
by: Farn, Hua, et al.
Published: (2024)
A KIM-compliant potfit for fitting sloppy interatomic potentials: Application to the EDIP model for silicon
by: Wen, Mingjian, et al.
Published: (2016)
by: Wen, Mingjian, et al.
Published: (2016)
Improving Narrative Classification and Explanation via Fine Tuned Language Models
by: Tyagi, Rishit, et al.
Published: (2025)
by: Tyagi, Rishit, et al.
Published: (2025)
OPERA: Online Data Pruning for Efficient Retrieval Model Adaptation
by: Fang, Haoyang, et al.
Published: (2026)
by: Fang, Haoyang, et al.
Published: (2026)
Intention-Adaptive LLM Fine-Tuning for Text Revision Generation
by: Liu, Zhexiong, et al.
Published: (2026)
by: Liu, Zhexiong, et al.
Published: (2026)
From Macro to Micro: Probing Dataset Diversity in Language Model Fine-Tuning
by: Li, Haoyu, et al.
Published: (2025)
by: Li, Haoyu, et al.
Published: (2025)
ProxSparse: Regularized Learning of Semi-Structured Sparsity Masks for Pretrained LLMs
by: Liu, Hongyi, et al.
Published: (2025)
by: Liu, Hongyi, et al.
Published: (2025)
BYOKG-RAG: Multi-Strategy Graph Retrieval for Knowledge Graph Question Answering
by: Mavromatis, Costas, et al.
Published: (2025)
by: Mavromatis, Costas, et al.
Published: (2025)
Efficient Ensemble for Fine-tuning Language Models on Multiple Datasets
by: Li, Dongyue, et al.
Published: (2025)
by: Li, Dongyue, et al.
Published: (2025)
Advancing Scientific Text Classification: Fine-Tuned Models with Dataset Expansion and Hard-Voting
by: Rostam, Zhyar Rzgar K, et al.
Published: (2025)
by: Rostam, Zhyar Rzgar K, et al.
Published: (2025)
Training Language Models to Generate Text with Citations via Fine-grained Rewards
by: Huang, Chengyu, et al.
Published: (2024)
by: Huang, Chengyu, et al.
Published: (2024)
Fine-Tuning Large Language Models for Scientific Text Classification: A Comparative Study
by: Rostam, Zhyar Rzgar K, et al.
Published: (2024)
by: Rostam, Zhyar Rzgar K, et al.
Published: (2024)
Leveraging Conditional Mutual Information to Improve Large Language Model Fine-Tuning For Classification
by: Sivakaran, Thanushon, et al.
Published: (2025)
by: Sivakaran, Thanushon, et al.
Published: (2025)
Overtrained Language Models Are Harder to Fine-Tune
by: Springer, Jacob Mitchell, et al.
Published: (2025)
by: Springer, Jacob Mitchell, et al.
Published: (2025)
Learning Fine-Grained Grounded Citations for Attributed Large Language Models
by: Huang, Lei, et al.
Published: (2024)
by: Huang, Lei, et al.
Published: (2024)
Prompting and Fine-Tuning Open-Sourced Large Language Models for Stance Classification
by: Cruickshank, Iain J., et al.
Published: (2023)
by: Cruickshank, Iain J., et al.
Published: (2023)
Understanding Silent Data Corruption in LLM Training
by: Ma, Jeffrey, et al.
Published: (2025)
by: Ma, Jeffrey, et al.
Published: (2025)
Open Materials Generation with Stochastic Interpolants
by: Hoellmer, Philipp, et al.
Published: (2025)
by: Hoellmer, Philipp, et al.
Published: (2025)
Latent Knowledge as a Predictor of Fact Acquisition in Fine-Tuned Large Language Models
by: Hier, Daniel B., et al.
Published: (2026)
by: Hier, Daniel B., et al.
Published: (2026)
A Large-Scale Dataset and Citation Intent Classification in Turkish with LLMs
by: Karaca, Kemal Sami, et al.
Published: (2025)
by: Karaca, Kemal Sami, et al.
Published: (2025)
Adapting Pretrained Language Models for Citation Classification via Self-Supervised Contrastive Learning
by: Li, Tong, et al.
Published: (2025)
by: Li, Tong, et al.
Published: (2025)
IndicLLMSuite: A Blueprint for Creating Pre-training and Fine-Tuning Datasets for Indian Languages
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2024)
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2024)
Similar Items
-
Pack of LLMs: Model Fusion at Test-Time via Perplexity Optimization
by: Mavromatis, Costas, et al.
Published: (2024) -
Extending Input Contexts of Language Models through Training on Segmented Sequences
by: Karypis, Petros, et al.
Published: (2023) -
SemPool: Simple, robust, and interpretable KG pooling for enhancing language models
by: Mavromatis, Costas, et al.
Published: (2024) -
Learning to Generate Answers with Citations via Factual Consistency Models
by: Aly, Rami, et al.
Published: (2024) -
GNN-RAG: Graph Neural Retrieval for Large Language Model Reasoning
by: Mavromatis, Costas, et al.
Published: (2024)