Neuron-Aware Data Selection In Instruction Tuning For Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xin, Wu, Junchao, Yang, Shu, Zhan, Runzhe, Wu, Zeyu, Yang, Min, Huang, Shujian, Chao, Lidia S., Wong, Derek F. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rethinking Prompt-based Debiasing in Large Language Models
von: Yang, Xinyi, et al.
Veröffentlicht: (2025)
von: Yang, Xinyi, et al.
Veröffentlicht: (2025)
Let's Focus on Neuron: Neuron-Level Supervised Fine-tuning for Large Language Model
von: Xu, Haoyun, et al.
Veröffentlicht: (2024)
von: Xu, Haoyun, et al.
Veröffentlicht: (2024)
A Survey on LLM-Generated Text Detection: Necessity, Methods, and Future Directions
von: Wu, Junchao, et al.
Veröffentlicht: (2023)
von: Wu, Junchao, et al.
Veröffentlicht: (2023)
Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore
von: Wu, Junchao, et al.
Veröffentlicht: (2024)
von: Wu, Junchao, et al.
Veröffentlicht: (2024)
RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
DetectRL: Benchmarking LLM-Generated Text Detection in Real-World Scenarios
von: Wu, Junchao, et al.
Veröffentlicht: (2024)
von: Wu, Junchao, et al.
Veröffentlicht: (2024)
Are Large Reasoning Models Good Translation Evaluators? Analysis and Performance Boost
von: Zhan, Runzhe, et al.
Veröffentlicht: (2025)
von: Zhan, Runzhe, et al.
Veröffentlicht: (2025)
Intrinsic Model Weaknesses: How Priming Attacks Unveil Vulnerabilities in Large Language Models
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
Path Drift in Large Reasoning Models:How First-Person Commitments Override Safety
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
Prefix Text as a Yarn: Eliciting Non-English Alignment in Foundation Language Model
von: Zhan, Runzhe, et al.
Veröffentlicht: (2024)
von: Zhan, Runzhe, et al.
Veröffentlicht: (2024)
Understanding and Mitigating Political Stance Cross-topic Generalization in Large Language Models
von: Zhang, Jiayi, et al.
Veröffentlicht: (2025)
von: Zhang, Jiayi, et al.
Veröffentlicht: (2025)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
von: Liu, Liangxin, et al.
Veröffentlicht: (2024)
von: Liu, Liangxin, et al.
Veröffentlicht: (2024)
Investigating CoT Monitorability in Large Reasoning Models
von: Yang, Shu, et al.
Veröffentlicht: (2025)
von: Yang, Shu, et al.
Veröffentlicht: (2025)
Understanding Aha Moments: from External Observations to Internal Mechanisms
von: Yang, Shu, et al.
Veröffentlicht: (2025)
von: Yang, Shu, et al.
Veröffentlicht: (2025)
Is Long-to-Short a Free Lunch? Investigating Inconsistency and Reasoning Efficiency in LRMs
von: Yang, Shu, et al.
Veröffentlicht: (2025)
von: Yang, Shu, et al.
Veröffentlicht: (2025)
Benchmarking the Detection of LLMs-Generated Modern Chinese Poetry
von: Wang, Shanshan, et al.
Veröffentlicht: (2025)
von: Wang, Shanshan, et al.
Veröffentlicht: (2025)
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
SCAR: Data Selection via Style Consistency-Aware Response Ranking for Efficient Instruction-Tuning of Large Language Models
von: Li, Zhuang, et al.
Veröffentlicht: (2024)
von: Li, Zhuang, et al.
Veröffentlicht: (2024)
Can Large Language Models Identify Implicit Suicidal Ideation? An Empirical Evaluation
von: Li, Tong, et al.
Veröffentlicht: (2025)
von: Li, Tong, et al.
Veröffentlicht: (2025)
A Two-Stage Prediction-Aware Contrastive Learning Framework for Multi-Intent NLU
von: Chen, Guanhua, et al.
Veröffentlicht: (2024)
von: Chen, Guanhua, et al.
Veröffentlicht: (2024)
Exposing the Cracks: Vulnerabilities of Retrieval-Augmented LLM-based Machine Translation
von: Sun, Yanming, et al.
Veröffentlicht: (2025)
von: Sun, Yanming, et al.
Veröffentlicht: (2025)
Fraud-R1 : A Multi-Round Benchmark for Assessing the Robustness of LLM Against Augmented Fraud and Phishing Inducements
von: Yang, Shu, et al.
Veröffentlicht: (2025)
von: Yang, Shu, et al.
Veröffentlicht: (2025)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
CommonIT: Commonality-Aware Instruction Tuning for Large Language Models via Data Partitions
von: Rao, Jun, et al.
Veröffentlicht: (2024)
von: Rao, Jun, et al.
Veröffentlicht: (2024)
Large-Scale Data Selection for Instruction Tuning
von: Ivison, Hamish, et al.
Veröffentlicht: (2025)
von: Ivison, Hamish, et al.
Veröffentlicht: (2025)
DPO-Tuned Large Language Models for Segmentation in Simultaneous Speech Translation
von: Yang, Zeyu, et al.
Veröffentlicht: (2025)
von: Yang, Zeyu, et al.
Veröffentlicht: (2025)
GraphGPT: Graph Instruction Tuning for Large Language Models
von: Tang, Jiabin, et al.
Veröffentlicht: (2023)
von: Tang, Jiabin, et al.
Veröffentlicht: (2023)
Federated Data-Efficient Instruction Tuning for Large Language Models
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
Importance-Aware Data Selection for Efficient LLM Instruction Tuning
von: Jiang, Tingyu, et al.
Veröffentlicht: (2025)
von: Jiang, Tingyu, et al.
Veröffentlicht: (2025)
Graph-oriented Instruction Tuning of Large Language Models for Generic Graph Mining
von: Tan, Yanchao, et al.
Veröffentlicht: (2024)
von: Tan, Yanchao, et al.
Veröffentlicht: (2024)
Select2Reason: Efficient Instruction-Tuning Data Selection for Long-CoT Reasoning
von: Yang, Cehao, et al.
Veröffentlicht: (2025)
von: Yang, Cehao, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Gradient Signal-to-Noise Data Selection for Instruction Tuning
von: Yuan, Zhihang, et al.
Veröffentlicht: (2026)
von: Yuan, Zhihang, et al.
Veröffentlicht: (2026)
CoEvol: Constructing Better Responses for Instruction Finetuning through Multi-Agent Cooperation
von: Li, Renhao, et al.
Veröffentlicht: (2024)
von: Li, Renhao, et al.
Veröffentlicht: (2024)
Eliciting the Translation Ability of Large Language Models via Multilingual Finetuning with Translation Instructions
von: Li, Jiahuan, et al.
Veröffentlicht: (2023)
von: Li, Jiahuan, et al.
Veröffentlicht: (2023)
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe
von: Xiao, Yuxin, et al.
Veröffentlicht: (2024)
von: Xiao, Yuxin, et al.
Veröffentlicht: (2024)
IAPT: Instruction-Aware Prompt Tuning for Large Language Models
von: Zhu, Wei, et al.
Veröffentlicht: (2024)
von: Zhu, Wei, et al.
Veröffentlicht: (2024)
Entropy-Based Data Selection for Language Models
von: Li, Hongming, et al.
Veröffentlicht: (2026)
von: Li, Hongming, et al.
Veröffentlicht: (2026)
MT-PATCHER: Selective and Extendable Knowledge Distillation from Large Language Models for Machine Translation
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning
von: Zhu, Wenhao, et al.
Veröffentlicht: (2025)
von: Zhu, Wenhao, et al.
Veröffentlicht: (2025)
Extroversion or Introversion? Controlling The Personality of Your Large Language Models
von: Chen, Yanquan, et al.
Veröffentlicht: (2024)
von: Chen, Yanquan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Rethinking Prompt-based Debiasing in Large Language Models
von: Yang, Xinyi, et al.
Veröffentlicht: (2025) -
Let's Focus on Neuron: Neuron-Level Supervised Fine-tuning for Large Language Model
von: Xu, Haoyun, et al.
Veröffentlicht: (2024) -
A Survey on LLM-Generated Text Detection: Necessity, Methods, and Future Directions
von: Wu, Junchao, et al.
Veröffentlicht: (2023) -
Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore
von: Wu, Junchao, et al.
Veröffentlicht: (2024) -
RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns
von: Chen, Xin, et al.
Veröffentlicht: (2025)