LLM-Powered Ensemble Learning for Paper Source Tracing: A GPU-Free Approach
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Kunlong, Wang, Junjun, Chen, Zhaoqun, Chen, Kunjin, Chen, Yitian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units
di: Chen, Jianhui, et al.
Pubblicazione: (2026)
di: Chen, Jianhui, et al.
Pubblicazione: (2026)
Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble
di: Zhang, Shun, et al.
Pubblicazione: (2024)
di: Zhang, Shun, et al.
Pubblicazione: (2024)
Understanding the planning of LLM agents: A survey
di: Huang, Xu, et al.
Pubblicazione: (2024)
di: Huang, Xu, et al.
Pubblicazione: (2024)
Train Small, Infer Large: Memory-Efficient LoRA Training for Large Language Models
di: Zhang, Jun, et al.
Pubblicazione: (2025)
di: Zhang, Jun, et al.
Pubblicazione: (2025)
PowerFlow: Unlocking the Dual Nature of LLMs via Principled Distribution Matching
di: Chen, Ruishuo, et al.
Pubblicazione: (2026)
di: Chen, Ruishuo, et al.
Pubblicazione: (2026)
Revisiting Entropy Regularization: Adaptive Coefficient Unlocks Its Potential for LLM Reinforcement Learning
di: Zhang, Xiaoyun, et al.
Pubblicazione: (2025)
di: Zhang, Xiaoyun, et al.
Pubblicazione: (2025)
DistiLLM-2: A Contrastive Approach Boosts the Distillation of LLMs
di: Ko, Jongwoo, et al.
Pubblicazione: (2025)
di: Ko, Jongwoo, et al.
Pubblicazione: (2025)
NOVER: Incentive Training for Language Models via Verifier-Free Reinforcement Learning
di: Liu, Wei, et al.
Pubblicazione: (2025)
di: Liu, Wei, et al.
Pubblicazione: (2025)
PortLLM: Personalizing Evolving Large Language Models with Training-Free and Portable Model Patches
di: Khan, Rana Muhammad Shahroz, et al.
Pubblicazione: (2024)
di: Khan, Rana Muhammad Shahroz, et al.
Pubblicazione: (2024)
Tracing Facts or just Copies? A critical investigation of the Competitions of Mechanisms in Large Language Models
di: Campregher, Dante, et al.
Pubblicazione: (2025)
di: Campregher, Dante, et al.
Pubblicazione: (2025)
Brain-Inspired Two-Stage Approach: Enhancing Mathematical Reasoning by Imitating Human Thought Processes
di: Chen, Yezeng, et al.
Pubblicazione: (2024)
di: Chen, Yezeng, et al.
Pubblicazione: (2024)
ACAR: Adaptive Complexity Routing for Multi-Model Ensembles with Auditable Decision Traces
di: Kumaresan, Ramchand
Pubblicazione: (2026)
di: Kumaresan, Ramchand
Pubblicazione: (2026)
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement
di: Zhao, Pu, et al.
Pubblicazione: (2024)
di: Zhao, Pu, et al.
Pubblicazione: (2024)
Simple Yet Effective: An Information-Theoretic Approach to Multi-LLM Uncertainty Quantification
di: Kruse, Maya, et al.
Pubblicazione: (2025)
di: Kruse, Maya, et al.
Pubblicazione: (2025)
Data Driven Optimization of GPU efficiency for Distributed LLM Adapter Serving
di: Agullo, Ferran, et al.
Pubblicazione: (2026)
di: Agullo, Ferran, et al.
Pubblicazione: (2026)
DASH: Fast Differentiable Architecture Search for Hybrid Attention in Minutes on a Single GPU
di: Chen, Weizhe, et al.
Pubblicazione: (2026)
di: Chen, Weizhe, et al.
Pubblicazione: (2026)
SciPIP: An LLM-based Scientific Paper Idea Proposer
di: Wang, Wenxiao, et al.
Pubblicazione: (2024)
di: Wang, Wenxiao, et al.
Pubblicazione: (2024)
LLM Pruning and Distillation in Practice: The Minitron Approach
di: Sreenivas, Sharath Turuvekere, et al.
Pubblicazione: (2024)
di: Sreenivas, Sharath Turuvekere, et al.
Pubblicazione: (2024)
Enhancing Annotated Bibliography Generation with LLM Ensembles
di: Bermejo, Sergio
Pubblicazione: (2024)
di: Bermejo, Sergio
Pubblicazione: (2024)
SEAL: Safety-enhanced Aligned LLM Fine-tuning via Bilevel Data Selection
di: Shen, Han, et al.
Pubblicazione: (2024)
di: Shen, Han, et al.
Pubblicazione: (2024)
LoX: Low-Rank Extrapolation Robustifies LLM Safety Against Fine-tuning
di: Perin, Gabriel J., et al.
Pubblicazione: (2025)
di: Perin, Gabriel J., et al.
Pubblicazione: (2025)
DFPE: A Diverse Fingerprint Ensemble for Enhancing LLM Performance
di: Cohen, Seffi, et al.
Pubblicazione: (2025)
di: Cohen, Seffi, et al.
Pubblicazione: (2025)
EMORL: Ensemble Multi-Objective Reinforcement Learning for Efficient and Flexible LLM Fine-Tuning
di: Kong, Lingxiao, et al.
Pubblicazione: (2025)
di: Kong, Lingxiao, et al.
Pubblicazione: (2025)
When Greedy Wins: Emergent Exploitation Bias in Meta-Bandit LLM Training
di: Chen, Sanxing, et al.
Pubblicazione: (2025)
di: Chen, Sanxing, et al.
Pubblicazione: (2025)
Model-GLUE: Democratized LLM Scaling for A Large Model Zoo in the Wild
di: Zhao, Xinyu, et al.
Pubblicazione: (2024)
di: Zhao, Xinyu, et al.
Pubblicazione: (2024)
TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories
di: Chen, Yen-Shan, et al.
Pubblicazione: (2026)
di: Chen, Yen-Shan, et al.
Pubblicazione: (2026)
OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning
di: Bao, Zhenghua, et al.
Pubblicazione: (2026)
di: Bao, Zhenghua, et al.
Pubblicazione: (2026)
Power Transformer Fault Prediction Based on Knowledge Graphs
di: Wang, Chao, et al.
Pubblicazione: (2024)
di: Wang, Chao, et al.
Pubblicazione: (2024)
Fine-Tune an SLM or Prompt an LLM? The Case of Generating Low-Code Workflows
di: Ayala, Orlando Marquez, et al.
Pubblicazione: (2025)
di: Ayala, Orlando Marquez, et al.
Pubblicazione: (2025)
Stratified GRPO: Handling Structural Heterogeneity in Reinforcement Learning of LLM Search Agents
di: Zhu, Mingkang, et al.
Pubblicazione: (2025)
di: Zhu, Mingkang, et al.
Pubblicazione: (2025)
Routing with Generated Data: Annotation-Free LLM Skill Estimation and Expert Selection
di: Niu, Tianyi, et al.
Pubblicazione: (2026)
di: Niu, Tianyi, et al.
Pubblicazione: (2026)
Probabilistic Consensus through Ensemble Validation: A Framework for LLM Reliability
di: Naik, Ninad
Pubblicazione: (2024)
di: Naik, Ninad
Pubblicazione: (2024)
TruthFlow: Truthful LLM Generation via Representation Flow Correction
di: Wang, Hanyu, et al.
Pubblicazione: (2025)
di: Wang, Hanyu, et al.
Pubblicazione: (2025)
Learn More, Forget Less: A Gradient-Aware Data Selection Approach for LLM
di: Liu, Yibai, et al.
Pubblicazione: (2025)
di: Liu, Yibai, et al.
Pubblicazione: (2025)
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
di: DeepSeek-AI, et al.
Pubblicazione: (2024)
di: DeepSeek-AI, et al.
Pubblicazione: (2024)
Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
A Multi-Power Law for Loss Curve Prediction Across Learning Rate Schedules
di: Luo, Kairong, et al.
Pubblicazione: (2025)
di: Luo, Kairong, et al.
Pubblicazione: (2025)
Mini-Giants: "Small" Language Models and Open Source Win-Win
di: Zhou, Zhengping, et al.
Pubblicazione: (2023)
di: Zhou, Zhengping, et al.
Pubblicazione: (2023)
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
di: Li, Ming, et al.
Pubblicazione: (2024)
di: Li, Ming, et al.
Pubblicazione: (2024)
A Comparative Study of Neurosymbolic AI Approaches to Interpretable Logical Reasoning
di: Chen, Michael K.
Pubblicazione: (2025)
di: Chen, Michael K.
Pubblicazione: (2025)
Documenti analoghi
-
Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units
di: Chen, Jianhui, et al.
Pubblicazione: (2026) -
Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble
di: Zhang, Shun, et al.
Pubblicazione: (2024) -
Understanding the planning of LLM agents: A survey
di: Huang, Xu, et al.
Pubblicazione: (2024) -
Train Small, Infer Large: Memory-Efficient LoRA Training for Large Language Models
di: Zhang, Jun, et al.
Pubblicazione: (2025) -
PowerFlow: Unlocking the Dual Nature of LLMs via Principled Distribution Matching
di: Chen, Ruishuo, et al.
Pubblicazione: (2026)