LEVI: Generalizable Fine-tuning via Layer-wise Ensemble of Different Views
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Roh, Yuji, Liu, Qingyun, Gui, Huan, Yuan, Zhe, Tang, Yujin, Whang, Steven Euijong, Liu, Liang, Bi, Shuchao, Hong, Lichan, Chi, Ed H., Zhao, Zhe |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
PFGuard: A Generative Framework with Privacy and Fairness Safeguards
par: Kim, Soyeon, et autres
Publié: (2024)
par: Kim, Soyeon, et autres
Publié: (2024)
Wisdom of Committee: Distilling from Foundation Model to Specialized Application Model
par: Liu, Zichang, et autres
Publié: (2024)
par: Liu, Zichang, et autres
Publié: (2024)
ReMem: Mutual Information-Aware Fine-tuning of Pretrained Vision Transformers for Effective Knowledge Distillation
par: Dong, Chengyu, et autres
Publié: (2025)
par: Dong, Chengyu, et autres
Publié: (2025)
Fair Class-Incremental Learning using Sample Weighting
par: Park, Jaeyoung, et autres
Publié: (2024)
par: Park, Jaeyoung, et autres
Publié: (2024)
DeFrame: Debiasing Large Language Models Against Framing Effects
par: Lim, Kahee, et autres
Publié: (2026)
par: Lim, Kahee, et autres
Publié: (2026)
T-CIL: Temperature Scaling using Adversarial Perturbation for Calibration in Class-Incremental Learning
par: Hwang, Seong-Hyeon, et autres
Publié: (2025)
par: Hwang, Seong-Hyeon, et autres
Publié: (2025)
GradMix: Gradient-based Selective Mixup for Robust Data Augmentation in Class-Incremental Learning
par: Kim, Minsu, et autres
Publié: (2025)
par: Kim, Minsu, et autres
Publié: (2025)
RC-Mixup: A Data Augmentation Strategy against Noisy Data for Regression Tasks
par: Hwang, Seong-Hyeon, et autres
Publié: (2024)
par: Hwang, Seong-Hyeon, et autres
Publié: (2024)
MIDAS: Misalignment-based Data Augmentation Strategy for Imbalanced Multimodal Learning
par: Hwang, Seong-Hyeon, et autres
Publié: (2025)
par: Hwang, Seong-Hyeon, et autres
Publié: (2025)
Stage-wise Fine-tuning for Graph-to-Text Generation
par: Wang, Qingyun, et autres
Publié: (2021)
par: Wang, Qingyun, et autres
Publié: (2021)
Efficient Layer-wise LLM Fine-tuning for Revision Intention Prediction
par: Liu, Zhexiong, et autres
Publié: (2025)
par: Liu, Zhexiong, et autres
Publié: (2025)
Harnessing Temporal Databases for Systematic Evaluation of Factual Time-Sensitive Question-Answering in Large Language Models
par: Kim, Soyeon, et autres
Publié: (2025)
par: Kim, Soyeon, et autres
Publié: (2025)
Classroom AI: Large Language Models as Grade-Specific Teachers
par: Oh, Jio, et autres
Publié: (2026)
par: Oh, Jio, et autres
Publié: (2026)
Explanation Multiplicity in SHAP: Characterization and Assessment
par: Hwang, Hyunseung, et autres
Publié: (2026)
par: Hwang, Hyunseung, et autres
Publié: (2026)
Balancing Fine-tuning and RAG: A Hybrid Strategy for Dynamic LLM Recommendation Updates
par: Meng, Changping, et autres
Publié: (2025)
par: Meng, Changping, et autres
Publié: (2025)
Falcon: Fair Active Learning using Multi-armed Bandits
par: Tae, Ki Hyun, et autres
Publié: (2024)
par: Tae, Ki Hyun, et autres
Publié: (2024)
HELENE: Hessian Layer-wise Clipping and Gradient Annealing for Accelerating Fine-tuning LLM with Zeroth-order Optimization
par: Zhao, Huaqin, et autres
Publié: (2024)
par: Zhao, Huaqin, et autres
Publié: (2024)
SHAP-based Explanations are Sensitive to Feature Representation
par: Hwang, Hyunseung, et autres
Publié: (2025)
par: Hwang, Hyunseung, et autres
Publié: (2025)
DialectLLM: A Dialect-Aware Dialog[ue] Generation Framework Beyond Standard American English
par: Oh, Jio, et autres
Publié: (2026)
par: Oh, Jio, et autres
Publié: (2026)
Layer-wise Swapping for Generalizable Multilingual Safety
par: Shin, Hyunseo, et autres
Publié: (2026)
par: Shin, Hyunseo, et autres
Publié: (2026)
GRASS: Gradient-based Adaptive Layer-wise Importance Sampling for Memory-efficient Large Language Model Fine-tuning
par: Tian, Kaiyuan, et autres
Publié: (2026)
par: Tian, Kaiyuan, et autres
Publié: (2026)
SlimGPT: Layer-wise Structured Pruning for Large Language Models
par: Ling, Gui, et autres
Publié: (2024)
par: Ling, Gui, et autres
Publié: (2024)
ERBench: An Entity-Relationship based Automatically Verifiable Hallucination Benchmark for Large Language Models
par: Oh, Jio, et autres
Publié: (2024)
par: Oh, Jio, et autres
Publié: (2024)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
par: Xu, Zhe
Publié: (2026)
par: Xu, Zhe
Publié: (2026)
You Only Fine-tune Once: Many-Shot In-Context Fine-Tuning for Large Language Models
par: He, Wenchong, et autres
Publié: (2025)
par: He, Wenchong, et autres
Publié: (2025)
LOTOS: Layer-wise Orthogonalization for Training Robust Ensembles
par: Ebrahimpour-Boroojeny, Ali, et autres
Publié: (2024)
par: Ebrahimpour-Boroojeny, Ali, et autres
Publié: (2024)
Step-wise Adaptive Integration of Supervised Fine-tuning and Reinforcement Learning for Task-Specific LLMs
par: Chen, Jack, et autres
Publié: (2025)
par: Chen, Jack, et autres
Publié: (2025)
Exploring Intrinsic Language-specific Subspaces in Fine-tuning Multilingual Neural Machine Translation
par: Cao, Zhe, et autres
Publié: (2024)
par: Cao, Zhe, et autres
Publié: (2024)
Layer-mediated tuning of spin and valley physics in stacked tetragonal altermagnetic bilayers
par: Tian, Jianke, et autres
Publié: (2026)
par: Tian, Jianke, et autres
Publié: (2026)
LaSM: Layer-wise Scaling Mechanism for Defending Pop-up Attack on GUI Agents
par: Yan, Zihe, et autres
Publié: (2025)
par: Yan, Zihe, et autres
Publié: (2025)
Explain Less, Understand More: Jargon Detection via Personalized Parameter-Efficient Fine-tuning
par: Wu, Bohao, et autres
Publié: (2025)
par: Wu, Bohao, et autres
Publié: (2025)
GPS-Gaussian: Generalizable Pixel-wise 3D Gaussian Splatting for Real-time Human Novel View Synthesis
par: Zheng, Shunyuan, et autres
Publié: (2023)
par: Zheng, Shunyuan, et autres
Publié: (2023)
Layer-wise Importance Matters: Less Memory for Better Performance in Parameter-efficient Fine-tuning of Large Language Models
par: Yao, Kai, et autres
Publié: (2024)
par: Yao, Kai, et autres
Publié: (2024)
Local Layer-wise Differential Privacy in Federated Learning
par: Li, Yunbo, et autres
Publié: (2026)
par: Li, Yunbo, et autres
Publié: (2026)
SEE: Continual Fine-tuning with Sequential Ensemble of Experts
par: Wang, Zhilin, et autres
Publié: (2025)
par: Wang, Zhilin, et autres
Publié: (2025)
A Layer-wise Analysis of Supervised Fine-Tuning
par: Zhao, Qinghua, et autres
Publié: (2026)
par: Zhao, Qinghua, et autres
Publié: (2026)
An Efficient 3D Convolutional Neural Network with Channel-wise, Spatial-grouped, and Temporal Convolutions
par: Wang, Zhe, et autres
Publié: (2025)
par: Wang, Zhe, et autres
Publié: (2025)
Talking with Tables for Better LLM Factual Data Interactions
par: Oh, Jio, et autres
Publié: (2024)
par: Oh, Jio, et autres
Publié: (2024)
Bridging the Gap: Unpacking the Hidden Challenges in Knowledge Distillation for Online Ranking Systems
par: Khani, Nikhil, et autres
Publié: (2024)
par: Khani, Nikhil, et autres
Publié: (2024)
Layer-wise LoRA fine-tuning: a similarity metric approach
par: Ogawa, Keith Ando, et autres
Publié: (2026)
par: Ogawa, Keith Ando, et autres
Publié: (2026)
Documents similaires
-
PFGuard: A Generative Framework with Privacy and Fairness Safeguards
par: Kim, Soyeon, et autres
Publié: (2024) -
Wisdom of Committee: Distilling from Foundation Model to Specialized Application Model
par: Liu, Zichang, et autres
Publié: (2024) -
ReMem: Mutual Information-Aware Fine-tuning of Pretrained Vision Transformers for Effective Knowledge Distillation
par: Dong, Chengyu, et autres
Publié: (2025) -
Fair Class-Incremental Learning using Sample Weighting
par: Park, Jaeyoung, et autres
Publié: (2024) -
DeFrame: Debiasing Large Language Models Against Framing Effects
par: Lim, Kahee, et autres
Publié: (2026)