Adaptive Self-Supervised Learning Strategies for Dynamic On-Device LLM Personalization
Fuente:
arXiv
Salvato in:
| Autori principali: | Mendoza, Rafael, Cruz, Isabella, Liu, Richard, Deshmukh, Aarav, Williams, David, Peng, Jesscia, Iyer, Rohan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Anomaly Detection in Human Language via Meta-Learning: A Few-Shot Approach
di: Singla, Saurav, et al.
Pubblicazione: (2025)
di: Singla, Saurav, et al.
Pubblicazione: (2025)
HUOZIIME: An On-Device LLM-enhanced Input Method for Deep Personalization
di: Shan, Baocai, et al.
Pubblicazione: (2026)
di: Shan, Baocai, et al.
Pubblicazione: (2026)
BeLLMan: Controlling LLM Congestion
di: Reddy, Tella Rajashekhar, et al.
Pubblicazione: (2025)
di: Reddy, Tella Rajashekhar, et al.
Pubblicazione: (2025)
Self-Assessment Tests are Unreliable Measures of LLM Personality
di: Gupta, Akshat, et al.
Pubblicazione: (2023)
di: Gupta, Akshat, et al.
Pubblicazione: (2023)
Now It Sounds Like You: Learning Personalized Vocabulary On Device
di: Wang, Sid, et al.
Pubblicazione: (2023)
di: Wang, Sid, et al.
Pubblicazione: (2023)
SignSpeak: Open-Source Time Series Classification for ASL Translation
di: Makkar, Aditya, et al.
Pubblicazione: (2024)
di: Makkar, Aditya, et al.
Pubblicazione: (2024)
OpenJarvis: Personal AI, On Personal Devices
di: Saad-Falcon, Jon, et al.
Pubblicazione: (2026)
di: Saad-Falcon, Jon, et al.
Pubblicazione: (2026)
PEToolLLM: Towards Personalized Tool Learning in Large Language Models
di: Xu, Qiancheng, et al.
Pubblicazione: (2025)
di: Xu, Qiancheng, et al.
Pubblicazione: (2025)
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models
di: Liu, Xiang, et al.
Pubblicazione: (2025)
di: Liu, Xiang, et al.
Pubblicazione: (2025)
M-GRPO: Stabilizing Self-Supervised Reinforcement Learning for Large Language Models with Momentum-Anchored Policy Optimization
di: Bai, Bizhe, et al.
Pubblicazione: (2025)
di: Bai, Bizhe, et al.
Pubblicazione: (2025)
RUVA: Personalized Transparent On-Device Graph Reasoning
di: Conte, Gabriele, et al.
Pubblicazione: (2026)
di: Conte, Gabriele, et al.
Pubblicazione: (2026)
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
di: Li, Zheng, et al.
Pubblicazione: (2025)
di: Li, Zheng, et al.
Pubblicazione: (2025)
Structure-BiEval: A Self-Supervised, Dual-Track Framework for Decoupling Structure and Content in LLM Evaluation for Web Information Systems
di: Zhao, Boxiang, et al.
Pubblicazione: (2026)
di: Zhao, Boxiang, et al.
Pubblicazione: (2026)
From Personal to Collective: On the Role of Local and Global Memory in LLM Personalization
di: Wang, Zehong, et al.
Pubblicazione: (2025)
di: Wang, Zehong, et al.
Pubblicazione: (2025)
SSL-SSAW: Self-Supervised Learning with Sigmoid Self-Attention Weighting for Question-Based Sign Language Translation
di: Liu, Zekang, et al.
Pubblicazione: (2025)
di: Liu, Zekang, et al.
Pubblicazione: (2025)
Self-Supervised Multimodal Learning: A Survey
di: Zong, Yongshuo, et al.
Pubblicazione: (2023)
di: Zong, Yongshuo, et al.
Pubblicazione: (2023)
Self-Harmony: Learning to Harmonize Self-Supervision and Self-Play in Test-Time Reinforcement Learning
di: Wang, Ru, et al.
Pubblicazione: (2025)
di: Wang, Ru, et al.
Pubblicazione: (2025)
When Babies Teach Babies: Can student knowledge sharing outperform Teacher-Guided Distillation on small datasets?
di: Iyer, Srikrishna
Pubblicazione: (2024)
di: Iyer, Srikrishna
Pubblicazione: (2024)
Personalized LLM Decoding via Contrasting Personal Preference
di: Bu, Hyungjune, et al.
Pubblicazione: (2025)
di: Bu, Hyungjune, et al.
Pubblicazione: (2025)
Differentiable Belief-based Opponent Shaping
di: Sane, Aarav G, et al.
Pubblicazione: (2026)
di: Sane, Aarav G, et al.
Pubblicazione: (2026)
GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation
di: Li, Sijia, et al.
Pubblicazione: (2026)
di: Li, Sijia, et al.
Pubblicazione: (2026)
Learning Dynamics of Meta-Learning in Small Model Pretraining
di: Africa, David Demitri, et al.
Pubblicazione: (2025)
di: Africa, David Demitri, et al.
Pubblicazione: (2025)
Adaptive-Solver Framework for Dynamic Strategy Selection in Large Language Model Reasoning
di: Zhou, Jianpeng, et al.
Pubblicazione: (2023)
di: Zhou, Jianpeng, et al.
Pubblicazione: (2023)
LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked
di: Phute, Mansi, et al.
Pubblicazione: (2023)
di: Phute, Mansi, et al.
Pubblicazione: (2023)
SPIN: Self-Supervised Prompt INjection
di: Zhou, Leon, et al.
Pubblicazione: (2024)
di: Zhou, Leon, et al.
Pubblicazione: (2024)
Personality as Relational Infrastructure: User Perceptions of Personality-Trait-Infused LLM Messaging
di: Hofer, Dominik P., et al.
Pubblicazione: (2026)
di: Hofer, Dominik P., et al.
Pubblicazione: (2026)
Self-Supervised Learning Based Handwriting Verification
di: Chauhan, Mihir, et al.
Pubblicazione: (2024)
di: Chauhan, Mihir, et al.
Pubblicazione: (2024)
Experiences Build Characters: The Linguistic Origins and Functional Impact of LLM Personality
di: Wang, Xi, et al.
Pubblicazione: (2026)
di: Wang, Xi, et al.
Pubblicazione: (2026)
Linear-Complexity Self-Supervised Learning for Speech Processing
di: Zhang, Shucong, et al.
Pubblicazione: (2024)
di: Zhang, Shucong, et al.
Pubblicazione: (2024)
Dynamic Generation of Personalities with Large Language Models
di: Liu, Jianzhi, et al.
Pubblicazione: (2024)
di: Liu, Jianzhi, et al.
Pubblicazione: (2024)
Never Start from Scratch: Expediting On-Device LLM Personalization via Explainable Model Selection
di: Wang, Haoming, et al.
Pubblicazione: (2025)
di: Wang, Haoming, et al.
Pubblicazione: (2025)
Code Execution as Grounded Supervision for LLM Reasoning
di: Jung, Dongwon, et al.
Pubblicazione: (2025)
di: Jung, Dongwon, et al.
Pubblicazione: (2025)
Adaptive Stopping for Multi-Turn LLM Reasoning
di: Zhou, Xiaofan, et al.
Pubblicazione: (2026)
di: Zhou, Xiaofan, et al.
Pubblicazione: (2026)
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning
di: Zhang, Kongcheng, et al.
Pubblicazione: (2025)
di: Zhang, Kongcheng, et al.
Pubblicazione: (2025)
Enabling On-Device Large Language Model Personalization with Self-Supervised Data Selection and Synthesis
di: Qin, Ruiyang, et al.
Pubblicazione: (2023)
di: Qin, Ruiyang, et al.
Pubblicazione: (2023)
An Exploration of Mamba for Speech Self-Supervised Models
di: Lin, Tzu-Quan, et al.
Pubblicazione: (2025)
di: Lin, Tzu-Quan, et al.
Pubblicazione: (2025)
Self-Supervised Prompt Optimization
di: Xiang, Jinyu, et al.
Pubblicazione: (2025)
di: Xiang, Jinyu, et al.
Pubblicazione: (2025)
A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
di: Liu, Zijun, et al.
Pubblicazione: (2023)
di: Liu, Zijun, et al.
Pubblicazione: (2023)
Benchmarking and Improving LLM Robustness for Personalized Generation
di: Okite, Chimaobi, et al.
Pubblicazione: (2025)
di: Okite, Chimaobi, et al.
Pubblicazione: (2025)
Learning Dynamics of LLM Finetuning
di: Ren, Yi, et al.
Pubblicazione: (2024)
di: Ren, Yi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Anomaly Detection in Human Language via Meta-Learning: A Few-Shot Approach
di: Singla, Saurav, et al.
Pubblicazione: (2025) -
HUOZIIME: An On-Device LLM-enhanced Input Method for Deep Personalization
di: Shan, Baocai, et al.
Pubblicazione: (2026) -
BeLLMan: Controlling LLM Congestion
di: Reddy, Tella Rajashekhar, et al.
Pubblicazione: (2025) -
Self-Assessment Tests are Unreliable Measures of LLM Personality
di: Gupta, Akshat, et al.
Pubblicazione: (2023) -
Now It Sounds Like You: Learning Personalized Vocabulary On Device
di: Wang, Sid, et al.
Pubblicazione: (2023)