Learn-to-learn on Arbitrary Textual Conditioning: A Hypernetwork-Driven Meta-Gated LLM
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ji, Luo, Qin, Qi, Xi, Ningyuan, Chen, Teng, Gu, Qingqing, Li, Hongyan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Practice of Post-Training on Llama-3 70B with Optimal Selection of Additional Language Mixture Ratio
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024)
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024)
LaMsS: When Large Language Models Meet Self-Skepticism
von: Wu, Yetao, et al.
Veröffentlicht: (2024)
von: Wu, Yetao, et al.
Veröffentlicht: (2024)
Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA
von: Chen, Teng, et al.
Veröffentlicht: (2026)
von: Chen, Teng, et al.
Veröffentlicht: (2026)
MeTHanol: Modularized Thinking Language Models with Intermediate Layer Thinking, Decoding and Bootstrapping Reasoning
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024)
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024)
Zhyper: Factorized Hypernetworks for Conditioned LLM Fine-Tuning
von: Abdalla, M. H. I., et al.
Veröffentlicht: (2025)
von: Abdalla, M. H. I., et al.
Veröffentlicht: (2025)
Large Language Model Can Be a Foundation for Hidden Rationale-Based Retrieval
von: Ji, Luo, et al.
Veröffentlicht: (2024)
von: Ji, Luo, et al.
Veröffentlicht: (2024)
Universal Hypernetworks for Arbitrary Models
von: Zhou, Xuanfeng
Veröffentlicht: (2026)
von: Zhou, Xuanfeng
Veröffentlicht: (2026)
Hypernetworks for Personalizing ASR to Atypical Speech
von: Müller-Eberstein, Max, et al.
Veröffentlicht: (2024)
von: Müller-Eberstein, Max, et al.
Veröffentlicht: (2024)
GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
Learning a Generative Meta-Model of LLM Activations
von: Luo, Grace, et al.
Veröffentlicht: (2026)
von: Luo, Grace, et al.
Veröffentlicht: (2026)
Node Level Graph Autoencoder: Unified Pretraining for Textual Graph Learning
von: Hu, Wenbin, et al.
Veröffentlicht: (2024)
von: Hu, Wenbin, et al.
Veröffentlicht: (2024)
Stabilizing Long-term Multi-turn Reinforcement Learning with Gated Rewards
von: Sun, Zetian, et al.
Veröffentlicht: (2025)
von: Sun, Zetian, et al.
Veröffentlicht: (2025)
HyperAdaLoRA: Accelerating LoRA Rank Allocation During Training via Hypernetworks without Sacrificing Performance
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
HyperEdit: Unlocking Instruction-based Text Editing in LLMs via Hypernetworks
von: Zeng, Yiming, et al.
Veröffentlicht: (2025)
von: Zeng, Yiming, et al.
Veröffentlicht: (2025)
Janus-Q: End-to-End Event-Driven Trading via Hierarchical-Gated Reward Modeling
von: Li, Xiang, et al.
Veröffentlicht: (2026)
von: Li, Xiang, et al.
Veröffentlicht: (2026)
Textual Gradients are a Flawed Metaphor for Automatic Prompt Optimization
von: Melcer, Daniel, et al.
Veröffentlicht: (2025)
von: Melcer, Daniel, et al.
Veröffentlicht: (2025)
FlowBot: Inducing LLM Workflows with Bilevel Optimization and Textual Gradients
von: Yu, Hongyeon, et al.
Veröffentlicht: (2026)
von: Yu, Hongyeon, et al.
Veröffentlicht: (2026)
SliM-LLM: Salience-Driven Mixed-Precision Quantization for Large Language Models
von: Huang, Wei, et al.
Veröffentlicht: (2024)
von: Huang, Wei, et al.
Veröffentlicht: (2024)
HYPEROFA: Expanding LLM Vocabulary to New Languages via Hypernetwork-Based Embedding Initialization
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
IDGenRec: LLM-RecSys Alignment with Textual ID Learning
von: Tan, Juntao, et al.
Veröffentlicht: (2024)
von: Tan, Juntao, et al.
Veröffentlicht: (2024)
HyperSteer: Activation Steering at Scale with Hypernetworks
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
PropMEND: Hypernetworks for Knowledge Propagation in LLMs
von: Liu, Zeyu Leo, et al.
Veröffentlicht: (2025)
von: Liu, Zeyu Leo, et al.
Veröffentlicht: (2025)
MetaRM: Shifted Distributions Alignment via Meta-Learning
von: Dou, Shihan, et al.
Veröffentlicht: (2024)
von: Dou, Shihan, et al.
Veröffentlicht: (2024)
HyperDAS: Towards Automating Mechanistic Interpretability with Hypernetworks
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
MAESTRO: Meta-learning Adaptive Estimation of Scalarization Trade-offs for Reward Optimization
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
Enhancing LLM Knowledge Learning through Generalization
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
EMSEdit: Efficient Multi-Step Meta-Learning-based Model Editing
von: Li, Xiaopeng, et al.
Veröffentlicht: (2025)
von: Li, Xiaopeng, et al.
Veröffentlicht: (2025)
From Descriptive to Prescriptive: Uncover the Social Value Alignment of LLM-based Agents
von: Qu, Jinxian, et al.
Veröffentlicht: (2026)
von: Qu, Jinxian, et al.
Veröffentlicht: (2026)
CARD: Towards Conditional Design of Multi-agent Topological Structures
von: Wu, Tongtong, et al.
Veröffentlicht: (2026)
von: Wu, Tongtong, et al.
Veröffentlicht: (2026)
Meta-Reinforcement Learning with Self-Reflection for Agentic Search
von: Xiao, Teng, et al.
Veröffentlicht: (2026)
von: Xiao, Teng, et al.
Veröffentlicht: (2026)
ChatTime: A Unified Multimodal Time Series Foundation Model Bridging Numerical and Textual Data
von: Wang, Chengsen, et al.
Veröffentlicht: (2024)
von: Wang, Chengsen, et al.
Veröffentlicht: (2024)
Dream to Chat: Model-based Reinforcement Learning on Dialogues with User Belief Modeling
von: Zhao, Yue, et al.
Veröffentlicht: (2025)
von: Zhao, Yue, et al.
Veröffentlicht: (2025)
Fast KVzip: Efficient and Accurate LLM Inference with Gated KV Eviction
von: Kim, Jang-Hyun, et al.
Veröffentlicht: (2026)
von: Kim, Jang-Hyun, et al.
Veröffentlicht: (2026)
Multi-Party Supervised Fine-tuning of Language Models for Multi-Party Dialogue Generation
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
MAML-en-LLM: Model Agnostic Meta-Training of LLMs for Improved In-Context Learning
von: Sinha, Sanchit, et al.
Veröffentlicht: (2024)
von: Sinha, Sanchit, et al.
Veröffentlicht: (2024)
Tokenization Multiplicity Leads to Arbitrary Price Variation in LLM-as-a-service
von: Chatzi, Ivi, et al.
Veröffentlicht: (2025)
von: Chatzi, Ivi, et al.
Veröffentlicht: (2025)
HyperTTS: Parameter Efficient Adaptation in Text to Speech using Hypernetworks
von: Li, Yingting, et al.
Veröffentlicht: (2024)
von: Li, Yingting, et al.
Veröffentlicht: (2024)
$\textbf{AGT$^{AO}$}$: Robust and Stabilized LLM Unlearning via Adversarial Gating Training with Adaptive Orthogonality
von: Li, Pengyu, et al.
Veröffentlicht: (2026)
von: Li, Pengyu, et al.
Veröffentlicht: (2026)
IGC: Integrating a Gated Calculator into an LLM to Solve Arithmetic Tasks Reliably and Efficiently
von: Dietz, Florian, et al.
Veröffentlicht: (2025)
von: Dietz, Florian, et al.
Veröffentlicht: (2025)
Faithful and Robust Local Interpretability for Textual Predictions
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2023)
von: Lopardo, Gianluigi, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
A Practice of Post-Training on Llama-3 70B with Optimal Selection of Additional Language Mixture Ratio
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024) -
LaMsS: When Large Language Models Meet Self-Skepticism
von: Wu, Yetao, et al.
Veröffentlicht: (2024) -
Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA
von: Chen, Teng, et al.
Veröffentlicht: (2026) -
MeTHanol: Modularized Thinking Language Models with Intermediate Layer Thinking, Decoding and Bootstrapping Reasoning
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024) -
Zhyper: Factorized Hypernetworks for Conditioned LLM Fine-Tuning
von: Abdalla, M. H. I., et al.
Veröffentlicht: (2025)