Gradient Transformer: Learning to Generate Updates for LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Binh-Nguyen, Tran, Khang, Phan, NhatHai, Khalil, Issa |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SGFusion: Stochastic Geographic Gradient Fusion in Federated Learning
by: Nguyen, Khoa, et al.
Published: (2025)
by: Nguyen, Khoa, et al.
Published: (2025)
PRPO: Paragraph-level Policy Optimization for Vision-Language Deepfake Detection
by: Nguyen, Tuan, et al.
Published: (2025)
by: Nguyen, Tuan, et al.
Published: (2025)
Program Structure-aware Language Models: Targeted Software Testing beyond Textual Semantics
by: Tran, Khang, et al.
Published: (2026)
by: Tran, Khang, et al.
Published: (2026)
PromSec: Prompt Optimization for Secure Generation of Functional Source Code with Large Language Models (LLMs)
by: Nazzal, Mahmoud, et al.
Published: (2024)
by: Nazzal, Mahmoud, et al.
Published: (2024)
Poison with Style: A Practical Poisoning Attack on Code Large Language Models
by: Tran, Khang, et al.
Published: (2026)
by: Tran, Khang, et al.
Published: (2026)
FairDP: Certified Fairness with Differential Privacy
by: Tran, Khang, et al.
Published: (2023)
by: Tran, Khang, et al.
Published: (2023)
A Client-level Assessment of Collaborative Backdoor Poisoning in Non-IID Federated Learning
by: Lai, Phung, et al.
Published: (2025)
by: Lai, Phung, et al.
Published: (2025)
ViGText: Deepfake Image Detection with Vision-Language Model Explanations and Graph Neural Networks
by: ALBarqawi, Ahmad, et al.
Published: (2025)
by: ALBarqawi, Ahmad, et al.
Published: (2025)
NOIR: Privacy-Preserving Generation of Code with Open-Source LLMs
by: Nguyen, Khoa, et al.
Published: (2026)
by: Nguyen, Khoa, et al.
Published: (2026)
Watermarking Degrades Alignment in Language Models: Analysis and Mitigation
by: Verma, Apurv, et al.
Published: (2025)
by: Verma, Apurv, et al.
Published: (2025)
MELCOT: A Hybrid Learning Architecture with Marginal Preservation for Matrix-Valued Regression
by: Tran, Khang, et al.
Published: (2025)
by: Tran, Khang, et al.
Published: (2025)
Fast Estimation of Wasserstein Distances via Regression on Sliced Wasserstein Distances
by: Nguyen, Khai, et al.
Published: (2025)
by: Nguyen, Khai, et al.
Published: (2025)
Attack On Prompt: Backdoor Attack in Prompt-Based Continual Learning
by: Nguyen, Trang, et al.
Published: (2024)
by: Nguyen, Trang, et al.
Published: (2024)
Towards Marginal Fairness Sliced Wasserstein Barycenter
by: Nguyen, Khai, et al.
Published: (2024)
by: Nguyen, Khai, et al.
Published: (2024)
Control, Optimal Transport and Neural Differential Equations in Supervised Learning
by: Phung, Minh-Nhat, et al.
Published: (2025)
by: Phung, Minh-Nhat, et al.
Published: (2025)
Leveraging Model Soups to Classify Intangible Cultural Heritage Images from the Mekong Delta
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
GROOT: Effective Design of Biological Sequences with Limited Experimental Data
by: Tran, Thanh V. T., et al.
Published: (2024)
by: Tran, Thanh V. T., et al.
Published: (2024)
Equivariant Neural Functional Networks for Transformers
by: Tran, Viet-Hoang, et al.
Published: (2024)
by: Tran, Viet-Hoang, et al.
Published: (2024)
Dendrograms of Mixing Measures for Softmax-Gated Gaussian Mixture of Experts: Consistency without Model Sweeps
by: Hai, Do Tien, et al.
Published: (2025)
by: Hai, Do Tien, et al.
Published: (2025)
Range-aware Positional Encoding via High-order Pretraining: Theory and Practice
by: Nguyen, Viet Anh, et al.
Published: (2024)
by: Nguyen, Viet Anh, et al.
Published: (2024)
Neural Collapse for Cross-entropy Class-Imbalanced Learning with Unconstrained ReLU Feature Model
by: Dang, Hien, et al.
Published: (2024)
by: Dang, Hien, et al.
Published: (2024)
Generative Conditional Distributions by Neural (Entropic) Optimal Transport
by: Nguyen, Bao, et al.
Published: (2024)
by: Nguyen, Bao, et al.
Published: (2024)
Demo: SGCode: A Flexible Prompt-Optimizing System for Secure Generation of Code
by: Ton, Khiem, et al.
Published: (2024)
by: Ton, Khiem, et al.
Published: (2024)
ECG-RAMBA: Zero-Shot ECG Generalization by Morphology-Rhythm Disentanglement and Long-Range Modeling
by: Nguyen, Hai Duong, et al.
Published: (2025)
by: Nguyen, Hai Duong, et al.
Published: (2025)
One-Prompt Strikes Back: Sparse Mixture of Experts for Prompt-based Continual Learning
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
MAGPrompt: Message-Adaptive Graph Prompt Tuning for Graph Neural Networks
by: Nguyen, Long D., et al.
Published: (2026)
by: Nguyen, Long D., et al.
Published: (2026)
Revisiting Incremental Stochastic Majorization-Minimization Algorithms with Applications to Mixture of Experts
by: Tran, TrungKhang, et al.
Published: (2026)
by: Tran, TrungKhang, et al.
Published: (2026)
Lightspeed Geometric Dataset Distance via Sliced Optimal Transport
by: Nguyen, Khai, et al.
Published: (2025)
by: Nguyen, Khai, et al.
Published: (2025)
A Relative Ignorability Framework for Decision-Relevant Observability in Control Theory and Reinforcement Learning
by: Bleile, MaryLena, et al.
Published: (2025)
by: Bleile, MaryLena, et al.
Published: (2025)
High-dimensional Many-to-many-to-many Mediation Analysis
by: Nguyen, Tien Dat, et al.
Published: (2026)
by: Nguyen, Tien Dat, et al.
Published: (2026)
Beyond Vanilla Variational Autoencoders: Detecting Posterior Collapse in Conditional and Hierarchical Variational Autoencoders
by: Dang, Hien, et al.
Published: (2023)
by: Dang, Hien, et al.
Published: (2023)
Accelerating Transformers with Spectrum-Preserving Token Merging
by: Tran, Hoai-Chau, et al.
Published: (2024)
by: Tran, Hoai-Chau, et al.
Published: (2024)
On Parameter Estimation in Deviated Gaussian Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Clustering-based Meta Bayesian Optimization with Theoretical Guarantee
by: Nguyen, Khoa, et al.
Published: (2025)
by: Nguyen, Khoa, et al.
Published: (2025)
FedX: Adaptive Model Decomposition and Quantization for IoT Federated Learning
by: Lai, Phung, et al.
Published: (2025)
by: Lai, Phung, et al.
Published: (2025)
A General Theory for Softmax Gating Multinomial Logistic Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2023)
by: Nguyen, Huy, et al.
Published: (2023)
HUANet: Hard-Constrained Unrolled ADMM for Constrained Convex Optimization
by: Tran, Trinh, et al.
Published: (2026)
by: Tran, Trinh, et al.
Published: (2026)
Leveraging Hierarchical Taxonomies in Prompt-based Continual Learning
by: Tran, Quyen, et al.
Published: (2024)
by: Tran, Quyen, et al.
Published: (2024)
Structure- and Stability-Preserving Learning of Port-Hamiltonian Systems
by: Nguyen, Binh, et al.
Published: (2026)
by: Nguyen, Binh, et al.
Published: (2026)
Adaptive Rainfall Forecasting from Multiple Geographical Models Using Matrix Profile and Ensemble Learning
by: Tran, Dung T., et al.
Published: (2025)
by: Tran, Dung T., et al.
Published: (2025)
Similar Items
-
SGFusion: Stochastic Geographic Gradient Fusion in Federated Learning
by: Nguyen, Khoa, et al.
Published: (2025) -
PRPO: Paragraph-level Policy Optimization for Vision-Language Deepfake Detection
by: Nguyen, Tuan, et al.
Published: (2025) -
Program Structure-aware Language Models: Targeted Software Testing beyond Textual Semantics
by: Tran, Khang, et al.
Published: (2026) -
PromSec: Prompt Optimization for Secure Generation of Functional Source Code with Large Language Models (LLMs)
by: Nazzal, Mahmoud, et al.
Published: (2024) -
Poison with Style: A Practical Poisoning Attack on Code Large Language Models
by: Tran, Khang, et al.
Published: (2026)