Equivariant Neural Functional Networks for Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Tran, Viet-Hoang, Vo, Thieu N., The, An Nguyen, Huu, Tho Tran, Nguyen-Nhat, Minh-Khoi, Tran, Thanh, Pham, Duy-Tung, Nguyen, Tan Minh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Equivariant Polynomial Functional Networks
by: Vo, Thieu N., et al.
Published: (2024)
by: Vo, Thieu N., et al.
Published: (2024)
A Clifford Algebraic Approach to E(n)-Equivariant High-order Graph Neural Networks
by: Tran, Viet-Hoang, et al.
Published: (2024)
by: Tran, Viet-Hoang, et al.
Published: (2024)
Monomial Matrix Group Equivariant Neural Functional Networks
by: Tran, Viet-Hoang, et al.
Published: (2024)
by: Tran, Viet-Hoang, et al.
Published: (2024)
Tree-Sliced Wasserstein Distance: A Geometric Perspective
by: Tran, Viet-Hoang, et al.
Published: (2024)
by: Tran, Viet-Hoang, et al.
Published: (2024)
Quasi-Equivariant Metanetworks
by: Tran, Viet-Hoang, et al.
Published: (2026)
by: Tran, Viet-Hoang, et al.
Published: (2026)
Revisiting LARS for Large Batch Training Generalization of Neural Networks
by: Do, Khoi, et al.
Published: (2023)
by: Do, Khoi, et al.
Published: (2023)
Revisiting Kernel Attention with Correlated Gaussian Process Representation
by: Bui, Long Minh, et al.
Published: (2025)
by: Bui, Long Minh, et al.
Published: (2025)
Dynamical Properties of Tokens in Self-Attention and Effects of Positional Encoding
by: Pham, Duy-Tung, et al.
Published: (2025)
by: Pham, Duy-Tung, et al.
Published: (2025)
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
by: Nguyen-Nhat, Minh-Khoi, et al.
Published: (2025)
by: Nguyen-Nhat, Minh-Khoi, et al.
Published: (2025)
Neural Collapse for Cross-entropy Class-Imbalanced Learning with Unconstrained ReLU Feature Model
by: Dang, Hien, et al.
Published: (2024)
by: Dang, Hien, et al.
Published: (2024)
Demystifying the Token Dynamics of Deep Selective State Space Models
by: Vo, Thieu N, et al.
Published: (2024)
by: Vo, Thieu N, et al.
Published: (2024)
Knowledge Abstraction for Knowledge-based Semantic Communication: A Generative Causality Invariant Approach
by: Nguyen, Minh-Duong, et al.
Published: (2025)
by: Nguyen, Minh-Duong, et al.
Published: (2025)
Spherical Tree-Sliced Wasserstein Distance
by: Tran, Viet-Hoang, et al.
Published: (2025)
by: Tran, Viet-Hoang, et al.
Published: (2025)
Beyond Vanilla Variational Autoencoders: Detecting Posterior Collapse in Conditional and Hierarchical Variational Autoencoders
by: Dang, Hien, et al.
Published: (2023)
by: Dang, Hien, et al.
Published: (2023)
Using Synthetic Data to estimate the True Error is theoretically and practically doable
by: Thanh, Hai Hoang, et al.
Published: (2025)
by: Thanh, Hai Hoang, et al.
Published: (2025)
Wasserstein Gaussianization and Efficient Variational Bayes for Robust Bayesian Synthetic Likelihood
by: Nguyen, Nhat-Minh, et al.
Published: (2023)
by: Nguyen, Nhat-Minh, et al.
Published: (2023)
Control, Optimal Transport and Neural Differential Equations in Supervised Learning
by: Phung, Minh-Nhat, et al.
Published: (2025)
by: Phung, Minh-Nhat, et al.
Published: (2025)
Distance-Based Tree-Sliced Wasserstein Distance
by: Tran, Hoang V., et al.
Published: (2025)
by: Tran, Hoang V., et al.
Published: (2025)
Enhancing Tropical Cyclone Path Forecasting with an Improved Transformer Network
by: Van Thanh, Nguyen, et al.
Published: (2025)
by: Van Thanh, Nguyen, et al.
Published: (2025)
iMoT: Inertial Motion Transformer for Inertial Navigation
by: Nguyen, Son Minh, et al.
Published: (2024)
by: Nguyen, Son Minh, et al.
Published: (2024)
Tree-Sliced Wasserstein Distance with Nonlinear Projection
by: Tran, Thanh, et al.
Published: (2025)
by: Tran, Thanh, et al.
Published: (2025)
Rethinking Multinomial Logistic Mixture of Experts with Sigmoid Gating Function
by: Pham, Tuan Minh, et al.
Published: (2026)
by: Pham, Tuan Minh, et al.
Published: (2026)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
MIST: Reliable Streaming Decision Trees for Online Class-Incremental Learning via McDiarmid Bound
by: Pham, Phu-Hoa, et al.
Published: (2026)
by: Pham, Phu-Hoa, et al.
Published: (2026)
Highly Efficient and Effective LLMs with Multi-Boolean Architectures
by: Tran, Ba-Hien, et al.
Published: (2025)
by: Tran, Ba-Hien, et al.
Published: (2025)
Leveraging Model Soups to Classify Intangible Cultural Heritage Images from the Mekong Delta
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
Revisiting Prefix-tuning: Statistical Benefits of Reparameterization among Prompts
by: Le, Minh, et al.
Published: (2024)
by: Le, Minh, et al.
Published: (2024)
Empowering Contrastive Federated Sequential Recommendation with LLMs
by: Nguyen, Thi Minh Chau, et al.
Published: (2026)
by: Nguyen, Thi Minh Chau, et al.
Published: (2026)
Understanding SAM's Robustness to Noisy Labels through Gradient Down-weighting
by: Luong, Hoang-Chau, et al.
Published: (2024)
by: Luong, Hoang-Chau, et al.
Published: (2024)
Unlocking Compositional Generalization in Continual Few-Shot Learning
by: Nguyen-Lam, Phu-Quy, et al.
Published: (2026)
by: Nguyen-Lam, Phu-Quy, et al.
Published: (2026)
Statistical Inference for Autoencoder-based Anomaly Detection after Representation Learning-based Domain Adaptation
by: Kiet, Tran Tuan, et al.
Published: (2025)
by: Kiet, Tran Tuan, et al.
Published: (2025)
FTIR, TGA/DSC, BET, and WCA analyses of materials, as well as effects of parameters on isolating cellulose fibers from reed
by: Le Thi Thanh, Huong, et al.
Published: (2024)
by: Le Thi Thanh, Huong, et al.
Published: (2024)
FTIR, BET, and DTA of materials and effects of parameters to isolate cellulose fibers from reed
by: Le Thi Thanh, Huong, et al.
Published: (2024)
by: Le Thi Thanh, Huong, et al.
Published: (2024)
MP-PINN: A Multi-Phase Physics-Informed Neural Network for Epidemic Forecasting
by: Nguyen, Thang, et al.
Published: (2024)
by: Nguyen, Thang, et al.
Published: (2024)
One-Prompt Strikes Back: Sparse Mixture of Experts for Prompt-based Continual Learning
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
Enhanced Multimodal Video Retrieval System: Integrating Query Expansion and Cross-modal Temporal Event Retrieval
by: Vo, Van-Thinh, et al.
Published: (2025)
by: Vo, Van-Thinh, et al.
Published: (2025)
Curvature-Aware Safety Restoration In LLMs Fine-Tuning
by: Bach, Thong, et al.
Published: (2025)
by: Bach, Thong, et al.
Published: (2025)
Building a temperature forecasting model for the city with the regression neural network (RNN)
by: Tran, Nguyen Phuc, et al.
Published: (2024)
by: Tran, Nguyen Phuc, et al.
Published: (2024)
CAMEx: Curvature-aware Merging of Experts
by: Nguyen, Dung V., et al.
Published: (2025)
by: Nguyen, Dung V., et al.
Published: (2025)
A Hyper-Transformer model for Controllable Pareto Front Learning with Split Feasibility Constraints
by: Tuan, Tran Anh, et al.
Published: (2024)
by: Tuan, Tran Anh, et al.
Published: (2024)
Similar Items
-
Equivariant Polynomial Functional Networks
by: Vo, Thieu N., et al.
Published: (2024) -
A Clifford Algebraic Approach to E(n)-Equivariant High-order Graph Neural Networks
by: Tran, Viet-Hoang, et al.
Published: (2024) -
Monomial Matrix Group Equivariant Neural Functional Networks
by: Tran, Viet-Hoang, et al.
Published: (2024) -
Tree-Sliced Wasserstein Distance: A Geometric Perspective
by: Tran, Viet-Hoang, et al.
Published: (2024) -
Quasi-Equivariant Metanetworks
by: Tran, Viet-Hoang, et al.
Published: (2026)