Equivariant Neural Functional Networks for Transformers
Fuente:
arXiv
Guardado en:
| Autores principales: | Tran, Viet-Hoang, Vo, Thieu N., The, An Nguyen, Huu, Tho Tran, Nguyen-Nhat, Minh-Khoi, Tran, Thanh, Pham, Duy-Tung, Nguyen, Tan Minh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Equivariant Polynomial Functional Networks
por: Vo, Thieu N., et al.
Publicado: (2024)
por: Vo, Thieu N., et al.
Publicado: (2024)
A Clifford Algebraic Approach to E(n)-Equivariant High-order Graph Neural Networks
por: Tran, Viet-Hoang, et al.
Publicado: (2024)
por: Tran, Viet-Hoang, et al.
Publicado: (2024)
Monomial Matrix Group Equivariant Neural Functional Networks
por: Tran, Viet-Hoang, et al.
Publicado: (2024)
por: Tran, Viet-Hoang, et al.
Publicado: (2024)
Tree-Sliced Wasserstein Distance: A Geometric Perspective
por: Tran, Viet-Hoang, et al.
Publicado: (2024)
por: Tran, Viet-Hoang, et al.
Publicado: (2024)
Quasi-Equivariant Metanetworks
por: Tran, Viet-Hoang, et al.
Publicado: (2026)
por: Tran, Viet-Hoang, et al.
Publicado: (2026)
Revisiting LARS for Large Batch Training Generalization of Neural Networks
por: Do, Khoi, et al.
Publicado: (2023)
por: Do, Khoi, et al.
Publicado: (2023)
Revisiting Kernel Attention with Correlated Gaussian Process Representation
por: Bui, Long Minh, et al.
Publicado: (2025)
por: Bui, Long Minh, et al.
Publicado: (2025)
Dynamical Properties of Tokens in Self-Attention and Effects of Positional Encoding
por: Pham, Duy-Tung, et al.
Publicado: (2025)
por: Pham, Duy-Tung, et al.
Publicado: (2025)
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
por: Nguyen-Nhat, Minh-Khoi, et al.
Publicado: (2025)
por: Nguyen-Nhat, Minh-Khoi, et al.
Publicado: (2025)
Neural Collapse for Cross-entropy Class-Imbalanced Learning with Unconstrained ReLU Feature Model
por: Dang, Hien, et al.
Publicado: (2024)
por: Dang, Hien, et al.
Publicado: (2024)
Demystifying the Token Dynamics of Deep Selective State Space Models
por: Vo, Thieu N, et al.
Publicado: (2024)
por: Vo, Thieu N, et al.
Publicado: (2024)
Knowledge Abstraction for Knowledge-based Semantic Communication: A Generative Causality Invariant Approach
por: Nguyen, Minh-Duong, et al.
Publicado: (2025)
por: Nguyen, Minh-Duong, et al.
Publicado: (2025)
Spherical Tree-Sliced Wasserstein Distance
por: Tran, Viet-Hoang, et al.
Publicado: (2025)
por: Tran, Viet-Hoang, et al.
Publicado: (2025)
Beyond Vanilla Variational Autoencoders: Detecting Posterior Collapse in Conditional and Hierarchical Variational Autoencoders
por: Dang, Hien, et al.
Publicado: (2023)
por: Dang, Hien, et al.
Publicado: (2023)
Using Synthetic Data to estimate the True Error is theoretically and practically doable
por: Thanh, Hai Hoang, et al.
Publicado: (2025)
por: Thanh, Hai Hoang, et al.
Publicado: (2025)
Wasserstein Gaussianization and Efficient Variational Bayes for Robust Bayesian Synthetic Likelihood
por: Nguyen, Nhat-Minh, et al.
Publicado: (2023)
por: Nguyen, Nhat-Minh, et al.
Publicado: (2023)
Control, Optimal Transport and Neural Differential Equations in Supervised Learning
por: Phung, Minh-Nhat, et al.
Publicado: (2025)
por: Phung, Minh-Nhat, et al.
Publicado: (2025)
Distance-Based Tree-Sliced Wasserstein Distance
por: Tran, Hoang V., et al.
Publicado: (2025)
por: Tran, Hoang V., et al.
Publicado: (2025)
Enhancing Tropical Cyclone Path Forecasting with an Improved Transformer Network
por: Van Thanh, Nguyen, et al.
Publicado: (2025)
por: Van Thanh, Nguyen, et al.
Publicado: (2025)
iMoT: Inertial Motion Transformer for Inertial Navigation
por: Nguyen, Son Minh, et al.
Publicado: (2024)
por: Nguyen, Son Minh, et al.
Publicado: (2024)
Tree-Sliced Wasserstein Distance with Nonlinear Projection
por: Tran, Thanh, et al.
Publicado: (2025)
por: Tran, Thanh, et al.
Publicado: (2025)
Rethinking Multinomial Logistic Mixture of Experts with Sigmoid Gating Function
por: Pham, Tuan Minh, et al.
Publicado: (2026)
por: Pham, Tuan Minh, et al.
Publicado: (2026)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
por: Le, Minh, et al.
Publicado: (2025)
por: Le, Minh, et al.
Publicado: (2025)
MIST: Reliable Streaming Decision Trees for Online Class-Incremental Learning via McDiarmid Bound
por: Pham, Phu-Hoa, et al.
Publicado: (2026)
por: Pham, Phu-Hoa, et al.
Publicado: (2026)
Highly Efficient and Effective LLMs with Multi-Boolean Architectures
por: Tran, Ba-Hien, et al.
Publicado: (2025)
por: Tran, Ba-Hien, et al.
Publicado: (2025)
Leveraging Model Soups to Classify Intangible Cultural Heritage Images from the Mekong Delta
por: Tran, Quoc-Khang, et al.
Publicado: (2026)
por: Tran, Quoc-Khang, et al.
Publicado: (2026)
Revisiting Prefix-tuning: Statistical Benefits of Reparameterization among Prompts
por: Le, Minh, et al.
Publicado: (2024)
por: Le, Minh, et al.
Publicado: (2024)
Empowering Contrastive Federated Sequential Recommendation with LLMs
por: Nguyen, Thi Minh Chau, et al.
Publicado: (2026)
por: Nguyen, Thi Minh Chau, et al.
Publicado: (2026)
Understanding SAM's Robustness to Noisy Labels through Gradient Down-weighting
por: Luong, Hoang-Chau, et al.
Publicado: (2024)
por: Luong, Hoang-Chau, et al.
Publicado: (2024)
Unlocking Compositional Generalization in Continual Few-Shot Learning
por: Nguyen-Lam, Phu-Quy, et al.
Publicado: (2026)
por: Nguyen-Lam, Phu-Quy, et al.
Publicado: (2026)
Statistical Inference for Autoencoder-based Anomaly Detection after Representation Learning-based Domain Adaptation
por: Kiet, Tran Tuan, et al.
Publicado: (2025)
por: Kiet, Tran Tuan, et al.
Publicado: (2025)
FTIR, TGA/DSC, BET, and WCA analyses of materials, as well as effects of parameters on isolating cellulose fibers from reed
por: Le Thi Thanh, Huong, et al.
Publicado: (2024)
por: Le Thi Thanh, Huong, et al.
Publicado: (2024)
FTIR, BET, and DTA of materials and effects of parameters to isolate cellulose fibers from reed
por: Le Thi Thanh, Huong, et al.
Publicado: (2024)
por: Le Thi Thanh, Huong, et al.
Publicado: (2024)
MP-PINN: A Multi-Phase Physics-Informed Neural Network for Epidemic Forecasting
por: Nguyen, Thang, et al.
Publicado: (2024)
por: Nguyen, Thang, et al.
Publicado: (2024)
One-Prompt Strikes Back: Sparse Mixture of Experts for Prompt-based Continual Learning
por: Le, Minh, et al.
Publicado: (2025)
por: Le, Minh, et al.
Publicado: (2025)
Enhanced Multimodal Video Retrieval System: Integrating Query Expansion and Cross-modal Temporal Event Retrieval
por: Vo, Van-Thinh, et al.
Publicado: (2025)
por: Vo, Van-Thinh, et al.
Publicado: (2025)
Curvature-Aware Safety Restoration In LLMs Fine-Tuning
por: Bach, Thong, et al.
Publicado: (2025)
por: Bach, Thong, et al.
Publicado: (2025)
Building a temperature forecasting model for the city with the regression neural network (RNN)
por: Tran, Nguyen Phuc, et al.
Publicado: (2024)
por: Tran, Nguyen Phuc, et al.
Publicado: (2024)
CAMEx: Curvature-aware Merging of Experts
por: Nguyen, Dung V., et al.
Publicado: (2025)
por: Nguyen, Dung V., et al.
Publicado: (2025)
A Hyper-Transformer model for Controllable Pareto Front Learning with Split Feasibility Constraints
por: Tuan, Tran Anh, et al.
Publicado: (2024)
por: Tuan, Tran Anh, et al.
Publicado: (2024)
Ejemplares similares
-
Equivariant Polynomial Functional Networks
por: Vo, Thieu N., et al.
Publicado: (2024) -
A Clifford Algebraic Approach to E(n)-Equivariant High-order Graph Neural Networks
por: Tran, Viet-Hoang, et al.
Publicado: (2024) -
Monomial Matrix Group Equivariant Neural Functional Networks
por: Tran, Viet-Hoang, et al.
Publicado: (2024) -
Tree-Sliced Wasserstein Distance: A Geometric Perspective
por: Tran, Viet-Hoang, et al.
Publicado: (2024) -
Quasi-Equivariant Metanetworks
por: Tran, Viet-Hoang, et al.
Publicado: (2026)