Equivariant Neural Functional Networks for Transformers
Fuente:
arXiv
Salvato in:
| Autori principali: | Tran, Viet-Hoang, Vo, Thieu N., The, An Nguyen, Huu, Tho Tran, Nguyen-Nhat, Minh-Khoi, Tran, Thanh, Pham, Duy-Tung, Nguyen, Tan Minh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Equivariant Polynomial Functional Networks
di: Vo, Thieu N., et al.
Pubblicazione: (2024)
di: Vo, Thieu N., et al.
Pubblicazione: (2024)
A Clifford Algebraic Approach to E(n)-Equivariant High-order Graph Neural Networks
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2024)
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2024)
Monomial Matrix Group Equivariant Neural Functional Networks
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2024)
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2024)
Tree-Sliced Wasserstein Distance: A Geometric Perspective
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2024)
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2024)
Quasi-Equivariant Metanetworks
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2026)
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2026)
Revisiting LARS for Large Batch Training Generalization of Neural Networks
di: Do, Khoi, et al.
Pubblicazione: (2023)
di: Do, Khoi, et al.
Pubblicazione: (2023)
Revisiting Kernel Attention with Correlated Gaussian Process Representation
di: Bui, Long Minh, et al.
Pubblicazione: (2025)
di: Bui, Long Minh, et al.
Pubblicazione: (2025)
Dynamical Properties of Tokens in Self-Attention and Effects of Positional Encoding
di: Pham, Duy-Tung, et al.
Pubblicazione: (2025)
di: Pham, Duy-Tung, et al.
Pubblicazione: (2025)
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
di: Nguyen-Nhat, Minh-Khoi, et al.
Pubblicazione: (2025)
di: Nguyen-Nhat, Minh-Khoi, et al.
Pubblicazione: (2025)
Neural Collapse for Cross-entropy Class-Imbalanced Learning with Unconstrained ReLU Feature Model
di: Dang, Hien, et al.
Pubblicazione: (2024)
di: Dang, Hien, et al.
Pubblicazione: (2024)
Demystifying the Token Dynamics of Deep Selective State Space Models
di: Vo, Thieu N, et al.
Pubblicazione: (2024)
di: Vo, Thieu N, et al.
Pubblicazione: (2024)
Knowledge Abstraction for Knowledge-based Semantic Communication: A Generative Causality Invariant Approach
di: Nguyen, Minh-Duong, et al.
Pubblicazione: (2025)
di: Nguyen, Minh-Duong, et al.
Pubblicazione: (2025)
Spherical Tree-Sliced Wasserstein Distance
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2025)
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2025)
Beyond Vanilla Variational Autoencoders: Detecting Posterior Collapse in Conditional and Hierarchical Variational Autoencoders
di: Dang, Hien, et al.
Pubblicazione: (2023)
di: Dang, Hien, et al.
Pubblicazione: (2023)
Using Synthetic Data to estimate the True Error is theoretically and practically doable
di: Thanh, Hai Hoang, et al.
Pubblicazione: (2025)
di: Thanh, Hai Hoang, et al.
Pubblicazione: (2025)
Wasserstein Gaussianization and Efficient Variational Bayes for Robust Bayesian Synthetic Likelihood
di: Nguyen, Nhat-Minh, et al.
Pubblicazione: (2023)
di: Nguyen, Nhat-Minh, et al.
Pubblicazione: (2023)
Control, Optimal Transport and Neural Differential Equations in Supervised Learning
di: Phung, Minh-Nhat, et al.
Pubblicazione: (2025)
di: Phung, Minh-Nhat, et al.
Pubblicazione: (2025)
Distance-Based Tree-Sliced Wasserstein Distance
di: Tran, Hoang V., et al.
Pubblicazione: (2025)
di: Tran, Hoang V., et al.
Pubblicazione: (2025)
Enhancing Tropical Cyclone Path Forecasting with an Improved Transformer Network
di: Van Thanh, Nguyen, et al.
Pubblicazione: (2025)
di: Van Thanh, Nguyen, et al.
Pubblicazione: (2025)
iMoT: Inertial Motion Transformer for Inertial Navigation
di: Nguyen, Son Minh, et al.
Pubblicazione: (2024)
di: Nguyen, Son Minh, et al.
Pubblicazione: (2024)
Tree-Sliced Wasserstein Distance with Nonlinear Projection
di: Tran, Thanh, et al.
Pubblicazione: (2025)
di: Tran, Thanh, et al.
Pubblicazione: (2025)
Rethinking Multinomial Logistic Mixture of Experts with Sigmoid Gating Function
di: Pham, Tuan Minh, et al.
Pubblicazione: (2026)
di: Pham, Tuan Minh, et al.
Pubblicazione: (2026)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
di: Le, Minh, et al.
Pubblicazione: (2025)
di: Le, Minh, et al.
Pubblicazione: (2025)
MIST: Reliable Streaming Decision Trees for Online Class-Incremental Learning via McDiarmid Bound
di: Pham, Phu-Hoa, et al.
Pubblicazione: (2026)
di: Pham, Phu-Hoa, et al.
Pubblicazione: (2026)
Highly Efficient and Effective LLMs with Multi-Boolean Architectures
di: Tran, Ba-Hien, et al.
Pubblicazione: (2025)
di: Tran, Ba-Hien, et al.
Pubblicazione: (2025)
Leveraging Model Soups to Classify Intangible Cultural Heritage Images from the Mekong Delta
di: Tran, Quoc-Khang, et al.
Pubblicazione: (2026)
di: Tran, Quoc-Khang, et al.
Pubblicazione: (2026)
Revisiting Prefix-tuning: Statistical Benefits of Reparameterization among Prompts
di: Le, Minh, et al.
Pubblicazione: (2024)
di: Le, Minh, et al.
Pubblicazione: (2024)
Empowering Contrastive Federated Sequential Recommendation with LLMs
di: Nguyen, Thi Minh Chau, et al.
Pubblicazione: (2026)
di: Nguyen, Thi Minh Chau, et al.
Pubblicazione: (2026)
Understanding SAM's Robustness to Noisy Labels through Gradient Down-weighting
di: Luong, Hoang-Chau, et al.
Pubblicazione: (2024)
di: Luong, Hoang-Chau, et al.
Pubblicazione: (2024)
Unlocking Compositional Generalization in Continual Few-Shot Learning
di: Nguyen-Lam, Phu-Quy, et al.
Pubblicazione: (2026)
di: Nguyen-Lam, Phu-Quy, et al.
Pubblicazione: (2026)
Statistical Inference for Autoencoder-based Anomaly Detection after Representation Learning-based Domain Adaptation
di: Kiet, Tran Tuan, et al.
Pubblicazione: (2025)
di: Kiet, Tran Tuan, et al.
Pubblicazione: (2025)
FTIR, TGA/DSC, BET, and WCA analyses of materials, as well as effects of parameters on isolating cellulose fibers from reed
di: Le Thi Thanh, Huong, et al.
Pubblicazione: (2024)
di: Le Thi Thanh, Huong, et al.
Pubblicazione: (2024)
FTIR, BET, and DTA of materials and effects of parameters to isolate cellulose fibers from reed
di: Le Thi Thanh, Huong, et al.
Pubblicazione: (2024)
di: Le Thi Thanh, Huong, et al.
Pubblicazione: (2024)
MP-PINN: A Multi-Phase Physics-Informed Neural Network for Epidemic Forecasting
di: Nguyen, Thang, et al.
Pubblicazione: (2024)
di: Nguyen, Thang, et al.
Pubblicazione: (2024)
One-Prompt Strikes Back: Sparse Mixture of Experts for Prompt-based Continual Learning
di: Le, Minh, et al.
Pubblicazione: (2025)
di: Le, Minh, et al.
Pubblicazione: (2025)
Enhanced Multimodal Video Retrieval System: Integrating Query Expansion and Cross-modal Temporal Event Retrieval
di: Vo, Van-Thinh, et al.
Pubblicazione: (2025)
di: Vo, Van-Thinh, et al.
Pubblicazione: (2025)
Curvature-Aware Safety Restoration In LLMs Fine-Tuning
di: Bach, Thong, et al.
Pubblicazione: (2025)
di: Bach, Thong, et al.
Pubblicazione: (2025)
Building a temperature forecasting model for the city with the regression neural network (RNN)
di: Tran, Nguyen Phuc, et al.
Pubblicazione: (2024)
di: Tran, Nguyen Phuc, et al.
Pubblicazione: (2024)
CAMEx: Curvature-aware Merging of Experts
di: Nguyen, Dung V., et al.
Pubblicazione: (2025)
di: Nguyen, Dung V., et al.
Pubblicazione: (2025)
A Hyper-Transformer model for Controllable Pareto Front Learning with Split Feasibility Constraints
di: Tuan, Tran Anh, et al.
Pubblicazione: (2024)
di: Tuan, Tran Anh, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Equivariant Polynomial Functional Networks
di: Vo, Thieu N., et al.
Pubblicazione: (2024) -
A Clifford Algebraic Approach to E(n)-Equivariant High-order Graph Neural Networks
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2024) -
Monomial Matrix Group Equivariant Neural Functional Networks
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2024) -
Tree-Sliced Wasserstein Distance: A Geometric Perspective
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2024) -
Quasi-Equivariant Metanetworks
di: Tran, Viet-Hoang, et al.
Pubblicazione: (2026)