multivariateGPT: a decoder-only transformer for multivariate categorical and numeric data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Loza, Andrew J., Kim, Jun Yup, Song, Shangzheng, Liu, Yihang, Sung, Joseph J. Y., Taylor, R Andrew, Shung, Dennis L. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Explicit Dropout: Deterministic Regularization for Transformer Architectures
von: Agrawal, Vidhi, et al.
Veröffentlicht: (2026)
von: Agrawal, Vidhi, et al.
Veröffentlicht: (2026)
Massive Redundancy in Gradient Transport Enables Sparse Online Learning
von: Merin, Aur Shalev
Veröffentlicht: (2026)
von: Merin, Aur Shalev
Veröffentlicht: (2026)
Optimizing Inference in Transformer-Based Models: A Multi-Method Benchmark
von: Ho, Siu Hang, et al.
Veröffentlicht: (2025)
von: Ho, Siu Hang, et al.
Veröffentlicht: (2025)
Is Cambodia the World's Largest Cashew Producer?
von: Chaya, Veasna, et al.
Veröffentlicht: (2024)
von: Chaya, Veasna, et al.
Veröffentlicht: (2024)
CellARC: Measuring Intelligence with Cellular Automata
von: Lžičař, Miroslav
Veröffentlicht: (2025)
von: Lžičař, Miroslav
Veröffentlicht: (2025)
torchsom: The Reference PyTorch Library for Self-Organizing Maps
von: Berthier, Louis, et al.
Veröffentlicht: (2025)
von: Berthier, Louis, et al.
Veröffentlicht: (2025)
H-Model: Dynamic Neural Architectures for Adaptive Processing
von: Hospodarchuk, Dmytro
Veröffentlicht: (2025)
von: Hospodarchuk, Dmytro
Veröffentlicht: (2025)
VDW-GNNs: Vector diffusion wavelets for geometric graph neural networks
von: Johnson, David R., et al.
Veröffentlicht: (2025)
von: Johnson, David R., et al.
Veröffentlicht: (2025)
DecompKAN: Decomposed Patch-KAN for Long-Term Time Series Forecasting
von: Mysore, Naveen
Veröffentlicht: (2026)
von: Mysore, Naveen
Veröffentlicht: (2026)
Scaling Laws in the Tiny Regime: How Small Models Change Their Mistakes
von: Alnemari, Mohammed, et al.
Veröffentlicht: (2026)
von: Alnemari, Mohammed, et al.
Veröffentlicht: (2026)
CVCM Track Circuits Pre-emptive Failure Diagnostics for Predictive Maintenance Using Deep Neural Networks
von: Mukherjee, Debdeep, et al.
Veröffentlicht: (2025)
von: Mukherjee, Debdeep, et al.
Veröffentlicht: (2025)
EchoLSTM: A Self-Reflective Recurrent Network for Stabilizing Long-Range Memory
von: K, Prasanth K, et al.
Veröffentlicht: (2025)
von: K, Prasanth K, et al.
Veröffentlicht: (2025)
Adaptive Negative Scheduling for Graph Contrastive Learning
von: Ali, Adnan, et al.
Veröffentlicht: (2026)
von: Ali, Adnan, et al.
Veröffentlicht: (2026)
The Impact of Data Characteristics on GNN Evaluation for Detecting Fake News
von: Karn, Isha, et al.
Veröffentlicht: (2025)
von: Karn, Isha, et al.
Veröffentlicht: (2025)
The Normalized Difference Layer: A Differentiable Spectral Index Formulation for Deep Learning
von: Lotfi, Ali, et al.
Veröffentlicht: (2026)
von: Lotfi, Ali, et al.
Veröffentlicht: (2026)
Predicting Traffic Accident Severity with Deep Neural Networks
von: Bibb, Meghan, et al.
Veröffentlicht: (2025)
von: Bibb, Meghan, et al.
Veröffentlicht: (2025)
Efficient Morphology-Control Co-Design via Stackelberg Proximal Policy Optimization
von: Dai, Yanning, et al.
Veröffentlicht: (2026)
von: Dai, Yanning, et al.
Veröffentlicht: (2026)
An in-depth look at approximation via deep and narrow neural networks
von: Dommel, Joris, et al.
Veröffentlicht: (2025)
von: Dommel, Joris, et al.
Veröffentlicht: (2025)
Neural Encoding for Image Recall: Human-Like Memory
von: Foussereau, Virgile, et al.
Veröffentlicht: (2024)
von: Foussereau, Virgile, et al.
Veröffentlicht: (2024)
Predictive Modeling of Maritime Radar Data Using Transformer Architecture
von: Qesaraku, Bjorna, et al.
Veröffentlicht: (2025)
von: Qesaraku, Bjorna, et al.
Veröffentlicht: (2025)
Revisiting Non-separable Binary Classification and its Applications in Anomaly Detection
von: Lau, Matthew, et al.
Veröffentlicht: (2023)
von: Lau, Matthew, et al.
Veröffentlicht: (2023)
Tricks and Plug-ins for Gradient Boosting with Transformers
von: Fang, Biyi, et al.
Veröffentlicht: (2025)
von: Fang, Biyi, et al.
Veröffentlicht: (2025)
Accurate typhoon intensity forecasts using a non-iterative spatiotemporal transformer model
von: Qu, Hongyu, et al.
Veröffentlicht: (2025)
von: Qu, Hongyu, et al.
Veröffentlicht: (2025)
PolyGLU: State-Conditional Activation Routing in Transformer Feed-Forward Networks
von: Medeiros, Daniel Nobrega
Veröffentlicht: (2026)
von: Medeiros, Daniel Nobrega
Veröffentlicht: (2026)
Revisiting GAN with Bayes-Optimal Discrimination
von: Naeini, Mohammadreza Tavasoli, et al.
Veröffentlicht: (2025)
von: Naeini, Mohammadreza Tavasoli, et al.
Veröffentlicht: (2025)
Complex-Valued Phase-Coherent Transformer
von: Hioki, Leona
Veröffentlicht: (2026)
von: Hioki, Leona
Veröffentlicht: (2026)
Optimized Gradient Clipping for Noisy Label Learning
von: Ye, Xichen, et al.
Veröffentlicht: (2024)
von: Ye, Xichen, et al.
Veröffentlicht: (2024)
Scalable, Technology-Agnostic Diagnosis and Predictive Maintenance for Point Machine using Deep Learning
von: Di Santi, Eduardo, et al.
Veröffentlicht: (2025)
von: Di Santi, Eduardo, et al.
Veröffentlicht: (2025)
Contract-Driven QoE Auditing for Speech and Singing Services: From MOS Regression to Service Graphs
von: Du, Wenzhang
Veröffentlicht: (2025)
von: Du, Wenzhang
Veröffentlicht: (2025)
Mathematical Foundations of Neural Tangents and Infinite-Width Networks
von: Mysore, Rachana, et al.
Veröffentlicht: (2025)
von: Mysore, Rachana, et al.
Veröffentlicht: (2025)
Rethinking Visual Intelligence: Insights from Video Pretraining
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
Closing the Theory-Practice Gap in Spiking Transformers via Effective Dimension
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
A Hybrid Inductive-Transductive Network for Traffic Flow Imputation on Unsampled Locations
von: Rahimiasl, Mohammadmahdi, et al.
Veröffentlicht: (2025)
von: Rahimiasl, Mohammadmahdi, et al.
Veröffentlicht: (2025)
JacNet: Learning Functions with Structured Jacobians
von: Lorraine, Jonathan, et al.
Veröffentlicht: (2024)
von: Lorraine, Jonathan, et al.
Veröffentlicht: (2024)
Pulse-Driven Neural Architecture: Learnable Oscillatory Dynamics for Robust Continuous-Time Sequence Processing
von: Sharma, Paras
Veröffentlicht: (2026)
von: Sharma, Paras
Veröffentlicht: (2026)
RCUKF: Data-Driven Modeling Meets Bayesian Estimation
von: Anurag, Kumar, et al.
Veröffentlicht: (2025)
von: Anurag, Kumar, et al.
Veröffentlicht: (2025)
PH-VAE: A Polynomial Hierarchical Variational Autoencoder Towards Disentangled Representation Learning
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
GraphNNK -- Graph Classification and Interpretability
von: Bolevic, Zeljko, et al.
Veröffentlicht: (2026)
von: Bolevic, Zeljko, et al.
Veröffentlicht: (2026)
Estimating the Event-Related Potential from Few EEG Trials
von: Nørskov, Anders Vestergaard, et al.
Veröffentlicht: (2025)
von: Nørskov, Anders Vestergaard, et al.
Veröffentlicht: (2025)
Temporal Functional Circuits: From Spline Plots to Faithful Explanations in KAN Forecasting
von: Mysore, Naveen
Veröffentlicht: (2026)
von: Mysore, Naveen
Veröffentlicht: (2026)
Ähnliche Einträge
-
Explicit Dropout: Deterministic Regularization for Transformer Architectures
von: Agrawal, Vidhi, et al.
Veröffentlicht: (2026) -
Massive Redundancy in Gradient Transport Enables Sparse Online Learning
von: Merin, Aur Shalev
Veröffentlicht: (2026) -
Optimizing Inference in Transformer-Based Models: A Multi-Method Benchmark
von: Ho, Siu Hang, et al.
Veröffentlicht: (2025) -
Is Cambodia the World's Largest Cashew Producer?
von: Chaya, Veasna, et al.
Veröffentlicht: (2024) -
CellARC: Measuring Intelligence with Cellular Automata
von: Lžičař, Miroslav
Veröffentlicht: (2025)