LLMs as High-Dimensional Nonlinear Autoregressive Models with Attention: Training, Alignment and Inference
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Krishnamurthy, Vikram |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
Large Language Model-informed ECG Dual Attention Network for Heart Failure Risk Prediction
von: Chen, Chen, et al.
Veröffentlicht: (2024)
von: Chen, Chen, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models for Wireless Symbol Detection via In-Context Learning
von: Abbas, Momin, et al.
Veröffentlicht: (2024)
von: Abbas, Momin, et al.
Veröffentlicht: (2024)
ECG-Expert-QA: A Benchmark for Evaluating Medical Large Language Models in Heart Disease Diagnosis
von: Wang, Xu, et al.
Veröffentlicht: (2025)
von: Wang, Xu, et al.
Veröffentlicht: (2025)
Semantic Soft Bootstrapping: Long Context Reasoning in LLMs without Reinforcement Learning
von: Mitra, Purbesh, et al.
Veröffentlicht: (2025)
von: Mitra, Purbesh, et al.
Veröffentlicht: (2025)
SuPreME: A Supervised Pre-training Framework for Multimodal ECG Representation Learning
von: Cai, Mingsheng, et al.
Veröffentlicht: (2025)
von: Cai, Mingsheng, et al.
Veröffentlicht: (2025)
Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization
von: Krishnamurthy, Vikram
Veröffentlicht: (2025)
von: Krishnamurthy, Vikram
Veröffentlicht: (2025)
Attention-aware Semantic Communications for Collaborative Inference
von: Im, Jiwoong, et al.
Veröffentlicht: (2024)
von: Im, Jiwoong, et al.
Veröffentlicht: (2024)
LoPT: Low-Rank Prompt Tuning for Parameter Efficient Language Models
von: Guo, Shouchang, et al.
Veröffentlicht: (2024)
von: Guo, Shouchang, et al.
Veröffentlicht: (2024)
WaveMind: Towards a Conversational EEG Foundation Model Aligned to Textual and Visual Modalities
von: Zeng, Ziyi, et al.
Veröffentlicht: (2025)
von: Zeng, Ziyi, et al.
Veröffentlicht: (2025)
Wavelet GPT: Wavelet Inspired Large Language Models
von: Verma, Prateek
Veröffentlicht: (2024)
von: Verma, Prateek
Veröffentlicht: (2024)
TRI-DEP: A Trimodal Comparative Study for Depression Detection Using Speech, Text, and EEG
von: Nurfidausi, Annisaa Fitri, et al.
Veröffentlicht: (2025)
von: Nurfidausi, Annisaa Fitri, et al.
Veröffentlicht: (2025)
Language Bias in Self-Supervised Learning For Automatic Speech Recognition
von: Storey, Edward, et al.
Veröffentlicht: (2025)
von: Storey, Edward, et al.
Veröffentlicht: (2025)
Reading Miscue Detection in Primary School through Automatic Speech Recognition
von: Gao, Lingyun, et al.
Veröffentlicht: (2024)
von: Gao, Lingyun, et al.
Veröffentlicht: (2024)
Attention-Driven Training-Free Efficiency Enhancement of Diffusion Models
von: Wang, Hongjie, et al.
Veröffentlicht: (2024)
von: Wang, Hongjie, et al.
Veröffentlicht: (2024)
NeuroHD-RA: Neural-distilled Hyperdimensional Model with Rhythm Alignment
von: He, ZhengXiao, et al.
Veröffentlicht: (2025)
von: He, ZhengXiao, et al.
Veröffentlicht: (2025)
AR-KAN: Autoregressive-Weight-Enhanced Kolmogorov-Arnold Network for Time Series Forecasting
von: Zeng, Chen, et al.
Veröffentlicht: (2025)
von: Zeng, Chen, et al.
Veröffentlicht: (2025)
Attention and Autoencoder Hybrid Model for Unsupervised Online Anomaly Detection
von: Najafi, Seyed Amirhossein, et al.
Veröffentlicht: (2024)
von: Najafi, Seyed Amirhossein, et al.
Veröffentlicht: (2024)
Prescriptive Agents based on RAG for Automated Maintenance (PARAM)
von: Harbola, Chitranshu, et al.
Veröffentlicht: (2025)
von: Harbola, Chitranshu, et al.
Veröffentlicht: (2025)
A Backdoor Approach with Inverted Labels Using Dirty Label-Flipping Attacks
von: Mengara, Orson
Veröffentlicht: (2024)
von: Mengara, Orson
Veröffentlicht: (2024)
Retrieval-Augmented Generation for Reliable Interpretation of Radio Regulations
von: Kassimi, Zakaria El, et al.
Veröffentlicht: (2025)
von: Kassimi, Zakaria El, et al.
Veröffentlicht: (2025)
Benchmarking Spatiotemporal Reasoning in LLMs and Reasoning Models: Capabilities and Challenges
von: Quan, Pengrui, et al.
Veröffentlicht: (2025)
von: Quan, Pengrui, et al.
Veröffentlicht: (2025)
Large-scale Training of Foundation Models for Wearable Biosignals
von: Abbaspourazad, Salar, et al.
Veröffentlicht: (2023)
von: Abbaspourazad, Salar, et al.
Veröffentlicht: (2023)
Attention-Aided MMSE for OFDM Channel Estimation: Learning Linear Filters with Attention
von: Ha, TaeJun, et al.
Veröffentlicht: (2025)
von: Ha, TaeJun, et al.
Veröffentlicht: (2025)
Cross-device Zero-shot Label Transfer via Alignment of Time Series Foundation Model Embeddings
von: Ravindra, Neal G., et al.
Veröffentlicht: (2025)
von: Ravindra, Neal G., et al.
Veröffentlicht: (2025)
FoME: A Foundation Model for EEG using Adaptive Temporal-Lateral Attention Scaling
von: Shi, Enze, et al.
Veröffentlicht: (2024)
von: Shi, Enze, et al.
Veröffentlicht: (2024)
Energy-Gated Attention: Spectral Salience as an Inductive Bias for Transformer Attention
von: Zeris, Athanasios
Veröffentlicht: (2026)
von: Zeris, Athanasios
Veröffentlicht: (2026)
Deep Time Warping for Multiple Time Series Alignment
von: Nourbakhsh, Alireza, et al.
Veröffentlicht: (2025)
von: Nourbakhsh, Alireza, et al.
Veröffentlicht: (2025)
Prompting Large Language Models for Training-Free Non-Intrusive Load Monitoring
von: Xue, Junyu, et al.
Veröffentlicht: (2025)
von: Xue, Junyu, et al.
Veröffentlicht: (2025)
Energy-Gated Attention and Wavelet Positional Encoding: Complementary Inductive Biases for Transformer Attention
von: Zeris, Athanasios
Veröffentlicht: (2026)
von: Zeris, Athanasios
Veröffentlicht: (2026)
Brant-X: A Unified Physiological Signal Alignment Framework
von: Zhang, Daoze, et al.
Veröffentlicht: (2024)
von: Zhang, Daoze, et al.
Veröffentlicht: (2024)
Distilling Calibration via Conformalized Credal Inference
von: Huang, Jiayi, et al.
Veröffentlicht: (2025)
von: Huang, Jiayi, et al.
Veröffentlicht: (2025)
Spatio-Temporal Attention Network for Epileptic Seizure Prediction
von: Li, Zan, et al.
Veröffentlicht: (2025)
von: Li, Zan, et al.
Veröffentlicht: (2025)
Zero-Forget Preservation of Semantic Communication Alignment in Distributed AI Networks
von: Hu, Jingzhi, et al.
Veröffentlicht: (2024)
von: Hu, Jingzhi, et al.
Veröffentlicht: (2024)
Fine-grained Contrastive Learning for ECG-Report Alignment with Waveform Enhancement
von: Li, Haitao, et al.
Veröffentlicht: (2025)
von: Li, Haitao, et al.
Veröffentlicht: (2025)
Density Adaptive Attention is All You Need: Robust Parameter-Efficient Fine-Tuning Across Multiple Modalities
von: Ioannides, Georgios, et al.
Veröffentlicht: (2024)
von: Ioannides, Georgios, et al.
Veröffentlicht: (2024)
Self-Alignment Learning to Improve Myocardial Infarction Detection from Single-Lead ECG
von: Jin, Jiarui, et al.
Veröffentlicht: (2025)
von: Jin, Jiarui, et al.
Veröffentlicht: (2025)
Integrating Biological and Machine Intelligence: Attention Mechanisms in Brain-Computer Interfaces
von: Wang, Jiyuan, et al.
Veröffentlicht: (2025)
von: Wang, Jiyuan, et al.
Veröffentlicht: (2025)
Pulse-PPG: An Open-Source Field-Trained PPG Foundation Model for Wearable Applications Across Lab and Field Settings
von: Saha, Mithun, et al.
Veröffentlicht: (2025)
von: Saha, Mithun, et al.
Veröffentlicht: (2025)
Calibrating Bayesian Learning via Regularization, Confidence Minimization, and Selective Inference
von: Huang, Jiayi, et al.
Veröffentlicht: (2024)
von: Huang, Jiayi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute
von: Liu, Sheng, et al.
Veröffentlicht: (2025) -
Large Language Model-informed ECG Dual Attention Network for Heart Failure Risk Prediction
von: Chen, Chen, et al.
Veröffentlicht: (2024) -
Leveraging Large Language Models for Wireless Symbol Detection via In-Context Learning
von: Abbas, Momin, et al.
Veröffentlicht: (2024) -
ECG-Expert-QA: A Benchmark for Evaluating Medical Large Language Models in Heart Disease Diagnosis
von: Wang, Xu, et al.
Veröffentlicht: (2025) -
Semantic Soft Bootstrapping: Long Context Reasoning in LLMs without Reinforcement Learning
von: Mitra, Purbesh, et al.
Veröffentlicht: (2025)