Hardness of Learning Regular Languages in the Next Symbol Prediction Setting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhattamishra, Satwik, Blunsom, Phil, Kanade, Varun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Separations in the Representational Capabilities of Transformers and Recurrent Architectures
von: Bhattamishra, Satwik, et al.
Veröffentlicht: (2024)
von: Bhattamishra, Satwik, et al.
Veröffentlicht: (2024)
Provably Learning Attention with Queries
von: Bhattamishra, Satwik, et al.
Veröffentlicht: (2026)
von: Bhattamishra, Satwik, et al.
Veröffentlicht: (2026)
Pause Tokens Strictly Increase the Expressivity of Constant-Depth Transformers
von: London, Charles, et al.
Veröffentlicht: (2025)
von: London, Charles, et al.
Veröffentlicht: (2025)
Discovering Interpretable Algorithms by Decompiling Transformers to RASP
von: Huang, Xinting, et al.
Veröffentlicht: (2026)
von: Huang, Xinting, et al.
Veröffentlicht: (2026)
Inshrinkerator: Compressing Deep Learning Training Checkpoints via Dynamic Quantization
von: Agrawal, Amey, et al.
Veröffentlicht: (2023)
von: Agrawal, Amey, et al.
Veröffentlicht: (2023)
A Formal Framework for Understanding Length Generalization in Transformers
von: Huang, Xinting, et al.
Veröffentlicht: (2024)
von: Huang, Xinting, et al.
Veröffentlicht: (2024)
Benefits and Limitations of Communication in Multi-Agent Reasoning
von: Rizvi-Martel, Michael, et al.
Veröffentlicht: (2025)
von: Rizvi-Martel, Michael, et al.
Veröffentlicht: (2025)
How Global Calibration Strengthens Multiaccuracy
von: Casacuberta, Sílvia, et al.
Veröffentlicht: (2025)
von: Casacuberta, Sílvia, et al.
Veröffentlicht: (2025)
The Transformer Cookbook
von: Yang, Andy, et al.
Veröffentlicht: (2025)
von: Yang, Andy, et al.
Veröffentlicht: (2025)
On the Hardness of Learning Regular Expressions
von: Attias, Idan, et al.
Veröffentlicht: (2025)
von: Attias, Idan, et al.
Veröffentlicht: (2025)
DEEDEE: Fast and Scalable Out-of-Distribution Dynamics Detection
von: Aljaafari, Tala, et al.
Veröffentlicht: (2025)
von: Aljaafari, Tala, et al.
Veröffentlicht: (2025)
Set Prediction for Next-Day Active Fire Forecasting
von: Bai, Yuchen, et al.
Veröffentlicht: (2026)
von: Bai, Yuchen, et al.
Veröffentlicht: (2026)
DyPP: Dynamic Parameter Prediction to Accelerate Convergence of Variational Quantum Algorithms
von: Kundu, Satwik, et al.
Veröffentlicht: (2023)
von: Kundu, Satwik, et al.
Veröffentlicht: (2023)
Security Concerns in Quantum Machine Learning as a Service
von: Kundu, Satwik, et al.
Veröffentlicht: (2024)
von: Kundu, Satwik, et al.
Veröffentlicht: (2024)
Neural Proposals, Symbolic Guarantees: Neuro-Symbolic Graph Generation with Hard Constraints
von: Geng, Chuqin, et al.
Veröffentlicht: (2026)
von: Geng, Chuqin, et al.
Veröffentlicht: (2026)
Explainable Machine Learning for Pediatric Dental Risk Stratification Using Socio-Demographic Determinants
von: Kanade, Manasi, et al.
Veröffentlicht: (2026)
von: Kanade, Manasi, et al.
Veröffentlicht: (2026)
Gating Enables Curvature: A Geometric Expressivity Gap in Attention
von: Bathula, Satwik, et al.
Veröffentlicht: (2026)
von: Bathula, Satwik, et al.
Veröffentlicht: (2026)
Inverse-Transpilation: Reverse-Engineering Quantum Compiler Optimization Passes from Circuit Snapshots
von: Kundu, Satwik, et al.
Veröffentlicht: (2025)
von: Kundu, Satwik, et al.
Veröffentlicht: (2025)
STIQ: Safeguarding Training and Inferencing of Quantum Neural Networks from Untrusted Cloud
von: Kundu, Satwik, et al.
Veröffentlicht: (2024)
von: Kundu, Satwik, et al.
Veröffentlicht: (2024)
Towards the Next Frontier in Speech Representation Learning Using Disentanglement
von: Krishna, Varun, et al.
Veröffentlicht: (2024)
von: Krishna, Varun, et al.
Veröffentlicht: (2024)
Beyond Regularity: Modeling Chaotic Mobility Patterns for Next Location Prediction
von: Wu, Yuqian, et al.
Veröffentlicht: (2025)
von: Wu, Yuqian, et al.
Veröffentlicht: (2025)
Adversarial Threats in Quantum Machine Learning: A Survey of Attacks and Defenses
von: Ghosh, Archisman, et al.
Veröffentlicht: (2025)
von: Ghosh, Archisman, et al.
Veröffentlicht: (2025)
AttackGNN: Red-Teaming GNNs in Hardware Security Using Reinforcement Learning
von: Gohil, Vasudev, et al.
Veröffentlicht: (2024)
von: Gohil, Vasudev, et al.
Veröffentlicht: (2024)
Robust Learning of Diverse Code Edits
von: Aggarwal, Tushar, et al.
Veröffentlicht: (2025)
von: Aggarwal, Tushar, et al.
Veröffentlicht: (2025)
BAM! Just Like That: Simple and Efficient Parameter Upcycling for Mixture of Experts
von: Zhang, Qizhen, et al.
Veröffentlicht: (2024)
von: Zhang, Qizhen, et al.
Veröffentlicht: (2024)
Recoverability Has a Law: The ERR Measure for Tool-Augmented Agents
von: Vuddanti, Sri Vatsa, et al.
Veröffentlicht: (2026)
von: Vuddanti, Sri Vatsa, et al.
Veröffentlicht: (2026)
How Reinforcement Learning After Next-Token Prediction Facilitates Learning
von: Tsilivis, Nikolaos, et al.
Veröffentlicht: (2025)
von: Tsilivis, Nikolaos, et al.
Veröffentlicht: (2025)
Agentic Risk-Aware Set-Based Engineering Design
von: Kumar, Varun, et al.
Veröffentlicht: (2026)
von: Kumar, Varun, et al.
Veröffentlicht: (2026)
Next-Latent Prediction Transformers Learn Compact World Models
von: Teoh, Jayden, et al.
Veröffentlicht: (2025)
von: Teoh, Jayden, et al.
Veröffentlicht: (2025)
RegMix: Adversarial Mutual and Generalization Regularization for Enhancing DNN Robustness
von: Liu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Liu, Zhenyu, et al.
Veröffentlicht: (2025)
Hard Regularization to Prevent Deep Online Clustering Collapse without Data Augmentation
von: Mahon, Louis, et al.
Veröffentlicht: (2023)
von: Mahon, Louis, et al.
Veröffentlicht: (2023)
Discovering Predictive Relational Object Symbols with Symbolic Attentive Layers
von: Ahmetoglu, Alper, et al.
Veröffentlicht: (2023)
von: Ahmetoglu, Alper, et al.
Veröffentlicht: (2023)
SetPINNs: Set-based Physics-informed Neural Networks
von: Nagda, Mayank, et al.
Veröffentlicht: (2024)
von: Nagda, Mayank, et al.
Veröffentlicht: (2024)
T-SHRED: Symbolic Regression for Regularization and Model Discovery with Transformer Shallow Recurrent Decoders
von: Yermakov, Alexey, et al.
Veröffentlicht: (2025)
von: Yermakov, Alexey, et al.
Veröffentlicht: (2025)
On the Hardness of Bandit Learning
von: Brukhim, Nataly, et al.
Veröffentlicht: (2025)
von: Brukhim, Nataly, et al.
Veröffentlicht: (2025)
Adaptively Private Next-Token Prediction of Large Language Models
von: Flemings, James, et al.
Veröffentlicht: (2024)
von: Flemings, James, et al.
Veröffentlicht: (2024)
Effects of Common Regularization Techniques on Open-Set Recognition
von: Rabin, Zachary, et al.
Veröffentlicht: (2024)
von: Rabin, Zachary, et al.
Veröffentlicht: (2024)
Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning
von: Luo, Qin-Wen, et al.
Veröffentlicht: (2026)
von: Luo, Qin-Wen, et al.
Veröffentlicht: (2026)
Efficient Learning Under Density Shift in Incremental Settings Using Cramér-Rao-Based Regularization
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
Evaluating Efficacy of Model Stealing Attacks and Defenses on Quantum Neural Networks
von: Kundu, Satwik, et al.
Veröffentlicht: (2024)
von: Kundu, Satwik, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Separations in the Representational Capabilities of Transformers and Recurrent Architectures
von: Bhattamishra, Satwik, et al.
Veröffentlicht: (2024) -
Provably Learning Attention with Queries
von: Bhattamishra, Satwik, et al.
Veröffentlicht: (2026) -
Pause Tokens Strictly Increase the Expressivity of Constant-Depth Transformers
von: London, Charles, et al.
Veröffentlicht: (2025) -
Discovering Interpretable Algorithms by Decompiling Transformers to RASP
von: Huang, Xinting, et al.
Veröffentlicht: (2026) -
Inshrinkerator: Compressing Deep Learning Training Checkpoints via Dynamic Quantization
von: Agrawal, Amey, et al.
Veröffentlicht: (2023)