On the Design Space Between Transformers and Recursive Neural Nets
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chowdhury, Jishnu Ray, Caragea, Cornelia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Investigating Recurrent Transformers with Dynamic Halt
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2024)
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2024)
Zero-Shot Verification-guided Chain of Thoughts
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2025)
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2025)
Zero-Shot Keyphrase Generation: Investigating Specialized Instructions and Multi-Sample Aggregation on Large Language Models
von: Mohan, Jayanth, et al.
Veröffentlicht: (2025)
von: Mohan, Jayanth, et al.
Veröffentlicht: (2025)
A Novel Cartography-Based Curriculum Learning Method Applied on RoNLI: The First Romanian Natural Language Inference Corpus
von: Poesina, Eduard, et al.
Veröffentlicht: (2024)
von: Poesina, Eduard, et al.
Veröffentlicht: (2024)
Unlocking Out-of-Distribution Generalization in Transformers via Recursive Latent Space Reasoning
von: Altabaa, Awni, et al.
Veröffentlicht: (2025)
von: Altabaa, Awni, et al.
Veröffentlicht: (2025)
LoopQ: Quantization for Recursive Transformers
von: Fang, Rui, et al.
Veröffentlicht: (2026)
von: Fang, Rui, et al.
Veröffentlicht: (2026)
Recursive Inference Machines for Neural Reasoning
von: Komisarczyk, Mieszko, et al.
Veröffentlicht: (2026)
von: Komisarczyk, Mieszko, et al.
Veröffentlicht: (2026)
MeSH: Memory-as-State-Highways for Recursive Transformers
von: Yu, Chengting, et al.
Veröffentlicht: (2025)
von: Yu, Chengting, et al.
Veröffentlicht: (2025)
Neural Collapse is Globally Optimal in Deep Regularized ResNets and Transformers
von: Súkeník, Peter, et al.
Veröffentlicht: (2025)
von: Súkeník, Peter, et al.
Veröffentlicht: (2025)
MultiMatch: Multihead Consistency Regularization Matching for Semi-Supervised Text Classification
von: Sirbu, Iustin, et al.
Veröffentlicht: (2025)
von: Sirbu, Iustin, et al.
Veröffentlicht: (2025)
What Makes Looped Transformers Perform Better Than Non-Recursive Ones
von: Gong, Zixuan, et al.
Veröffentlicht: (2025)
von: Gong, Zixuan, et al.
Veröffentlicht: (2025)
Convolutional Neural Nets vs Vision Transformers: A SpaceNet Case Study with Balanced vs Imbalanced Regimes
von: Gothi, Akshar
Veröffentlicht: (2025)
von: Gothi, Akshar
Veröffentlicht: (2025)
Artificial Intelligence-Driven Network-on-Chip Design Space Exploration: Neural Network Architectures for Design
von: N, Amogh Anshu, et al.
Veröffentlicht: (2025)
von: N, Amogh Anshu, et al.
Veröffentlicht: (2025)
Looping Back to Move Forward: Recursive Transformers for Efficient and Flexible Large Multimodal Models
von: Xu, Ruihan, et al.
Veröffentlicht: (2026)
von: Xu, Ruihan, et al.
Veröffentlicht: (2026)
Predictive Modeling and Explainable AI for Veterinary Safety Profiles, Residue Assessment, and Health Outcomes Using Real-World Data and Physicochemical Properties
von: Sholehrasa, Hossein, et al.
Veröffentlicht: (2025)
von: Sholehrasa, Hossein, et al.
Veröffentlicht: (2025)
Design Space Exploration of Hybrid Quantum Neural Networks for Chronic Kidney Disease
von: Kashif, Muhammad, et al.
Veröffentlicht: (2026)
von: Kashif, Muhammad, et al.
Veröffentlicht: (2026)
Provable Generalization in Overparameterized Neural Nets
von: Dhingra, Aviral
Veröffentlicht: (2025)
von: Dhingra, Aviral
Veröffentlicht: (2025)
Artificial Neural Nets and the Representation of Human Concepts
von: Freiesleben, Timo
Veröffentlicht: (2023)
von: Freiesleben, Timo
Veröffentlicht: (2023)
Decoupling Search and Learning in Neural Net Training
von: Vegesna, Akshay, et al.
Veröffentlicht: (2025)
von: Vegesna, Akshay, et al.
Veröffentlicht: (2025)
TabDistill: Distilling Transformers into Neural Nets for Few-Shot Tabular Classification
von: Dissanayake, Pasan, et al.
Veröffentlicht: (2025)
von: Dissanayake, Pasan, et al.
Veröffentlicht: (2025)
Lost in the Middle at Birth: An Exact Theory of Transformer Position Bias
von: Chowdhury, Borun D
Veröffentlicht: (2026)
von: Chowdhury, Borun D
Veröffentlicht: (2026)
Evaluating Large Language Models for Stance Detection on Financial Targets from SEC Filing Reports and Earnings Call Transcripts
von: Gyawali, Nikesh, et al.
Veröffentlicht: (2025)
von: Gyawali, Nikesh, et al.
Veröffentlicht: (2025)
Modality-Decoupled Online Recursive Editing
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
Interaction Locality in Hierarchical Recursive Reasoning
von: Miyanishi, Yosuke, et al.
Veröffentlicht: (2026)
von: Miyanishi, Yosuke, et al.
Veröffentlicht: (2026)
Recursive Deep Inverse Reinforcement Learning
von: Ghanem, Paul, et al.
Veröffentlicht: (2025)
von: Ghanem, Paul, et al.
Veröffentlicht: (2025)
Neural Decompiling of Tracr Transformers
von: Thurnherr, Hannes, et al.
Veröffentlicht: (2024)
von: Thurnherr, Hannes, et al.
Veröffentlicht: (2024)
Recursive Backwards Q-Learning in Deterministic Environments
von: Diekhoff, Jan, et al.
Veröffentlicht: (2024)
von: Diekhoff, Jan, et al.
Veröffentlicht: (2024)
Spectral Transformer Neural Processes
von: Chen, Xianhe, et al.
Veröffentlicht: (2026)
von: Chen, Xianhe, et al.
Veröffentlicht: (2026)
Transformers are Graph Neural Networks
von: Joshi, Chaitanya K.
Veröffentlicht: (2025)
von: Joshi, Chaitanya K.
Veröffentlicht: (2025)
Less is More: Recursive Reasoning with Tiny Networks
von: Jolicoeur-Martineau, Alexia
Veröffentlicht: (2025)
von: Jolicoeur-Martineau, Alexia
Veröffentlicht: (2025)
Test-time Adaptation of Tiny Recursive Models
von: McGovern, Ronan Killian
Veröffentlicht: (2025)
von: McGovern, Ronan Killian
Veröffentlicht: (2025)
Symmetry in Neural Network Parameter Spaces
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
NeuralGrok: Accelerate Grokking by Neural Gradient Transformation
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2025)
Exploring the Design Space of Transition Matching
von: Singer, Uriel, et al.
Veröffentlicht: (2025)
von: Singer, Uriel, et al.
Veröffentlicht: (2025)
Tiled Bit Networks: Sub-Bit Neural Network Compression Through Reuse of Learnable Binary Vectors
von: Gorbett, Matt, et al.
Veröffentlicht: (2024)
von: Gorbett, Matt, et al.
Veröffentlicht: (2024)
HardNet: Hard-Constrained Neural Networks with Universal Approximation Guarantees
von: Min, Youngjae, et al.
Veröffentlicht: (2024)
von: Min, Youngjae, et al.
Veröffentlicht: (2024)
PolyNet: Learning Diverse Solution Strategies for Neural Combinatorial Optimization
von: Hottung, André, et al.
Veröffentlicht: (2024)
von: Hottung, André, et al.
Veröffentlicht: (2024)
CoFrNets: Interpretable Neural Architecture Inspired by Continued Fractions
von: Puri, Isha, et al.
Veröffentlicht: (2025)
von: Puri, Isha, et al.
Veröffentlicht: (2025)
When Do Neural Nets Outperform Boosted Trees on Tabular Data?
von: McElfresh, Duncan, et al.
Veröffentlicht: (2023)
von: McElfresh, Duncan, et al.
Veröffentlicht: (2023)
SnareNet: Flexible Repair Layers for Neural Networks with Hard Constraints
von: Chu, Ya-Chi, et al.
Veröffentlicht: (2026)
von: Chu, Ya-Chi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Investigating Recurrent Transformers with Dynamic Halt
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2024) -
Zero-Shot Verification-guided Chain of Thoughts
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2025) -
Zero-Shot Keyphrase Generation: Investigating Specialized Instructions and Multi-Sample Aggregation on Large Language Models
von: Mohan, Jayanth, et al.
Veröffentlicht: (2025) -
A Novel Cartography-Based Curriculum Learning Method Applied on RoNLI: The First Romanian Natural Language Inference Corpus
von: Poesina, Eduard, et al.
Veröffentlicht: (2024) -
Unlocking Out-of-Distribution Generalization in Transformers via Recursive Latent Space Reasoning
von: Altabaa, Awni, et al.
Veröffentlicht: (2025)