Gespeichert in:
| Hauptverfasser: | Boža, Vladimír, Macko, Vladimír |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2505.11076 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Two Sparse Matrices are Better than One: Sparsifying Neural Networks with Double Sparse Factorization
von: Boža, Vladimír, et al.
Veröffentlicht: (2024)
von: Boža, Vladimír, et al.
Veröffentlicht: (2024)
MACKO: Sparse Matrix-Vector Multiplication for Low Sparsity
von: Macko, Vladimír, et al.
Veröffentlicht: (2025)
von: Macko, Vladimír, et al.
Veröffentlicht: (2025)
Fast and Effective Weight Update for Pruned Large Language Models
von: Boža, Vladimír
Veröffentlicht: (2024)
von: Boža, Vladimír
Veröffentlicht: (2024)
Attention and Compression is all you need for Controllably Efficient Language Models
von: Prakash, Jatin, et al.
Veröffentlicht: (2025)
von: Prakash, Jatin, et al.
Veröffentlicht: (2025)
One protein is all you need
von: Bushuiev, Anton, et al.
Veröffentlicht: (2024)
von: Bushuiev, Anton, et al.
Veröffentlicht: (2024)
Kolmogorov GAM Networks are all you need!
von: Polson, Sarah, et al.
Veröffentlicht: (2025)
von: Polson, Sarah, et al.
Veröffentlicht: (2025)
KV-weights are all you need for skipless transformers
von: Graef, Nils
Veröffentlicht: (2024)
von: Graef, Nils
Veröffentlicht: (2024)
Tabular Data: Is Deep Learning all you need?
von: Zabërgja, Guri, et al.
Veröffentlicht: (2024)
von: Zabërgja, Guri, et al.
Veröffentlicht: (2024)
Image compositing is all you need for data augmentation
von: Shermaine, Ang Jia Ning, et al.
Veröffentlicht: (2025)
von: Shermaine, Ang Jia Ning, et al.
Veröffentlicht: (2025)
Large Language Models aren't all that you need
von: Holla, Kiran Voderhobli, et al.
Veröffentlicht: (2024)
von: Holla, Kiran Voderhobli, et al.
Veröffentlicht: (2024)
Experts are all you need: A Composable Framework for Large Language Model Inference
von: Sridharan, Shrihari, et al.
Veröffentlicht: (2025)
von: Sridharan, Shrihari, et al.
Veröffentlicht: (2025)
Cross-Modal Safety Alignment: Is textual unlearning all you need?
von: Chakraborty, Trishna, et al.
Veröffentlicht: (2024)
von: Chakraborty, Trishna, et al.
Veröffentlicht: (2024)
Attention is all you need for boosting graph convolutional neural network
von: Wu, Yinwei
Veröffentlicht: (2024)
von: Wu, Yinwei
Veröffentlicht: (2024)
Linear attention is (maybe) all you need (to understand transformer optimization)
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2023)
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2023)
Graph is all you need? Lightweight data-agnostic neural architecture search without training
von: Huang, Zhenhan, et al.
Veröffentlicht: (2024)
von: Huang, Zhenhan, et al.
Veröffentlicht: (2024)
Simulation-based inference with scattering representations: scattering is all you need
von: Lin, Kiyam, et al.
Veröffentlicht: (2024)
von: Lin, Kiyam, et al.
Veröffentlicht: (2024)
Neural Operator: Is data all you need to model the world? An insight into the paradigm of data-driven scientific ML
von: Viswanath, Hrishikesh, et al.
Veröffentlicht: (2023)
von: Viswanath, Hrishikesh, et al.
Veröffentlicht: (2023)
DC is all you need: describing ReLU from a signal processing standpoint
von: Kechris, Christodoulos, et al.
Veröffentlicht: (2024)
von: Kechris, Christodoulos, et al.
Veröffentlicht: (2024)
Anti-concentration is (almost) all you need
von: Heinrich, Markus, et al.
Veröffentlicht: (2025)
von: Heinrich, Markus, et al.
Veröffentlicht: (2025)
Is attention all you need in medical image analysis? A review
von: Papanastasiou, Giorgos, et al.
Veröffentlicht: (2023)
von: Papanastasiou, Giorgos, et al.
Veröffentlicht: (2023)
Few Labels are all you need: A Weakly Supervised Framework for Appliance Localization in Smart-Meter Series
von: Petralia, Adrien, et al.
Veröffentlicht: (2025)
von: Petralia, Adrien, et al.
Veröffentlicht: (2025)
Attention is all you need for an improved CNN-based flash flood susceptibility modeling. The case of the ungauged Rheraya watershed, Morocco
von: Elghouat, Akram, et al.
Veröffentlicht: (2024)
von: Elghouat, Akram, et al.
Veröffentlicht: (2024)
Signformer is all you need: Towards Edge AI for Sign Language
von: Yang, Eta
Veröffentlicht: (2024)
von: Yang, Eta
Veröffentlicht: (2024)
Slim attention: cut your context memory in half without loss -- K-cache is all you need for MHA
von: Graef, Nils, et al.
Veröffentlicht: (2025)
von: Graef, Nils, et al.
Veröffentlicht: (2025)
Block removal for large language models through constrained binary optimization
von: Jansen, David, et al.
Veröffentlicht: (2026)
von: Jansen, David, et al.
Veröffentlicht: (2026)
1 bit is all we need: binary normalized neural networks
von: Cabral, Eduardo Lobo Lustoda, et al.
Veröffentlicht: (2025)
von: Cabral, Eduardo Lobo Lustoda, et al.
Veröffentlicht: (2025)
Why you don't overfit, and don't need Bayes if you only train for one epoch
von: Aitchison, Laurence
Veröffentlicht: (2024)
von: Aitchison, Laurence
Veröffentlicht: (2024)
Visual cognition in multimodal large language models
von: Buschoff, Luca M. Schulze, et al.
Veröffentlicht: (2023)
von: Buschoff, Luca M. Schulze, et al.
Veröffentlicht: (2023)
Hypothesis generation and updating in large language models
von: Xiong, Hua-Dong
Veröffentlicht: (2026)
von: Xiong, Hua-Dong
Veröffentlicht: (2026)
Representation in large language models
von: Yetman, Cameron
Veröffentlicht: (2025)
von: Yetman, Cameron
Veröffentlicht: (2025)
100 instances is all you need: predicting the success of a new LLM on unseen data by testing on a few instances
von: Pacchiardi, Lorenzo, et al.
Veröffentlicht: (2024)
von: Pacchiardi, Lorenzo, et al.
Veröffentlicht: (2024)
A Bayesian Optimization approach for calibrating large-scale activity-based transport models
von: Agriesti, Serio, et al.
Veröffentlicht: (2023)
von: Agriesti, Serio, et al.
Veröffentlicht: (2023)
Compressed models are NOT miniature versions of large models
von: Rai, Rohit Raj, et al.
Veröffentlicht: (2024)
von: Rai, Rohit Raj, et al.
Veröffentlicht: (2024)
Amortizing intractable inference in large language models
von: Hu, Edward J., et al.
Veröffentlicht: (2023)
von: Hu, Edward J., et al.
Veröffentlicht: (2023)
Are nuclear masks all you need for improved out-of-domain generalisation? A closer look at cancer classification in histopathology
von: Tomar, Dhananjay, et al.
Veröffentlicht: (2024)
von: Tomar, Dhananjay, et al.
Veröffentlicht: (2024)
Alignment faking in large language models
von: Greenblatt, Ryan, et al.
Veröffentlicht: (2024)
von: Greenblatt, Ryan, et al.
Veröffentlicht: (2024)
AI-AI Bias: large language models favor communications generated by large language models
von: Laurito, Walter, et al.
Veröffentlicht: (2024)
von: Laurito, Walter, et al.
Veröffentlicht: (2024)
Training microrobots to swim by a large language model
von: Xu, Zhuoqun, et al.
Veröffentlicht: (2024)
von: Xu, Zhuoqun, et al.
Veröffentlicht: (2024)
Compression is all you need: Modeling Mathematics
von: Aksenov, Vitaly, et al.
Veröffentlicht: (2026)
von: Aksenov, Vitaly, et al.
Veröffentlicht: (2026)
Layer-wise dynamic rank for compressing large language models
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Two Sparse Matrices are Better than One: Sparsifying Neural Networks with Double Sparse Factorization
von: Boža, Vladimír, et al.
Veröffentlicht: (2024) -
MACKO: Sparse Matrix-Vector Multiplication for Low Sparsity
von: Macko, Vladimír, et al.
Veröffentlicht: (2025) -
Fast and Effective Weight Update for Pruned Large Language Models
von: Boža, Vladimír
Veröffentlicht: (2024) -
Attention and Compression is all you need for Controllably Efficient Language Models
von: Prakash, Jatin, et al.
Veröffentlicht: (2025) -
One protein is all you need
von: Bushuiev, Anton, et al.
Veröffentlicht: (2024)