Saved in:
| Main Authors: | Borobia, Hector, Seguí-Mas, Elies, Tormo-Carbó, Guillermina |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.01192 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
by: Borobia, Hector, et al.
Published: (2026)
by: Borobia, Hector, et al.
Published: (2026)
Where Should LoRA Go? Component-Type Placement in Hybrid Language Models
by: Borobia, Hector, et al.
Published: (2026)
by: Borobia, Hector, et al.
Published: (2026)
Component-Aware Self-Speculative Decoding in Hybrid Language Models
by: Borobia, Hector, et al.
Published: (2026)
by: Borobia, Hector, et al.
Published: (2026)
Functional Component Ablation Reveals Specialization Patterns in Hybrid Language Model Architectures
by: Borobia, Hector, et al.
Published: (2026)
by: Borobia, Hector, et al.
Published: (2026)
Emergence of Quantised Representations Isolated to Anisotropic Functions
by: Bird, George
Published: (2025)
by: Bird, George
Published: (2025)
The Affine Divergence: Aligning Activation Updates Beyond Normalisation
by: Bird, George
Published: (2025)
by: Bird, George
Published: (2025)
Binarized Neural Networks Converge Toward Algorithmic Simplicity: Empirical Support for the Learning-as-Compression Hypothesis
by: Sakabe, Eduardo Y., et al.
Published: (2025)
by: Sakabe, Eduardo Y., et al.
Published: (2025)
The Inclusion Depth of Pattern Languages: An Open Problem in Algorithmic Learning Theory
by: Luo, Wei
Published: (2026)
by: Luo, Wei
Published: (2026)
Multidimensional Adaptive Coefficient for Inference Trajectory Optimization in Flow and Diffusion
by: Lee, Dohoon, et al.
Published: (2024)
by: Lee, Dohoon, et al.
Published: (2024)
Information-Theoretic Quality Metric of Low-Dimensional Embeddings
by: Gutiérrez-Bernal, Sebastián, et al.
Published: (2025)
by: Gutiérrez-Bernal, Sebastián, et al.
Published: (2025)
Prompting Neural-Guided Equation Discovery Based on Residuals
by: Brugger, Jannis, et al.
Published: (2025)
by: Brugger, Jannis, et al.
Published: (2025)
Towards a Neural Lambda Calculus: Neurosymbolic AI Applied to the Foundations of Functional Programming
by: Flach, João, et al.
Published: (2023)
by: Flach, João, et al.
Published: (2023)
Tricks and Plug-ins for Gradient Boosting with Transformers
by: Fang, Biyi, et al.
Published: (2025)
by: Fang, Biyi, et al.
Published: (2025)
Does Editing Provide Evidence for Localization?
by: Wang, Zihao, et al.
Published: (2025)
by: Wang, Zihao, et al.
Published: (2025)
Hybrid Quantum-Classical Mixture of Experts: Unlocking Topological Advantage via Interference-Based Routing
by: Heddad, Reda, et al.
Published: (2025)
by: Heddad, Reda, et al.
Published: (2025)
Nyström $M$-Hilbert-Schmidt Independence Criterion
by: Kalinke, Florian, et al.
Published: (2023)
by: Kalinke, Florian, et al.
Published: (2023)
Maximum Mean Discrepancy on Exponential Windows for Online Change Detection
by: Kalinke, Florian, et al.
Published: (2022)
by: Kalinke, Florian, et al.
Published: (2022)
From Provable Correctness to Probabilistic Generation: A Comparative Review of Program Synthesis Paradigms
by: Kobaladze, Zurabi, et al.
Published: (2025)
by: Kobaladze, Zurabi, et al.
Published: (2025)
Graded Transformers
by: Shaska Sr, Tony
Published: (2025)
by: Shaska Sr, Tony
Published: (2025)
Theoretical Analysis of Positional Encodings in Transformer Models: Impact on Expressiveness and Generalization
by: Li, Yin
Published: (2025)
by: Li, Yin
Published: (2025)
On the Compatibility of Generative AI and Generative Linguistics
by: Portelance, Eva, et al.
Published: (2024)
by: Portelance, Eva, et al.
Published: (2024)
Privacy-Preserving Explainable AIoT Application via SHAP Entropy Regularization
by: Sharma, Dilli Prasad, et al.
Published: (2025)
by: Sharma, Dilli Prasad, et al.
Published: (2025)
Internalizing Tools as Morphisms in Graded Transformers
by: Shaska, Tony
Published: (2025)
by: Shaska, Tony
Published: (2025)
Information-Theoretic Measures in AI: A Practical Decision Guide
by: Papadopoulos, Nikolaos Al., et al.
Published: (2026)
by: Papadopoulos, Nikolaos Al., et al.
Published: (2026)
Decoding Rewards in Competitive Games: Inverse Game Theory with Entropy Regularization
by: Liao, Junyi, et al.
Published: (2026)
by: Liao, Junyi, et al.
Published: (2026)
Entropy-Reservoir Bregman Projection: An Information-Geometric Unification of Model Collapse
by: Chen, Jingwei
Published: (2025)
by: Chen, Jingwei
Published: (2025)
STACHE: Local Black-Box Explanations for Reinforcement Learning Policies
by: Elashkin, Andrew, et al.
Published: (2025)
by: Elashkin, Andrew, et al.
Published: (2025)
The Serial Scaling Hypothesis
by: Liu, Yuxi, et al.
Published: (2025)
by: Liu, Yuxi, et al.
Published: (2025)
Probabilistic Digital Twins of Users: Latent Representation Learning with Statistically Validated Semantics
by: David, Daniel
Published: (2025)
by: David, Daniel
Published: (2025)
Spiking Sequence Machines and Transformers
by: Bose, Joy
Published: (2026)
by: Bose, Joy
Published: (2026)
Attention-Based Fusion of IQ and FFT Spectrograms with AoA Features for GNSS Jammer Localization
by: Heublein, Lucas, et al.
Published: (2025)
by: Heublein, Lucas, et al.
Published: (2025)
Taming Polysemanticity in LLMs: Provable Feature Recovery via Sparse Autoencoders
by: Chen, Siyu, et al.
Published: (2025)
by: Chen, Siyu, et al.
Published: (2025)
Long-Term Electricity Demand Prediction Using Non-negative Tensor Factorization and Genetic Algorithm-Driven Temporal Modeling
by: Masaki, Toma, et al.
Published: (2025)
by: Masaki, Toma, et al.
Published: (2025)
SATORIS-N: Spectral Analysis based Traffic Observation Recovery via Informed Subspaces and Nuclear-norm minimization
by: Mohanty, Sampad, et al.
Published: (2026)
by: Mohanty, Sampad, et al.
Published: (2026)
Leibniz's Monadology as Foundation for the Artificial Age Score: A Formal Architecture for Al Memory Evaluation
by: Kayadibi, Seyma Yaman
Published: (2025)
by: Kayadibi, Seyma Yaman
Published: (2025)
On the Optimal Representation Efficiency of Barlow Twins: An Information-Geometric Interpretation
by: Zhang, Di
Published: (2025)
by: Zhang, Di
Published: (2025)
HybridVFL: Disentangled Feature Learning for Edge-Enabled Vertical Federated Multimodal Classification
by: Anoosha, Mostafa, et al.
Published: (2025)
by: Anoosha, Mostafa, et al.
Published: (2025)
Laws of Learning Dynamics and the Core of Learners
by: Jung, Inkee, et al.
Published: (2026)
by: Jung, Inkee, et al.
Published: (2026)
Deep Neural Networks with General Activations: Super-Convergence in Sobolev Norms
by: Yang, Yahong, et al.
Published: (2025)
by: Yang, Yahong, et al.
Published: (2025)
GDNSQ: Gradual Differentiable Noise Scale Quantization for Low-bit Neural Networks
by: Salishev, Sergey, et al.
Published: (2025)
by: Salishev, Sergey, et al.
Published: (2025)
Similar Items
-
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
by: Borobia, Hector, et al.
Published: (2026) -
Where Should LoRA Go? Component-Type Placement in Hybrid Language Models
by: Borobia, Hector, et al.
Published: (2026) -
Component-Aware Self-Speculative Decoding in Hybrid Language Models
by: Borobia, Hector, et al.
Published: (2026) -
Functional Component Ablation Reveals Specialization Patterns in Hybrid Language Model Architectures
by: Borobia, Hector, et al.
Published: (2026) -
Emergence of Quantised Representations Isolated to Anisotropic Functions
by: Bird, George
Published: (2025)