Neural Sum-of-Squares: Certifying the Nonnegativity of Polynomials with Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Pelleriti, Nico, Spiegel, Christoph, Liu, Shiwei, Martínez-Rubio, David, Zimmer, Max, Pokutta, Sebastian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Agentic Researcher: A Practical Guide to AI-Assisted Research in Mathematics and Machine Learning
by: Zimmer, Max, et al.
Published: (2026)
by: Zimmer, Max, et al.
Published: (2026)
Sparse Model Soups: A Recipe for Improved Pruning via Model Averaging
by: Zimmer, Max, et al.
Published: (2023)
by: Zimmer, Max, et al.
Published: (2023)
PERP: Rethinking the Prune-Retrain Paradigm in the Era of LLMs
by: Zimmer, Max, et al.
Published: (2023)
by: Zimmer, Max, et al.
Published: (2023)
Approximating Latent Manifolds in Neural Networks via Vanishing Ideals
by: Pelleriti, Nico, et al.
Published: (2025)
by: Pelleriti, Nico, et al.
Published: (2025)
Computational Algebra with Attention: Transformer Oracles for Border Basis Algorithms
by: Kera, Hiroshi, et al.
Published: (2025)
by: Kera, Hiroshi, et al.
Published: (2025)
Compression-aware Training of Neural Networks using Frank-Wolfe
by: Zimmer, Max, et al.
Published: (2022)
by: Zimmer, Max, et al.
Published: (2022)
A Free Lunch in LLM Compression: Revisiting Retraining after Pruning
by: Wagner, Moritz, et al.
Published: (2025)
by: Wagner, Moritz, et al.
Published: (2025)
On the Byzantine-Resilience of Distillation-Based Federated Learning
by: Roux, Christophe, et al.
Published: (2024)
by: Roux, Christophe, et al.
Published: (2024)
SparseSwaps: Tractable LLM Pruning Mask Refinement at Scale
by: Zimmer, Max, et al.
Published: (2025)
by: Zimmer, Max, et al.
Published: (2025)
Neural Discovery in Mathematics: Do Machines Dream of Colored Planes?
by: Mundinger, Konrad, et al.
Published: (2025)
by: Mundinger, Konrad, et al.
Published: (2025)
From Associations to Activations: Comparing Behavioral and Hidden-State Semantic Geometry in LLMs
by: Schiekiera, Louis, et al.
Published: (2026)
by: Schiekiera, Louis, et al.
Published: (2026)
Interpretability Guarantees with Merlin-Arthur Classifiers
by: Wäldchen, Stephan, et al.
Published: (2022)
by: Wäldchen, Stephan, et al.
Published: (2022)
Neural Parameter Regression for Explicit Representations of PDE Solution Operators
by: Mundinger, Konrad, et al.
Published: (2024)
by: Mundinger, Konrad, et al.
Published: (2024)
Capturing Temporal Dynamics in Large-Scale Canopy Tree Height Estimation
by: Pauls, Jan, et al.
Published: (2025)
by: Pauls, Jan, et al.
Published: (2025)
Sum of Squares Circuits
by: Loconte, Lorenzo, et al.
Published: (2024)
by: Loconte, Lorenzo, et al.
Published: (2024)
Certifying Global Robustness for Deep Neural Networks
by: Li, You, et al.
Published: (2024)
by: Li, You, et al.
Published: (2024)
Estimating Canopy Height at Scale
by: Pauls, Jan, et al.
Published: (2024)
by: Pauls, Jan, et al.
Published: (2024)
Transformers Don't In-Context Learn Least Squares Regression
by: Hill, Joshua, et al.
Published: (2025)
by: Hill, Joshua, et al.
Published: (2025)
ECHOSAT: Estimating Canopy Height Over Space And Time
by: Pauls, Jan, et al.
Published: (2026)
by: Pauls, Jan, et al.
Published: (2026)
Polynomial Selection in Spectral Graph Neural Networks: An Error-Sum of Function Slices Approach
by: Li, Guoming, et al.
Published: (2024)
by: Li, Guoming, et al.
Published: (2024)
ZeroS: Zero-Sum Linear Attention for Efficient Transformers
by: Lu, Jiecheng, et al.
Published: (2026)
by: Lu, Jiecheng, et al.
Published: (2026)
When Does Sparsity Mitigate the Curse of Depth in LLMs
by: Muhtar, Dilxat, et al.
Published: (2026)
by: Muhtar, Dilxat, et al.
Published: (2026)
Beyond Short Steps in Frank-Wolfe Algorithms
by: Martínez-Rubio, David, et al.
Published: (2025)
by: Martínez-Rubio, David, et al.
Published: (2025)
Transformer-Squared: Self-adaptive LLMs
by: Sun, Qi, et al.
Published: (2025)
by: Sun, Qi, et al.
Published: (2025)
Polynormer: Polynomial-Expressive Graph Transformer in Linear Time
by: Deng, Chenhui, et al.
Published: (2024)
by: Deng, Chenhui, et al.
Published: (2024)
BlockCert: Certified Blockwise Extraction of Transformer Mechanisms
by: Andric, Sandro
Published: (2025)
by: Andric, Sandro
Published: (2025)
Neural Collapse is Globally Optimal in Deep Regularized ResNets and Transformers
by: Súkeník, Peter, et al.
Published: (2025)
by: Súkeník, Peter, et al.
Published: (2025)
FairProof : Confidential and Certifiable Fairness for Neural Networks
by: Yadav, Chhavi, et al.
Published: (2024)
by: Yadav, Chhavi, et al.
Published: (2024)
Certified Signed Graph Unlearning
by: Zhao, Junpeng, et al.
Published: (2025)
by: Zhao, Junpeng, et al.
Published: (2025)
Certifying Counterfactual Bias in LLMs
by: Chaudhary, Isha, et al.
Published: (2024)
by: Chaudhary, Isha, et al.
Published: (2024)
How to Square Tensor Networks and Circuits Without Squaring Them
by: Loconte, Lorenzo, et al.
Published: (2025)
by: Loconte, Lorenzo, et al.
Published: (2025)
SCFormer: Structured Channel-wise Transformer with Cumulative Historical State for Multivariate Time Series Forecasting
by: Guo, Shiwei, et al.
Published: (2025)
by: Guo, Shiwei, et al.
Published: (2025)
Certified Policy Optimisation for Nested Causal Bandits via PAC-Bayes Risk
by: Woydt, Tim, et al.
Published: (2026)
by: Woydt, Tim, et al.
Published: (2026)
Efficient Certified Reasoning for Binarized Neural Networks
by: Yang, Jiong, et al.
Published: (2025)
by: Yang, Jiong, et al.
Published: (2025)
Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN
by: Li, Pengxiang, et al.
Published: (2024)
by: Li, Pengxiang, et al.
Published: (2024)
Sum-of-Squares Programming for Ma-Trudinger-Wang Regularity of Optimal Transport Maps
by: Shivakumar, Sachin, et al.
Published: (2024)
by: Shivakumar, Sachin, et al.
Published: (2024)
Tensor Polynomial Additive Model
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
Neural Concept Verifier: Scaling Prover-Verifier Games via Concept Encodings
by: Turan, Berkant, et al.
Published: (2025)
by: Turan, Berkant, et al.
Published: (2025)
Ordinary Least Squares is a Special Case of Transformer
by: Tan, Xiaojun, et al.
Published: (2026)
by: Tan, Xiaojun, et al.
Published: (2026)
KNARsack: Teaching Neural Algorithmic Reasoners to Solve Pseudo-Polynomial Problems
by: Požgaj, Stjepan, et al.
Published: (2025)
by: Požgaj, Stjepan, et al.
Published: (2025)
Similar Items
-
The Agentic Researcher: A Practical Guide to AI-Assisted Research in Mathematics and Machine Learning
by: Zimmer, Max, et al.
Published: (2026) -
Sparse Model Soups: A Recipe for Improved Pruning via Model Averaging
by: Zimmer, Max, et al.
Published: (2023) -
PERP: Rethinking the Prune-Retrain Paradigm in the Era of LLMs
by: Zimmer, Max, et al.
Published: (2023) -
Approximating Latent Manifolds in Neural Networks via Vanishing Ideals
by: Pelleriti, Nico, et al.
Published: (2025) -
Computational Algebra with Attention: Transformer Oracles for Border Basis Algorithms
by: Kera, Hiroshi, et al.
Published: (2025)