Scaling Laws for Neural Material Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Trikha, Akshay, Chu, Kyle, Gosai, Advait, Szachta, Parker, Weiner, Eric |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Fusing Rewards and Preferences in Reinforcement Learning
por: Khorasani, Sadegh, et al.
Publicado: (2025)
por: Khorasani, Sadegh, et al.
Publicado: (2025)
Versatile Ordering Network: An Attention-based Neural Network for Ordering Across Scales and Quality Metrics
por: Yu, Zehua, et al.
Publicado: (2024)
por: Yu, Zehua, et al.
Publicado: (2024)
The Data Efficiency Frontier of Financial Foundation Models: Scaling Laws from Continued Pretraining
por: Ponnock, Jesse
Publicado: (2025)
por: Ponnock, Jesse
Publicado: (2025)
2Mamba2Furious: Linear in Complexity, Competitive in Accuracy
por: Mongaras, Gabriel, et al.
Publicado: (2026)
por: Mongaras, Gabriel, et al.
Publicado: (2026)
Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales
por: Salfati, Samuel
Publicado: (2026)
por: Salfati, Samuel
Publicado: (2026)
CoxSE: Exploring the Potential of Self-Explaining Neural Networks with Cox Proportional Hazards Model for Survival Analysis
por: Alabdallah, Abdallah, et al.
Publicado: (2024)
por: Alabdallah, Abdallah, et al.
Publicado: (2024)
Self-Expanding Neural Networks
por: Mitchell, Rupert, et al.
Publicado: (2023)
por: Mitchell, Rupert, et al.
Publicado: (2023)
Discrete Latent Structure in Neural Networks
por: Niculae, Vlad, et al.
Publicado: (2023)
por: Niculae, Vlad, et al.
Publicado: (2023)
Implicit Regularization and Generalization in Overparameterized Neural Networks
por: Johannsen, Zeran
Publicado: (2026)
por: Johannsen, Zeran
Publicado: (2026)
Kolmogorov Arnold Networks and Multi-Layer Perceptrons: A Paradigm Shift in Neural Modelling
por: Gaonkar, Aradhya, et al.
Publicado: (2026)
por: Gaonkar, Aradhya, et al.
Publicado: (2026)
Strengthening the Internal Adversarial Robustness in Lifted Neural Networks
por: Zach, Christopher
Publicado: (2025)
por: Zach, Christopher
Publicado: (2025)
Sparse Concept Anchoring for Interpretable and Controllable Neural Representations
por: Fraser, Sandy, et al.
Publicado: (2025)
por: Fraser, Sandy, et al.
Publicado: (2025)
The Bayesian Confidence (BACON) Estimator for Deep Neural Networks
por: Kee, Patrick D., et al.
Publicado: (2024)
por: Kee, Patrick D., et al.
Publicado: (2024)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
por: Hedar, Abdel-Rahman, et al.
Publicado: (2024)
por: Hedar, Abdel-Rahman, et al.
Publicado: (2024)
Learning Useful Representations of Recurrent Neural Network Weight Matrices
por: Herrmann, Vincent, et al.
Publicado: (2024)
por: Herrmann, Vincent, et al.
Publicado: (2024)
Expressivity of Graph Neural Networks Through the Lens of Adversarial Robustness
por: Campi, Francesco, et al.
Publicado: (2023)
por: Campi, Francesco, et al.
Publicado: (2023)
Normalization Layer Per-Example Gradients are Sufficient to Predict Gradient Noise Scale in Transformers
por: Gray, Gavia, et al.
Publicado: (2024)
por: Gray, Gavia, et al.
Publicado: (2024)
ReBoot: Encrypted Training of Deep Neural Networks with CKKS Bootstrapping
por: Pirillo, Alberto, et al.
Publicado: (2025)
por: Pirillo, Alberto, et al.
Publicado: (2025)
Multiple Token Divergence: Measuring and Steering In-Context Computation Density
por: Herrmann, Vincent, et al.
Publicado: (2025)
por: Herrmann, Vincent, et al.
Publicado: (2025)
Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs
por: Easley, Eric, et al.
Publicado: (2026)
por: Easley, Eric, et al.
Publicado: (2026)
Playing Hex and Counter Wargames using Reinforcement Learning and Recurrent Neural Networks
por: Palma, Guilherme, et al.
Publicado: (2025)
por: Palma, Guilherme, et al.
Publicado: (2025)
Exploring Neural Granger Causality with xLSTMs: Unveiling Temporal Dependencies in Complex Data
por: Poonia, Harsh, et al.
Publicado: (2025)
por: Poonia, Harsh, et al.
Publicado: (2025)
Getting ViT in Shape: Scaling Laws for Compute-Optimal Model Design
por: Alabdulmohsin, Ibrahim, et al.
Publicado: (2023)
por: Alabdulmohsin, Ibrahim, et al.
Publicado: (2023)
Closing the Curvature Gap: Full Transformer Hessians and Their Implications for Scaling Laws
por: Petrov, Egor, et al.
Publicado: (2025)
por: Petrov, Egor, et al.
Publicado: (2025)
Understanding Boolean Function Learnability on Deep Neural Networks: PAC Learning Meets Neurosymbolic Models
por: Nicolau, Marcio, et al.
Publicado: (2020)
por: Nicolau, Marcio, et al.
Publicado: (2020)
Uncertainty Quantification in Multivariable Regression for Material Property Prediction with Bayesian Neural Networks
por: Li, Longze, et al.
Publicado: (2023)
por: Li, Longze, et al.
Publicado: (2023)
Graph Neural Network Based Action Ranking for Planning
por: Mangannavar, Rajesh, et al.
Publicado: (2024)
por: Mangannavar, Rajesh, et al.
Publicado: (2024)
Pruning Spurious Subgraphs for Graph Out-of-Distribution Generalization
por: Yao, Tianjun, et al.
Publicado: (2025)
por: Yao, Tianjun, et al.
Publicado: (2025)
FluidWorld: Reaction-Diffusion Dynamics as a Predictive Substrate for World Models
por: Polly, Fabien
Publicado: (2026)
por: Polly, Fabien
Publicado: (2026)
Pre-trained Models Perform the Best When Token Distributions Follow Zipf's Law
por: He, Yanjin, et al.
Publicado: (2025)
por: He, Yanjin, et al.
Publicado: (2025)
Downsized and Compromised?: Assessing the Faithfulness of Model Compression
por: Kamal, Moumita, et al.
Publicado: (2025)
por: Kamal, Moumita, et al.
Publicado: (2025)
Seasoning Generative Models for a Generalization Aftertaste
por: Husain, Hisham, et al.
Publicado: (2026)
por: Husain, Hisham, et al.
Publicado: (2026)
Explanations Based on Item Response Theory (eXirt): A Model-Specific Method to Explain Tree-Ensemble Model in Trust Perspective
por: Ribeiro, José, et al.
Publicado: (2022)
por: Ribeiro, José, et al.
Publicado: (2022)
Measuring IIA Violations in Similarity Choices with Bayesian Models
por: Corrêa, Hugo Sales, et al.
Publicado: (2025)
por: Corrêa, Hugo Sales, et al.
Publicado: (2025)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
por: Yousaf, Iqra
Publicado: (2024)
por: Yousaf, Iqra
Publicado: (2024)
Tempered Calculus for ML: Application to Hyperbolic Model Embedding
por: Nock, Richard, et al.
Publicado: (2024)
por: Nock, Richard, et al.
Publicado: (2024)
CASSANDRA: Programmatic and Probabilistic Learning and Inference for Stochastic World Modeling
por: Lymperopoulos, Panagiotis, et al.
Publicado: (2026)
por: Lymperopoulos, Panagiotis, et al.
Publicado: (2026)
TFMAdapter: Lightweight Instance-Level Adaptation of Foundation Models for Forecasting with Covariates
por: Dange, Afrin, et al.
Publicado: (2025)
por: Dange, Afrin, et al.
Publicado: (2025)
Securing Reliability: A Brief Overview on Enhancing In-Context Learning for Foundation Models
por: Huang, Yunpeng, et al.
Publicado: (2024)
por: Huang, Yunpeng, et al.
Publicado: (2024)
HGCN(O): A Self-Tuning GCN HyperModel Toolkit for Outcome Prediction in Event-Sequence Data
por: Wang, Fang, et al.
Publicado: (2025)
por: Wang, Fang, et al.
Publicado: (2025)
Ejemplares similares
-
Fusing Rewards and Preferences in Reinforcement Learning
por: Khorasani, Sadegh, et al.
Publicado: (2025) -
Versatile Ordering Network: An Attention-based Neural Network for Ordering Across Scales and Quality Metrics
por: Yu, Zehua, et al.
Publicado: (2024) -
The Data Efficiency Frontier of Financial Foundation Models: Scaling Laws from Continued Pretraining
por: Ponnock, Jesse
Publicado: (2025) -
2Mamba2Furious: Linear in Complexity, Competitive in Accuracy
por: Mongaras, Gabriel, et al.
Publicado: (2026) -
Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales
por: Salfati, Samuel
Publicado: (2026)