stable-pretraining-v1: Foundation Model Research Made Simple
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Balestriero, Randall, Van Assel, Hugues, BuGhanem, Sami, Maes, Lucas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PyTorch-based Geometric Learning with Non-CUDA Processing Units: Experiences from Intel Gaudi-v2 HPUs
von: Bu, Fanchen, et al.
Veröffentlicht: (2025)
von: Bu, Fanchen, et al.
Veröffentlicht: (2025)
Adversarial Feature Map Pruning for Backdoor
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
SWE-Arena: An Interactive Platform for Evaluating Foundation Models in Software Engineering
von: Zhao, Zhimin
Veröffentlicht: (2025)
von: Zhao, Zhimin
Veröffentlicht: (2025)
On the Workflows and Smells of Leaderboard Operations (LBOps): An Exploratory Study of Foundation Model Leaderboards
von: Zhao, Zhimin, et al.
Veröffentlicht: (2024)
von: Zhao, Zhimin, et al.
Veröffentlicht: (2024)
Ditch the Denoiser: Emergence of Noise Robustness in Self-Supervised Learning from Data Curriculum
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
Towards MLOps: A DevOps Tools Recommender System for Machine Learning System
von: Shah, Pir Sami Ullah, et al.
Veröffentlicht: (2024)
von: Shah, Pir Sami Ullah, et al.
Veröffentlicht: (2024)
Foundation Model Engineering: Engineering Foundation Models Just as Engineering Software
von: Ran, Dezhi, et al.
Veröffentlicht: (2024)
von: Ran, Dezhi, et al.
Veröffentlicht: (2024)
Deploying Geospatial Foundation Models in the Real World: Lessons from WorldCereal
von: Butsko, Christina, et al.
Veröffentlicht: (2025)
von: Butsko, Christina, et al.
Veröffentlicht: (2025)
Automating the Enterprise with Foundation Models
von: Wornow, Michael, et al.
Veröffentlicht: (2024)
von: Wornow, Michael, et al.
Veröffentlicht: (2024)
Evaluating the Process Modeling Abilities of Large Language Models -- Preliminary Foundations and Results
von: Fettke, Peter, et al.
Veröffentlicht: (2025)
von: Fettke, Peter, et al.
Veröffentlicht: (2025)
Shapley-Guided Neural Repair Approach via Derivative-Free Optimization
von: Sun, Xinyu, et al.
Veröffentlicht: (2026)
von: Sun, Xinyu, et al.
Veröffentlicht: (2026)
Perspective of Software Engineering Researchers on Machine Learning Practices Regarding Research, Review, and Education
von: Mojica-Hanke, Anamaria, et al.
Veröffentlicht: (2024)
von: Mojica-Hanke, Anamaria, et al.
Veröffentlicht: (2024)
LLMs in Coding and their Impact on the Commercial Software Engineering Landscape
von: Belozerov, Vladislav, et al.
Veröffentlicht: (2025)
von: Belozerov, Vladislav, et al.
Veröffentlicht: (2025)
Polynomiogram: An Integrated Framework for Root Visualization and Generative Art
von: Nguyen, Hoang Duc, et al.
Veröffentlicht: (2025)
von: Nguyen, Hoang Duc, et al.
Veröffentlicht: (2025)
Which Programming Language and Model Work Best With LLM-as-a-Judge For Code Retrieval?
von: Roberts, Lucas, et al.
Veröffentlicht: (2025)
von: Roberts, Lucas, et al.
Veröffentlicht: (2025)
Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection
von: Khajezade, Mohamad, et al.
Veröffentlicht: (2026)
von: Khajezade, Mohamad, et al.
Veröffentlicht: (2026)
The Impact of Hyperparameters on Large Language Model Inference Performance: An Evaluation of vLLM and HuggingFace Pipelines
von: Martinez, Matias
Veröffentlicht: (2024)
von: Martinez, Matias
Veröffentlicht: (2024)
The Hitchhikers Guide to Production-ready Trustworthy Foundation Model powered Software (FMware)
von: Vasilevski, Kirill, et al.
Veröffentlicht: (2025)
von: Vasilevski, Kirill, et al.
Veröffentlicht: (2025)
WONDERBREAD: A Benchmark for Evaluating Multimodal Foundation Models on Business Process Management Tasks
von: Wornow, Michael, et al.
Veröffentlicht: (2024)
von: Wornow, Michael, et al.
Veröffentlicht: (2024)
More Rigorous Software Engineering Would Improve Reproducibility in Machine Learning Research
von: Wolter, Moritz, et al.
Veröffentlicht: (2025)
von: Wolter, Moritz, et al.
Veröffentlicht: (2025)
ExplainFuzz: Explainable and Constraint-Conditioned Test Generation with Probabilistic Circuits
von: Baiget, Annaëlle, et al.
Veröffentlicht: (2026)
von: Baiget, Annaëlle, et al.
Veröffentlicht: (2026)
FROAV: A Framework for RAG Observation and Agent Verification -- Lowering the Barrier to LLM Agent Research
von: Lin, Tzu-Hsuan, et al.
Veröffentlicht: (2026)
von: Lin, Tzu-Hsuan, et al.
Veröffentlicht: (2026)
Joint Embedding vs Reconstruction: Provable Benefits of Latent Space Prediction for Self Supervised Learning
von: Van Assel, Hugues, et al.
Veröffentlicht: (2025)
von: Van Assel, Hugues, et al.
Veröffentlicht: (2025)
RAMBO: Enhancing RAG-based Repository-Level Method Body Completion
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2024)
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2024)
LEANCODE: Understanding Models Better for Code Simplification of Pre-trained Large Language Models
von: Wang, Yan, et al.
Veröffentlicht: (2025)
von: Wang, Yan, et al.
Veröffentlicht: (2025)
Applying Large Language Models to Issue Classification: Revisiting with Extended Data and New Models
von: Aracena, Gabriel, et al.
Veröffentlicht: (2025)
von: Aracena, Gabriel, et al.
Veröffentlicht: (2025)
Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation
von: Jacopin, Éric
Veröffentlicht: (2026)
von: Jacopin, Éric
Veröffentlicht: (2026)
An Efficient Model Maintenance Approach for MLOps
von: Majidi, Forough, et al.
Veröffentlicht: (2024)
von: Majidi, Forough, et al.
Veröffentlicht: (2024)
Calibration and Correctness of Language Models for Code
von: Spiess, Claudio, et al.
Veröffentlicht: (2024)
von: Spiess, Claudio, et al.
Veröffentlicht: (2024)
How Do Model Export Formats Impact the Development of ML-Enabled Systems? A Case Study on Model Integration
von: Parida, Shreyas Kumar, et al.
Veröffentlicht: (2025)
von: Parida, Shreyas Kumar, et al.
Veröffentlicht: (2025)
Code Graph Model (CGM): A Graph-Integrated Large Language Model for Repository-Level Software Engineering Tasks
von: Tao, Hongyuan, et al.
Veröffentlicht: (2025)
von: Tao, Hongyuan, et al.
Veröffentlicht: (2025)
Understanding Robustness of Model Editing in Code LLMs
von: Chhetri, Vinaik, et al.
Veröffentlicht: (2025)
von: Chhetri, Vinaik, et al.
Veröffentlicht: (2025)
Comparative Analysis of AWS Model Deployment Services
von: Bagai, Rahul
Veröffentlicht: (2024)
von: Bagai, Rahul
Veröffentlicht: (2024)
Expert-Driven Monitoring of Operational ML Models
von: Leest, Joran, et al.
Veröffentlicht: (2024)
von: Leest, Joran, et al.
Veröffentlicht: (2024)
How do Machine Learning Models Change?
von: Castaño, Joel, et al.
Veröffentlicht: (2024)
von: Castaño, Joel, et al.
Veröffentlicht: (2024)
Model Cascading for Code: A Cascaded Black-Box Multi-Model Framework for Cost-Efficient Code Completion with Self-Testing
von: Chen, Boyuan, et al.
Veröffentlicht: (2024)
von: Chen, Boyuan, et al.
Veröffentlicht: (2024)
Concolic Testing on Individual Fairness of Neural Network Models
von: Huang, Ming-I, et al.
Veröffentlicht: (2025)
von: Huang, Ming-I, et al.
Veröffentlicht: (2025)
Towards a Small Language Model Lifecycle Framework
von: Miraghaei, Parsa, et al.
Veröffentlicht: (2025)
von: Miraghaei, Parsa, et al.
Veröffentlicht: (2025)
Integrating Large Language Models for Automated Structural Analysis
von: Liang, Haoran, et al.
Veröffentlicht: (2025)
von: Liang, Haoran, et al.
Veröffentlicht: (2025)
Data Requirement Goal Modeling for Machine Learning Systems
von: Yamani, Asma, et al.
Veröffentlicht: (2025)
von: Yamani, Asma, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PyTorch-based Geometric Learning with Non-CUDA Processing Units: Experiences from Intel Gaudi-v2 HPUs
von: Bu, Fanchen, et al.
Veröffentlicht: (2025) -
Adversarial Feature Map Pruning for Backdoor
von: Huang, Dong, et al.
Veröffentlicht: (2023) -
SWE-Arena: An Interactive Platform for Evaluating Foundation Models in Software Engineering
von: Zhao, Zhimin
Veröffentlicht: (2025) -
On the Workflows and Smells of Leaderboard Operations (LBOps): An Exploratory Study of Foundation Model Leaderboards
von: Zhao, Zhimin, et al.
Veröffentlicht: (2024) -
Ditch the Denoiser: Emergence of Noise Robustness in Self-Supervised Learning from Data Curriculum
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)