GenAI for Systems: Recurring Challenges and Design Principles from Software to Silicon
Fuente:
arXiv
Saved in:
| Main Authors: | Tschand, Arya, Wang, Chenyu, Wan, Zishen, Cheng, Andrew, Cristescu, Ioana, He, Kevin, Huang, Howard, Ingare, Alexander, Kangaslahti, Akseli, Kangaslahti, Sara, Lebryk, Theo, Lin, Hongjin, Ma, Jeffrey Jian, Meterez, Alexandru, Mohri, Clara, Morwani, Depen, Qin, Sunny, Rinberg, Roy, Rodriguez-Diaz, Paula, Taliotis, Alyssa Mia, Fathi, Pernille Undrum, Zhao, Rosie, Zhou, Todd, Reddi, Vijay Janapa |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dynamic Targeting of Satellite Observations Using Supplemental Geostationary Satellite Data and Hierarchical Planning
by: Kangaslahti, Akseli, et al.
Published: (2026)
by: Kangaslahti, Akseli, et al.
Published: (2026)
Anytime Pretraining: Horizon-Free Learning-Rate Schedules with Weight Averaging
by: Meterez, Alexandru, et al.
Published: (2026)
by: Meterez, Alexandru, et al.
Published: (2026)
Network-Based Interventions for HIV Prevention via Cascade-Aware Suppression of Transmission
by: Kangaslahti, Akseli, et al.
Published: (2026)
by: Kangaslahti, Akseli, et al.
Published: (2026)
Continuous Language Model Interpolation for Dynamic and Controllable Text Generation
by: Kangaslahti, Sara, et al.
Published: (2024)
by: Kangaslahti, Sara, et al.
Published: (2024)
A Simplified Analysis of SGD for Linear Regression with Weight Averaging
by: Meterez, Alexandru, et al.
Published: (2025)
by: Meterez, Alexandru, et al.
Published: (2025)
Seesaw: Accelerating Training by Balancing Learning Rate and Batch Size Scheduling
by: Meterez, Alexandru, et al.
Published: (2025)
by: Meterez, Alexandru, et al.
Published: (2025)
The Recurrent Transformer: Greater Effective Depth and Efficient Decoding
by: Oncescu, Costin-Andrei, et al.
Published: (2026)
by: Oncescu, Costin-Andrei, et al.
Published: (2026)
Policy-Embedded Graph Expansion: Networked HIV Testing with Diffusion-Driven Network Samples
by: Kangaslahti, Akseli, et al.
Published: (2026)
by: Kangaslahti, Akseli, et al.
Published: (2026)
Deconstructing What Makes a Good Optimizer for Language Models
by: Zhao, Rosie, et al.
Published: (2024)
by: Zhao, Rosie, et al.
Published: (2024)
Hidden Breakthroughs in Language Model Training
by: Kangaslahti, Sara, et al.
Published: (2025)
by: Kangaslahti, Sara, et al.
Published: (2025)
Learning-Based Planning for Improving Science Return of Earth Observation Satellites
by: Breitfeld, Abigail, et al.
Published: (2025)
by: Breitfeld, Abigail, et al.
Published: (2025)
TinyTorch: Building Machine Learning Systems from First Principles
by: Reddi, Vijay Janapa
Published: (2026)
by: Reddi, Vijay Janapa
Published: (2026)
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
by: Tschand, Arya, et al.
Published: (2025)
by: Tschand, Arya, et al.
Published: (2025)
Feature emergence via margin maximization: case studies in algebraic tasks
by: Morwani, Depen, et al.
Published: (2023)
by: Morwani, Depen, et al.
Published: (2023)
Beyond Implicit Bias: The Insignificance of SGD Noise in Online Learning
by: Vyas, Nikhil, et al.
Published: (2023)
by: Vyas, Nikhil, et al.
Published: (2023)
Inverse Depth Scaling From Most Layers Being Similar
by: Liu, Yizhou, et al.
Published: (2026)
by: Liu, Yizhou, et al.
Published: (2026)
Fast Forwarding Low-Rank Training
by: Rahamim, Adir, et al.
Published: (2024)
by: Rahamim, Adir, et al.
Published: (2024)
Connections between Schedule-Free Optimizers, AdEMAMix, and Accelerated SGD Variants
by: Morwani, Depen, et al.
Published: (2025)
by: Morwani, Depen, et al.
Published: (2025)
The Potential of Second-Order Optimization for LLMs: A Study with Full Gauss-Newton
by: Abreu, Natalie, et al.
Published: (2025)
by: Abreu, Natalie, et al.
Published: (2025)
Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions
by: Kong, Lingkai, et al.
Published: (2026)
by: Kong, Lingkai, et al.
Published: (2026)
Generative AI Agents in Autonomous Machines: A Safety Perspective
by: Jabbour, Jason, et al.
Published: (2024)
by: Jabbour, Jason, et al.
Published: (2024)
SOAP: Improving and Stabilizing Shampoo using Adam
by: Vyas, Nikhil, et al.
Published: (2024)
by: Vyas, Nikhil, et al.
Published: (2024)
Slm-mux: Orchestrating small language models for reasoning
by: Wang, Chenyu, et al.
Published: (2025)
by: Wang, Chenyu, et al.
Published: (2025)
The Magnificent Seven Challenges and Opportunities in Domain-Specific Accelerator Design for Autonomous Systems
by: Neuman, Sabrina M., et al.
Published: (2024)
by: Neuman, Sabrina M., et al.
Published: (2024)
A New Perspective on Shampoo's Preconditioner
by: Morwani, Depen, et al.
Published: (2024)
by: Morwani, Depen, et al.
Published: (2024)
Analyzing Political Text at Scale with Online Tensor LDA
by: Kangaslahti, Sara, et al.
Published: (2025)
by: Kangaslahti, Sara, et al.
Published: (2025)
Surface Modification for III-V Selective Area Molecular Beam Epitaxy of Non-Selective Mask Materials
by: García, Ashlee M., et al.
Published: (2026)
by: García, Ashlee M., et al.
Published: (2026)
LOTION: Smoothing the Optimization Landscape for Quantized Training
by: Kwun, Mujin, et al.
Published: (2025)
by: Kwun, Mujin, et al.
Published: (2025)
FedStaleWeight: Buffered Asynchronous Federated Learning with Fair Aggregation via Staleness Reweighting
by: Ma, Jeffrey, et al.
Published: (2024)
by: Ma, Jeffrey, et al.
Published: (2024)
Boomerang Distillation Enables Zero-Shot Model Size Interpolation
by: Kangaslahti, Sara, et al.
Published: (2025)
by: Kangaslahti, Sara, et al.
Published: (2025)
Adam or Gauss-Newton? A Comparative Study In Terms of Basis Alignment and SGD Noise
by: Liu, Bingbin, et al.
Published: (2025)
by: Liu, Bingbin, et al.
Published: (2025)
Generative AI in Embodied Systems: System-Level Analysis of Performance, Efficiency and Scalability
by: Wan, Zishen, et al.
Published: (2025)
by: Wan, Zishen, et al.
Published: (2025)
QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture
by: Prakash, Shvetank, et al.
Published: (2025)
by: Prakash, Shvetank, et al.
Published: (2025)
Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining
by: Zhao, Rosie, et al.
Published: (2025)
by: Zhao, Rosie, et al.
Published: (2025)
SocratiQ: A Generative AI-Powered Learning Companion for Personalized Education and Broader Accessibility
by: Jabbour, Jason, et al.
Published: (2025)
by: Jabbour, Jason, et al.
Published: (2025)
Lifetime-Aware Design for Item-Level Intelligence at the Extreme Edge
by: Prakash, Shvetank, et al.
Published: (2025)
by: Prakash, Shvetank, et al.
Published: (2025)
How Does Critical Batch Size Scale in Pre-training?
by: Zhang, Hanlin, et al.
Published: (2024)
by: Zhang, Hanlin, et al.
Published: (2024)
Tabula: Efficiently Computing Nonlinear Activation Functions for Secure Neural Network Inference
by: Lam, Maximilian, et al.
Published: (2022)
by: Lam, Maximilian, et al.
Published: (2022)
QuArch: A Question-Answering Dataset for AI Agents in Computer Architecture
by: Prakash, Shvetank, et al.
Published: (2025)
by: Prakash, Shvetank, et al.
Published: (2025)
Rigorous Evaluation of Microarchitectural Side-Channels with Statistical Model Checking
by: Li, Weihang, et al.
Published: (2025)
by: Li, Weihang, et al.
Published: (2025)
Similar Items
-
Dynamic Targeting of Satellite Observations Using Supplemental Geostationary Satellite Data and Hierarchical Planning
by: Kangaslahti, Akseli, et al.
Published: (2026) -
Anytime Pretraining: Horizon-Free Learning-Rate Schedules with Weight Averaging
by: Meterez, Alexandru, et al.
Published: (2026) -
Network-Based Interventions for HIV Prevention via Cascade-Aware Suppression of Transmission
by: Kangaslahti, Akseli, et al.
Published: (2026) -
Continuous Language Model Interpolation for Dynamic and Controllable Text Generation
by: Kangaslahti, Sara, et al.
Published: (2024) -
A Simplified Analysis of SGD for Linear Regression with Weight Averaging
by: Meterez, Alexandru, et al.
Published: (2025)