Saved in:
| Main Authors: | Hosseini, Ryien, Simini, Filippo, Vishwanath, Venkatram, Willett, Rebecca, Hoffmann, Henry |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2503.01720 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sketch-Augmented Features Improve Learning Long-Range Dependencies in Graph Neural Networks
by: Hosseini, Ryien, et al.
Published: (2025)
by: Hosseini, Ryien, et al.
Published: (2025)
A Deep Probabilistic Framework for Continuous Time Dynamic Graph Generation
by: Hosseini, Ryien, et al.
Published: (2024)
by: Hosseini, Ryien, et al.
Published: (2024)
Observation, Not Prediction: Conversation-Level Disaggregated Scheduling for Agentic Serving
by: Ding, Jianru, et al.
Published: (2026)
by: Ding, Jianru, et al.
Published: (2026)
LExI: Layer-Adaptive Active Experts for Efficient MoE Model Inference
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
Extending $μ$P: Spectral Conditions for Feature Learning Across Optimizers
by: Gupta, Akshita, et al.
Published: (2026)
by: Gupta, Akshita, et al.
Published: (2026)
Swimba: Switch Mamba Model Scales State Space Models
by: Du, Zhixu, et al.
Published: (2026)
by: Du, Zhixu, et al.
Published: (2026)
Scalable and Consistent Graph Neural Networks for Distributed Mesh-based Data-driven Modeling
by: Barwey, Shivam, et al.
Published: (2024)
by: Barwey, Shivam, et al.
Published: (2024)
ReLU Neural Networks with Linear Layers are Biased Towards Single- and Multi-Index Models
by: Parkinson, Suzanna, et al.
Published: (2023)
by: Parkinson, Suzanna, et al.
Published: (2023)
PreLoRA: Hybrid Pre-training of Vision Transformers with Full Training and Low-Rank Adapters
by: Thapa, Krishu K, et al.
Published: (2025)
by: Thapa, Krishu K, et al.
Published: (2025)
BaKlaVa -- Budgeted Allocation of KV cache for Long-context Inference
by: Gulhan, Ahmed Burak, et al.
Published: (2025)
by: Gulhan, Ahmed Burak, et al.
Published: (2025)
PagedEviction: Structured Block-wise KV Cache Pruning for Efficient Large Language Model Inference
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
Mesh-based Super-Resolution of Fluid Flows with Multiscale Graph Neural Networks
by: Barwey, Shivam, et al.
Published: (2024)
by: Barwey, Shivam, et al.
Published: (2024)
MoE-Inference-Bench: Performance Evaluation of Mixture of Expert Large Language and Vision Models
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
How do simple rotations affect the implicit bias of Adam?
by: DePavia, Adela, et al.
Published: (2025)
by: DePavia, Adela, et al.
Published: (2025)
Embed and Emulate: Contrastive representations for simulation-based inference
by: Jiang, Ruoxi, et al.
Published: (2024)
by: Jiang, Ruoxi, et al.
Published: (2024)
LLM-Inference-Bench: Inference Benchmarking of Large Language Models on AI Accelerators
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2024)
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2024)
Integrating Uncertainty Awareness into Conformalized Quantile Regression
by: Rossellini, Raphael, et al.
Published: (2023)
by: Rossellini, Raphael, et al.
Published: (2023)
Stabilizing black-box model selection with the inflated argmax
by: Adrian, Melissa, et al.
Published: (2024)
by: Adrian, Melissa, et al.
Published: (2024)
Data Assimilation with Machine Learning Surrogate Models: A Case Study with FourCastNet
by: Adrian, Melissa, et al.
Published: (2024)
by: Adrian, Melissa, et al.
Published: (2024)
Auto-differentiable data assimilation: Co-learning of states, dynamics, and filtering algorithms
by: Adrian, Melissa, et al.
Published: (2026)
by: Adrian, Melissa, et al.
Published: (2026)
Building a stable classifier with the inflated argmax
by: Soloff, Jake A., et al.
Published: (2024)
by: Soloff, Jake A., et al.
Published: (2024)
Bagging Provides Assumption-free Stability
by: Soloff, Jake A., et al.
Published: (2023)
by: Soloff, Jake A., et al.
Published: (2023)
Depth Separation in Norm-Bounded Infinite-Width Neural Networks
by: Parkinson, Suzanna, et al.
Published: (2024)
by: Parkinson, Suzanna, et al.
Published: (2024)
Learning Paths for Dynamic Measure Transport: A Control Perspective
by: Maurais, Aimee, et al.
Published: (2025)
by: Maurais, Aimee, et al.
Published: (2025)
Training neural operators to preserve invariant measures of chaotic attractors
by: Jiang, Ruoxi, et al.
Published: (2023)
by: Jiang, Ruoxi, et al.
Published: (2023)
Beyond Ensemble Averages: Leveraging Climate Model Ensembles for Subseasonal Forecasting
by: Orlova, Elena, et al.
Published: (2022)
by: Orlova, Elena, et al.
Published: (2022)
A Model-Guided Neural Network Method for the Inverse Scattering Problem
by: Tsang, Olivia, et al.
Published: (2025)
by: Tsang, Olivia, et al.
Published: (2025)
Solving Inverse Problems with Deep Linear Neural Networks: Global Convergence Guarantees for Gradient Descent with Weight Decay
by: Laus, Hannah, et al.
Published: (2025)
by: Laus, Hannah, et al.
Published: (2025)
Accelerating PDE Surrogates via RL-Guided Mesh Optimization
by: Meng, Yang, et al.
Published: (2026)
by: Meng, Yang, et al.
Published: (2026)
Assumption-free stability for ranking problems
by: Liang, Ruiting, et al.
Published: (2025)
by: Liang, Ruiting, et al.
Published: (2025)
Mean-Field Langevin Dynamics for Signed Measures via a Bilevel Approach
by: Wang, Guillaume, et al.
Published: (2024)
by: Wang, Guillaume, et al.
Published: (2024)
Deep Stochastic Mechanics
by: Orlova, Elena, et al.
Published: (2023)
by: Orlova, Elena, et al.
Published: (2023)
Can a calibration metric be both testable and actionable?
by: Rossellini, Raphael, et al.
Published: (2025)
by: Rossellini, Raphael, et al.
Published: (2025)
Minimax Rates for Hyperbolic Hierarchical Learning
by: Rawal, Divit, et al.
Published: (2026)
by: Rawal, Divit, et al.
Published: (2026)
A Generalized Tikhonov Layer for Interpretable-by-design Graph Neural Networks
by: Tremblay, Nicolas, et al.
Published: (2026)
by: Tremblay, Nicolas, et al.
Published: (2026)
Discount Model Search for Quality Diversity Optimization in High-Dimensional Measure Spaces
by: Tjanaka, Bryon, et al.
Published: (2026)
by: Tjanaka, Bryon, et al.
Published: (2026)
Statistical Guarantees in Synthetic Data through Conformal Adversarial Generation
by: Vishwakarma, Rahul, et al.
Published: (2025)
by: Vishwakarma, Rahul, et al.
Published: (2025)
Hierarchical Implicit Neural Emulators
by: Jiang, Ruoxi, et al.
Published: (2025)
by: Jiang, Ruoxi, et al.
Published: (2025)
Comparing Methods for Bias Mitigation in Graph Neural Networks
by: Hoffmann, Barbara, et al.
Published: (2025)
by: Hoffmann, Barbara, et al.
Published: (2025)
Certified Guidance for Planning with Deep Generative Models
by: Giacomarra, Francesco, et al.
Published: (2025)
by: Giacomarra, Francesco, et al.
Published: (2025)
Similar Items
-
Sketch-Augmented Features Improve Learning Long-Range Dependencies in Graph Neural Networks
by: Hosseini, Ryien, et al.
Published: (2025) -
A Deep Probabilistic Framework for Continuous Time Dynamic Graph Generation
by: Hosseini, Ryien, et al.
Published: (2024) -
Observation, Not Prediction: Conversation-Level Disaggregated Scheduling for Agentic Serving
by: Ding, Jianru, et al.
Published: (2026) -
LExI: Layer-Adaptive Active Experts for Efficient MoE Model Inference
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025) -
Extending $μ$P: Spectral Conditions for Feature Learning Across Optimizers
by: Gupta, Akshita, et al.
Published: (2026)