Position: AI Scaling: From Up to Down and Out
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yunke, Li, Yanxi, Xu, Chang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HiCI: Hierarchical Construction-Integration for Long-Context Attention
by: Zeng, Xiangyu, et al.
Published: (2026)
by: Zeng, Xiangyu, et al.
Published: (2026)
Scaling Down to Scale Up: A Cost-Benefit Analysis of Replacing OpenAI's LLM with Open Source SLMs in Production
by: Irugalbandara, Chandra, et al.
Published: (2023)
by: Irugalbandara, Chandra, et al.
Published: (2023)
R$^3$L: Reflect-then-Retry Reinforcement Learning with Language-Guided Exploration, Pivotal Credit, and Positive Amplification
by: Shi, Weijie, et al.
Published: (2026)
by: Shi, Weijie, et al.
Published: (2026)
EPR-GAIL: An EPR-Enhanced Hierarchical Imitation Learning Framework to Simulate Complex User Consumption Behaviors
by: Feng, Tao, et al.
Published: (2025)
by: Feng, Tao, et al.
Published: (2025)
Causality Enhanced Origin-Destination Flow Prediction in Data-Scarce Cities
by: Feng, Tao, et al.
Published: (2025)
by: Feng, Tao, et al.
Published: (2025)
Provable Scaling Laws for the Test-Time Compute of Large Language Models
by: Chen, Yanxi, et al.
Published: (2024)
by: Chen, Yanxi, et al.
Published: (2024)
Negative as Positive: Enhancing Out-of-distribution Generalization for Graph Contrastive Learning
by: Wang, Zixu, et al.
Published: (2024)
by: Wang, Zixu, et al.
Published: (2024)
Scaling Up Data Parallelism in Decentralized Deep Learning
by: Xie, Bing, et al.
Published: (2025)
by: Xie, Bing, et al.
Published: (2025)
Scaling Up Bayesian DAG Sampling
by: Nikzad, Daniele, et al.
Published: (2025)
by: Nikzad, Daniele, et al.
Published: (2025)
Taming False Positives in Out-of-Distribution Detection with Human Feedback
by: Vishwakarma, Harit, et al.
Published: (2024)
by: Vishwakarma, Harit, et al.
Published: (2024)
MIDUS: Memory-Infused Depth Up-Scaling
by: Kim, Taero, et al.
Published: (2025)
by: Kim, Taero, et al.
Published: (2025)
Scaling Up Probabilistic Circuits by Latent Variable Distillation
by: Liu, Anji, et al.
Published: (2022)
by: Liu, Anji, et al.
Published: (2022)
Position Paper: From Edge AI to Adaptive Edge AI
by: Pittorino, Fabrizio, et al.
Published: (2026)
by: Pittorino, Fabrizio, et al.
Published: (2026)
Score-based Conditional Out-of-Distribution Augmentation for Graph Covariate Shift
by: Wang, Bohan, et al.
Published: (2024)
by: Wang, Bohan, et al.
Published: (2024)
COUNTDOWN: Contextually Sparse Activation Filtering Out Unnecessary Weights in Down Projection
by: Cheon, Jaewon, et al.
Published: (2025)
by: Cheon, Jaewon, et al.
Published: (2025)
Boosting Single Positive Multi-label Classification with Generalized Robust Loss
by: Chen, Yanxi, et al.
Published: (2024)
by: Chen, Yanxi, et al.
Published: (2024)
Position: Weight Space Should Be a First-Class Generative AI Modality
by: Wang, Zhangyang, et al.
Published: (2026)
by: Wang, Zhangyang, et al.
Published: (2026)
On the Entropy Dynamics in Reinforcement Fine-Tuning of Large Language Models
by: Wang, Shumin, et al.
Published: (2026)
by: Wang, Shumin, et al.
Published: (2026)
Beyond Model Collapse: Scaling Up with Synthesized Data Requires Verification
by: Feng, Yunzhen, et al.
Published: (2024)
by: Feng, Yunzhen, et al.
Published: (2024)
EE-LLM: Large-Scale Training and Inference of Early-Exit Large Language Models with 3D Parallelism
by: Chen, Yanxi, et al.
Published: (2023)
by: Chen, Yanxi, et al.
Published: (2023)
Designing Algorithms Empowered by Language Models: An Analytical Framework, Case Studies, and Insights
by: Chen, Yanxi, et al.
Published: (2024)
by: Chen, Yanxi, et al.
Published: (2024)
Enhancing Distribution and Label Consistency for Graph Out-of-Distribution Generalization
by: Wang, Song, et al.
Published: (2025)
by: Wang, Song, et al.
Published: (2025)
AutoChemSchematic AI: Agentic Physics-Aware Automation for Chemical Manufacturing Scale-Up
by: Srinivas, Sakhinana Sagar, et al.
Published: (2025)
by: Srinivas, Sakhinana Sagar, et al.
Published: (2025)
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2024)
by: Lee, Hojoon, et al.
Published: (2024)
Breaking Down Financial News Impact: A Novel AI Approach with Geometric Hypergraphs
by: Harit, Anoushka, et al.
Published: (2024)
by: Harit, Anoushka, et al.
Published: (2024)
Investigating Out-of-Distribution Generalization of GNNs: An Architecture Perspective
by: Guo, Kai, et al.
Published: (2024)
by: Guo, Kai, et al.
Published: (2024)
Abstracting Geo-specific Terrains to Scale Up Reinforcement Learning
by: Ustun, Volkan, et al.
Published: (2025)
by: Ustun, Volkan, et al.
Published: (2025)
Position: agentic AI orchestration should be Bayes-consistent
by: Papamarkou, Theodore, et al.
Published: (2026)
by: Papamarkou, Theodore, et al.
Published: (2026)
Consistency-Guided Temperature Scaling Using Style and Content Information for Out-of-Domain Calibration
by: Choi, Wonjeong, et al.
Published: (2024)
by: Choi, Wonjeong, et al.
Published: (2024)
On-Policy RL Meets Off-Policy Experts: Harmonizing Supervised Fine-Tuning and Reinforcement Learning via Dynamic Weighting
by: Zhang, Wenhao, et al.
Published: (2025)
by: Zhang, Wenhao, et al.
Published: (2025)
Moving Out: Physically-grounded Human-AI Collaboration
by: Kang, Xuhui, et al.
Published: (2025)
by: Kang, Xuhui, et al.
Published: (2025)
Sparsity and Out-of-Distribution Generalization
by: Aaronson, Scott, et al.
Published: (2026)
by: Aaronson, Scott, et al.
Published: (2026)
LESA: Learnable LLM Layer Scaling-Up
by: Yang, Yifei, et al.
Published: (2025)
by: Yang, Yifei, et al.
Published: (2025)
Position: We Need An Algorithmic Understanding of Generative AI
by: Eberle, Oliver, et al.
Published: (2025)
by: Eberle, Oliver, et al.
Published: (2025)
Generative Risk Minimization for Out-of-Distribution Generalization on Graphs
by: Wang, Song, et al.
Published: (2025)
by: Wang, Song, et al.
Published: (2025)
WATS: Calibrating Graph Neural Networks with Wavelet-Aware Temperature Scaling
by: Li, Xiaoyang, et al.
Published: (2025)
by: Li, Xiaoyang, et al.
Published: (2025)
Structure-Adaptive Conformal Inference for Large-Scale Out-of-Distribution Testing
by: Sun, Rongyi, et al.
Published: (2026)
by: Sun, Rongyi, et al.
Published: (2026)
From Abstract to Actionable: Pairwise Shapley Values for Explainable AI
by: Xu, Jiaxin, et al.
Published: (2025)
by: Xu, Jiaxin, et al.
Published: (2025)
Positional-aware Spatio-Temporal Network for Large-Scale Traffic Prediction
by: Chen, Runfei
Published: (2026)
by: Chen, Runfei
Published: (2026)
EE-Tuning: An Economical yet Scalable Solution for Tuning Early-Exit Large Language Models
by: Pan, Xuchen, et al.
Published: (2024)
by: Pan, Xuchen, et al.
Published: (2024)
Similar Items
-
HiCI: Hierarchical Construction-Integration for Long-Context Attention
by: Zeng, Xiangyu, et al.
Published: (2026) -
Scaling Down to Scale Up: A Cost-Benefit Analysis of Replacing OpenAI's LLM with Open Source SLMs in Production
by: Irugalbandara, Chandra, et al.
Published: (2023) -
R$^3$L: Reflect-then-Retry Reinforcement Learning with Language-Guided Exploration, Pivotal Credit, and Positive Amplification
by: Shi, Weijie, et al.
Published: (2026) -
EPR-GAIL: An EPR-Enhanced Hierarchical Imitation Learning Framework to Simulate Complex User Consumption Behaviors
by: Feng, Tao, et al.
Published: (2025) -
Causality Enhanced Origin-Destination Flow Prediction in Data-Scarce Cities
by: Feng, Tao, et al.
Published: (2025)