Get RICH or Die Scaling: Profitably Trading Inference Compute for Robustness
Fuente:
arXiv
Saved in:
| Main Authors: | McDonald, Tavish, Lei, Bo, Fort, Stanislav, Kailkhura, Bhavya, Bartoldson, Brian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adversarial Robustness Limits via Scaling-Law and Human-Alignment Studies
by: Bartoldson, Brian R., et al.
Published: (2024)
by: Bartoldson, Brian R., et al.
Published: (2024)
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
by: Geiping, Jonas, et al.
Published: (2025)
by: Geiping, Jonas, et al.
Published: (2025)
AegisLLM: Scaling Agentic Systems for Self-Reflective Defense in LLM Security
by: Cai, Zikui, et al.
Published: (2025)
by: Cai, Zikui, et al.
Published: (2025)
Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts
by: Zheng, Haizhong, et al.
Published: (2025)
by: Zheng, Haizhong, et al.
Published: (2025)
Speculative Diffusion Decoding: Accelerating Language Generation through Diffusion
by: Christopher, Jacob K, et al.
Published: (2024)
by: Christopher, Jacob K, et al.
Published: (2024)
Double Visual Defense: Adversarial Pre-training and Instruction Tuning for Improving Vision-Language Model Robustness
by: Wang, Zeyu, et al.
Published: (2025)
by: Wang, Zeyu, et al.
Published: (2025)
Certifiably-Robust Federated Adversarial Learning via Randomized Smoothing
by: Chen, Cheng, et al.
Published: (2021)
by: Chen, Cheng, et al.
Published: (2021)
Improving Robustness In Sparse Autoencoders via Masked Regularization
by: Narayanaswamy, Vivek, et al.
Published: (2026)
by: Narayanaswamy, Vivek, et al.
Published: (2026)
Trajectory Balance with Asynchrony: Decoupling Exploration and Learning for Fast, Scalable LLM Post-Training
by: Bartoldson, Brian, et al.
Published: (2025)
by: Bartoldson, Brian, et al.
Published: (2025)
A Note on Implementation Errors in Recent Adaptive Attacks Against Multi-Resolution Self-Ensembles
by: Fort, Stanislav
Published: (2025)
by: Fort, Stanislav
Published: (2025)
LongCoT: Benchmarking Long-Horizon Chain-of-Thought Reasoning
by: Motwani, Sumeet Ramesh, et al.
Published: (2026)
by: Motwani, Sumeet Ramesh, et al.
Published: (2026)
Transformers Can Do Arithmetic with the Right Embeddings
by: McLeish, Sean, et al.
Published: (2024)
by: McLeish, Sean, et al.
Published: (2024)
Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence
by: McLeish, Sean, et al.
Published: (2025)
by: McLeish, Sean, et al.
Published: (2025)
FedCluster: Boosting the Convergence of Federated Learning via Cluster-Cycling
by: Chen, Cheng, et al.
Published: (2020)
by: Chen, Cheng, et al.
Published: (2020)
Training Dynamics of Transformers to Recognize Word Co-occurrence via Gradient Flow Analysis
by: Yang, Hongru, et al.
Published: (2024)
by: Yang, Hongru, et al.
Published: (2024)
Mixture of Robust Experts (MoRE):A Robust Denoising Method towards multiple perturbations
by: Cheng, Hao, et al.
Published: (2021)
by: Cheng, Hao, et al.
Published: (2021)
Recursive Self-Aggregation Unlocks Deep Thinking in Large Language Models
by: Venkatraman, Siddarth, et al.
Published: (2025)
by: Venkatraman, Siddarth, et al.
Published: (2025)
Ensemble everything everywhere: Multi-scale aggregation for adversarial robustness
by: Fort, Stanislav, et al.
Published: (2024)
by: Fort, Stanislav, et al.
Published: (2024)
End-to-End Mesh Optimization of a Hybrid Deep Learning Black-Box PDE Solver
by: Ma, Shaocong, et al.
Published: (2024)
by: Ma, Shaocong, et al.
Published: (2024)
Trading Inference-Time Compute for Adversarial Robustness
by: Zaremba, Wojciech, et al.
Published: (2025)
by: Zaremba, Wojciech, et al.
Published: (2025)
Forecasting Fails: Unveiling Evasion Attacks in Weather Prediction Models
by: Arif, Huzaifa, et al.
Published: (2025)
by: Arif, Huzaifa, et al.
Published: (2025)
A Comedy of Estimators: On KL Regularization in RL Training of LLMs
by: Shah, Vedant, et al.
Published: (2025)
by: Shah, Vedant, et al.
Published: (2025)
Constrained Discrete Diffusion
by: Cardei, Michael, et al.
Published: (2025)
by: Cardei, Michael, et al.
Published: (2025)
ProtAlign: Contrastive learning paradigm for Sequence and structure alignment
by: Ranganath, Aditya, et al.
Published: (2026)
by: Ranganath, Aditya, et al.
Published: (2026)
LLM Unlearning Reveals a Stronger-Than-Expected Coreset Effect in Current Benchmarks
by: Pal, Soumyadeep, et al.
Published: (2025)
by: Pal, Soumyadeep, et al.
Published: (2025)
DeepZero: Scaling up Zeroth-Order Optimization for Deep Model Training
by: Chen, Aochuan, et al.
Published: (2023)
by: Chen, Aochuan, et al.
Published: (2023)
Solving adversarial examples requires solving exponential misalignment
by: Salvatore, Alessandro, et al.
Published: (2026)
by: Salvatore, Alessandro, et al.
Published: (2026)
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
by: Duan, Jinhao, et al.
Published: (2025)
by: Duan, Jinhao, et al.
Published: (2025)
Low-rank finetuning for LLMs: A fairness perspective
by: Das, Saswat, et al.
Published: (2024)
by: Das, Saswat, et al.
Published: (2024)
Near Optimal Decision Trees in a SPLIT Second
by: Babbar, Varun, et al.
Published: (2025)
by: Babbar, Varun, et al.
Published: (2025)
Interpretable Generalized Additive Models for Datasets with Missing Values
by: McTavish, Hayden, et al.
Published: (2024)
by: McTavish, Hayden, et al.
Published: (2024)
floq: Training Critics via Flow-Matching for Scaling Compute in Value-Based RL
by: Agrawalla, Bhavya, et al.
Published: (2025)
by: Agrawalla, Bhavya, et al.
Published: (2025)
A Bayesian Approach to Robust Inverse Reinforcement Learning
by: Wei, Ran, et al.
Published: (2023)
by: Wei, Ran, et al.
Published: (2023)
Deep Latent Force Models: ODE-based Process Convolutions for Bayesian Deep Learning
by: Baldwin-McDonald, Thomas, et al.
Published: (2023)
by: Baldwin-McDonald, Thomas, et al.
Published: (2023)
StockBench: Can LLM Agents Trade Stocks Profitably In Real-world Markets?
by: Chen, Yanxu, et al.
Published: (2025)
by: Chen, Yanxu, et al.
Published: (2025)
Log-Concave Coupling for Sampling Neural Net Posteriors
by: McDonald, Curtis, et al.
Published: (2024)
by: McDonald, Curtis, et al.
Published: (2024)
Robot Arm Control via Cognitive Map Learners
by: McDonald, Nathan, et al.
Published: (2026)
by: McDonald, Nathan, et al.
Published: (2026)
Active Learning Enables Extrapolation in Molecular Generative Models
by: Antoniuk, Evan R., et al.
Published: (2025)
by: Antoniuk, Evan R., et al.
Published: (2025)
Leveraging Hierarchical Feature Sharing for Efficient Dataset Condensation
by: Zheng, Haizhong, et al.
Published: (2023)
by: Zheng, Haizhong, et al.
Published: (2023)
GRNFormer: A Biologically-Guided Framework for Integrating Gene Regulatory Networks into RNA Foundation Models
by: Qiu, Mufan, et al.
Published: (2025)
by: Qiu, Mufan, et al.
Published: (2025)
Similar Items
-
Adversarial Robustness Limits via Scaling-Law and Human-Alignment Studies
by: Bartoldson, Brian R., et al.
Published: (2024) -
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
by: Geiping, Jonas, et al.
Published: (2025) -
AegisLLM: Scaling Agentic Systems for Self-Reflective Defense in LLM Security
by: Cai, Zikui, et al.
Published: (2025) -
Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts
by: Zheng, Haizhong, et al.
Published: (2025) -
Speculative Diffusion Decoding: Accelerating Language Generation through Diffusion
by: Christopher, Jacob K, et al.
Published: (2024)