The Achilles Heel of AI: Fundamentals of Risk-Aware Training Data for High-Consequence Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Cook, Dave, Klawa, Tim |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning
di: Singh, Joykirat, et al.
Pubblicazione: (2024)
di: Singh, Joykirat, et al.
Pubblicazione: (2024)
Achilles' Heels: Vulnerable Record Identification in Synthetic Data Publishing
di: Meeus, Matthieu, et al.
Pubblicazione: (2023)
di: Meeus, Matthieu, et al.
Pubblicazione: (2023)
Generative AI Training and Copyright Law
di: Stober, Sebastian, et al.
Pubblicazione: (2025)
di: Stober, Sebastian, et al.
Pubblicazione: (2025)
Cyborg Data: Merging Human with AI Generated Training Data
di: North, Kai, et al.
Pubblicazione: (2025)
di: North, Kai, et al.
Pubblicazione: (2025)
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
di: Moya, Christian, et al.
Pubblicazione: (2026)
di: Moya, Christian, et al.
Pubblicazione: (2026)
Enhancing Smart Farming Through Federated Learning: A Secure, Scalable, and Efficient Approach for AI-Driven Agriculture
di: Janga, Ritesh, et al.
Pubblicazione: (2025)
di: Janga, Ritesh, et al.
Pubblicazione: (2025)
Data Warmup: Complexity-Aware Curricula for Efficient Diffusion Training
di: Lin, Jinhong, et al.
Pubblicazione: (2026)
di: Lin, Jinhong, et al.
Pubblicazione: (2026)
Scaling Laws for Pre-training Agents and World Models
di: Pearce, Tim, et al.
Pubblicazione: (2024)
di: Pearce, Tim, et al.
Pubblicazione: (2024)
Uncovering Gradient Inversion Risks in Practical Language Model Training
di: Feng, Xinguo, et al.
Pubblicazione: (2025)
di: Feng, Xinguo, et al.
Pubblicazione: (2025)
AI-CARE: Carbon-Aware Reporting Evaluation Metric for AI Models
di: Santosh, KC, et al.
Pubblicazione: (2026)
di: Santosh, KC, et al.
Pubblicazione: (2026)
Fundamental Limits of Deep Learning-Based Binary Classifiers Trained with Hinge Loss
di: Getu, Tilahun M., et al.
Pubblicazione: (2023)
di: Getu, Tilahun M., et al.
Pubblicazione: (2023)
Multimodal Visual-Tactile Representation Learning through Self-Supervised Contrastive Pre-Training
di: Dave, Vedant, et al.
Pubblicazione: (2024)
di: Dave, Vedant, et al.
Pubblicazione: (2024)
QuAIL: Quality-Aware Inertial Learning for Robust Training under Data Corruption
di: Sabella, Mattia, et al.
Pubblicazione: (2026)
di: Sabella, Mattia, et al.
Pubblicazione: (2026)
Utility-Aware Data Pricing: Token-Level Quality and Empirical Training Gain for LLMs
di: Xu, Minghui, et al.
Pubblicazione: (2026)
di: Xu, Minghui, et al.
Pubblicazione: (2026)
A Note on Shumailov et al. (2024): `AI Models Collapse When Trained on Recursively Generated Data'
di: Borji, Ali
Pubblicazione: (2024)
di: Borji, Ali
Pubblicazione: (2024)
Conflict-Aware Adversarial Training
di: Xue, Zhiyu, et al.
Pubblicazione: (2024)
di: Xue, Zhiyu, et al.
Pubblicazione: (2024)
Fast-DataShapley: Neural Modeling for Training Data Valuation
di: Sun, Haifeng, et al.
Pubblicazione: (2025)
di: Sun, Haifeng, et al.
Pubblicazione: (2025)
Diffusion-Modeled Reinforcement Learning for Carbon and Risk-Aware Microgrid Optimization
di: Zhao, Yunyi, et al.
Pubblicazione: (2025)
di: Zhao, Yunyi, et al.
Pubblicazione: (2025)
Model-Based Epistemic Variance of Values for Risk-Aware Policy Optimization
di: Luis, Carlos E., et al.
Pubblicazione: (2023)
di: Luis, Carlos E., et al.
Pubblicazione: (2023)
Noise-Aware Training of Layout-Aware Language Models
di: Sarkhel, Ritesh, et al.
Pubblicazione: (2024)
di: Sarkhel, Ritesh, et al.
Pubblicazione: (2024)
OAT-Rephrase: Optimization-Aware Training Data Rephrasing for Zeroth-Order LLM Fine-Tuning
di: Long, Jikai, et al.
Pubblicazione: (2025)
di: Long, Jikai, et al.
Pubblicazione: (2025)
Achilles' Heel of Mamba: Essential difficulties of the Mamba architecture demonstrated by synthetic data
di: Chen, Tianyi, et al.
Pubblicazione: (2025)
di: Chen, Tianyi, et al.
Pubblicazione: (2025)
Compute-Optimal Quantization-Aware Training
di: Dremov, Aleksandr, et al.
Pubblicazione: (2025)
di: Dremov, Aleksandr, et al.
Pubblicazione: (2025)
Load-Aware Training Scheduling for Model Circulation-based Decentralized Federated Learning
di: Kainuma, Haruki, et al.
Pubblicazione: (2025)
di: Kainuma, Haruki, et al.
Pubblicazione: (2025)
Provable Training Data Identification for Large Language Models
di: Liu, Zhenlong, et al.
Pubblicazione: (2025)
di: Liu, Zhenlong, et al.
Pubblicazione: (2025)
Native Reasoning Models: Training Language Models to Reason on Unverifiable Data
di: Wang, Yuanfu, et al.
Pubblicazione: (2026)
di: Wang, Yuanfu, et al.
Pubblicazione: (2026)
Improved Sample Complexity For Diffusion Model Training Without Empirical Risk Minimizer Access
di: Gaur, Mudit, et al.
Pubblicazione: (2025)
di: Gaur, Mudit, et al.
Pubblicazione: (2025)
Membership Privacy Risks of Sharpness Aware Minimization
di: Kim, Young In, et al.
Pubblicazione: (2023)
di: Kim, Young In, et al.
Pubblicazione: (2023)
Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility
di: Wang, Mengxuan, et al.
Pubblicazione: (2026)
di: Wang, Mengxuan, et al.
Pubblicazione: (2026)
Symmetry-Aware Transformer Training for Automated Planning
di: Fritzsche, Markus, et al.
Pubblicazione: (2025)
di: Fritzsche, Markus, et al.
Pubblicazione: (2025)
Topology-Aware Revival for Efficient Sparse Training
di: Jin, Meiling, et al.
Pubblicazione: (2026)
di: Jin, Meiling, et al.
Pubblicazione: (2026)
Reasoning-Aware Training for Time Series Forecasting
di: Ahamed, Md Atik, et al.
Pubblicazione: (2026)
di: Ahamed, Md Atik, et al.
Pubblicazione: (2026)
DAFA: Distance-Aware Fair Adversarial Training
di: Lee, Hyungyu, et al.
Pubblicazione: (2024)
di: Lee, Hyungyu, et al.
Pubblicazione: (2024)
Task-Aware Parameter-Efficient Fine-Tuning of Large Pre-Trained Models at the Edge
di: Hu, Senkang, et al.
Pubblicazione: (2025)
di: Hu, Senkang, et al.
Pubblicazione: (2025)
The Achilles' Heel of LLMs: How Altering a Handful of Neurons Can Cripple Language Abilities
di: Qin, Zixuan, et al.
Pubblicazione: (2025)
di: Qin, Zixuan, et al.
Pubblicazione: (2025)
Certified Policy Optimisation for Nested Causal Bandits via PAC-Bayes Risk
di: Woydt, Tim, et al.
Pubblicazione: (2026)
di: Woydt, Tim, et al.
Pubblicazione: (2026)
The Strain of Success: A Predictive Model for Injury Risk Mitigation and Team Success in Soccer
di: Everett, Gregory, et al.
Pubblicazione: (2024)
di: Everett, Gregory, et al.
Pubblicazione: (2024)
Hephaestus: Improving Fundamental Agent Capabilities of Large Language Models through Continual Pre-Training
di: Zhuang, Yuchen, et al.
Pubblicazione: (2025)
di: Zhuang, Yuchen, et al.
Pubblicazione: (2025)
Ergodic Risk Measures: Towards a Risk-Aware Foundation for Continual Reinforcement Learning
di: Rojas, Juan Sebastian, et al.
Pubblicazione: (2025)
di: Rojas, Juan Sebastian, et al.
Pubblicazione: (2025)
A Quantized VAE-MLP Botnet Detection Model: A Systematic Evaluation of Quantization-Aware Training and Post-Training Quantization Strategies
di: Wasswa, Hassan, et al.
Pubblicazione: (2025)
di: Wasswa, Hassan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning
di: Singh, Joykirat, et al.
Pubblicazione: (2024) -
Achilles' Heels: Vulnerable Record Identification in Synthetic Data Publishing
di: Meeus, Matthieu, et al.
Pubblicazione: (2023) -
Generative AI Training and Copyright Law
di: Stober, Sebastian, et al.
Pubblicazione: (2025) -
Cyborg Data: Merging Human with AI Generated Training Data
di: North, Kai, et al.
Pubblicazione: (2025) -
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
di: Moya, Christian, et al.
Pubblicazione: (2026)