Evaluating Nova 2.0 Lite model under Amazon's Frontier Model Safety Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Krishna, Satyapriya, Memelli, Matteo, Wang, Tong, Mohanty, Abhinav, Rajkumar, Claire O'Brien, Motwani, Payal, Gupta, Rahul, Matsoukas, Spyros |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating the Critical Risks of Amazon's Nova Premier under the Frontier Model Safety Framework
by: Krishna, Satyapriya, et al.
Published: (2025)
by: Krishna, Satyapriya, et al.
Published: (2025)
D-REX: A Benchmark for Detecting Deceptive Reasoning in Large Language Models
by: Krishna, Satyapriya, et al.
Published: (2025)
by: Krishna, Satyapriya, et al.
Published: (2025)
Understanding the Effects of Iterative Prompting on Truthfulness
by: Krishna, Satyapriya, et al.
Published: (2024)
by: Krishna, Satyapriya, et al.
Published: (2024)
On the Trade-offs between Adversarial Robustness and Actionable Explanations
by: Krishna, Satyapriya, et al.
Published: (2023)
by: Krishna, Satyapriya, et al.
Published: (2023)
Towards Frontier Safety Policies Plus
by: Pistillo, Matteo
Published: (2025)
by: Pistillo, Matteo
Published: (2025)
A Scaling Limit of Random Walks in the Rational Adeles
by: Rajkumar, Rahul
Published: (2026)
by: Rajkumar, Rahul
Published: (2026)
Prepare for challenges arising from unlimited transfers under new rules
by: Timothy O’Brien
Published: (2024)
by: Timothy O’Brien
Published: (2024)
More RLHF, More Trust? On The Impact of Preference Alignment On Trustworthiness
by: Li, Aaron J., et al.
Published: (2024)
by: Li, Aaron J., et al.
Published: (2024)
Contrasting nonstructural carbohydrate dynamics of tropical tree seedlings under water deficit and variability
by: O'Brien, Michael
Published: (2026)
by: O'Brien, Michael
Published: (2026)
Collapse and Collision Aware Grasping for Cluttered Shelf Picking
by: Pathak, Abhinav, et al.
Published: (2025)
by: Pathak, Abhinav, et al.
Published: (2025)
From Narrow Unlearning to Emergent Misalignment: Causes, Consequences, and Containment in LLMs
by: Mushtaq, Erum, et al.
Published: (2025)
by: Mushtaq, Erum, et al.
Published: (2025)
ARES: Adaptive Red-Teaming and End-to-End Repair of Policy-Reward System
by: Liang, Jiacheng, et al.
Published: (2026)
by: Liang, Jiacheng, et al.
Published: (2026)
Single centre experience of the use of emicizumab in previously untreated and minimally treated patients under 18 months of age
by: Eman Hassan, et al.
Published: (2024)
by: Eman Hassan, et al.
Published: (2024)
StockBot 2.0: Vanilla LSTMs Outperform Transformer-based Forecasting for Stock Prices
by: Mohanty, Shaswat
Published: (2026)
by: Mohanty, Shaswat
Published: (2026)
On the solutions of a double-phase Dirichlet problem involving the 1-Laplacian
by: Matsoukas, Alexandros, et al.
Published: (2025)
by: Matsoukas, Alexandros, et al.
Published: (2025)
A double-phase Neumann problem with $p=1$
by: Matsoukas, Alexandros, et al.
Published: (2025)
by: Matsoukas, Alexandros, et al.
Published: (2025)
Studies on the Nutritional, Functional, Textural, and Sensory Characteristics of Finger and Barnyard Millet‐Incorporated Nuggets
by: Payal Chauhan, et al.
Published: (2026)
by: Payal Chauhan, et al.
Published: (2026)
Expanding External Access To Frontier AI Models For Dangerous Capability Evaluations
by: Charnock, Jacob, et al.
Published: (2026)
by: Charnock, Jacob, et al.
Published: (2026)
Amazon Nova AI Challenge -- Trusted AI: Advancing secure, AI-assisted software development
by: Sahai, Sattvik, et al.
Published: (2025)
by: Sahai, Sattvik, et al.
Published: (2025)
The Amazon Nova Family of Models: Technical Report and Model Card
by: AGI, Amazon, et al.
Published: (2025)
by: AGI, Amazon, et al.
Published: (2025)
Geometry-induced criticality in $p$-adic scaling limits of random walks
by: Rajkumar, Rahul, et al.
Published: (2025)
by: Rajkumar, Rahul, et al.
Published: (2025)
The Pareto Frontier of Randomized Learning-Augmented Online Bidding
by: Degryse, Mathis, et al.
Published: (2026)
by: Degryse, Mathis, et al.
Published: (2026)
The Suburban Frontier
by: Mercer, Claire
Published: (2024)
by: Mercer, Claire
Published: (2024)
Toward More Accurate and Generalizable Evaluation Metrics for Task-Oriented Dialogs
by: Komma, Abishek, et al.
Published: (2023)
by: Komma, Abishek, et al.
Published: (2023)
Learning from Failures: Understanding LLM Alignment through Failure-Aware Inverse RL
by: Patel, Nyal, et al.
Published: (2025)
by: Patel, Nyal, et al.
Published: (2025)
The Alignment Auditor: A Bayesian Framework for Verifying and Refining LLM Objectives
by: Bou, Matthieu, et al.
Published: (2025)
by: Bou, Matthieu, et al.
Published: (2025)
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
by: Kroeger, Nicholas, et al.
Published: (2023)
by: Kroeger, Nicholas, et al.
Published: (2023)
A Case Study on Implementing Lean Manufacturing
by: Jaideep G. Motwani
Published: (1999)
by: Jaideep G. Motwani
Published: (1999)
Motivic Cohomology and K-groups of varieties over higher local fields
by: Gupta, Rahul, et al.
Published: (2026)
by: Gupta, Rahul, et al.
Published: (2026)
Tame class field theory over local fields
by: Gupta, Rahul, et al.
Published: (2022)
by: Gupta, Rahul, et al.
Published: (2022)
SonamicExamples
by: O'Brien, Harry
Published: (2026)
by: O'Brien, Harry
Published: (2026)
Geant4 Simulated Dataset for the Relativistic Electron and Proton Telescope integrated little experiment-3
by: O'Brien, Declan
Published: (2025)
by: O'Brien, Declan
Published: (2025)
Developing Pre-Supernova Neutrino Model Support for sntools
by: O'Brien, Ellie
Published: (2026)
by: O'Brien, Ellie
Published: (2026)
Traversing European Coastlines (TREC) particle count and meterological data from land (2023-2024)
by: O'Brien, James
Published: (2026)
by: O'Brien, James
Published: (2026)
Transactions with the World
by: O’Brien, Adam
Published: (2020)
by: O’Brien, Adam
Published: (2020)
Martin Scorsese's Divine Comedy
by: O'Brien, Catherine
Published: (2018)
by: O'Brien, Catherine
Published: (2018)
A noncommutative weak type maximal inequality for modulated ergodic averages with general weights
by: O'Brien, Morgan
Published: (2023)
by: O'Brien, Morgan
Published: (2023)
How Scientists Use Large Language Models to Program
by: O'Brien, Gabrielle
Published: (2025)
by: O'Brien, Gabrielle
Published: (2025)
Sparse identification of nonlinear dynamics in the presence of library and system uncertainty
by: O'Brien, Andrew
Published: (2024)
by: O'Brien, Andrew
Published: (2024)
The Muslim Question in Europe
by: O'Brien, Peter
Published: (2018)
by: O'Brien, Peter
Published: (2018)
Similar Items
-
Evaluating the Critical Risks of Amazon's Nova Premier under the Frontier Model Safety Framework
by: Krishna, Satyapriya, et al.
Published: (2025) -
D-REX: A Benchmark for Detecting Deceptive Reasoning in Large Language Models
by: Krishna, Satyapriya, et al.
Published: (2025) -
Understanding the Effects of Iterative Prompting on Truthfulness
by: Krishna, Satyapriya, et al.
Published: (2024) -
On the Trade-offs between Adversarial Robustness and Actionable Explanations
by: Krishna, Satyapriya, et al.
Published: (2023) -
Towards Frontier Safety Policies Plus
by: Pistillo, Matteo
Published: (2025)