An Independent Safety Evaluation of Kimi K2.5
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yong, Zheng-Xin, Mahajan, Parv, Wang, Andy, Caspary, Ida, Yestekov, Yernat, Che, Zora, Levy, Mosh, Najt, Elle, Murphy, Dennis, Kulkarni, Prashant, McKinney, Lev, Nishimura-Gasparian, Kei, Potham, Ram, Lynch, Aengus, Chen, Michael L. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating LLM Agent Adherence to Hierarchical Safety Principles: A Lightweight Benchmark for Probing Foundational Controllability Components
von: Potham, Ram
Veröffentlicht: (2025)
von: Potham, Ram
Veröffentlicht: (2025)
The Persistent Vulnerability of Aligned AI Systems
von: Lynch, Aengus
Veröffentlicht: (2026)
von: Lynch, Aengus
Veröffentlicht: (2026)
GULPS: Two-Qubit Gate Synthesis via Linear Programming for Heterogeneous Instruction Sets
von: McKinney, Evan, et al.
Veröffentlicht: (2025)
von: McKinney, Evan, et al.
Veröffentlicht: (2025)
Model-Based Soft Maximization of Suitable Metrics of Long-Term Human Power
von: Heitzig, Jobst, et al.
Veröffentlicht: (2025)
von: Heitzig, Jobst, et al.
Veröffentlicht: (2025)
Corrigibility as a Singular Target: A Vision for Inherently Reliable Foundation Models
von: Potham, Ram, et al.
Veröffentlicht: (2025)
von: Potham, Ram, et al.
Veröffentlicht: (2025)
Towards Understanding Specification Gaming in Reasoning Models
von: Nishimura-Gasparian, Kei, et al.
Veröffentlicht: (2026)
von: Nishimura-Gasparian, Kei, et al.
Veröffentlicht: (2026)
SLEIGHT-Bench: A Benchmark of Evasion Attacks Against Agent Monitors
von: Najt, Elle, et al.
Veröffentlicht: (2026)
von: Najt, Elle, et al.
Veröffentlicht: (2026)
Postcolonialism and Migration in French Comics
von: McKinney, Mark
Veröffentlicht: (2025)
von: McKinney, Mark
Veröffentlicht: (2025)
Grandmothering While Black: A Twenty‐First‐Century Story of Love, Coercion, and Survival. By Lashawnda L.Pittman. University of California Press, Oakland, California, 2023. 336 pp. $92.04 (hardcover). ISBN: 978‐0‐52‐038995‐3; $29.95 (paperback). ISBN: 978‐0‐52‐038996‐0; $29.95 (ebook). ISBN: 978‐0‐52‐038997‐7
von: Elliana McKinney
Veröffentlicht: (2025)
von: Elliana McKinney
Veröffentlicht: (2025)
Schools Inquiring About Seven-Day School Rerecording of Public and Instructional Television Programs.
von: McKinney, Eleanor
Veröffentlicht: (1975)
von: McKinney, Eleanor
Veröffentlicht: (1975)
Same Task, More Tokens: the Impact of Input Length on the Reasoning Performance of Large Language Models
von: Levy, Mosh, et al.
Veröffentlicht: (2024)
von: Levy, Mosh, et al.
Veröffentlicht: (2024)
Knowledge Navigator: LLM-guided Browsing Framework for Exploratory Search in Scientific Literature
von: Katz, Uri, et al.
Veröffentlicht: (2024)
von: Katz, Uri, et al.
Veröffentlicht: (2024)
Transpose Attack: Stealing Datasets with Bidirectional Training
von: Amit, Guy, et al.
Veröffentlicht: (2023)
von: Amit, Guy, et al.
Veröffentlicht: (2023)
Humans Perceive Wrong Narratives from AI Reasoning Texts
von: Levy, Mosh, et al.
Veröffentlicht: (2025)
von: Levy, Mosh, et al.
Veröffentlicht: (2025)
Water conservation surveys of New South Wales
von: McKinney, Hugh Giffen
Veröffentlicht: (1896)
von: McKinney, Hugh Giffen
Veröffentlicht: (1896)
Evolution of erect marine bryozoan faunas : repeated succes of unilaminate species
von: McKinney, F.K
Veröffentlicht: (1986)
von: McKinney, F.K
Veröffentlicht: (1986)
Created from nafta : the structure, function, and significance of the treatys related institutions / Joseph A. McKinney
von: McKinney, Joseph A
von: McKinney, Joseph A
Media Utilization in the Classroom.
von: Bowie, Melvin McKinney
Veröffentlicht: (1985)
von: Bowie, Melvin McKinney
Veröffentlicht: (1985)
Conceptual and Practical Matters: The Challenges and Benefits of Conducting Educational Research Using Historical Data. Sage Research Methods Cases Part 2
von: Stephen J. McKinney
Veröffentlicht: (2017)
von: Stephen J. McKinney
Veröffentlicht: (2017)
The Contribution of Iona and Peter Opie to Children's Literature.
von: McKinney, Barbara J.
Veröffentlicht: (1996)
von: McKinney, Barbara J.
Veröffentlicht: (1996)
Another Degree? What For?
von: McKinney, Eleanor R.
Veröffentlicht: (1969)
von: McKinney, Eleanor R.
Veröffentlicht: (1969)
MAEBE: Multi-Agent Emergent Behavior Framework
von: Erisken, Sinem, et al.
Veröffentlicht: (2025)
von: Erisken, Sinem, et al.
Veröffentlicht: (2025)
Urbanización y competencia por el suelo en el Área Metropolitana de Mendoza (1990-2020): Dinámicas de la interfase urbano-rural en Guaymallén
von: María Belén Najt Ruiz
Veröffentlicht: (2025)
von: María Belén Najt Ruiz
Veröffentlicht: (2025)
Transferability Ranking of Adversarial Examples
von: Levy, Mosh, et al.
Veröffentlicht: (2022)
von: Levy, Mosh, et al.
Veröffentlicht: (2022)
State over Tokens: Characterizing the Role of Reasoning Tokens
von: Levy, Mosh, et al.
Veröffentlicht: (2025)
von: Levy, Mosh, et al.
Veröffentlicht: (2025)
Early Signs of Steganographic Capabilities in Frontier LLMs
von: Zolkowski, Artur, et al.
Veröffentlicht: (2025)
von: Zolkowski, Artur, et al.
Veröffentlicht: (2025)
Optimizing Neuro-Fuzzy and Colonial Competition Algorithms for Skin Cancer Diagnosis in Dermatoscopic Images
von: Khaleghpour, Hamideh, et al.
Veröffentlicht: (2025)
von: Khaleghpour, Hamideh, et al.
Veröffentlicht: (2025)
Unified AI for Accurate Audio Anomaly Detection
von: Khaleghpour, Hamideh, et al.
Veröffentlicht: (2025)
von: Khaleghpour, Hamideh, et al.
Veröffentlicht: (2025)
Leakage Safe Graph Features for Interpretable Fraud Detection in Temporal Transaction Networks
von: Khaleghpour, Hamideh, et al.
Veröffentlicht: (2026)
von: Khaleghpour, Hamideh, et al.
Veröffentlicht: (2026)
Student Conceptions of Group Work: Visual Research into LIS Student Group Work Using the Draw-and-Write Technique
von: McKinney, Pamela, et al.
Veröffentlicht: (2018)
von: McKinney, Pamela, et al.
Veröffentlicht: (2018)
What is Ethical: AIHED Driving Humans or Human-Driven AIHED? A Conceptual Framework enabling the Ethos of AI-driven Higher education
von: Mahajan, Prashant
Veröffentlicht: (2025)
von: Mahajan, Prashant
Veröffentlicht: (2025)
From diagnostic errors to diagnostic excellence in emergency care: Time to flip the script
von: Prashant Mahajan
Veröffentlicht: (2024)
von: Prashant Mahajan
Veröffentlicht: (2024)
AI Family Integration Index (AFII): Benchmarking a New Global Readiness for AI as Family
von: Mahajan, Prashant
Veröffentlicht: (2025)
von: Mahajan, Prashant
Veröffentlicht: (2025)
Kimi-Audio Technical Report
von: KimiTeam, et al.
Veröffentlicht: (2025)
von: KimiTeam, et al.
Veröffentlicht: (2025)
Kimi-VL Technical Report
von: Kimi Team, et al.
Veröffentlicht: (2025)
von: Kimi Team, et al.
Veröffentlicht: (2025)
Compared to What? Baselines and Metrics for Counterfactual Prompting
von: Yang, Zihao, et al.
Veröffentlicht: (2026)
von: Yang, Zihao, et al.
Veröffentlicht: (2026)
和Kimi谈AI们
von: Lee, Ian
Veröffentlicht: (2026)
von: Lee, Ian
Veröffentlicht: (2026)
Record long-distance movement of a Deer Mouse, Peromyscus maniculatus, in a New England montane boreal forest
von: Wood, Connor M., et al.
Veröffentlicht: (2015)
von: Wood, Connor M., et al.
Veröffentlicht: (2015)
Leadership & professional development: Mitigating assessment bias as a hospitalist
von: Christina M. McKinney, et al.
Veröffentlicht: (2026)
von: Christina M. McKinney, et al.
Veröffentlicht: (2026)
STEM and the City: A Report on STEM Education in the Great American Urban Public School System. Second Edition
von: Clair Berube, et al.
Veröffentlicht: (2025)
von: Clair Berube, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Evaluating LLM Agent Adherence to Hierarchical Safety Principles: A Lightweight Benchmark for Probing Foundational Controllability Components
von: Potham, Ram
Veröffentlicht: (2025) -
The Persistent Vulnerability of Aligned AI Systems
von: Lynch, Aengus
Veröffentlicht: (2026) -
GULPS: Two-Qubit Gate Synthesis via Linear Programming for Heterogeneous Instruction Sets
von: McKinney, Evan, et al.
Veröffentlicht: (2025) -
Model-Based Soft Maximization of Suitable Metrics of Long-Term Human Power
von: Heitzig, Jobst, et al.
Veröffentlicht: (2025) -
Corrigibility as a Singular Target: A Vision for Inherently Reliable Foundation Models
von: Potham, Ram, et al.
Veröffentlicht: (2025)