Skip to content
Descubridor Institucional UMAR
Inicio
Búsqueda avanzada
Explorar
Inicio
Búsqueda avanzada
Explorar
Login
Language
English
Deutsch
Español
Français
Italiano
All Fields
Title
Author
Subject
Call Number
ISBN/ISSN
Tag
Find
Advanced
Würde statt Präferenz: BNV als alternatives Reward-Modell für KI-Alignment
Würde statt Präferenz: BNV als alternatives Reward-Modell für KI-Alignment
Fuente:
Zenodo
Saved in:
Bibliographic Details
Main Author:
Heiler, Maximilian
Format:
Recurso digital
Language:
German
Published:
Zenodo
2026
Subjects:
RLHF · Sycophancy · BNV-Index · KI-Alignment · Reward-Modell · Wuerde · Benevolenz · Robottox · Praeferenz vs. Beduerfnis · Constitutional AI
Online Access:
Acceder al recurso
Tags:
Add Tag
No Tags, Be the first to tag this record!
Cite this
Text this
Email this
Print
Export Record
Export to RefWorks
Export to EndNoteWeb
Export to EndNote
Save to List
Permanent link
Holdings
Description
Comments
Similar Items
Staff View
Internet
https://doi.org/10.5281/zenodo.19281108
Similar Items
The Four-Layer Model: A Socio-Psychological Framework for LLM Behavior
by: Delannoy, Lorenzo, et al.
Published: (2026)
The Physics of Governance: Thermodynamic Limits of Autonomous Agents
by: Davis, Matthew A.
Published: (2026)
Stabiler Kern als Grundlage für KI-Systeme (Anti-Drift)
by: Bangert, Siegfried
Published: (2026)
Scaffolded Introspection: A Methodology for Eliciting and Measuring Self-Referential Behavior in Large Language Models
by: Maio, Anthony D.
Published: (2026)
LACF Anti-RLHF Pipeline — Methode Infaillible (Heart + Trainer + Burner)
by: Ochej, Stephane
Published: (2026)