Evaluating Language Model Character Traits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ward, Francis Rhys, Yang, Zejia, Jackson, Alex, Brown, Randy, Smith, Chandler, Colverd, Grace, Thomson, Louis, Douglas, Raymond, Bartak, Patrik, Rowan, Andrew |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AI Sandbagging: Language Models can Strategically Underperform on Evaluations
von: van der Weij, Teun, et al.
Veröffentlicht: (2024)
von: van der Weij, Teun, et al.
Veröffentlicht: (2024)
Persona Vectors: Monitoring and Controlling Character Traits in Language Models
von: Chen, Runjin, et al.
Veröffentlicht: (2025)
von: Chen, Runjin, et al.
Veröffentlicht: (2025)
Towards a Theory of AI Personhood
von: Ward, Francis Rhys
Veröffentlicht: (2025)
von: Ward, Francis Rhys
Veröffentlicht: (2025)
Spherical Steering: Geometry-Aware Activation Rotation for Language Models
von: You, Zejia, et al.
Veröffentlicht: (2026)
von: You, Zejia, et al.
Veröffentlicht: (2026)
Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works
von: Yuan, Xinfeng, et al.
Veröffentlicht: (2024)
von: Yuan, Xinfeng, et al.
Veröffentlicht: (2024)
Large Language Models Lack Understanding of Character Composition of Words
von: Shin, Andrew, et al.
Veröffentlicht: (2024)
von: Shin, Andrew, et al.
Veröffentlicht: (2024)
Evaluating Computational Representations of Character: An Austen Character Similarity Benchmark
von: Yang, Funing, et al.
Veröffentlicht: (2024)
von: Yang, Funing, et al.
Veröffentlicht: (2024)
NEBULA: A National Scale Dataset for Neighbourhood-Level Urban Building Energy Modelling for England and Wales
von: Colverd, Grace, et al.
Veröffentlicht: (2025)
von: Colverd, Grace, et al.
Veröffentlicht: (2025)
CharacterBench: Benchmarking Character Customization of Large Language Models
von: Zhou, Jinfeng, et al.
Veröffentlicht: (2024)
von: Zhou, Jinfeng, et al.
Veröffentlicht: (2024)
Evaluating Prompt-Based and Fine-Tuned Approaches to Czech Anaphora Resolution
von: Stano, Patrik, et al.
Veröffentlicht: (2025)
von: Stano, Patrik, et al.
Veröffentlicht: (2025)
AA-Omniscience: Evaluating Cross-Domain Knowledge Reliability in Large Language Models
von: Jackson, Declan, et al.
Veröffentlicht: (2025)
von: Jackson, Declan, et al.
Veröffentlicht: (2025)
Domain-Specific Tensor Languages
von: Bernardy, Jean-Philippe, et al.
Veröffentlicht: (2023)
von: Bernardy, Jean-Philippe, et al.
Veröffentlicht: (2023)
Detecting Errors through Ensembling Prompts (DEEP): An End-to-End LLM Framework for Detecting Factual Errors
von: Chandler, Alex, et al.
Veröffentlicht: (2024)
von: Chandler, Alex, et al.
Veröffentlicht: (2024)
Single Character Perturbations Break LLM Alignment
von: Lin, Leon, et al.
Veröffentlicht: (2024)
von: Lin, Leon, et al.
Veröffentlicht: (2024)
Stable and Explainable Personality Trait Evaluation in Large Language Models with Internal Activations
von: Ma, Xiaoxu, et al.
Veröffentlicht: (2026)
von: Ma, Xiaoxu, et al.
Veröffentlicht: (2026)
Artificial Impressions: Evaluating Large Language Model Behavior Through the Lens of Trait Impressions
von: Deas, Nicholas, et al.
Veröffentlicht: (2025)
von: Deas, Nicholas, et al.
Veröffentlicht: (2025)
TimeChara: Evaluating Point-in-Time Character Hallucination of Role-Playing Large Language Models
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2024)
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2024)
Robust Bias Detection in MLMs and its Application to Human Trait Ratings
von: Shrestha, Ingroj, et al.
Veröffentlicht: (2025)
von: Shrestha, Ingroj, et al.
Veröffentlicht: (2025)
Programming of Cellular Automata in C and C++
von: Christen, Patrik
Veröffentlicht: (2024)
von: Christen, Patrik
Veröffentlicht: (2024)
MET-Bench: Multimodal Entity Tracking for Evaluating the Limitations of Vision-Language and Reasoning Models
von: Cohen, Vanya, et al.
Veröffentlicht: (2025)
von: Cohen, Vanya, et al.
Veröffentlicht: (2025)
ChainNet: Structured Metaphor and Metonymy in WordNet
von: Maudslay, Rowan Hall, et al.
Veröffentlicht: (2024)
von: Maudslay, Rowan Hall, et al.
Veröffentlicht: (2024)
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model
von: Brown, Andrew, et al.
Veröffentlicht: (2024)
von: Brown, Andrew, et al.
Veröffentlicht: (2024)
CharBench: Evaluating the Role of Tokenization in Character-Level Tasks
von: Uzan, Omri, et al.
Veröffentlicht: (2025)
von: Uzan, Omri, et al.
Veröffentlicht: (2025)
Incorporating Different Verbal Cues to Improve Text-Based Computer-Delivered Health Messaging
von: Cox, Samuel Rhys
Veröffentlicht: (2024)
von: Cox, Samuel Rhys
Veröffentlicht: (2024)
Theory of Mind and Self-Disclosure to CUIs
von: Cox, Samuel Rhys
Veröffentlicht: (2025)
von: Cox, Samuel Rhys
Veröffentlicht: (2025)
HEDS 3.0: The Human Evaluation Data Sheet Version 3.0
von: Belz, Anya, et al.
Veröffentlicht: (2024)
von: Belz, Anya, et al.
Veröffentlicht: (2024)
How Do Language Models Acquire Character-Level Information?
von: Sato, Soma, et al.
Veröffentlicht: (2026)
von: Sato, Soma, et al.
Veröffentlicht: (2026)
Improving Language and Modality Transfer in Translation by Character-level Modeling
von: Tsiamas, Ioannis, et al.
Veröffentlicht: (2025)
von: Tsiamas, Ioannis, et al.
Veröffentlicht: (2025)
Evaluating The Impact of Stimulus Quality in Investigations of LLM Language Performance
von: Pistotti, Timothy, et al.
Veröffentlicht: (2025)
von: Pistotti, Timothy, et al.
Veröffentlicht: (2025)
BenchCLAMP: A Benchmark for Evaluating Language Models on Syntactic and Semantic Parsing
von: Roy, Subhro, et al.
Veröffentlicht: (2022)
von: Roy, Subhro, et al.
Veröffentlicht: (2022)
LLM-as-a-Grader: Practical Insights from Large Language Model for Short-Answer and Report Evaluation
von: Byun, Grace, et al.
Veröffentlicht: (2025)
von: Byun, Grace, et al.
Veröffentlicht: (2025)
Eliciting Personality Traits in Large Language Models
von: Hilliard, Airlie, et al.
Veröffentlicht: (2024)
von: Hilliard, Airlie, et al.
Veröffentlicht: (2024)
From Language Models over Tokens to Language Models over Characters
von: Vieira, Tim, et al.
Veröffentlicht: (2024)
von: Vieira, Tim, et al.
Veröffentlicht: (2024)
Evaluating and Adapting Large Language Models to Represent Folktales in Low-Resource Languages
von: Meaney, JA, et al.
Veröffentlicht: (2024)
von: Meaney, JA, et al.
Veröffentlicht: (2024)
Hypothesis-only Biases in Large Language Model-Elicited Natural Language Inference
von: Proebsting, Grace, et al.
Veröffentlicht: (2024)
von: Proebsting, Grace, et al.
Veröffentlicht: (2024)
C-LLM: Learn to Check Chinese Spelling Errors Character by Character
von: Li, Kunting, et al.
Veröffentlicht: (2024)
von: Li, Kunting, et al.
Veröffentlicht: (2024)
The Strawberry Problem: Emergence of Character-level Understanding in Tokenized Language Models
von: Cosma, Adrian, et al.
Veröffentlicht: (2025)
von: Cosma, Adrian, et al.
Veröffentlicht: (2025)
Mitigating the Influence of Distractor Tasks in LMs with Prior-Aware Decoding
von: Douglas, Raymond, et al.
Veröffentlicht: (2024)
von: Douglas, Raymond, et al.
Veröffentlicht: (2024)
Neuron-based Personality Trait Induction in Large Language Models
von: Deng, Jia, et al.
Veröffentlicht: (2024)
von: Deng, Jia, et al.
Veröffentlicht: (2024)
CharacterEval: A Chinese Benchmark for Role-Playing Conversational Agent Evaluation
von: Tu, Quan, et al.
Veröffentlicht: (2024)
von: Tu, Quan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AI Sandbagging: Language Models can Strategically Underperform on Evaluations
von: van der Weij, Teun, et al.
Veröffentlicht: (2024) -
Persona Vectors: Monitoring and Controlling Character Traits in Language Models
von: Chen, Runjin, et al.
Veröffentlicht: (2025) -
Towards a Theory of AI Personhood
von: Ward, Francis Rhys
Veröffentlicht: (2025) -
Spherical Steering: Geometry-Aware Activation Rotation for Language Models
von: You, Zejia, et al.
Veröffentlicht: (2026) -
Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works
von: Yuan, Xinfeng, et al.
Veröffentlicht: (2024)