Value Portrait: Assessing Language Models' Values through Psychometrically and Ecologically Valid Items
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Han, Jongwook, Choi, Dongmin, Song, Woojung, Lee, Eun-Ju, Jo, Yohan |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Human Psychometric Questionnaires Mischaracterize LLM Behavior
par: Song, Woojung, et autres
Publié: (2025)
par: Song, Woojung, et autres
Publié: (2025)
Dialogue Systems for Emotional Support via Value Reinforcement
par: Kim, Juhee, et autres
Publié: (2025)
par: Kim, Juhee, et autres
Publié: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
par: Peters, Sydney, et autres
Publié: (2025)
par: Peters, Sydney, et autres
Publié: (2025)
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
par: Lim, Sungjib, et autres
Publié: (2025)
par: Lim, Sungjib, et autres
Publié: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
par: Fadli, Samih
Publié: (2025)
par: Fadli, Samih
Publié: (2025)
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
par: Yao, Ben, et autres
Publié: (2025)
par: Yao, Ben, et autres
Publié: (2025)
The Generation Gap: Exploring Age Bias in the Value Systems of Large Language Models
par: Liu, Siyang, et autres
Publié: (2024)
par: Liu, Siyang, et autres
Publié: (2024)
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
par: Karinshak, Elise, et autres
Publié: (2024)
par: Karinshak, Elise, et autres
Publié: (2024)
Homogeneous Keys, Heterogeneous Values: Exploiting Local KV Cache Asymmetry for Long-Context LLMs
par: Cui, Wanyun, et autres
Publié: (2025)
par: Cui, Wanyun, et autres
Publié: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
par: Collado-Montañez, Jaime, et autres
Publié: (2025)
par: Collado-Montañez, Jaime, et autres
Publié: (2025)
ConPET: Continual Parameter-Efficient Tuning for Large Language Models
par: Song, Chenyang, et autres
Publié: (2023)
par: Song, Chenyang, et autres
Publié: (2023)
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
par: Park, Seungcheol, et autres
Publié: (2025)
par: Park, Seungcheol, et autres
Publié: (2025)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
par: Han, Xudong, et autres
Publié: (2025)
par: Han, Xudong, et autres
Publié: (2025)
Quantifying Data Contamination in Psychometric Evaluations of LLMs
par: Han, Jongwook, et autres
Publié: (2025)
par: Han, Jongwook, et autres
Publié: (2025)
Surprisingly Fragile: Assessing and Addressing Prompt Instability in Multimodal Foundation Models
par: Stewart, Ian, et autres
Publié: (2024)
par: Stewart, Ian, et autres
Publié: (2024)
A Multi-Pass Large Language Model Framework for Precise and Efficient Radiology Report Error Detection
par: Kim, Songsoo, et autres
Publié: (2025)
par: Kim, Songsoo, et autres
Publié: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
par: Ashuach, Tomer, et autres
Publié: (2025)
par: Ashuach, Tomer, et autres
Publié: (2025)
Adaptive Steering and Remasking for Safe Generation in Diffusion Language Models
par: Lee, Yejin, et autres
Publié: (2026)
par: Lee, Yejin, et autres
Publié: (2026)
RUQuant: Towards Refining Uniform Quantization for Large Language Models
par: Liu, Han, et autres
Publié: (2026)
par: Liu, Han, et autres
Publié: (2026)
A Multi-Task Benchmark for Abusive Language Detection in Low-Resource Settings
par: Gaim, Fitsum, et autres
Publié: (2025)
par: Gaim, Fitsum, et autres
Publié: (2025)
Does Localization Inform Unlearning? A Rigorous Examination of Local Parameter Attribution for Knowledge Unlearning in Language Models
par: Lee, Hwiyeong, et autres
Publié: (2025)
par: Lee, Hwiyeong, et autres
Publié: (2025)
Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents
par: Wang, Renxi, et autres
Publié: (2024)
par: Wang, Renxi, et autres
Publié: (2024)
SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models
par: Liu, Han, et autres
Publié: (2026)
par: Liu, Han, et autres
Publié: (2026)
Neither Valid nor Reliable? Investigating the Use of LLMs as Judges
par: Chehbouni, Khaoula, et autres
Publié: (2025)
par: Chehbouni, Khaoula, et autres
Publié: (2025)
Reliable Part-of-Speech Tagging of Historical Corpora through Set-Valued Prediction
par: Heid, Stefan, et autres
Publié: (2020)
par: Heid, Stefan, et autres
Publié: (2020)
SentiCSE: A Sentiment-aware Contrastive Sentence Embedding Framework with Sentiment-guided Textual Similarity
par: Kim, Jaemin, et autres
Publié: (2024)
par: Kim, Jaemin, et autres
Publié: (2024)
Multi-Method Validation of Large Language Model Medical Translation Across High- and Low-Resource Languages
par: Anyaegbuna, Chukwuebuka, et autres
Publié: (2026)
par: Anyaegbuna, Chukwuebuka, et autres
Publié: (2026)
"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations
par: Choi, Jonghyeon, et autres
Publié: (2025)
par: Choi, Jonghyeon, et autres
Publié: (2025)
Value Lens: Using Large Language Models to Understand Human Values
par: Fernández, Eduardo de la Cruz, et autres
Publié: (2025)
par: Fernández, Eduardo de la Cruz, et autres
Publié: (2025)
Precise Length Control in Large Language Models
par: Butcher, Bradley, et autres
Publié: (2024)
par: Butcher, Bradley, et autres
Publié: (2024)
What Drives Performance in Multilingual Language Models?
par: Nezhad, Sina Bagheri, et autres
Publié: (2024)
par: Nezhad, Sina Bagheri, et autres
Publié: (2024)
Large Language Models for Biomedical Article Classification
par: Proboszcz, Jakub, et autres
Publié: (2026)
par: Proboszcz, Jakub, et autres
Publié: (2026)
Action-Item-Driven Summarization of Long Meeting Transcripts
par: Golia, Logan, et autres
Publié: (2023)
par: Golia, Logan, et autres
Publié: (2023)
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
par: Park, Seungcheol, et autres
Publié: (2023)
par: Park, Seungcheol, et autres
Publié: (2023)
SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading
par: Dinh, Tu Anh, et autres
Publié: (2024)
par: Dinh, Tu Anh, et autres
Publié: (2024)
Strategy Adaptation in Large Language Model Werewolf Agents
par: Nakamori, Fuya, et autres
Publié: (2025)
par: Nakamori, Fuya, et autres
Publié: (2025)
PL-Guard: Benchmarking Language Model Safety for Polish
par: Krasnodębska, Aleksandra, et autres
Publié: (2025)
par: Krasnodębska, Aleksandra, et autres
Publié: (2025)
Socially Responsible Data for Large Multilingual Language Models
par: Smart, Andrew, et autres
Publié: (2024)
par: Smart, Andrew, et autres
Publié: (2024)
A RoBERTa-Based Functional Syntax Annotation Model for Chinese Texts
par: Xiaohui, Han, et autres
Publié: (2025)
par: Xiaohui, Han, et autres
Publié: (2025)
Token Erasure as a Footprint of Implicit Vocabulary Items in LLMs
par: Feucht, Sheridan, et autres
Publié: (2024)
par: Feucht, Sheridan, et autres
Publié: (2024)
Documents similaires
-
Human Psychometric Questionnaires Mischaracterize LLM Behavior
par: Song, Woojung, et autres
Publié: (2025) -
Dialogue Systems for Emotional Support via Value Reinforcement
par: Kim, Juhee, et autres
Publié: (2025) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
par: Peters, Sydney, et autres
Publié: (2025) -
Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
par: Lim, Sungjib, et autres
Publié: (2025) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
par: Fadli, Samih
Publié: (2025)