Gespeichert in:
| Hauptverfasser: | Schelb, Julian, Borin, Orr, Garcia, David, Spitz, Andreas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2503.10229 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Assessing In-context Learning and Fine-tuning for Topic Classification of German Web Data
von: Schelb, Julian, et al.
Veröffentlicht: (2024)
von: Schelb, Julian, et al.
Veröffentlicht: (2024)
Loci Similes: A Benchmark for Extracting Intertextualities in Latin Literature
von: Schelb, Julian, et al.
Veröffentlicht: (2026)
von: Schelb, Julian, et al.
Veröffentlicht: (2026)
Only a Little to the Left: A Theory-grounded Measure of Political Bias in Large Language Models
von: Faulborn, Mats, et al.
Veröffentlicht: (2025)
von: Faulborn, Mats, et al.
Veröffentlicht: (2025)
PsychoLex: Unveiling the Psychological Mind of Large Language Models
von: Abbasi, Mohammad Amin, et al.
Veröffentlicht: (2024)
von: Abbasi, Mohammad Amin, et al.
Veröffentlicht: (2024)
Quantifying the Risks of Tool-assisted Rephrasing to Linguistic Diversity
von: Wang, Mengying, et al.
Veröffentlicht: (2024)
von: Wang, Mengying, et al.
Veröffentlicht: (2024)
Do Psychometric Tests Work for Large Language Models? Evaluation of Tests on Sexism, Racism, and Morality
von: Jung, Jana, et al.
Veröffentlicht: (2025)
von: Jung, Jana, et al.
Veröffentlicht: (2025)
Generative Psycho-Lexical Approach for Constructing Value Systems in Large Language Models
von: Ye, Haoran, et al.
Veröffentlicht: (2025)
von: Ye, Haoran, et al.
Veröffentlicht: (2025)
Tag-Pag: A Dedicated Tool for Systematic Web Page Annotations
von: Pogrebnjak, Anton, et al.
Veröffentlicht: (2025)
von: Pogrebnjak, Anton, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models with Psychometrics
von: Li, Yuan, et al.
Veröffentlicht: (2024)
von: Li, Yuan, et al.
Veröffentlicht: (2024)
Preference Learning Unlocks LLMs' Psycho-Counseling Skills
von: Zhang, Mian, et al.
Veröffentlicht: (2025)
von: Zhang, Mian, et al.
Veröffentlicht: (2025)
Psychometric Predictive Power of Large Language Models
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
Who is ChatGPT? Benchmarking LLMs' Psychological Portrayal Using PsychoBench
von: Huang, Jen-tse, et al.
Veröffentlicht: (2023)
von: Huang, Jen-tse, et al.
Veröffentlicht: (2023)
Statistical Hypothesis Testing for Auditing Robustness in Language Models
von: Rauba, Paulius, et al.
Veröffentlicht: (2025)
von: Rauba, Paulius, et al.
Veröffentlicht: (2025)
Psycho-linguistic Experiment on Universal Semantic Components of Verbal Humor: System Description and Annotation
von: Mikhalkova, Elena, et al.
Veröffentlicht: (2024)
von: Mikhalkova, Elena, et al.
Veröffentlicht: (2024)
Do LLMs Give Psychometrically Plausible Responses in Educational Assessments?
von: Säuberli, Andreas, et al.
Veröffentlicht: (2025)
von: Säuberli, Andreas, et al.
Veröffentlicht: (2025)
Psychometric Personality Shaping Modulates Capabilities and Safety in Language Models
von: Fitz, Stephen, et al.
Veröffentlicht: (2025)
von: Fitz, Stephen, et al.
Veröffentlicht: (2025)
Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization
von: Sun, Yudao, et al.
Veröffentlicht: (2025)
von: Sun, Yudao, et al.
Veröffentlicht: (2025)
Test-Time Fairness and Robustness in Large Language Models
von: Cotta, Leonardo, et al.
Veröffentlicht: (2024)
von: Cotta, Leonardo, et al.
Veröffentlicht: (2024)
Psychometric Alignment: Capturing Human Knowledge Distributions via Language Models
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
Automatic Generation and Evaluation of Reading Comprehension Test Items with Large Language Models
von: Säuberli, Andreas, et al.
Veröffentlicht: (2024)
von: Säuberli, Andreas, et al.
Veröffentlicht: (2024)
Adaptive Testing for LLM Evaluation: A Psychometric Alternative to Static Benchmarks
von: Li, Peiyu, et al.
Veröffentlicht: (2025)
von: Li, Peiyu, et al.
Veröffentlicht: (2025)
United States Politicians' Tone Became More Negative with 2016 Primary Campaigns
von: Külz, Jonathan, et al.
Veröffentlicht: (2022)
von: Külz, Jonathan, et al.
Veröffentlicht: (2022)
Measuring Human and AI Values Based on Generative Psychometrics with Large Language Models
von: Ye, Haoran, et al.
Veröffentlicht: (2024)
von: Ye, Haoran, et al.
Veröffentlicht: (2024)
Large Language Model Psychometrics: A Systematic Review of Evaluation, Validation, and Enhancement
von: Ye, Haoran, et al.
Veröffentlicht: (2025)
von: Ye, Haoran, et al.
Veröffentlicht: (2025)
The Impact of Image Resolution on Biomedical Multimodal Large Language Models
von: Chen, Liangyu, et al.
Veröffentlicht: (2025)
von: Chen, Liangyu, et al.
Veröffentlicht: (2025)
The Expressions of Depression and Anxiety in Chinese Psycho-counseling: Usage of First-person Singular Pronoun and Negative Emotional Words
von: Ma, Lizhi, et al.
Veröffentlicht: (2025)
von: Ma, Lizhi, et al.
Veröffentlicht: (2025)
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
von: Wen, Yuchen, et al.
Veröffentlicht: (2024)
von: Wen, Yuchen, et al.
Veröffentlicht: (2024)
Defining and Evaluating Visual Language Models' Basic Spatial Abilities: A Perspective from Psychometrics
von: Xu, Wenrui, et al.
Veröffentlicht: (2025)
von: Xu, Wenrui, et al.
Veröffentlicht: (2025)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
VideoAgent: Long-form Video Understanding with Large Language Model as Agent
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
Camouflage is all you need: Evaluating and Enhancing Language Model Robustness Against Camouflage Adversarial Attacks
von: Huertas-García, Álvaro, et al.
Veröffentlicht: (2024)
von: Huertas-García, Álvaro, et al.
Veröffentlicht: (2024)
A Novel Psychometrics-Based Approach to Developing Professional Competency Benchmark for Large Language Models
von: Kardanova, Elena, et al.
Veröffentlicht: (2024)
von: Kardanova, Elena, et al.
Veröffentlicht: (2024)
Mixed-R1: Unified Reward Perspective For Reasoning Capability in Multimodal Large Language Models
von: Xu, Shilin, et al.
Veröffentlicht: (2025)
von: Xu, Shilin, et al.
Veröffentlicht: (2025)
Synchronic and Diachronic Aspects of Kanashi
von: Saxena, Anju, et al.
Veröffentlicht: (2022)
von: Saxena, Anju, et al.
Veröffentlicht: (2022)
You don't need a personality test to know these models are unreliable: Assessing the Reliability of Large Language Models on Psychometric Instruments
von: Shu, Bangzhao, et al.
Veröffentlicht: (2023)
von: Shu, Bangzhao, et al.
Veröffentlicht: (2023)
Do GPT Language Models Suffer From Split Personality Disorder? The Advent Of Substrate-Free Psychometrics
von: Romero, Peter, et al.
Veröffentlicht: (2024)
von: Romero, Peter, et al.
Veröffentlicht: (2024)
Search-R3: Unifying Reasoning and Embedding in Large Language Models
von: Gui, Yuntao, et al.
Veröffentlicht: (2025)
von: Gui, Yuntao, et al.
Veröffentlicht: (2025)
PsychoGAT: A Novel Psychological Measurement Paradigm through Interactive Fiction Games with LLM Agents
von: Yang, Qisen, et al.
Veröffentlicht: (2024)
von: Yang, Qisen, et al.
Veröffentlicht: (2024)
On Non-interactive Evaluation of Animal Communication Translators
von: Paradise, Orr, et al.
Veröffentlicht: (2025)
von: Paradise, Orr, et al.
Veröffentlicht: (2025)
Langformers: Unified NLP Pipelines for Language Models
von: Lamsal, Rabindra, et al.
Veröffentlicht: (2025)
von: Lamsal, Rabindra, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Assessing In-context Learning and Fine-tuning for Topic Classification of German Web Data
von: Schelb, Julian, et al.
Veröffentlicht: (2024) -
Loci Similes: A Benchmark for Extracting Intertextualities in Latin Literature
von: Schelb, Julian, et al.
Veröffentlicht: (2026) -
Only a Little to the Left: A Theory-grounded Measure of Political Bias in Large Language Models
von: Faulborn, Mats, et al.
Veröffentlicht: (2025) -
PsychoLex: Unveiling the Psychological Mind of Large Language Models
von: Abbasi, Mohammad Amin, et al.
Veröffentlicht: (2024) -
Quantifying the Risks of Tool-assisted Rephrasing to Linguistic Diversity
von: Wang, Mengying, et al.
Veröffentlicht: (2024)