Feedback Forensics: A Toolkit to Measure AI Personality
Fuente:
arXiv
Saved in:
| Main Authors: | Findeis, Arduin, Kaufmann, Timo, Hüllermeier, Eyke, Mullins, Robert |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Inverse Constitutional AI: Compressing Preferences into Principles
by: Findeis, Arduin, et al.
Published: (2024)
by: Findeis, Arduin, et al.
Published: (2024)
Can External Validation Tools Improve Annotation Quality for LLM-as-a-Judge?
by: Findeis, Arduin, et al.
Published: (2025)
by: Findeis, Arduin, et al.
Published: (2025)
Problem Solving Through Human-AI Preference-Based Cooperation
by: Dutta, Subhabrata, et al.
Published: (2024)
by: Dutta, Subhabrata, et al.
Published: (2024)
OCALM: Object-Centric Assessment with Language Models
by: Kaufmann, Timo, et al.
Published: (2024)
by: Kaufmann, Timo, et al.
Published: (2024)
AI Steerability 360: A Toolkit for Steering Large Language Models
by: Miehling, Erik, et al.
Published: (2026)
by: Miehling, Erik, et al.
Published: (2026)
Antithetic Sampling for Top-k Shapley Identification
by: Kolpaczki, Patrick, et al.
Published: (2025)
by: Kolpaczki, Patrick, et al.
Published: (2025)
A Survey of AI-generated Text Forensic Systems: Detection, Attribution, and Characterization
by: Kumarage, Tharindu, et al.
Published: (2024)
by: Kumarage, Tharindu, et al.
Published: (2024)
Calibrated Preference Learning: The Case of Label Ranking
by: Thies, Santo M. A. R., et al.
Published: (2026)
by: Thies, Santo M. A. R., et al.
Published: (2026)
Controllable and explainable personality sliders for LLMs at inference time
by: Hoppe, Florian, et al.
Published: (2026)
by: Hoppe, Florian, et al.
Published: (2026)
A Survey of Reinforcement Learning from Human Feedback
by: Kaufmann, Timo, et al.
Published: (2023)
by: Kaufmann, Timo, et al.
Published: (2023)
Personalizing LLMs with Binary Feedback: A Preference-Corrected Optimization Framework
by: Ma, Xilai, et al.
Published: (2026)
by: Ma, Xilai, et al.
Published: (2026)
Sketch: A Toolkit for Streamlining LLM Operations
by: Jiang, Xin, et al.
Published: (2024)
by: Jiang, Xin, et al.
Published: (2024)
Personalized Language Modeling from Personalized Human Feedback
by: Li, Xinyu, et al.
Published: (2024)
by: Li, Xinyu, et al.
Published: (2024)
An Automatic Question Usability Evaluation Toolkit
by: Moore, Steven, et al.
Published: (2024)
by: Moore, Steven, et al.
Published: (2024)
ForensicsData: A Digital Forensics Dataset for Large Language Models
by: Chakir, Youssef, et al.
Published: (2025)
by: Chakir, Youssef, et al.
Published: (2025)
Citekit: A Modular Toolkit for Large Language Model Citation Generation
by: Shen, Jiajun, et al.
Published: (2024)
by: Shen, Jiajun, et al.
Published: (2024)
Improved Personalized Headline Generation via Denoising Fake Interests from Implicit Feedback
by: Liu, Kejin, et al.
Published: (2025)
by: Liu, Kejin, et al.
Published: (2025)
Aligning to Illusions: Choice Blindness in Human and AI Feedback
by: Wu, Wenbin
Published: (2026)
by: Wu, Wenbin
Published: (2026)
Learning Personalized Agents from Human Feedback
by: Liang, Kaiqu, et al.
Published: (2026)
by: Liang, Kaiqu, et al.
Published: (2026)
Large language models struggle with ethnographic text annotation
by: Goodall, Leonardo S., et al.
Published: (2026)
by: Goodall, Leonardo S., et al.
Published: (2026)
Self-Assessment Tests are Unreliable Measures of LLM Personality
by: Gupta, Akshat, et al.
Published: (2023)
by: Gupta, Akshat, et al.
Published: (2023)
OntoAligner: A Comprehensive Modular and Robust Python Toolkit for Ontology Alignment
by: Giglou, Hamed Babaei, et al.
Published: (2025)
by: Giglou, Hamed Babaei, et al.
Published: (2025)
WalledEval: A Comprehensive Safety Evaluation Toolkit for Large Language Models
by: Gupta, Prannaya, et al.
Published: (2024)
by: Gupta, Prannaya, et al.
Published: (2024)
findsylls: A Language-Agnostic Toolkit for Syllable-Level Speech Tokenization and Embedding
by: Martínez, Héctor Javier Vázquez
Published: (2026)
by: Martínez, Héctor Javier Vázquez
Published: (2026)
LMFlow: An Extensible Toolkit for Finetuning and Inference of Large Foundation Models
by: Diao, Shizhe, et al.
Published: (2023)
by: Diao, Shizhe, et al.
Published: (2023)
Measuring AI Reasoning: A Guide for Researchers
by: Nwadike, Munachiso Samuel, et al.
Published: (2026)
by: Nwadike, Munachiso Samuel, et al.
Published: (2026)
HalluCiteChecker: A Lightweight Toolkit for Hallucinated Citation Detection and Verification in the Era of AI Scientists
by: Sakai, Yusuke, et al.
Published: (2026)
by: Sakai, Yusuke, et al.
Published: (2026)
EasyDistill: A Comprehensive Toolkit for Effective Knowledge Distillation of Large Language Models
by: Wang, Chengyu, et al.
Published: (2025)
by: Wang, Chengyu, et al.
Published: (2025)
SinaTools: Open Source Toolkit for Arabic Natural Language Processing
by: Hammouda, Tymaa, et al.
Published: (2024)
by: Hammouda, Tymaa, et al.
Published: (2024)
UltraFeedback: Boosting Language Models with Scaled AI Feedback
by: Cui, Ganqu, et al.
Published: (2023)
by: Cui, Ganqu, et al.
Published: (2023)
Memoria: A Scalable Agentic Memory Framework for Personalized Conversational AI
by: Sarin, Samarth, et al.
Published: (2025)
by: Sarin, Samarth, et al.
Published: (2025)
From Feedback to Checklists: Grounded Evaluation of AI-Generated Clinical Notes
by: Zhou, Karen, et al.
Published: (2025)
by: Zhou, Karen, et al.
Published: (2025)
AI PERSONA: Towards Life-long Personalization of LLMs
by: Wang, Tiannan, et al.
Published: (2024)
by: Wang, Tiannan, et al.
Published: (2024)
TravelAgent: An AI Assistant for Personalized Travel Planning
by: Chen, Aili, et al.
Published: (2024)
by: Chen, Aili, et al.
Published: (2024)
Humanity in AI: Detecting the Personality of Large Language Models
by: Zhan, Baohua, et al.
Published: (2024)
by: Zhan, Baohua, et al.
Published: (2024)
IPO-Mine: A Toolkit and Dataset for Section-Structured Analysis of Long, Multimodal IPO Documents
by: Galarnyk, Michael, et al.
Published: (2026)
by: Galarnyk, Michael, et al.
Published: (2026)
The Responsible Development of Automated Student Feedback with Generative AI
by: Lindsay, Euan D, et al.
Published: (2023)
by: Lindsay, Euan D, et al.
Published: (2023)
Learning from Natural Language Feedback for Personalized Question Answering
by: Salemi, Alireza, et al.
Published: (2025)
by: Salemi, Alireza, et al.
Published: (2025)
SciEvalKit: An Open-source Evaluation Toolkit for Scientific General Intelligence
by: Wang, Yiheng, et al.
Published: (2025)
by: Wang, Yiheng, et al.
Published: (2025)
Fabricator: An Open Source Toolkit for Generating Labeled Training Data with Teacher LLMs
by: Golde, Jonas, et al.
Published: (2023)
by: Golde, Jonas, et al.
Published: (2023)
Similar Items
-
Inverse Constitutional AI: Compressing Preferences into Principles
by: Findeis, Arduin, et al.
Published: (2024) -
Can External Validation Tools Improve Annotation Quality for LLM-as-a-Judge?
by: Findeis, Arduin, et al.
Published: (2025) -
Problem Solving Through Human-AI Preference-Based Cooperation
by: Dutta, Subhabrata, et al.
Published: (2024) -
OCALM: Object-Centric Assessment with Language Models
by: Kaufmann, Timo, et al.
Published: (2024) -
AI Steerability 360: A Toolkit for Steering Large Language Models
by: Miehling, Erik, et al.
Published: (2026)