Measuring Scalar Constructs in Social Science with LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Licht, Hauke, Sarkar, Rupak, Wu, Patrick Y., Goel, Pranav, Stoehr, Niklas, Ash, Elliott, Hoyle, Alexander Miserlis |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Natural Language Decompositions of Implicit Content Enable Better Text Representations
by: Hoyle, Alexander, et al.
Published: (2023)
by: Hoyle, Alexander, et al.
Published: (2023)
Computational emotion analysis with multimodal LLMs: Current evidence on an emerging methodological opportunity
by: Licht, Hauke
Published: (2025)
by: Licht, Hauke
Published: (2025)
How Persuasive is Your Context?
by: Nguyen, Tu, et al.
Published: (2025)
by: Nguyen, Tu, et al.
Published: (2025)
The Medium Is Not the Message: Deconfounding Document Embeddings via Linear Concept Erasure
by: Fan, Yu, et al.
Published: (2025)
by: Fan, Yu, et al.
Published: (2025)
Whose Preferences? Differences in Fairness Preferences and Their Impact on the Fairness of AI Utilizing Human Feedback
by: Lerner, Emilia Agis, et al.
Published: (2024)
by: Lerner, Emilia Agis, et al.
Published: (2024)
World Models for Math Story Problems
by: Opedal, Andreas, et al.
Published: (2023)
by: Opedal, Andreas, et al.
Published: (2023)
Can Reasoning Help Large Language Models Capture Human Annotator Disagreement?
by: Ni, Jingwei, et al.
Published: (2025)
by: Ni, Jingwei, et al.
Published: (2025)
Read the Paper, Write the Code: Agentic Reproduction of Social-Science Results
by: Kohler, Benjamin, et al.
Published: (2026)
by: Kohler, Benjamin, et al.
Published: (2026)
Modeling Motivated Reasoning in Law: Evaluating Strategic Role Conditioning in LLM Summarization
by: Cho, Eunjung, et al.
Published: (2025)
by: Cho, Eunjung, et al.
Published: (2025)
Localizing Paragraph Memorization in Language Models
by: Stoehr, Niklas, et al.
Published: (2024)
by: Stoehr, Niklas, et al.
Published: (2024)
Agentic Insight Generation in VSM Simulations
by: Selak, Micha, et al.
Published: (2026)
by: Selak, Micha, et al.
Published: (2026)
LePaRD: A Large-Scale Dataset of Judges Citing Precedents
by: Mahari, Robert, et al.
Published: (2023)
by: Mahari, Robert, et al.
Published: (2023)
Unsupervised Contrast-Consistent Ranking with Language Models
by: Stoehr, Niklas, et al.
Published: (2023)
by: Stoehr, Niklas, et al.
Published: (2023)
Co-DETECT: Collaborative Discovery of Edge Cases in Text Classification
by: Xiong, Chenfei, et al.
Published: (2025)
by: Xiong, Chenfei, et al.
Published: (2025)
Context versus Prior Knowledge in Language Models
by: Du, Kevin, et al.
Published: (2024)
by: Du, Kevin, et al.
Published: (2024)
Aligning Large Language Models with Diverse Political Viewpoints
by: Stammbach, Dominik, et al.
Published: (2024)
by: Stammbach, Dominik, et al.
Published: (2024)
Translating Legalese: Enhancing Public Understanding of Court Opinions with Legal Summarizers
by: Ash, Elliott, et al.
Published: (2023)
by: Ash, Elliott, et al.
Published: (2023)
Where Do People Tell Stories Online? Story Detection Across Online Communities
by: Antoniak, Maria, et al.
Published: (2023)
by: Antoniak, Maria, et al.
Published: (2023)
Activation Scaling for Steering and Interpreting Language Models
by: Stoehr, Niklas, et al.
Published: (2024)
by: Stoehr, Niklas, et al.
Published: (2024)
Codebook LLMs: Evaluating LLMs as Measurement Tools for Political Science Concepts
by: Halterman, Andrew, et al.
Published: (2024)
by: Halterman, Andrew, et al.
Published: (2024)
ELEPHANT: Measuring and understanding social sycophancy in LLMs
by: Cheng, Myra, et al.
Published: (2025)
by: Cheng, Myra, et al.
Published: (2025)
A SMART Mnemonic Sounds like "Glue Tonic": Mixing LLMs with Student Feedback to Make Mnemonic Learning Stick
by: Balepur, Nishant, et al.
Published: (2024)
by: Balepur, Nishant, et al.
Published: (2024)
Variational Best-of-N Alignment
by: Amini, Afra, et al.
Published: (2024)
by: Amini, Afra, et al.
Published: (2024)
Media Slant is Contagious
by: Widmer, Philine, et al.
Published: (2022)
by: Widmer, Philine, et al.
Published: (2022)
Large Language Models Can Be a Viable Substitute for Expert Political Surveys When a Shock Disrupts Traditional Measurement Approaches
by: Wu, Patrick Y.
Published: (2025)
by: Wu, Patrick Y.
Published: (2025)
ProxAnn: Use-Oriented Evaluations of Topic Models and Document Clustering
by: Hoyle, Alexander, et al.
Published: (2025)
by: Hoyle, Alexander, et al.
Published: (2025)
Towards Faithful and Robust LLM Specialists for Evidence-Based Question-Answering
by: Schimanski, Tobias, et al.
Published: (2024)
by: Schimanski, Tobias, et al.
Published: (2024)
Controllable Context Sensitivity and the Knob Behind It
by: Minder, Julian, et al.
Published: (2024)
by: Minder, Julian, et al.
Published: (2024)
Fake News Detection After LLM Laundering: Measurement and Explanation
by: Das, Rupak Kumar, et al.
Published: (2025)
by: Das, Rupak Kumar, et al.
Published: (2025)
Understanding Common Ground Misalignment in Goal-Oriented Dialog: A Case-Study with Ubuntu Chat Logs
by: Sarkar, Rupak, et al.
Published: (2025)
by: Sarkar, Rupak, et al.
Published: (2025)
TopicGPT: A Prompt-based Topic Modeling Framework
by: Pham, Chau Minh, et al.
Published: (2023)
by: Pham, Chau Minh, et al.
Published: (2023)
Position: Enough of Scaling LLMs! Lets Focus on Downscaling
by: Goel, Yash, et al.
Published: (2025)
by: Goel, Yash, et al.
Published: (2025)
CUTE: Measuring LLMs' Understanding of Their Tokens
by: Edman, Lukas, et al.
Published: (2024)
by: Edman, Lukas, et al.
Published: (2024)
Social Science Meets LLMs: How Reliable Are Large Language Models in Social Simulations?
by: Huang, Yue, et al.
Published: (2024)
by: Huang, Yue, et al.
Published: (2024)
SemEval-2017 Task 4: Sentiment Analysis in Twitter using BERT
by: Das, Rupak Kumar, et al.
Published: (2024)
by: Das, Rupak Kumar, et al.
Published: (2024)
AFaCTA: Assisting the Annotation of Factual Claim Detection with Reliable LLM Annotators
by: Ni, Jingwei, et al.
Published: (2024)
by: Ni, Jingwei, et al.
Published: (2024)
RoboPhD: Self-Improving Text-to-SQL Through Autonomous Agent Evolution
by: Borthwick, Andrew, et al.
Published: (2026)
by: Borthwick, Andrew, et al.
Published: (2026)
LLM-Measure: Generating Valid, Consistent, and Reproducible Text-Based Measures for Social Science Research
by: Yang, Yi, et al.
Published: (2024)
by: Yang, Yi, et al.
Published: (2024)
Toward Responsible and Epistemically Grounded Multilingual LLMs for Computational Social Science and Humanities
by: Zaghouani, Wajdi
Published: (2026)
by: Zaghouani, Wajdi
Published: (2026)
Safer Reasoning Traces: Measuring and Mitigating Chain-of-Thought Leakage in LLMs
by: Ahrend, Patrick, et al.
Published: (2026)
by: Ahrend, Patrick, et al.
Published: (2026)
Similar Items
-
Natural Language Decompositions of Implicit Content Enable Better Text Representations
by: Hoyle, Alexander, et al.
Published: (2023) -
Computational emotion analysis with multimodal LLMs: Current evidence on an emerging methodological opportunity
by: Licht, Hauke
Published: (2025) -
How Persuasive is Your Context?
by: Nguyen, Tu, et al.
Published: (2025) -
The Medium Is Not the Message: Deconfounding Document Embeddings via Linear Concept Erasure
by: Fan, Yu, et al.
Published: (2025) -
Whose Preferences? Differences in Fairness Preferences and Their Impact on the Fairness of AI Utilizing Human Feedback
by: Lerner, Emilia Agis, et al.
Published: (2024)