REALM: A Dataset of Real-World LLM Use Cases
Fuente:
arXiv
Salvato in:
| Autori principali: | Cheng, Jingwen, Ghate, Kshitish, Hua, Wenyue, Wang, William Yang, Shen, Hong, Fang, Fei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Who's in Charge? Disempowerment Patterns in Real-World LLM Usage
di: Sharma, Mrinank, et al.
Pubblicazione: (2026)
di: Sharma, Mrinank, et al.
Pubblicazione: (2026)
Towards Apples to Apples for AI Evaluations: From Real-World Use Cases to Evaluation Scenarios
di: Choong, Yee-Yin, et al.
Pubblicazione: (2026)
di: Choong, Yee-Yin, et al.
Pubblicazione: (2026)
Beyond One-Way Influence: Bidirectional Opinion Dynamics in Multi-Turn Human-LLM Interactions
di: Jiang, Yuyang, et al.
Pubblicazione: (2025)
di: Jiang, Yuyang, et al.
Pubblicazione: (2025)
Multi-Agent Comedy Club: Investigating Community Discussion Effects on LLM Humor Generation
di: Hong, Shiwei, et al.
Pubblicazione: (2026)
di: Hong, Shiwei, et al.
Pubblicazione: (2026)
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
di: Ghosh, Himel, et al.
Pubblicazione: (2026)
di: Ghosh, Himel, et al.
Pubblicazione: (2026)
If Eleanor Rigby Had Met ChatGPT: A Study on Loneliness in a Post-LLM World
di: de Wynter, Adrian
Pubblicazione: (2024)
di: de Wynter, Adrian
Pubblicazione: (2024)
LLM Agents for Education: Advances and Applications
di: Chu, Zhendong, et al.
Pubblicazione: (2025)
di: Chu, Zhendong, et al.
Pubblicazione: (2025)
DiMA: An LLM-Powered Ride-Hailing Assistant at DiDi
di: Ning, Yansong, et al.
Pubblicazione: (2025)
di: Ning, Yansong, et al.
Pubblicazione: (2025)
LLM Novice Uplift on Dual-Use, In Silico Biology Tasks
di: Zhang, Chen Bo Calvin, et al.
Pubblicazione: (2026)
di: Zhang, Chen Bo Calvin, et al.
Pubblicazione: (2026)
Exploring Communication Strategies for Collaborative LLM Agents in Mathematical Problem-Solving
di: Zhang, Liang, et al.
Pubblicazione: (2025)
di: Zhang, Liang, et al.
Pubblicazione: (2025)
Minion: A Technology Probe to Explore How Users Negotiate Harmful Value Conflicts with AI Companions
di: Fan, Xianzhe, et al.
Pubblicazione: (2024)
di: Fan, Xianzhe, et al.
Pubblicazione: (2024)
Social Welfare Function Leaderboard: When LLM Agents Allocate Social Welfare
di: Shi, Zhengliang, et al.
Pubblicazione: (2025)
di: Shi, Zhengliang, et al.
Pubblicazione: (2025)
Psychometric Comparability of LLM-Based Digital Twins
di: Zhang, Yufei, et al.
Pubblicazione: (2025)
di: Zhang, Yufei, et al.
Pubblicazione: (2025)
Human-Centred LLM Privacy Audits: Findings and Frictions
di: Staufer, Dimitri, et al.
Pubblicazione: (2026)
di: Staufer, Dimitri, et al.
Pubblicazione: (2026)
LLM Social Simulations Are a Promising Research Method
di: Anthis, Jacy Reese, et al.
Pubblicazione: (2025)
di: Anthis, Jacy Reese, et al.
Pubblicazione: (2025)
MythTriage: Scalable Detection of Opioid Use Disorder Myths on a Video-Sharing Platform
di: Jung, Hayoung, et al.
Pubblicazione: (2025)
di: Jung, Hayoung, et al.
Pubblicazione: (2025)
Toxic HallucinAItions: Perturbing Prompts and Tracing LLM Circuits
di: Shimgekar, Soorya Ram, et al.
Pubblicazione: (2026)
di: Shimgekar, Soorya Ram, et al.
Pubblicazione: (2026)
Between Rules and Reality: On the Context Sensitivity of LLM Moral Judgment
di: Sauter, Adrian, et al.
Pubblicazione: (2026)
di: Sauter, Adrian, et al.
Pubblicazione: (2026)
TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated Stories
di: Bhagat, Kirti, et al.
Pubblicazione: (2025)
di: Bhagat, Kirti, et al.
Pubblicazione: (2025)
Real or Robotic? Assessing Whether LLMs Accurately Simulate Qualities of Human Responses in Dialogue
di: Ivey, Jonathan, et al.
Pubblicazione: (2024)
di: Ivey, Jonathan, et al.
Pubblicazione: (2024)
From Prompts to Constructs: A Dual-Validity Framework for LLM Research in Psychology
di: Lin, Zhicheng
Pubblicazione: (2025)
di: Lin, Zhicheng
Pubblicazione: (2025)
LearnLens: LLM-Enabled Personalised, Curriculum-Grounded Feedback with Educators in the Loop
di: Zhao, Runcong, et al.
Pubblicazione: (2025)
di: Zhao, Runcong, et al.
Pubblicazione: (2025)
Synthetic Reader Panels: Tournament-Based Ideation with LLM Personas for Autonomous Publishing
di: Zimmerman, Fred
Pubblicazione: (2026)
di: Zimmerman, Fred
Pubblicazione: (2026)
An Empirical Investigation of Gender Stereotype Representation in Large Language Models: The Italian Case
di: Giachino, Gioele, et al.
Pubblicazione: (2025)
di: Giachino, Gioele, et al.
Pubblicazione: (2025)
Alignment Drift in Long-Term Human-LLM Interaction: A Mechanism-Oriented Framework
di: Yao, Xintong
Pubblicazione: (2026)
di: Yao, Xintong
Pubblicazione: (2026)
Comprehensive Study on Sentiment Analysis: From Rule-based to modern LLM based system
di: Gupta, Shailja, et al.
Pubblicazione: (2024)
di: Gupta, Shailja, et al.
Pubblicazione: (2024)
Enhancing Mathematics Learning for Hard-of-Hearing Students Through Real-Time Palestinian Sign Language Recognition: A New Dataset
di: Khandaqji, Fidaa, et al.
Pubblicazione: (2025)
di: Khandaqji, Fidaa, et al.
Pubblicazione: (2025)
Bidirectional Human-AI Alignment in Education for Trustworthy Learning Environments
di: Shen, Hua
Pubblicazione: (2025)
di: Shen, Hua
Pubblicazione: (2025)
Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature Review
di: Pang, Rock Yuren, et al.
Pubblicazione: (2025)
di: Pang, Rock Yuren, et al.
Pubblicazione: (2025)
Causal Stories from Sensor Traces: Auditing Epistemic Overreach in LLM-Generated Personal Sensing Explanations
di: Zhu, Shanshan, et al.
Pubblicazione: (2026)
di: Zhu, Shanshan, et al.
Pubblicazione: (2026)
MentalChat16K: A Benchmark Dataset for Conversational Mental Health Assistance
di: Xu, Jia, et al.
Pubblicazione: (2025)
di: Xu, Jia, et al.
Pubblicazione: (2025)
From tools to thieves: Measuring and understanding public perceptions of AI through crowdsourced metaphors
di: Cheng, Myra, et al.
Pubblicazione: (2025)
di: Cheng, Myra, et al.
Pubblicazione: (2025)
Benchmarking LLM Tool-Use in the Wild
di: Yu, Peijie, et al.
Pubblicazione: (2026)
di: Yu, Peijie, et al.
Pubblicazione: (2026)
Exploring the Human-LLM Synergy in Advancing Theory-driven Qualitative Analysis
di: Meng, Han, et al.
Pubblicazione: (2024)
di: Meng, Han, et al.
Pubblicazione: (2024)
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
di: Si, Chenglei, et al.
Pubblicazione: (2025)
di: Si, Chenglei, et al.
Pubblicazione: (2025)
STAR: SocioTechnical Approach to Red Teaming Language Models
di: Weidinger, Laura, et al.
Pubblicazione: (2024)
di: Weidinger, Laura, et al.
Pubblicazione: (2024)
Epistemic Alignment: A Mediating Framework for User-LLM Knowledge Delivery
di: Clark, Nicholas, et al.
Pubblicazione: (2025)
di: Clark, Nicholas, et al.
Pubblicazione: (2025)
When LLMs Can't Help: Real-World Evaluation of LLMs in Nutrition
di: Li, Karen Jia-Hui, et al.
Pubblicazione: (2025)
di: Li, Karen Jia-Hui, et al.
Pubblicazione: (2025)
Documenting Deployment with Fabric: A Repository of Real-World AI Governance
di: Jorgensen, Mackenzie, et al.
Pubblicazione: (2025)
di: Jorgensen, Mackenzie, et al.
Pubblicazione: (2025)
Case Study of GAI for Generating Novel Images for Real-World Embroidery
di: Glazko, Kate, et al.
Pubblicazione: (2025)
di: Glazko, Kate, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Who's in Charge? Disempowerment Patterns in Real-World LLM Usage
di: Sharma, Mrinank, et al.
Pubblicazione: (2026) -
Towards Apples to Apples for AI Evaluations: From Real-World Use Cases to Evaluation Scenarios
di: Choong, Yee-Yin, et al.
Pubblicazione: (2026) -
Beyond One-Way Influence: Bidirectional Opinion Dynamics in Multi-Turn Human-LLM Interactions
di: Jiang, Yuyang, et al.
Pubblicazione: (2025) -
Multi-Agent Comedy Club: Investigating Community Discussion Effects on LLM Humor Generation
di: Hong, Shiwei, et al.
Pubblicazione: (2026) -
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
di: Ghosh, Himel, et al.
Pubblicazione: (2026)