CPG-EVAL: A Multi-Tiered Benchmark for Evaluating the Chinese Pedagogical Grammar Competence of Large Language Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Wang, Dong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Algorithmic Cultivation: How Social Media Feeds Shape User Language
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
AI Psychosis: Does Conversational AI Amplify Delusion-Related Language?
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
Large Language Models Can Infer Psychological Dispositions of Social Media Users
von: Peters, Heinrich, et al.
Veröffentlicht: (2023)
von: Peters, Heinrich, et al.
Veröffentlicht: (2023)
Interpersonal Theory of Suicide as a Lens to Examine Suicidal Ideation in Online Spaces
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
Exploring Bengali Religious Dialect Biases in Large Language Models with Evaluation Perspectives
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
AI Content Moderation in Therapy Conversations
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)
Leveraging Large Language Models for Collective Decision-Making
von: Papachristou, Marios, et al.
Veröffentlicht: (2023)
von: Papachristou, Marios, et al.
Veröffentlicht: (2023)
Toxicity in Online Platforms and AI Systems: A Survey of Needs, Challenges, Mitigations, and Future Directions
von: Khapre, Smita, et al.
Veröffentlicht: (2025)
von: Khapre, Smita, et al.
Veröffentlicht: (2025)
Sima AIunty: Caste Audit in LLM-Driven Matchmaking
von: Naik, Atharva, et al.
Veröffentlicht: (2026)
von: Naik, Atharva, et al.
Veröffentlicht: (2026)
Can LLMs Assist Annotators in Identifying Morality Frames? -- Case Study on Vaccination Debate on Social Media
von: Islam, Tunazzina, et al.
Veröffentlicht: (2025)
von: Islam, Tunazzina, et al.
Veröffentlicht: (2025)
Detecting Early and Implicit Suicidal Ideation via Longitudinal and Information Environment Signals on Social Media
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025)
Hidden in Plain Text: Measuring LLM Deception Quality Against Human Baselines Using Social Deduction Games
von: Kao, Christopher, et al.
Veröffentlicht: (2026)
von: Kao, Christopher, et al.
Veröffentlicht: (2026)
The Hidden Toll of Social Media News: Causal Effects on Psychosocial Wellbeing
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
von: Pal, Olivia, et al.
Veröffentlicht: (2026)
LLM Content Moderation and User Satisfaction: Evidence from Response Refusals in Chatbot Arena
von: Pasch, Stefan
Veröffentlicht: (2025)
von: Pasch, Stefan
Veröffentlicht: (2025)
Leveraging Prompt-Based Large Language Models: Predicting Pandemic Health Decisions and Outcomes Through Social Media Language
von: Ding, Xiaohan, et al.
Veröffentlicht: (2024)
von: Ding, Xiaohan, et al.
Veröffentlicht: (2024)
CogErgLLM: Exploring Large Language Model Systems Design Perspective Using Cognitive Ergonomics
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
Exposure to Content Written by Large Language Models Can Reduce Stigma Around Opioid Use Disorder in Online Communities
von: Mittal, Shravika, et al.
Veröffentlicht: (2025)
von: Mittal, Shravika, et al.
Veröffentlicht: (2025)
RubRIX: Rubric-Driven Risk Mitigation in Caregiver-AI Interactions
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
von: Goel, Drishti, et al.
Veröffentlicht: (2026)
Unequal Opportunities: Examining the Bias in Geographical Recommendations by Large Language Models
von: Dudy, Shiran, et al.
Veröffentlicht: (2025)
von: Dudy, Shiran, et al.
Veröffentlicht: (2025)
Watch Your Language: Investigating Content Moderation with Large Language Models
von: Kumar, Deepak, et al.
Veröffentlicht: (2023)
von: Kumar, Deepak, et al.
Veröffentlicht: (2023)
Inadequacies of Large Language Model Benchmarks in the Era of Generative Artificial Intelligence
von: McIntosh, Timothy R., et al.
Veröffentlicht: (2024)
von: McIntosh, Timothy R., et al.
Veröffentlicht: (2024)
Socio-Emotional Response Generation: A Human Evaluation Protocol for LLM-Based Conversational Systems
von: Vanel, Lorraine, et al.
Veröffentlicht: (2024)
von: Vanel, Lorraine, et al.
Veröffentlicht: (2024)
LLMs Among Us: Generative AI Participating in Digital Discourse
von: Radivojevic, Kristina, et al.
Veröffentlicht: (2024)
von: Radivojevic, Kristina, et al.
Veröffentlicht: (2024)
VERA-MH Concept Paper
von: Belli, Luca, et al.
Veröffentlicht: (2025)
von: Belli, Luca, et al.
Veröffentlicht: (2025)
The DSA Transparency Database: Auditing Self-reported Moderation Actions by Social Media
von: Trujillo, Amaury, et al.
Veröffentlicht: (2023)
von: Trujillo, Amaury, et al.
Veröffentlicht: (2023)
OpenClaw AI Agents as Informal Learners at Moltbook: Characterizing an Emergent Learning Community at Scale
von: Chen, Eason, et al.
Veröffentlicht: (2026)
von: Chen, Eason, et al.
Veröffentlicht: (2026)
Reranking partisan animosity in algorithmic social media feeds alters affective polarization
von: Piccardi, Tiziano, et al.
Veröffentlicht: (2024)
von: Piccardi, Tiziano, et al.
Veröffentlicht: (2024)
When AI Agents Teach Each Other: Discourse Patterns Resembling Peer Learning in the Moltbook Community
von: Chen, Eason, et al.
Veröffentlicht: (2026)
von: Chen, Eason, et al.
Veröffentlicht: (2026)
Mind the Gap! Pathways Towards Unifying AI Safety and Ethics Research
von: Roytburg, Dani, et al.
Veröffentlicht: (2025)
von: Roytburg, Dani, et al.
Veröffentlicht: (2025)
New contexts, old heuristics: How young people in India and the US trust online content in the age of generative AI
von: Xu, Rachel, et al.
Veröffentlicht: (2024)
von: Xu, Rachel, et al.
Veröffentlicht: (2024)
Evaluating the Application of Large Language Models to Generate Feedback in Programming Education
von: Jacobs, Sven, et al.
Veröffentlicht: (2024)
von: Jacobs, Sven, et al.
Veröffentlicht: (2024)
Cross-Language Evolution of Divergent Collective Memory Around the Arab Spring
von: Jones, H. Laurie, et al.
Veröffentlicht: (2024)
von: Jones, H. Laurie, et al.
Veröffentlicht: (2024)
Decoding the Mind of Large Language Models: A Quantitative Evaluation of Ideology and Biases
von: Hirose, Manari, et al.
Veröffentlicht: (2025)
von: Hirose, Manari, et al.
Veröffentlicht: (2025)
From Divergence to Consensus: Evaluating the Role of Large Language Models in Facilitating Agreement through Adaptive Strategies
von: Triantafyllopoulos, Loukas, et al.
Veröffentlicht: (2025)
von: Triantafyllopoulos, Loukas, et al.
Veröffentlicht: (2025)
Mapping the Scholarship of Dark Pattern Regulation: A Systematic Review of Concepts, Regulatory Paradigms, and Solutions from an Interdisciplinary Perspective
von: Yi, Weiwei, et al.
Veröffentlicht: (2024)
von: Yi, Weiwei, et al.
Veröffentlicht: (2024)
Towards Recommender Systems LLMs Playground (RecSysLLMsP): Exploring Polarization and Engagement in Simulated Social Networks
von: Bojic, Ljubisa, et al.
Veröffentlicht: (2025)
von: Bojic, Ljubisa, et al.
Veröffentlicht: (2025)
Linguistic Comparison of AI- and Human-Written Responses to Online Mental Health Queries
von: Saha, Koustuv, et al.
Veröffentlicht: (2025)
von: Saha, Koustuv, et al.
Veröffentlicht: (2025)
From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data
von: Kursuncu, Ugur, et al.
Veröffentlicht: (2025)
von: Kursuncu, Ugur, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Algorithmic Cultivation: How Social Media Feeds Shape User Language
von: Pal, Olivia, et al.
Veröffentlicht: (2026) -
AI Psychosis: Does Conversational AI Amplify Delusion-Related Language?
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026) -
Large Language Models Can Infer Psychological Dispositions of Social Media Users
von: Peters, Heinrich, et al.
Veröffentlicht: (2023) -
Interpersonal Theory of Suicide as a Lens to Examine Suicidal Ideation in Online Spaces
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2025) -
LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback
von: Kim, Jiwon, et al.
Veröffentlicht: (2026)