Wikibench: Community-Driven Data Curation for AI Evaluation on Wikipedia
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kuo, Tzu-Sheng, Halfaker, Aaron, Cheng, Zirui, Kim, Jiwoo, Wu, Meng-Hsin, Wu, Tongshuang, Holstein, Kenneth, Zhu, Haiyi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Botender: Supporting Communities in Collaboratively Designing AI Agents through Case-Based Provocations
von: Kuo, Tzu-Sheng, et al.
Veröffentlicht: (2025)
von: Kuo, Tzu-Sheng, et al.
Veröffentlicht: (2025)
PolicyCraft: Supporting Collaborative and Participatory Policy Design through Case-Grounded Deliberation
von: Kuo, Tzu-Sheng, et al.
Veröffentlicht: (2024)
von: Kuo, Tzu-Sheng, et al.
Veröffentlicht: (2024)
The Situate AI Guidebook: Co-Designing a Toolkit to Support Multi-Stakeholder Early-stage Deliberations Around Public Sector AI Proposals
von: Kawakami, Anna, et al.
Veröffentlicht: (2024)
von: Kawakami, Anna, et al.
Veröffentlicht: (2024)
Not Everyone Wins with LLMs: Behavioral Patterns and Pedagogical Implications for AI Literacy in Programmatic Data Science
von: Ma, Qianou, et al.
Veröffentlicht: (2025)
von: Ma, Qianou, et al.
Veröffentlicht: (2025)
Studying Up Public Sector AI: How Networks of Power Relations Shape Agency Decisions Around AI Design and Use
von: Kawakami, Anna, et al.
Veröffentlicht: (2024)
von: Kawakami, Anna, et al.
Veröffentlicht: (2024)
Effects of algorithmic flagging on fairness: quasi-experimental evidence from Wikipedia
von: TeBlunthuis, Nathan, et al.
Veröffentlicht: (2020)
von: TeBlunthuis, Nathan, et al.
Veröffentlicht: (2020)
How to Teach Programming in the AI Era? Using LLMs as a Teachable Agent for Debugging
von: Ma, Qianou, et al.
Veröffentlicht: (2023)
von: Ma, Qianou, et al.
Veröffentlicht: (2023)
Prototyping Multimodal GenAI Real-Time Agents with Counterfactual Replays and Hybrid Wizard-of-Oz
von: Gmeiner, Frederic, et al.
Veröffentlicht: (2025)
von: Gmeiner, Frederic, et al.
Veröffentlicht: (2025)
Understanding, Challenging, and Demystifying Perceptions of Gig Worker Vulnerabilities
von: de Jong, Sander, et al.
Veröffentlicht: (2025)
von: de Jong, Sander, et al.
Veröffentlicht: (2025)
PolicyPad: Collaborative Prototyping of LLM Policies
von: Feng, K. J. Kevin, et al.
Veröffentlicht: (2025)
von: Feng, K. J. Kevin, et al.
Veröffentlicht: (2025)
What Should We Engineer in Prompts? Training Humans in Requirement-Driven LLM Use
von: Ma, Qianou, et al.
Veröffentlicht: (2024)
von: Ma, Qianou, et al.
Veröffentlicht: (2024)
Improving Automated Feedback Systems for Tutor Training in Low-Resource Scenarios through Data Augmentation
von: Xu, Chentianye, et al.
Veröffentlicht: (2025)
von: Xu, Chentianye, et al.
Veröffentlicht: (2025)
Don't be Fooled: The Misinformation Effect of Explanations in Human-AI Collaboration
von: Spitzer, Philipp, et al.
Veröffentlicht: (2024)
von: Spitzer, Philipp, et al.
Veröffentlicht: (2024)
Gig2Gether: Data-sharing to Empower, Unify and Demystify Gig Work
von: Hsieh, Jane, et al.
Veröffentlicht: (2025)
von: Hsieh, Jane, et al.
Veröffentlicht: (2025)
Making the Right Thing: Bridging HCI and Responsible AI in Early-Stage AI Concept Selection
von: Jung, Ji-Youn, et al.
Veröffentlicht: (2025)
von: Jung, Ji-Youn, et al.
Veröffentlicht: (2025)
Summaries, Highlights, and Action items: Design, implementation and evaluation of an LLM-powered meeting recap system
von: Asthana, Sumit, et al.
Veröffentlicht: (2023)
von: Asthana, Sumit, et al.
Veröffentlicht: (2023)
From Prompts to Reflection: Designing Reflective Play for GenAI Literacy
von: Ma, Qianou, et al.
Veröffentlicht: (2025)
von: Ma, Qianou, et al.
Veröffentlicht: (2025)
Supporting AI-Augmented Meta-Decision Making with InDecision
von: Castañeda, Chance, et al.
Veröffentlicht: (2025)
von: Castañeda, Chance, et al.
Veröffentlicht: (2025)
Simulating Couple Conflict: Designing A Multi-Agent System for Therapy Training and Practice
von: Wang, Canwen, et al.
Veröffentlicht: (2026)
von: Wang, Canwen, et al.
Veröffentlicht: (2026)
Orbit: A Framework for Designing and Evaluating Multi-objective Rankers
von: Yang, Chenyang, et al.
Veröffentlicht: (2024)
von: Yang, Chenyang, et al.
Veröffentlicht: (2024)
SPHERE: An Evaluation Card for Human-AI Systems
von: Ma, Qianou, et al.
Veröffentlicht: (2025)
von: Ma, Qianou, et al.
Veröffentlicht: (2025)
WeAudit: Scaffolding User Auditors and AI Practitioners in Auditing Generative AI
von: Deng, Wesley Hanwen, et al.
Veröffentlicht: (2025)
von: Deng, Wesley Hanwen, et al.
Veröffentlicht: (2025)
AI Mismatches: Identifying Potential Algorithmic Harms Before AI Development
von: Saxena, Devansh, et al.
Veröffentlicht: (2025)
von: Saxena, Devansh, et al.
Veröffentlicht: (2025)
Evidotes: Integrating Scientific Evidence and Anecdotes to Support Uncertainties Triggered by Peer Health Posts
von: Bali, Shreya, et al.
Veröffentlicht: (2026)
von: Bali, Shreya, et al.
Veröffentlicht: (2026)
ComplLLM: Fine-tuning LLMs to Discover Complementary Signals for Decision-making
von: Guo, Ziyang, et al.
Veröffentlicht: (2026)
von: Guo, Ziyang, et al.
Veröffentlicht: (2026)
Exploring the Potential of Metacognitive Support Agents for Human-AI Co-Creation
von: Gmeiner, Frederic, et al.
Veröffentlicht: (2025)
von: Gmeiner, Frederic, et al.
Veröffentlicht: (2025)
Development of Mental Models in Human-AI Collaboration: A Conceptual Framework
von: Holstein, Joshua, et al.
Veröffentlicht: (2025)
von: Holstein, Joshua, et al.
Veröffentlicht: (2025)
LLMs in Wikipedia: Investigating How LLMs Impact Participation in Knowledge Communities
von: Zhou, Moyan, et al.
Veröffentlicht: (2025)
von: Zhou, Moyan, et al.
Veröffentlicht: (2025)
Funding AI for Good: A Call for Meaningful Engagement
von: Lin, Hongjin, et al.
Veröffentlicht: (2025)
von: Lin, Hongjin, et al.
Veröffentlicht: (2025)
PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI
von: Deng, Wesley Hanwen, et al.
Veröffentlicht: (2026)
von: Deng, Wesley Hanwen, et al.
Veröffentlicht: (2026)
When Thinking Pays Off: Incentive Alignment for Human-AI Collaboration
von: Holstein, Joshua, et al.
Veröffentlicht: (2025)
von: Holstein, Joshua, et al.
Veröffentlicht: (2025)
Burst: Collaborative Curation in Connected Social Media Communities
von: Zhang, Yutong, et al.
Veröffentlicht: (2025)
von: Zhang, Yutong, et al.
Veröffentlicht: (2025)
Believing vs. Achieving -- The Disconnect between Efficacy Beliefs and Collaborative Outcomes
von: Spitzer, Philipp, et al.
Veröffentlicht: (2026)
von: Spitzer, Philipp, et al.
Veröffentlicht: (2026)
RECAP: An End-to-End Platform for Capturing, Replaying, and Analyzing AI-Assisted Programming Interactions
von: He, Keyu, et al.
Veröffentlicht: (2026)
von: He, Keyu, et al.
Veröffentlicht: (2026)
LLMs as Workers in Human-Computational Algorithms? Replicating Crowdsourcing Pipelines with LLMs
von: Wu, Tongshuang, et al.
Veröffentlicht: (2023)
von: Wu, Tongshuang, et al.
Veröffentlicht: (2023)
ORES-Inspect: A technology probe for machine learning audits on enwiki
von: Levonian, Zachary, et al.
Veröffentlicht: (2024)
von: Levonian, Zachary, et al.
Veröffentlicht: (2024)
Behavioral Indicators of Overreliance During Interaction with Conversational Language Models
von: Liu, Chang, et al.
Veröffentlicht: (2026)
von: Liu, Chang, et al.
Veröffentlicht: (2026)
JumpStarter: Human-AI Planning with Task-Structured Context Curation
von: Zhang, Xuanming, et al.
Veröffentlicht: (2024)
von: Zhang, Xuanming, et al.
Veröffentlicht: (2024)
SpreadLine: Visualizing Egocentric Dynamic Influence
von: Kuo, Yun-Hsin, et al.
Veröffentlicht: (2024)
von: Kuo, Yun-Hsin, et al.
Veröffentlicht: (2024)
"I Just Don't Want My Work Being Fed Into The AI Blender": Queer Artists on Refusing and Resisting Generative AI
von: Taylor, Jordan, et al.
Veröffentlicht: (2026)
von: Taylor, Jordan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Botender: Supporting Communities in Collaboratively Designing AI Agents through Case-Based Provocations
von: Kuo, Tzu-Sheng, et al.
Veröffentlicht: (2025) -
PolicyCraft: Supporting Collaborative and Participatory Policy Design through Case-Grounded Deliberation
von: Kuo, Tzu-Sheng, et al.
Veröffentlicht: (2024) -
The Situate AI Guidebook: Co-Designing a Toolkit to Support Multi-Stakeholder Early-stage Deliberations Around Public Sector AI Proposals
von: Kawakami, Anna, et al.
Veröffentlicht: (2024) -
Not Everyone Wins with LLMs: Behavioral Patterns and Pedagogical Implications for AI Literacy in Programmatic Data Science
von: Ma, Qianou, et al.
Veröffentlicht: (2025) -
Studying Up Public Sector AI: How Networks of Power Relations Shape Agency Decisions Around AI Design and Use
von: Kawakami, Anna, et al.
Veröffentlicht: (2024)