"We Have No Idea How Models will Behave in Production until Production": How Engineers Operationalize Machine Learning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Shankar, Shreya, Garcia, Rolando, Hellerstein, Joseph M, Parameswaran, Aditya G |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Steering Semantic Data Processing With DocWrangler
par: Shankar, Shreya, et autres
Publié: (2025)
par: Shankar, Shreya, et autres
Publié: (2025)
RAG Without the Lag: Interactive Debugging for Retrieval-Augmented Generation Pipelines
par: Lauro, Quentin Romero, et autres
Publié: (2025)
par: Lauro, Quentin Romero, et autres
Publié: (2025)
Rethinking Dataset Discovery with DataScout
par: Lin, Rachel, et autres
Publié: (2025)
par: Lin, Rachel, et autres
Publié: (2025)
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human Preferences
par: Shankar, Shreya, et autres
Publié: (2024)
par: Shankar, Shreya, et autres
Publié: (2024)
Reactive Writers: How Co-Writing with AI Changes How We Engage with Ideas
par: Bhat, Advait, et autres
Publié: (2026)
par: Bhat, Advait, et autres
Publié: (2026)
How Far Are We? The Triumphs and Trials of Generative AI in Learning Software Engineering
par: Choudhuri, Rudrajit, et autres
Publié: (2023)
par: Choudhuri, Rudrajit, et autres
Publié: (2023)
How Do We Evaluate Experiences in Immersive Environments?
par: Li, Xiang, et autres
Publié: (2026)
par: Li, Xiang, et autres
Publié: (2026)
Adapting to LLMs: How Insiders and Outsiders Reshape Scientific Knowledge Production
par: Xu, Huimin, et autres
Publié: (2025)
par: Xu, Huimin, et autres
Publié: (2025)
From Pen to Prompt: How Creative Writers Integrate AI into their Writing Practice
par: Guo, Alicia, et autres
Publié: (2024)
par: Guo, Alicia, et autres
Publié: (2024)
"If the Machine Is As Good As Me, Then What Use Am I?" -- How the Use of ChatGPT Changes Young Professionals' Perception of Productivity and Accomplishment
par: Kobiella, Charlotte, et autres
Publié: (2024)
par: Kobiella, Charlotte, et autres
Publié: (2024)
"Having Confidence in My Confidence Intervals": How Data Users Engage with Privacy-Protected Wikipedia Data
par: Triedman, Harold, et autres
Publié: (2025)
par: Triedman, Harold, et autres
Publié: (2025)
What We Augment When We Augment Visualizations: A Design Elicitation Study of How We Visually Express Data Relationships
par: Guo, Grace, et autres
Publié: (2024)
par: Guo, Grace, et autres
Publié: (2024)
Knowledge Prompting: How Knowledge Engineers Use Large Language Models
par: Koutsiana, Elisavet, et autres
Publié: (2024)
par: Koutsiana, Elisavet, et autres
Publié: (2024)
Designing with Culture: How Social Norms Shape Trust and Preference in Health Chatbots
par: Wadhwa, Arpita, et autres
Publié: (2025)
par: Wadhwa, Arpita, et autres
Publié: (2025)
Conceptualization, Operationalization, and Measurement of Machine Companionship: A Scoping Review
par: Banks, Jaime, et autres
Publié: (2025)
par: Banks, Jaime, et autres
Publié: (2025)
Sixteen Years of Phishing User Studies: What Have We Learned?
par: Baki, Shahryar, et autres
Publié: (2021)
par: Baki, Shahryar, et autres
Publié: (2021)
Show Me How: Benefits and Challenges of Agent-Augmented Counterfactual Explanations for Non-Expert Users
par: Bhattacharya, Aditya, et autres
Publié: (2025)
par: Bhattacharya, Aditya, et autres
Publié: (2025)
"Having Lunch Now": Understanding How Users Engage with a Proactive Agent for Daily Planning and Self-Reflection
par: Abbas, Adnan, et autres
Publié: (2025)
par: Abbas, Adnan, et autres
Publié: (2025)
Whether We Care, How We Reason: The Dual Role of Anthropomorphism and Moral Foundations in Robot Abuse
par: Yang, Fan, et autres
Publié: (2026)
par: Yang, Fan, et autres
Publié: (2026)
Deconstructing Categorization in Visualization Recommendation: A Taxonomy and Comparative Study
par: Lee, Doris Jung-Lin, et autres
Publié: (2021)
par: Lee, Doris Jung-Lin, et autres
Publié: (2021)
"We do use it, but not how hearing people think": How the Deaf and Hard of Hearing Community Uses Large Language Model Tools
par: Huffman, Shuxu, et autres
Publié: (2024)
par: Huffman, Shuxu, et autres
Publié: (2024)
How is the Pilot Doing: VTOL Pilot Workload Estimation by Multimodal Machine Learning on Psycho-physiological Signals
par: Park, Jong Hoon, et autres
Publié: (2024)
par: Park, Jong Hoon, et autres
Publié: (2024)
Expedient Assistance and Consequential Misunderstanding: Envisioning an Operationalized Mutual Theory of Mind
par: Weisz, Justin D., et autres
Publié: (2024)
par: Weisz, Justin D., et autres
Publié: (2024)
How Do We Research Human-Robot Interaction in the Age of Large Language Models? A Systematic Review
par: Wang, Yufeng, et autres
Publié: (2026)
par: Wang, Yufeng, et autres
Publié: (2026)
AI as We Describe It: How Large Language Models and Their Applications in Health are Represented Across Channels of Public Discourse
par: Zhou, Jiawei, et autres
Publié: (2025)
par: Zhou, Jiawei, et autres
Publié: (2025)
Jupybara: Operationalizing a Design Space for Actionable Data Analysis and Storytelling with LLMs
par: Wang, Huichen Will, et autres
Publié: (2025)
par: Wang, Huichen Will, et autres
Publié: (2025)
Explorable Ideas: Externalizing Ideas as Explorable Environments
par: Jung, Euijun, et autres
Publié: (2025)
par: Jung, Euijun, et autres
Publié: (2025)
Teaching Machine Learning Through Cricket: A Practical Engineering Education Approach
par: Ameen, Mohd Ruhul, et autres
Publié: (2025)
par: Ameen, Mohd Ruhul, et autres
Publié: (2025)
How Problematic Writer-AI Interactions (Rather than Problematic AI) Hinder Writers' Idea Generation
par: Umarova, Khonzoda, et autres
Publié: (2025)
par: Umarova, Khonzoda, et autres
Publié: (2025)
"Cold, Calculated, and Condescending": How AI Identifies and Explains Ableism Compared to Disabled People
par: Phutane, Mahika, et autres
Publié: (2024)
par: Phutane, Mahika, et autres
Publié: (2024)
UXR Point of View on Product Feature Prioritization Prior To Multi-Million Engineering Commitments
par: Lau, Jonas, et autres
Publié: (2025)
par: Lau, Jonas, et autres
Publié: (2025)
The Bidirectional Relationship Between XAI and Regulation: Operationalizing XAI for the AI Act
par: Hummel, Anton, et autres
Publié: (2025)
par: Hummel, Anton, et autres
Publié: (2025)
Navigating the Rashomon Effect: How Personalization Can Help Adjust Interpretable Machine Learning Models to Individual Users
par: Rosenberger, Julian, et autres
Publié: (2025)
par: Rosenberger, Julian, et autres
Publié: (2025)
Scaffolding Creativity: How Divergent and Convergent LLM Personas Shape Human Machine Creative Problem-Solving
par: Rosenbaum, Alon, et autres
Publié: (2025)
par: Rosenbaum, Alon, et autres
Publié: (2025)
"Over-the-Hood" AI Inclusivity Bugs and How 3 AI Product Teams Found and Fixed Them
par: Anderson, Andrew, et autres
Publié: (2025)
par: Anderson, Andrew, et autres
Publié: (2025)
Seamful XAI: Operationalizing Seamful Design in Explainable AI
par: Ehsan, Upol, et autres
Publié: (2022)
par: Ehsan, Upol, et autres
Publié: (2022)
"It's trained by non-disabled people": Evaluating How Image Quality Affects Product Captioning with Vision-Language Models
par: Garg, Kapil, et autres
Publié: (2025)
par: Garg, Kapil, et autres
Publié: (2025)
How Problematic are Suspenseful Interactions?
par: Uhde, Alarith
Publié: (2025)
par: Uhde, Alarith
Publié: (2025)
How AI Ideas Affect the Creativity, Diversity, and Evolution of Human Ideas: Evidence From a Large, Dynamic Experiment
par: Ashkinaze, Joshua, et autres
Publié: (2024)
par: Ashkinaze, Joshua, et autres
Publié: (2024)
Learning to Live with AI: How Students Develop AI Literacy Through Naturalistic ChatGPT Interaction
par: Ammari, Tawfiq, et autres
Publié: (2026)
par: Ammari, Tawfiq, et autres
Publié: (2026)
Documents similaires
-
Steering Semantic Data Processing With DocWrangler
par: Shankar, Shreya, et autres
Publié: (2025) -
RAG Without the Lag: Interactive Debugging for Retrieval-Augmented Generation Pipelines
par: Lauro, Quentin Romero, et autres
Publié: (2025) -
Rethinking Dataset Discovery with DataScout
par: Lin, Rachel, et autres
Publié: (2025) -
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human Preferences
par: Shankar, Shreya, et autres
Publié: (2024) -
Reactive Writers: How Co-Writing with AI Changes How We Engage with Ideas
par: Bhat, Advait, et autres
Publié: (2026)