Towards Avoiding the Data Mess: Industry Insights from Data Mesh Implementations
Fuente:
arXiv
Saved in:
| Main Authors: | Bode, Jan, Kühl, Niklas, Kreuzberger, Dominik, Hirschl, Sebastian, Holtmann, Carsten |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards a Problem-Oriented Domain Adaptation Framework for Machine Learning
by: Spitzer, Philipp, et al.
Published: (2025)
by: Spitzer, Philipp, et al.
Published: (2025)
Data-Centric Artificial Intelligence
by: Jakubik, Johannes, et al.
Published: (2022)
by: Jakubik, Johannes, et al.
Published: (2022)
Utilizing Data Fingerprints for Privacy-Preserving Algorithm Selection in Time Series Classification: Performance and Uncertainty Estimation on Unseen Datasets
by: Böcking, Lars, et al.
Published: (2024)
by: Böcking, Lars, et al.
Published: (2024)
Human Preferences in Large Language Model Latent Space: A Technical Analysis on the Reliability of Synthetic Data in Voting Outcome Prediction
by: Ball, Sarah, et al.
Published: (2025)
by: Ball, Sarah, et al.
Published: (2025)
Data Quality Challenges in Retrieval-Augmented Generation
by: Müller, Leopold, et al.
Published: (2025)
by: Müller, Leopold, et al.
Published: (2025)
Solver-Aided Expansion of Loops to Avoid Generate-and-Test
by: Dewally, Niklas, et al.
Published: (2025)
by: Dewally, Niklas, et al.
Published: (2025)
Integrating Causal Machine Learning into Clinical Decision Support Systems: Insights from Literature and Practice
by: Zipperling, Domenique, et al.
Published: (2026)
by: Zipperling, Domenique, et al.
Published: (2026)
Towards Human-Understandable Multi-Dimensional Concept Discovery
by: Grobrügge, Arne, et al.
Published: (2025)
by: Grobrügge, Arne, et al.
Published: (2025)
Implications of the AI Act for Non-Discrimination Law and Algorithmic Fairness
by: Deck, Luca, et al.
Published: (2024)
by: Deck, Luca, et al.
Published: (2024)
Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs
by: Zhou, Wei, et al.
Published: (2026)
by: Zhou, Wei, et al.
Published: (2026)
Normative Common Ground Replication (NormCoRe): Replication-by-Translation for Studying Norms in Multi-Agent AI
by: Deck, Luca, et al.
Published: (2026)
by: Deck, Luca, et al.
Published: (2026)
Investigating the Role of Explainability and AI Literacy in User Compliance
by: Kühl, Niklas, et al.
Published: (2024)
by: Kühl, Niklas, et al.
Published: (2024)
A Critical Survey on Fairness Benefits of Explainable AI
by: Deck, Luca, et al.
Published: (2023)
by: Deck, Luca, et al.
Published: (2023)
The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?
by: Hägele, Alexander, et al.
Published: (2026)
by: Hägele, Alexander, et al.
Published: (2026)
PromptPilot: Improving Human-AI Collaboration Through LLM-Enhanced Prompt Engineering
by: Gutheil, Niklas, et al.
Published: (2025)
by: Gutheil, Niklas, et al.
Published: (2025)
MessIRve: A Large-Scale Spanish Information Retrieval Dataset
by: Valentini, Francisco, et al.
Published: (2024)
by: Valentini, Francisco, et al.
Published: (2024)
Towards Autonomous Business Intelligence via Data-to-Insight Discovery Agent
by: Wu, Dongming, et al.
Published: (2026)
by: Wu, Dongming, et al.
Published: (2026)
CollaFuse: Navigating Limited Resources and Privacy in Collaborative Generative AI
by: Zipperling, Domenique, et al.
Published: (2024)
by: Zipperling, Domenique, et al.
Published: (2024)
A Multivocal Literature Review on Privacy and Fairness in Federated Learning
by: Balbierer, Beatrice, et al.
Published: (2024)
by: Balbierer, Beatrice, et al.
Published: (2024)
Beyond the Data Mesh Illusion: Designing Modern AI-augmented Lakehouses to Bridge the Gap Between Theory and Practice
by: Angélil, Oliver, et al.
Published: (2026)
by: Angélil, Oliver, et al.
Published: (2026)
A Multi-Level Strategy for Deepfake Content Moderation under EU Regulation
by: Förster, Max-Paul, et al.
Published: (2025)
by: Förster, Max-Paul, et al.
Published: (2025)
Protect and Extend -- Using GANs for Synthetic Data Generation of Time-Series Medical Records
by: Ashrafi, Navid, et al.
Published: (2024)
by: Ashrafi, Navid, et al.
Published: (2024)
Efficient Parking Search using Shared Fleet Data
by: Strauß, Niklas, et al.
Published: (2024)
by: Strauß, Niklas, et al.
Published: (2024)
A Survey of AI Reliance
by: Eckhardt, Sven, et al.
Published: (2024)
by: Eckhardt, Sven, et al.
Published: (2024)
Generating Reliable Synthetic Clinical Trial Data: The Role of Hyperparameter Optimization and Domain Constraints
by: Hahn, Waldemar, et al.
Published: (2025)
by: Hahn, Waldemar, et al.
Published: (2025)
Complementarity in Human-AI Collaboration: Concept, Sources, and Evidence
by: Hemmer, Patrick, et al.
Published: (2024)
by: Hemmer, Patrick, et al.
Published: (2024)
Transferring Domain Knowledge with (X)AI-Based Learning Systems
by: Spitzer, Philipp, et al.
Published: (2024)
by: Spitzer, Philipp, et al.
Published: (2024)
A Semantic Approach for Big Data Exploration in Industry 4.0
by: Berges, Idoia, et al.
Published: (2024)
by: Berges, Idoia, et al.
Published: (2024)
CollaFuse: Collaborative Diffusion Models
by: Allmendinger, Simeon, et al.
Published: (2024)
by: Allmendinger, Simeon, et al.
Published: (2024)
Quanda: An Interpretability Toolkit for Training Data Attribution Evaluation and Beyond
by: Bareeva, Dilyara, et al.
Published: (2024)
by: Bareeva, Dilyara, et al.
Published: (2024)
Unsupervised Speech Enhancement using Data-defined Priors
by: Klement, Dominik, et al.
Published: (2025)
by: Klement, Dominik, et al.
Published: (2025)
The Evolution of LLM Adoption in Industry Data Curation Practices
by: Qian, Crystal, et al.
Published: (2024)
by: Qian, Crystal, et al.
Published: (2024)
From Classification to Clinical Insights: Towards Analyzing and Reasoning About Mobile and Behavioral Health Data With Large Language Models
by: Englhardt, Zachary, et al.
Published: (2023)
by: Englhardt, Zachary, et al.
Published: (2023)
Insights into the Unknown: Federated Data Diversity Analysis on Molecular Data
by: Bujotzek, Markus, et al.
Published: (2025)
by: Bujotzek, Markus, et al.
Published: (2025)
Enabling Secure and Ephemeral AI Workloads in Data Mesh Environments
by: Patel, Chinkit, et al.
Published: (2025)
by: Patel, Chinkit, et al.
Published: (2025)
Better Neural PDE Solvers Through Data-Free Mesh Movers
by: Hu, Peiyan, et al.
Published: (2023)
by: Hu, Peiyan, et al.
Published: (2023)
Towards Flexible Spectrum Access: Data-Driven Insights into Spectrum Demand
by: Alkadamani, Mohamad, et al.
Published: (2026)
by: Alkadamani, Mohamad, et al.
Published: (2026)
What Matters in Data Curation for Multimodal Reasoning? Insights from the DCVLR Challenge
by: Shin, Yosub, et al.
Published: (2026)
by: Shin, Yosub, et al.
Published: (2026)
MeshFleet: Filtered and Annotated 3D Vehicle Dataset for Domain Specific Generative Modeling
by: Boborzi, Damian, et al.
Published: (2025)
by: Boborzi, Damian, et al.
Published: (2025)
InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Data Agents
by: Zhu, Zhenghao, et al.
Published: (2025)
by: Zhu, Zhenghao, et al.
Published: (2025)
Similar Items
-
Towards a Problem-Oriented Domain Adaptation Framework for Machine Learning
by: Spitzer, Philipp, et al.
Published: (2025) -
Data-Centric Artificial Intelligence
by: Jakubik, Johannes, et al.
Published: (2022) -
Utilizing Data Fingerprints for Privacy-Preserving Algorithm Selection in Time Series Classification: Performance and Uncertainty Estimation on Unseen Datasets
by: Böcking, Lars, et al.
Published: (2024) -
Human Preferences in Large Language Model Latent Space: A Technical Analysis on the Reliability of Synthetic Data in Voting Outcome Prediction
by: Ball, Sarah, et al.
Published: (2025) -
Data Quality Challenges in Retrieval-Augmented Generation
by: Müller, Leopold, et al.
Published: (2025)