Compliance Rating Scheme: A Data Provenance Framework for Generative AI Datasets
Fuente:
arXiv
Saved in:
| Main Authors: | Bohacek, Matyas, Echavarri, Ignacio Vilanova |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dataset of News Articles with Provenance Metadata for Media Relevance Assessment
by: Peterka, Tomas, et al.
Published: (2025)
by: Peterka, Tomas, et al.
Published: (2025)
GenAI Confessions: Black-box Membership Inference for Generative Image Models
by: Bohacek, Matyas, et al.
Published: (2025)
by: Bohacek, Matyas, et al.
Published: (2025)
DAOnt: A Formal Ontology for EU Data Act Compliance
by: Leyva-Sánchez, Sheyla, et al.
Published: (2026)
by: Leyva-Sánchez, Sheyla, et al.
Published: (2026)
Towards a Framework for Supporting the Ethical and Regulatory Certification of AI Systems
by: Kovac, Fabian, et al.
Published: (2025)
by: Kovac, Fabian, et al.
Published: (2025)
Nepotistically Trained Generative-AI Models Collapse
by: Bohacek, Matyas, et al.
Published: (2023)
by: Bohacek, Matyas, et al.
Published: (2023)
AI Evaluation Should Require Standardized Item-Level Data Releases
by: Jiang, Han, et al.
Published: (2026)
by: Jiang, Han, et al.
Published: (2026)
Large Language Models and Provenance Metadata for Determining the Relevance of Images and Videos in News Stories
by: Peterka, Tomas, et al.
Published: (2025)
by: Peterka, Tomas, et al.
Published: (2025)
Data Stewardship Decoded: Mapping Its Diverse Manifestations and Emerging Relevance at a time of AI
by: Verhulst, Stefaan
Published: (2025)
by: Verhulst, Stefaan
Published: (2025)
Synthetic Human Action Video Data Generation with Pose Transfer
by: Knapp, Vaclav, et al.
Published: (2025)
by: Knapp, Vaclav, et al.
Published: (2025)
ONION: A Multi-Layered Framework for Participatory ER Design
by: Makovska, Viktoriia, et al.
Published: (2025)
by: Makovska, Viktoriia, et al.
Published: (2025)
CREDAL: Close Reading of Data Models
by: Fletcher, George, et al.
Published: (2025)
by: Fletcher, George, et al.
Published: (2025)
Scorecards for Synthetic Medical Data Evaluation and Reporting
by: Zamzmi, Ghada, et al.
Published: (2024)
by: Zamzmi, Ghada, et al.
Published: (2024)
Human Action CLIPs: Detecting AI-generated Human Motion
by: Bohacek, Matyas, et al.
Published: (2024)
by: Bohacek, Matyas, et al.
Published: (2024)
A Standardized Machine-readable Dataset Documentation Format for Responsible AI
by: Jain, Nitisha, et al.
Published: (2024)
by: Jain, Nitisha, et al.
Published: (2024)
Large Language Models and Synthetic Data for Monitoring Dataset Mentions in Research Papers
by: Solatorio, Aivin V., et al.
Published: (2025)
by: Solatorio, Aivin V., et al.
Published: (2025)
Parajudica: An RDF-Based Reasoner and Metamodel for Multi-Framework Context-Dependent Data Compliance Assessments
by: Moreau, Luc, et al.
Published: (2025)
by: Moreau, Luc, et al.
Published: (2025)
Can Pose Transfer Models Generate Realistic Human Motion?
by: Knapp, Vaclav, et al.
Published: (2025)
by: Knapp, Vaclav, et al.
Published: (2025)
The Impact of AI-Generated Text on the Internet
by: Dolezal, Jonas, et al.
Published: (2026)
by: Dolezal, Jonas, et al.
Published: (2026)
Harnessing Structured Knowledge: A Concept Map-Based Approach for High-Quality Multiple Choice Question Generation with Effective Distractors
by: Scaria, Nicy, et al.
Published: (2025)
by: Scaria, Nicy, et al.
Published: (2025)
Who is Responsible When AI Fails? Mapping Causes, Entities, and Consequences of AI Privacy and Ethical Incidents
by: Hadan, Hilda, et al.
Published: (2025)
by: Hadan, Hilda, et al.
Published: (2025)
RailEstate: An Interactive System for Metro Linked Property Trends
by: Chang, Chen-Wei, et al.
Published: (2025)
by: Chang, Chen-Wei, et al.
Published: (2025)
The Climate Change Knowledge Graph: Supporting Climate Services
by: Ceriani, Miguel, et al.
Published: (2026)
by: Ceriani, Miguel, et al.
Published: (2026)
Trajectory Data Management and Mining: A Survey from Deep Learning to the LLM Era
by: Chen, Wei, et al.
Published: (2024)
by: Chen, Wei, et al.
Published: (2024)
Evaluating the Generalization Ability of Spatiotemporal Model in Urban Scenario
by: Wang, Hongjun, et al.
Published: (2024)
by: Wang, Hongjun, et al.
Published: (2024)
A Fine-Tuned BERT Classifier for Personal-Letter Titles in Late-Ming and Early-Qing Collected Works
by: Luo, Queenie
Published: (2026)
by: Luo, Queenie
Published: (2026)
Enabling the Reuse of Personal Data in Research: A Classification Model for Legal Compliance
by: Noguera, Eduard Mata i, et al.
Published: (2025)
by: Noguera, Eduard Mata i, et al.
Published: (2025)
Semiring Provenance for Lightweight Description Logics
by: Bourgaux, Camille, et al.
Published: (2023)
by: Bourgaux, Camille, et al.
Published: (2023)
Data Authenticity, Consent, & Provenance for AI are all broken: what will it take to fix them?
by: Longpre, Shayne, et al.
Published: (2024)
by: Longpre, Shayne, et al.
Published: (2024)
World-POI: Global Point-of-Interest Data Enriched from Foursquare and OpenStreetMap as Tabular and Graph Data
by: Amiri, Hossein, et al.
Published: (2025)
by: Amiri, Hossein, et al.
Published: (2025)
From Reactive to Proactive: A Multi-Regulatory Empirical Analysis of 480 AI Incidents and a Data-Driven Governance Compliance Framework
by: Mumtaz, Ummara, et al.
Published: (2026)
by: Mumtaz, Ummara, et al.
Published: (2026)
A semantic approach to mapping the Provenance Ontology to Basic Formal Ontology
by: Prudhomme, Tim, et al.
Published: (2024)
by: Prudhomme, Tim, et al.
Published: (2024)
Towards an AI-Driven Video-Based American Sign Language Dictionary: Exploring Design and Usage Experience with Learners
by: Hassan, Saad, et al.
Published: (2025)
by: Hassan, Saad, et al.
Published: (2025)
Reviewing Uses of Regulatory Compliance Monitoring
by: Klessascheck, Finn, et al.
Published: (2024)
by: Klessascheck, Finn, et al.
Published: (2024)
Understanding Help-Seeking Behavior of Students Using LLMs vs. Web Search for Writing SQL Queries
by: Kumar, Harsh, et al.
Published: (2024)
by: Kumar, Harsh, et al.
Published: (2024)
Do Not Trust Licenses You See: Dataset Compliance Requires Massive-Scale AI-Powered Lifecycle Tracing
by: Kim, Jaekyeom, et al.
Published: (2025)
by: Kim, Jaekyeom, et al.
Published: (2025)
FlexiDataGen: An Adaptive LLM Framework for Dynamic Semantic Dataset Generation in Sensitive Domains
by: Jelodar, Hamed, et al.
Published: (2025)
by: Jelodar, Hamed, et al.
Published: (2025)
Generating Skyline Datasets for Data Science Models
by: Wang, Mengying, et al.
Published: (2025)
by: Wang, Mengying, et al.
Published: (2025)
Towards Contextual Sensitive Data Detection
by: Telkamp, Liang, et al.
Published: (2025)
by: Telkamp, Liang, et al.
Published: (2025)
Authenticity Debt and the Synthetic Content Threat Landscape: A Layered Framework for Trust, Provenance, and IP Governance in the Generative AI Era
by: Sengupta, Shubhashis, et al.
Published: (2026)
by: Sengupta, Shubhashis, et al.
Published: (2026)
Integrating AI's Carbon Footprint into Risk Management Frameworks: Strategies and Tools for Sustainable Compliance in Banking Sector
by: Tkachenko, Nataliya
Published: (2024)
by: Tkachenko, Nataliya
Published: (2024)
Similar Items
-
Dataset of News Articles with Provenance Metadata for Media Relevance Assessment
by: Peterka, Tomas, et al.
Published: (2025) -
GenAI Confessions: Black-box Membership Inference for Generative Image Models
by: Bohacek, Matyas, et al.
Published: (2025) -
DAOnt: A Formal Ontology for EU Data Act Compliance
by: Leyva-Sánchez, Sheyla, et al.
Published: (2026) -
Towards a Framework for Supporting the Ethical and Regulatory Certification of AI Systems
by: Kovac, Fabian, et al.
Published: (2025) -
Nepotistically Trained Generative-AI Models Collapse
by: Bohacek, Matyas, et al.
Published: (2023)