The State of Data Curation at NeurIPS: An Assessment of Dataset Development Practices in the Datasets and Benchmarks Track
Fuente:
arXiv
Saved in:
| Main Authors: | Bhardwaj, Eshta, Gujral, Harshit, Wu, Siyi, Zogheib, Ciara, Maharaj, Tegan, Becker, Christoph |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Machine Learning Data Practices through a Data Curation Lens: An Evaluation Framework
by: Bhardwaj, Eshta, et al.
Published: (2024)
by: Bhardwaj, Eshta, et al.
Published: (2024)
Evaluating Structured Documentation as a Tool for Reflexivity in Dataset Development
by: Bhardwaj, Eshta, et al.
Published: (2026)
by: Bhardwaj, Eshta, et al.
Published: (2026)
A Systematic Review of NeurIPS Dataset Management Practices
by: Wu, Yiwei, et al.
Published: (2024)
by: Wu, Yiwei, et al.
Published: (2024)
Exploring the Viability of the Updated World3 Model for Examining the Impact of Computing on Planetary Boundaries
by: Guliyeva, Nara, et al.
Published: (2025)
by: Guliyeva, Nara, et al.
Published: (2025)
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
by: Vishwarupe, Varad, et al.
Published: (2026)
by: Vishwarupe, Varad, et al.
Published: (2026)
Limits to AI Growth: The Ecological and Social Consequences of Scaling
by: Bhardwaj, Eshta, et al.
Published: (2025)
by: Bhardwaj, Eshta, et al.
Published: (2025)
Results of the Big ANN: NeurIPS'23 competition
by: Simhadri, Harsha Vardhan, et al.
Published: (2024)
by: Simhadri, Harsha Vardhan, et al.
Published: (2024)
The NeurIPS 2023 Machine Learning for Audio Workshop: Affective Audio Benchmarks and Novel Data
by: Baird, Alice, et al.
Published: (2024)
by: Baird, Alice, et al.
Published: (2024)
NeurIPS 2023 LLM Efficiency Fine-tuning Competition
by: Saroufim, Mark, et al.
Published: (2025)
by: Saroufim, Mark, et al.
Published: (2025)
RMIT-ADM+S at the MMU-RAG NeurIPS 2025 Competition
by: Ran, Kun, et al.
Published: (2026)
by: Ran, Kun, et al.
Published: (2026)
NeurIPS should lead scientific consensus on AI policy
by: Bommasani, Rishi
Published: (2025)
by: Bommasani, Rishi
Published: (2025)
The nextAI Solution to the NeurIPS 2023 LLM Efficiency Challenge
by: Park, Gyuwon, et al.
Published: (2026)
by: Park, Gyuwon, et al.
Published: (2026)
Aditya-Ranjan1234/mindgames_NeurIPS2025: Revac_8 Initial Release – NeurIPS 2025 Winning Agent
by: mihiraryaa, et al.
Published: (2026)
by: mihiraryaa, et al.
Published: (2026)
NeurIPS: Neuro-anatomical Inductive Priors for Sphere-based Brain Decoding
by: Yu, Sijin, et al.
Published: (2026)
by: Yu, Sijin, et al.
Published: (2026)
NeurIPS 2023 Competition: Privacy Preserving Federated Learning Document VQA
by: Tobaben, Marlon, et al.
Published: (2024)
by: Tobaben, Marlon, et al.
Published: (2024)
NeurIPS 2024 ML4CFD Competition: Results and Retrospective Analysis
by: Yagoubi, Mouadh, et al.
Published: (2025)
by: Yagoubi, Mouadh, et al.
Published: (2025)
First-Place Solution to NeurIPS 2024 Invisible Watermark Removal Challenge
by: Shamshad, Fahad, et al.
Published: (2025)
by: Shamshad, Fahad, et al.
Published: (2025)
Code for APE: A Data-Centric Benchmark for Efficient LLM Adaptation in Text Summarization (NeurIPS 2025 Submission)
by: Anonymous Author(s)
Published: (2025)
by: Anonymous Author(s)
Published: (2025)
Results of the NeurIPS 2023 Neural MMO Competition on Multi-task Reinforcement Learning
by: Suárez, Joseph, et al.
Published: (2025)
by: Suárez, Joseph, et al.
Published: (2025)
TIEG-Youpu Solution for NeurIPS 2022 WikiKG90Mv2-LSC
by: Nie, Feng, et al.
Published: (2026)
by: Nie, Feng, et al.
Published: (2026)
Usefulness of LLMs as an Author Checklist Assistant for Scientific Papers: NeurIPS'24 Experiment
by: Goldberg, Alexander, et al.
Published: (2024)
by: Goldberg, Alexander, et al.
Published: (2024)
Are we making progress in unlearning? Findings from the first NeurIPS unlearning competition
by: Triantafillou, Eleni, et al.
Published: (2024)
by: Triantafillou, Eleni, et al.
Published: (2024)
NeurIPS 2024 Ariel Data Challenge: Characterisation of Exoplanetary Atmospheres Using a Data-Centric Approach
by: Blanchard, Jeremie, et al.
Published: (2025)
by: Blanchard, Jeremie, et al.
Published: (2025)
Limits at a Distance: Design Directions to Address Psychological Distance in Policy Decisions Affecting Planetary Boundaries
by: Bhardwaj, Eshta, et al.
Published: (2025)
by: Bhardwaj, Eshta, et al.
Published: (2025)
"Near Data" and "Far Data" for Urban Sustainability: How Do Community Advocates Envision Data Intermediaries?
by: Qiao, Han, et al.
Published: (2025)
by: Qiao, Han, et al.
Published: (2025)
NeurIPS 2025 E2LM Competition : Early Training Evaluation of Language Models
by: Yagoubi, Mouadh, et al.
Published: (2025)
by: Yagoubi, Mouadh, et al.
Published: (2025)
Machine Learning Research Has Outpaced Its Communication Norms and NeurIPS Should Act
by: Rangarajan, Ajay Mandyam, et al.
Published: (2026)
by: Rangarajan, Ajay Mandyam, et al.
Published: (2026)
MedAI: Evaluating TxAgent's Therapeutic Agentic Reasoning in the NeurIPS CURE-Bench Competition
by: Cofala, Tim, et al.
Published: (2025)
by: Cofala, Tim, et al.
Published: (2025)
datasets for NeurIPS 2024 paper "Learning diverse causally emergent representations from time series data"
by: McSharry, David
Published: (2025)
by: McSharry, David
Published: (2025)
Compound Deception in Elite Peer Review: A Failure Mode Taxonomy of 100 Fabricated Citations at NeurIPS 2025
by: Ansari, Samar
Published: (2026)
by: Ansari, Samar
Published: (2026)
NeurIPS 2024 ML4CFD Competition: Harnessing Machine Learning for Computational Fluid Dynamics in Airfoil Design
by: Yagoubi, Mouadh, et al.
Published: (2024)
by: Yagoubi, Mouadh, et al.
Published: (2024)
A Taxonomy of Challenges to Curating Fair Datasets
by: Zhao, Dora, et al.
Published: (2024)
by: Zhao, Dora, et al.
Published: (2024)
Beyond Predictive Algorithms in Child Welfare
by: Moon, Erina Seh-Young, et al.
Published: (2024)
by: Moon, Erina Seh-Young, et al.
Published: (2024)
Code and experiment data from the NeurIPS 2025 paper "Classical Planning with LLM-Generated Heuristics: Challenging the State of the Art with Python Code"
by: Corrêa, Augusto B., et al.
Published: (2025)
by: Corrêa, Augusto B., et al.
Published: (2025)
Intersections Between Government Data and AI Strategies: A Case Study of Technology Policies in Canada's Federal Service
by: Kaushar Mahetaji, et al.
Published: (2025)
by: Kaushar Mahetaji, et al.
Published: (2025)
Bringing the People Back In: Contesting Benchmark Machine Learning Datasets
by: Denton, Remi, et al.
Published: (2020)
by: Denton, Remi, et al.
Published: (2020)
What do model reports say about their ChemBio benchmark evaluations? Comparing recent releases to the STREAM framework
by: Reed, Tom, et al.
Published: (2025)
by: Reed, Tom, et al.
Published: (2025)
The Impact of Student Writing Assessment Literacy on Psychological Factors: An Ordinal Logistic Regression Analysis
by: Cao, Siyi, et al.
Published: (2025)
by: Cao, Siyi, et al.
Published: (2025)
REDDIX-NET: A Novel Dataset and Benchmark for Moderating Online Explicit Services
by: Sathvik, MSVPJ, et al.
Published: (2025)
by: Sathvik, MSVPJ, et al.
Published: (2025)
A Highly Granular Temporary Migration Dataset Derived From Mobile Phone Data in Senegal
by: Blanchard, Paul, et al.
Published: (2024)
by: Blanchard, Paul, et al.
Published: (2024)
Similar Items
-
Machine Learning Data Practices through a Data Curation Lens: An Evaluation Framework
by: Bhardwaj, Eshta, et al.
Published: (2024) -
Evaluating Structured Documentation as a Tool for Reflexivity in Dataset Development
by: Bhardwaj, Eshta, et al.
Published: (2026) -
A Systematic Review of NeurIPS Dataset Management Practices
by: Wu, Yiwei, et al.
Published: (2024) -
Exploring the Viability of the Updated World3 Model for Examining the Impact of Computing on Planetary Boundaries
by: Guliyeva, Nara, et al.
Published: (2025) -
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
by: Vishwarupe, Varad, et al.
Published: (2026)