Saved in:
| Main Authors: | Tan, Jingwen, Rajbahadur, Gopi Krishnan, Li, Zi, Song, Xiangfu, Lin, Jianshan, Li, Dan, Zheng, Zibin, Hassan, Ahmed E. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.00106 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Hugging Face to GitHub: Tracing License Drift in the Open-Source AI Ecosystem
by: Jewitt, James, et al.
Published: (2025)
by: Jewitt, James, et al.
Published: (2025)
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
by: Jewitt, James, et al.
Published: (2026)
by: Jewitt, James, et al.
Published: (2026)
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
by: Ebrahimi, Amir M., et al.
Published: (2026)
by: Ebrahimi, Amir M., et al.
Published: (2026)
Studying the Impact of TensorFlow and PyTorch Bindings on Machine Learning Software Quality
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Model Context Protocol (MCP) Tool Descriptions Are Smelly! Towards Improving AI Agent Efficiency with Augmented MCP Tool Descriptions
by: Hasan, Mohammed Mehedi, et al.
Published: (2026)
by: Hasan, Mohammed Mehedi, et al.
Published: (2026)
Data Quality Antipatterns for Software Analytics
by: Bhatia, Aaditya, et al.
Published: (2024)
by: Bhatia, Aaditya, et al.
Published: (2024)
From Cool Demos to Production-Ready FMware: Core Challenges and a Technology Roadmap
by: Rajbahadur, Gopi Krishnan, et al.
Published: (2024)
by: Rajbahadur, Gopi Krishnan, et al.
Published: (2024)
Model Context Protocol (MCP) at First Glance: Studying the Security and Maintainability of MCP Servers
by: Hasan, Mohammed Mehedi, et al.
Published: (2025)
by: Hasan, Mohammed Mehedi, et al.
Published: (2025)
An Empirical Study of Testing Practices in Open Source AI Agent Frameworks and Agentic Applications
by: Hasan, Mohammed Mehedi, et al.
Published: (2025)
by: Hasan, Mohammed Mehedi, et al.
Published: (2025)
Building an Open AIBOM Standard in the Wild
by: Rajbahadur, Gopi Krishnan, et al.
Published: (2025)
by: Rajbahadur, Gopi Krishnan, et al.
Published: (2025)
Implementing AI Bill of Materials (AI BOM) with SPDX 3.0: A Comprehensive Guide to Creating AI and Dataset Bill of Materials
by: Bennet, Karen, et al.
Published: (2025)
by: Bennet, Karen, et al.
Published: (2025)
Keeping Deep Learning Models in Check: A History-Based Approach to Mitigate Overfitting
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
LiCoEval: Evaluating LLMs on License Compliance in Code Generation
by: Xu, Weiwei, et al.
Published: (2024)
by: Xu, Weiwei, et al.
Published: (2024)
Hidden Licensing Risks in the LLMware Ecosystem
by: Wang, Bo, et al.
Published: (2026)
by: Wang, Bo, et al.
Published: (2026)
On the Possibility of Breaking Copyleft Licenses When Reusing Code Generated by ChatGPT
by: Colombo, Gaia, et al.
Published: (2025)
by: Colombo, Gaia, et al.
Published: (2025)
SimClone: Detecting Tabular Data Clones using Value Similarity
by: Yang, Xu, et al.
Published: (2024)
by: Yang, Xu, et al.
Published: (2024)
OSS License Identification at Scale: A Comprehensive Dataset Using World of Code
by: Jahanshahi, Mahmoud, et al.
Published: (2024)
by: Jahanshahi, Mahmoud, et al.
Published: (2024)
Rethinking Software Engineering in the Foundation Model Era: A Curated Catalogue of Challenges in the Development of Trustworthy FMware
by: Hassan, Ahmed E., et al.
Published: (2024)
by: Hassan, Ahmed E., et al.
Published: (2024)
An Exploratory Investigation into Code License Infringements in Large Language Model Training Datasets
by: Katzy, Jonathan, et al.
Published: (2024)
by: Katzy, Jonathan, et al.
Published: (2024)
Cracks in The Stack: Hidden Vulnerabilities and Licensing Risks in LLM Pre-Training Datasets
by: Jahanshahi, Mahmoud, et al.
Published: (2025)
by: Jahanshahi, Mahmoud, et al.
Published: (2025)
SPICE: An Automated SWE-Bench Labeling Pipeline for Issue Clarity, Test Coverage, and Effort Estimation
by: Oliva, Gustavo A., et al.
Published: (2025)
by: Oliva, Gustavo A., et al.
Published: (2025)
The Hitchhikers Guide to Production-ready Trustworthy Foundation Model powered Software (FMware)
by: Vasilevski, Kirill, et al.
Published: (2025)
by: Vasilevski, Kirill, et al.
Published: (2025)
Catch the Butterfly: Peeking into the Terms and Conflicts among SPDX Licenses
by: Liu, Tao, et al.
Published: (2024)
by: Liu, Tao, et al.
Published: (2024)
Developers' Perspectives on Software Licensing: Current Practices, Challenges, and Tools
by: Wintersgill, Nathan, et al.
Published: (2025)
by: Wintersgill, Nathan, et al.
Published: (2025)
Open Source at a Crossroads: The Future of Licensing Driven by Monetization
by: Kula, Raula Gaikovina, et al.
Published: (2025)
by: Kula, Raula Gaikovina, et al.
Published: (2025)
Comentarios breves sobre la GNU General Public License v3
by: Malcolm Bain
Published: (2009)
by: Malcolm Bain
Published: (2009)
Open Source, Hidden Costs: A Systematic Literature Review on OSS License Management
by: Li, Boyuan, et al.
Published: (2025)
by: Li, Boyuan, et al.
Published: (2025)
RepoForge: Training a SOTA Fast-thinking SWE Agent with an End-to-End Data Curation Pipeline Synergizing SFT and RL at Scale
by: Chen, Zhilong, et al.
Published: (2025)
by: Chen, Zhilong, et al.
Published: (2025)
On the Standardization of Behavioral Use Clauses and Their Adoption for Responsible Licensing of AI
by: McDuff, Daniel, et al.
Published: (2024)
by: McDuff, Daniel, et al.
Published: (2024)
An Empirical Analysis of Machine Learning Model and Dataset Documentation, Supply Chain, and Licensing Challenges on Hugging Face
by: Stalnaker, Trevor, et al.
Published: (2025)
by: Stalnaker, Trevor, et al.
Published: (2025)
The Case for Contextual Copyleft: Licensing Open Source Training Data and Generative AI
by: Shanklin, Grant, et al.
Published: (2025)
by: Shanklin, Grant, et al.
Published: (2025)
DevLicOps: A Framework for Mitigating Licensing Risks in AI-Generated Code
by: Sharma, Pratyush Nidhi, et al.
Published: (2025)
by: Sharma, Pratyush Nidhi, et al.
Published: (2025)
Developer Perspectives on Licensing and Copyright Issues Arising from Generative AI for Software Development
by: Stalnaker, Trevor, et al.
Published: (2024)
by: Stalnaker, Trevor, et al.
Published: (2024)
Small Changes, Big Trouble: Demystifying and Parsing License Variants for Incompatibility Detection in the PyPI Ecosystem
by: Xu, Weiwei, et al.
Published: (2025)
by: Xu, Weiwei, et al.
Published: (2025)
"The Law Doesn't Work Like a Computer": Exploring Software Licensing Issues Faced by Legal Practitioners
by: Wintersgill, Nathan, et al.
Published: (2024)
by: Wintersgill, Nathan, et al.
Published: (2024)
CodeGenLink: A Tool to Find the Likely Origin and License of Automatically Generated Code
by: Bifolco, Daniele, et al.
Published: (2025)
by: Bifolco, Daniele, et al.
Published: (2025)
Does the Order of Fine-tuning Matter and Why?
by: Chen, Qihong, et al.
Published: (2024)
by: Chen, Qihong, et al.
Published: (2024)
SmartReco: Detecting Read-Only Reentrancy via Fine-Grained Cross-DApp Analysis
by: Zhang, Jingwen, et al.
Published: (2024)
by: Zhang, Jingwen, et al.
Published: (2024)
Textual analysis of End User License Agreement for red-flagging potentially malicious software
by: Khan, Behraj, et al.
Published: (2024)
by: Khan, Behraj, et al.
Published: (2024)
Software Engineering and Foundation Models: Insights from Industry Blogs Using a Jury of Foundation Models
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Similar Items
-
From Hugging Face to GitHub: Tracing License Drift in the Open-Source AI Ecosystem
by: Jewitt, James, et al.
Published: (2025) -
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
by: Jewitt, James, et al.
Published: (2026) -
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
by: Ebrahimi, Amir M., et al.
Published: (2026) -
Studying the Impact of TensorFlow and PyTorch Bindings on Machine Learning Software Quality
by: Li, Hao, et al.
Published: (2024) -
Model Context Protocol (MCP) Tool Descriptions Are Smelly! Towards Improving AI Agent Efficiency with Augmented MCP Tool Descriptions
by: Hasan, Mohammed Mehedi, et al.
Published: (2026)