Do Not Trust Licenses You See: Dataset Compliance Requires Massive-Scale AI-Powered Lifecycle Tracing
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Jaekyeom, Sohn, Sungryull, Jo, Gerrard Jeongwon, Choi, Jihoon, Bae, Kyunghoon, Lee, Hwayoung, Park, Yongmin, Lee, Honglak |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AutoGuide: Automated Generation and Selection of Context-Aware Guidelines for Large Language Model Agents
di: Fu, Yao, et al.
Pubblicazione: (2024)
di: Fu, Yao, et al.
Pubblicazione: (2024)
Scaling Web Agent Training through Automatic Data Generation and Fine-grained Evaluation
di: Logeswaran, Lajanugen, et al.
Pubblicazione: (2026)
di: Logeswaran, Lajanugen, et al.
Pubblicazione: (2026)
Auto-Intent: Automated Intent Discovery and Self-Exploration for Large Language Model Web Agents
di: Kim, Jaekyeom, et al.
Pubblicazione: (2024)
di: Kim, Jaekyeom, et al.
Pubblicazione: (2024)
Scalable Video-to-Dataset Generation for Cross-Platform Mobile Agents
di: Jang, Yunseok, et al.
Pubblicazione: (2025)
di: Jang, Yunseok, et al.
Pubblicazione: (2025)
Interactive and Expressive Code-Augmented Planning with Large Language Models
di: Liu, Anthony Z., et al.
Pubblicazione: (2024)
di: Liu, Anthony Z., et al.
Pubblicazione: (2024)
Gaming the Judge: Unfaithful Chain-of-Thought Can Undermine Agent Evaluation
di: Khalifa, Muhammad, et al.
Pubblicazione: (2026)
di: Khalifa, Muhammad, et al.
Pubblicazione: (2026)
Instruction Matters: A Simple yet Effective Task Selection for Optimized Instruction Tuning of Specific Tasks
di: Lee, Changho, et al.
Pubblicazione: (2024)
di: Lee, Changho, et al.
Pubblicazione: (2024)
Significantly improving zero-shot X-ray pathology classification via fine-tuning pre-trained image-text encoders
di: Jang, Jongseong, et al.
Pubblicazione: (2022)
di: Jang, Jongseong, et al.
Pubblicazione: (2022)
Selective LoRA for Visual Tokens and Attention Heads
di: Luo, Tiange, et al.
Pubblicazione: (2025)
di: Luo, Tiange, et al.
Pubblicazione: (2025)
EXAONE Deep: Reasoning Enhanced Language Models
di: Bae, Kyunghoon, et al.
Pubblicazione: (2025)
di: Bae, Kyunghoon, et al.
Pubblicazione: (2025)
EXAONE 3.5: Series of Large Language Models for Real-world Use Cases
di: An, Soyoung, et al.
Pubblicazione: (2024)
di: An, Soyoung, et al.
Pubblicazione: (2024)
LGAI-EMBEDDING-Preview Technical Report
di: Choi, Jooyoung, et al.
Pubblicazione: (2025)
di: Choi, Jooyoung, et al.
Pubblicazione: (2025)
SafeDPO: A Simple Approach to Direct Preference Optimization with Enhanced Safety
di: Kim, Geon-Hyeong, et al.
Pubblicazione: (2025)
di: Kim, Geon-Hyeong, et al.
Pubblicazione: (2025)
Small Language Models Need Strong Verifiers to Self-Correct Reasoning
di: Zhang, Yunxiang, et al.
Pubblicazione: (2024)
di: Zhang, Yunxiang, et al.
Pubblicazione: (2024)
EXAONE 3.0 7.8B Instruction Tuned Language Model
di: An, Soyoung, et al.
Pubblicazione: (2024)
di: An, Soyoung, et al.
Pubblicazione: (2024)
EXAONE 4.0: Unified Large Language Models Integrating Non-reasoning and Reasoning Modes
di: Bae, Kyunghoon, et al.
Pubblicazione: (2025)
di: Bae, Kyunghoon, et al.
Pubblicazione: (2025)
AI Trust Reshaping Administrative Burdens: Understanding Trust-Burden Dynamics in LLM-Assisted Benefits Systems
di: Jo, Jeongwon, et al.
Pubblicazione: (2025)
di: Jo, Jeongwon, et al.
Pubblicazione: (2025)
Process Reward Models That Think
di: Khalifa, Muhammad, et al.
Pubblicazione: (2025)
di: Khalifa, Muhammad, et al.
Pubblicazione: (2025)
LicenseGPT: A Fine-tuned Foundation Model for Publicly Available Dataset License Compliance
di: Tan, Jingwen, et al.
Pubblicazione: (2024)
di: Tan, Jingwen, et al.
Pubblicazione: (2024)
Substrates (Acyl‐CoA and Diacylglycerol) Entry and Products (CoA and Triacylglycerol) Egress Pathways in DGAT1
di: Hwayoung Lee, et al.
Pubblicazione: (2025)
di: Hwayoung Lee, et al.
Pubblicazione: (2025)
"Hey, Did You See This?"
di: Nelson, Cathy Jo
Pubblicazione: (2011)
di: Nelson, Cathy Jo
Pubblicazione: (2011)
Deep Exploration of Cross-Lingual Zero-Shot Generalization in Instruction Tuning
di: Han, Janghoon, et al.
Pubblicazione: (2024)
di: Han, Janghoon, et al.
Pubblicazione: (2024)
MolMole: Molecule Mining from Scientific Literature
di: Research, LG AI, et al.
Pubblicazione: (2025)
di: Research, LG AI, et al.
Pubblicazione: (2025)
Clear Preferences Leave Traces: Reference Model-Guided Sampling for Preference Learning
di: Diwan, Nirav, et al.
Pubblicazione: (2025)
di: Diwan, Nirav, et al.
Pubblicazione: (2025)
Kinetically Engineered Lithiophilic Dual‐Metal Layers for Dendrite‐Free, High‐Energy Anode‐Less All‐Solid‐State Batteries
di: Jihoon Oh, et al.
Pubblicazione: (2025)
di: Jihoon Oh, et al.
Pubblicazione: (2025)
FinDER: Financial Dataset for Question Answering and Evaluating Retrieval-Augmented Generation
di: Choi, Chanyeol, et al.
Pubblicazione: (2025)
di: Choi, Chanyeol, et al.
Pubblicazione: (2025)
MLRC-Bench: Can Language Agents Solve Machine Learning Research Challenges?
di: Zhang, Yunxiang, et al.
Pubblicazione: (2025)
di: Zhang, Yunxiang, et al.
Pubblicazione: (2025)
Distributional Alignment as a Criterion for Designing Task Vectors in In-Context Learning
di: Kwon, Jihoon, et al.
Pubblicazione: (2026)
di: Kwon, Jihoon, et al.
Pubblicazione: (2026)
Do Pulsar Timing Datasets Favor Massive Gravity?
di: Choi, Chris, et al.
Pubblicazione: (2025)
di: Choi, Chris, et al.
Pubblicazione: (2025)
Levelling up learning in higher education: Gamification of in‐video components to supercharge student learning and achievement in video‐based learning
di: Jeongwon Lee, et al.
Pubblicazione: (2025)
di: Jeongwon Lee, et al.
Pubblicazione: (2025)
CovMatch: Cross-Covariance Guided Multimodal Dataset Distillation with Trainable Text Encoder
di: Lee, Yongmin, et al.
Pubblicazione: (2025)
di: Lee, Yongmin, et al.
Pubblicazione: (2025)
Deep Generative Design for Mass Production
di: Kim, Jihoon, et al.
Pubblicazione: (2024)
di: Kim, Jihoon, et al.
Pubblicazione: (2024)
SelMatch: Effectively Scaling Up Dataset Distillation via Selection-Based Initialization and Partial Updates by Trajectory Matching
di: Lee, Yongmin, et al.
Pubblicazione: (2024)
di: Lee, Yongmin, et al.
Pubblicazione: (2024)
All‐Solid‐State Batteries with Extremely Low N/P Ratio Operating at Low Stack Pressure
di: Jihoon Oh, et al.
Pubblicazione: (2024)
di: Jihoon Oh, et al.
Pubblicazione: (2024)
Non-convergence of the rotating stratified flows toward the quasi-geostrophic dynamics
di: Jo, Min Jun, et al.
Pubblicazione: (2022)
di: Jo, Min Jun, et al.
Pubblicazione: (2022)
Language Model Can Do Knowledge Tracing: Simple but Effective Method to Integrate Language Model and Knowledge Tracing Task
di: Lee, Unggi, et al.
Pubblicazione: (2024)
di: Lee, Unggi, et al.
Pubblicazione: (2024)
Geometric Embedding Alignment via Curvature Matching in Transfer Learning
di: Ko, Sung Moon, et al.
Pubblicazione: (2025)
di: Ko, Sung Moon, et al.
Pubblicazione: (2025)
Remember Your Trace: Memory-Guided Long-Horizon Agentic Framework for Consistent and Hierarchical Repository-Level Code Documentation
di: Bae, Suyoung, et al.
Pubblicazione: (2026)
di: Bae, Suyoung, et al.
Pubblicazione: (2026)
Licensing Open Government Data
di: Lee, Jyh-An
Pubblicazione: (2025)
di: Lee, Jyh-An
Pubblicazione: (2025)
Do You "Trust" This Visualization? An Inventory to Measure Trust in Visualizations
di: Wang, Huichen Will, et al.
Pubblicazione: (2025)
di: Wang, Huichen Will, et al.
Pubblicazione: (2025)
Documenti analoghi
-
AutoGuide: Automated Generation and Selection of Context-Aware Guidelines for Large Language Model Agents
di: Fu, Yao, et al.
Pubblicazione: (2024) -
Scaling Web Agent Training through Automatic Data Generation and Fine-grained Evaluation
di: Logeswaran, Lajanugen, et al.
Pubblicazione: (2026) -
Auto-Intent: Automated Intent Discovery and Self-Exploration for Large Language Model Web Agents
di: Kim, Jaekyeom, et al.
Pubblicazione: (2024) -
Scalable Video-to-Dataset Generation for Cross-Platform Mobile Agents
di: Jang, Yunseok, et al.
Pubblicazione: (2025) -
Interactive and Expressive Code-Augmented Planning with Large Language Models
di: Liu, Anthony Z., et al.
Pubblicazione: (2024)