SHIELD: Evaluation and Defense Strategies for Copyright Compliance in LLM Text Generation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Liu, Xiaoze, Sun, Ting, Xu, Tianyang, Wu, Feijie, Wang, Cunxiang, Wang, Xiaoqian, Gao, Jing |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
SUV: Scalable Large Language Model Copyright Compliance with Regularized Selective Unlearning
par: Xu, Tianyang, et autres
Publié: (2025)
par: Xu, Tianyang, et autres
Publié: (2025)
Evaluating the Factuality of Large Language Models using Large-Scale Knowledge Graphs
par: Liu, Xiaoze, et autres
Publié: (2024)
par: Liu, Xiaoze, et autres
Publié: (2024)
Beyond English: Unveiling Multilingual Bias in LLM Copyright Compliance
par: Chen, Yupeng, et autres
Publié: (2025)
par: Chen, Yupeng, et autres
Publié: (2025)
Towards Federated RLHF with Aggregated Client Preference for LLMs
par: Wu, Feijie, et autres
Publié: (2024)
par: Wu, Feijie, et autres
Publié: (2024)
Evaluating LLM-Generated Legal Explanations for Regulatory Compliance in Social Media Influencer Marketing
par: Gui, Haoyang, et autres
Publié: (2025)
par: Gui, Haoyang, et autres
Publié: (2025)
The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems
par: Liu, Xiaoze, et autres
Publié: (2026)
par: Liu, Xiaoze, et autres
Publié: (2026)
Attacks, Defenses and Evaluations for LLM Conversation Safety: A Survey
par: Dong, Zhichen, et autres
Publié: (2024)
par: Dong, Zhichen, et autres
Publié: (2024)
Towards Dynamic Theory of Mind: Evaluating LLM Adaptation to Temporal Evolution of Human States
par: Xiao, Yang, et autres
Publié: (2025)
par: Xiao, Yang, et autres
Publié: (2025)
Safeguarding Decentralized Social Media: LLM Agents for Automating Community Rule Compliance
par: La Cava, Lucio, et autres
Publié: (2024)
par: La Cava, Lucio, et autres
Publié: (2024)
Bridging the Copyright Gap: Do Large Vision-Language Models Recognize and Respect Copyrighted Content?
par: Xu, Naen, et autres
Publié: (2025)
par: Xu, Naen, et autres
Publié: (2025)
Leveraging Explainable AI for LLM Text Attribution: Differentiating Human-Written and Multiple LLMs-Generated Text
par: Najjar, Ayat, et autres
Publié: (2025)
par: Najjar, Ayat, et autres
Publié: (2025)
PhantomHunter: Detecting Unseen Privately-Tuned LLM-Generated Text via Family-Aware Learning
par: Shi, Yuhui, et autres
Publié: (2025)
par: Shi, Yuhui, et autres
Publié: (2025)
When the Same Coefficients Reach Different Places: Asymmetric Realizability in Transplanting Tokenizers across Large Language Models
par: Liu, Xiaoze, et autres
Publié: (2025)
par: Liu, Xiaoze, et autres
Publié: (2025)
HPSS: Heuristic Prompting Strategy Search for LLM Evaluators
par: Wen, Bosi, et autres
Publié: (2025)
par: Wen, Bosi, et autres
Publié: (2025)
On the Detectability of LLM-Generated Text: What Exactly Is LLM-Generated Text?
par: Geng, Mingmeng, et autres
Publié: (2025)
par: Geng, Mingmeng, et autres
Publié: (2025)
SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning
par: Wang, Lichao, et autres
Publié: (2026)
par: Wang, Lichao, et autres
Publié: (2026)
Multilingual Prompting for Improving LLM Generation Diversity
par: Wang, Qihan, et autres
Publié: (2025)
par: Wang, Qihan, et autres
Publié: (2025)
Exploring Safety Alignment Evaluation of LLMs in Chinese Mental Health Dialogues via LLM-as-Judge
par: Cai, Yunna, et autres
Publié: (2025)
par: Cai, Yunna, et autres
Publié: (2025)
Agree to Disagree? A Meta-Evaluation of LLM Misgendering
par: Subramonian, Arjun, et autres
Publié: (2025)
par: Subramonian, Arjun, et autres
Publié: (2025)
Gaps or Hallucinations? Gazing into Machine-Generated Legal Analysis for Fine-grained Text Evaluations
par: Hou, Abe Bohan, et autres
Publié: (2024)
par: Hou, Abe Bohan, et autres
Publié: (2024)
Do Proactive Agents Really Need an LLM to Decide When to Wake and What to Anchor?
par: Liu, Xiaoze, et autres
Publié: (2026)
par: Liu, Xiaoze, et autres
Publié: (2026)
SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales
par: Xu, Tianyang, et autres
Publié: (2024)
par: Xu, Tianyang, et autres
Publié: (2024)
The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas
par: Wu, Ya, et autres
Publié: (2025)
par: Wu, Ya, et autres
Publié: (2025)
LLM Defenses Are Not Robust to Multi-Turn Human Jailbreaks Yet
par: Li, Nathaniel, et autres
Publié: (2024)
par: Li, Nathaniel, et autres
Publié: (2024)
RAVEL: Reasoning Agents for Validating and Evaluating LLM Text Synthesis
par: Feng, Andrew Zhuoer, et autres
Publié: (2026)
par: Feng, Andrew Zhuoer, et autres
Publié: (2026)
Translators as Invisible Teachers of AI: Copyright, Translation Memory, and the Political Economy of Linguistic Data
par: Yamada, Masaru
Publié: (2026)
par: Yamada, Masaru
Publié: (2026)
Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM
par: Zhang, Chi, et autres
Publié: (2025)
par: Zhang, Chi, et autres
Publié: (2025)
Handling Students Dropouts in an LLM-driven Interactive Online Course Using Language Models
par: Wang, Yuanchun, et autres
Publié: (2025)
par: Wang, Yuanchun, et autres
Publié: (2025)
Usable XAI: 10 Strategies Towards Exploiting Explainability in the LLM Era
par: Wu, Xuansheng, et autres
Publié: (2024)
par: Wu, Xuansheng, et autres
Publié: (2024)
LM$^2$otifs : An Explainable Framework for Machine-Generated Texts Detection
par: Zheng, Xu, et autres
Publié: (2025)
par: Zheng, Xu, et autres
Publié: (2025)
Human or LLM as Standardized Patients? A Comparative Study for Medical Education
par: Zhang, Bingquan, et autres
Publié: (2025)
par: Zhang, Bingquan, et autres
Publié: (2025)
The Better Angels of Machine Personality: How Personality Relates to LLM Safety
par: Zhang, Jie, et autres
Publié: (2024)
par: Zhang, Jie, et autres
Publié: (2024)
FedBiOT: LLM Local Fine-tuning in Federated Learning without Full Model
par: Wu, Feijie, et autres
Publié: (2024)
par: Wu, Feijie, et autres
Publié: (2024)
Explainable Ethical Assessment on Human Behaviors by Generating Conflicting Social Norms
par: Sun, Yuxi, et autres
Publié: (2025)
par: Sun, Yuxi, et autres
Publié: (2025)
Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement
par: Kersting, Nicholas S., et autres
Publié: (2026)
par: Kersting, Nicholas S., et autres
Publié: (2026)
Accept or Deny? Evaluating LLM Fairness and Performance in Loan Approval across Table-to-Text Serialization Approaches
par: Azime, Israel Abebe, et autres
Publié: (2025)
par: Azime, Israel Abebe, et autres
Publié: (2025)
Research on Violent Text Detection System Based on BERT-fasttext Model
par: Yang, Yongsheng, et autres
Publié: (2024)
par: Yang, Yongsheng, et autres
Publié: (2024)
Evaluation of LLM Vulnerabilities to Being Misused for Personalized Disinformation Generation
par: Zugecova, Aneta, et autres
Publié: (2024)
par: Zugecova, Aneta, et autres
Publié: (2024)
From Text to Multimodality: Exploring the Evolution and Impact of Large Language Models in Medical Practice
par: Niu, Qian, et autres
Publié: (2024)
par: Niu, Qian, et autres
Publié: (2024)
Alignment Whack-a-Mole : Finetuning Activates Verbatim Recall of Copyrighted Books in Large Language Models
par: Liu, Xinyue, et autres
Publié: (2026)
par: Liu, Xinyue, et autres
Publié: (2026)
Documents similaires
-
SUV: Scalable Large Language Model Copyright Compliance with Regularized Selective Unlearning
par: Xu, Tianyang, et autres
Publié: (2025) -
Evaluating the Factuality of Large Language Models using Large-Scale Knowledge Graphs
par: Liu, Xiaoze, et autres
Publié: (2024) -
Beyond English: Unveiling Multilingual Bias in LLM Copyright Compliance
par: Chen, Yupeng, et autres
Publié: (2025) -
Towards Federated RLHF with Aggregated Client Preference for LLMs
par: Wu, Feijie, et autres
Publié: (2024) -
Evaluating LLM-Generated Legal Explanations for Regulatory Compliance in Social Media Influencer Marketing
par: Gui, Haoyang, et autres
Publié: (2025)