AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
Fuente:
arXiv
Salvato in:
| Autori principali: | Lu, Ximing, Sclar, Melanie, Hallinan, Skyler, Mireshghallah, Niloofar, Liu, Jiacheng, Han, Seungju, Ettinger, Allyson, Jiang, Liwei, Chandu, Khyathi, Dziri, Nouha, Choi, Yejin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
di: Jiang, Liwei, et al.
Pubblicazione: (2024)
di: Jiang, Liwei, et al.
Pubblicazione: (2024)
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
di: Han, Seungju, et al.
Pubblicazione: (2024)
di: Han, Seungju, et al.
Pubblicazione: (2024)
The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage
di: Hallinan, Skyler, et al.
Pubblicazione: (2025)
di: Hallinan, Skyler, et al.
Pubblicazione: (2025)
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
di: Qiu, Linlu, et al.
Pubblicazione: (2023)
di: Qiu, Linlu, et al.
Pubblicazione: (2023)
A Roadmap to Pluralistic Alignment
di: Sorensen, Taylor, et al.
Pubblicazione: (2024)
di: Sorensen, Taylor, et al.
Pubblicazione: (2024)
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
di: Naseh, Ali, et al.
Pubblicazione: (2025)
di: Naseh, Ali, et al.
Pubblicazione: (2025)
WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild
di: Lin, Bill Yuchen, et al.
Pubblicazione: (2024)
di: Lin, Bill Yuchen, et al.
Pubblicazione: (2024)
Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
di: Sclar, Melanie, et al.
Pubblicazione: (2023)
di: Sclar, Melanie, et al.
Pubblicazione: (2023)
Multi-Attribute Constraint Satisfaction via Language Model Rewriting
di: Baheti, Ashutosh, et al.
Pubblicazione: (2024)
di: Baheti, Ashutosh, et al.
Pubblicazione: (2024)
StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
di: Fisher, Jillian, et al.
Pubblicazione: (2024)
di: Fisher, Jillian, et al.
Pubblicazione: (2024)
Prismatic Synthesis: Gradient-based Data Diversification Boosts Generalization in LLM Reasoning
di: Jung, Jaehun, et al.
Pubblicazione: (2025)
di: Jung, Jaehun, et al.
Pubblicazione: (2025)
RewardBench: Evaluating Reward Models for Language Modeling
di: Lambert, Nathan, et al.
Pubblicazione: (2024)
di: Lambert, Nathan, et al.
Pubblicazione: (2024)
What Makes it Ok to Set a Fire? Iterative Self-distillation of Contexts and Rationales for Disambiguating Defeasible Social and Moral Situations
di: Rao, Kavel, et al.
Pubblicazione: (2023)
di: Rao, Kavel, et al.
Pubblicazione: (2023)
Certainly Uncertain: A Benchmark and Metric for Multimodal Epistemic and Aleatoric Awareness
di: Chandu, Khyathi Raghavi, et al.
Pubblicazione: (2024)
di: Chandu, Khyathi Raghavi, et al.
Pubblicazione: (2024)
RESTOR: Knowledge Recovery in Machine Unlearning
di: Rezaei, Keivan, et al.
Pubblicazione: (2024)
di: Rezaei, Keivan, et al.
Pubblicazione: (2024)
L3GO: Language Agents with Chain-of-3D-Thoughts for Generating Unconventional Objects
di: Yamada, Yutaro, et al.
Pubblicazione: (2024)
di: Yamada, Yutaro, et al.
Pubblicazione: (2024)
The Art of Saying No: Contextual Noncompliance in Language Models
di: Brahman, Faeze, et al.
Pubblicazione: (2024)
di: Brahman, Faeze, et al.
Pubblicazione: (2024)
Position: Privacy Is Not Just Memorization!
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025)
Trust No Bot: Discovering Personal Disclosures in Human-LLM Conversations in the Wild
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2024)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2024)
Smaller Language Models are Better Black-box Machine-Generated Text Detectors
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
To Err is AI : A Case Study Informing LLM Flaw Reporting Practices
di: McGregor, Sean, et al.
Pubblicazione: (2024)
di: McGregor, Sean, et al.
Pubblicazione: (2024)
CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation
di: Chen, Tong, et al.
Pubblicazione: (2024)
di: Chen, Tong, et al.
Pubblicazione: (2024)
Spectrum Tuning: Post-Training for Distributional Coverage and In-Context Steerability
di: Sorensen, Taylor, et al.
Pubblicazione: (2025)
di: Sorensen, Taylor, et al.
Pubblicazione: (2025)
Alpaca against Vicuna: Using LLMs to Uncover Memorization of LLMs
di: Kassem, Aly M., et al.
Pubblicazione: (2024)
di: Kassem, Aly M., et al.
Pubblicazione: (2024)
Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
di: Ravichander, Abhilasha, et al.
Pubblicazione: (2025)
di: Ravichander, Abhilasha, et al.
Pubblicazione: (2025)
Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
di: Sorensen, Taylor, et al.
Pubblicazione: (2023)
di: Sorensen, Taylor, et al.
Pubblicazione: (2023)
Surfacing Semantic Orthogonality Across Model Safety Benchmarks: A Multi-Dimensional Analysis
di: Bennion, Jonathan, et al.
Pubblicazione: (2025)
di: Bennion, Jonathan, et al.
Pubblicazione: (2025)
Tailoring Self-Rationalizers with Multi-Reward Distillation
di: Ramnath, Sahana, et al.
Pubblicazione: (2023)
di: Ramnath, Sahana, et al.
Pubblicazione: (2023)
Anonymous Authentication using Attribute-based Encryption
di: Oualha, Nouha
Pubblicazione: (2025)
di: Oualha, Nouha
Pubblicazione: (2025)
Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)
di: Jiang, Liwei, et al.
Pubblicazione: (2025)
di: Jiang, Liwei, et al.
Pubblicazione: (2025)
Reinforcement Learning Improves Traversal of Hierarchical Knowledge in LLMs
di: Zhang, Renfei, et al.
Pubblicazione: (2025)
di: Zhang, Renfei, et al.
Pubblicazione: (2025)
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
di: Zhou, Jijie, et al.
Pubblicazione: (2025)
di: Zhou, Jijie, et al.
Pubblicazione: (2025)
HAICOSYSTEM: An Ecosystem for Sandboxing Safety Risks in Human-AI Interactions
di: Zhou, Xuhui, et al.
Pubblicazione: (2024)
di: Zhou, Xuhui, et al.
Pubblicazione: (2024)
Agent Lumos: Unified and Modular Training for Open-Source Language Agents
di: Yin, Da, et al.
Pubblicazione: (2023)
di: Yin, Da, et al.
Pubblicazione: (2023)
Selective "Selective Prediction": Reducing Unnecessary Abstention in Vision-Language Reasoning
di: Srinivasan, Tejas, et al.
Pubblicazione: (2024)
di: Srinivasan, Tejas, et al.
Pubblicazione: (2024)
CULTURE-GEN: Revealing Global Cultural Perception in Language Models through Natural Language Prompting
di: Li, Huihan, et al.
Pubblicazione: (2024)
di: Li, Huihan, et al.
Pubblicazione: (2024)
Finding Flawed Fictions: Evaluating Complex Reasoning in Language Models via Plot Hole Detection
di: Ahuja, Kabir, et al.
Pubblicazione: (2025)
di: Ahuja, Kabir, et al.
Pubblicazione: (2025)
Experimental Contexts Can Facilitate Robust Semantic Property Inference in Language Models, but Inconsistently
di: Misra, Kanishka, et al.
Pubblicazione: (2024)
di: Misra, Kanishka, et al.
Pubblicazione: (2024)
When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models
di: Li, Yanhong, et al.
Pubblicazione: (2024)
di: Li, Yanhong, et al.
Pubblicazione: (2024)
Audio-Based Linguistic Feature Extraction for Enhancing Multi-lingual and Low-Resource Text-to-Speech
di: Kim, Youngjae, et al.
Pubblicazione: (2024)
di: Kim, Youngjae, et al.
Pubblicazione: (2024)
Documenti analoghi
-
WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
di: Jiang, Liwei, et al.
Pubblicazione: (2024) -
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
di: Han, Seungju, et al.
Pubblicazione: (2024) -
The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage
di: Hallinan, Skyler, et al.
Pubblicazione: (2025) -
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
di: Qiu, Linlu, et al.
Pubblicazione: (2023) -
A Roadmap to Pluralistic Alignment
di: Sorensen, Taylor, et al.
Pubblicazione: (2024)