AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
Fuente:
arXiv
Guardado en:
| Autores principales: | Lu, Ximing, Sclar, Melanie, Hallinan, Skyler, Mireshghallah, Niloofar, Liu, Jiacheng, Han, Seungju, Ettinger, Allyson, Jiang, Liwei, Chandu, Khyathi, Dziri, Nouha, Choi, Yejin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
por: Jiang, Liwei, et al.
Publicado: (2024)
por: Jiang, Liwei, et al.
Publicado: (2024)
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
por: Han, Seungju, et al.
Publicado: (2024)
por: Han, Seungju, et al.
Publicado: (2024)
The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage
por: Hallinan, Skyler, et al.
Publicado: (2025)
por: Hallinan, Skyler, et al.
Publicado: (2025)
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
por: Qiu, Linlu, et al.
Publicado: (2023)
por: Qiu, Linlu, et al.
Publicado: (2023)
A Roadmap to Pluralistic Alignment
por: Sorensen, Taylor, et al.
Publicado: (2024)
por: Sorensen, Taylor, et al.
Publicado: (2024)
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
por: Naseh, Ali, et al.
Publicado: (2025)
por: Naseh, Ali, et al.
Publicado: (2025)
WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild
por: Lin, Bill Yuchen, et al.
Publicado: (2024)
por: Lin, Bill Yuchen, et al.
Publicado: (2024)
Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
por: Sclar, Melanie, et al.
Publicado: (2023)
por: Sclar, Melanie, et al.
Publicado: (2023)
Multi-Attribute Constraint Satisfaction via Language Model Rewriting
por: Baheti, Ashutosh, et al.
Publicado: (2024)
por: Baheti, Ashutosh, et al.
Publicado: (2024)
StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
por: Fisher, Jillian, et al.
Publicado: (2024)
por: Fisher, Jillian, et al.
Publicado: (2024)
Prismatic Synthesis: Gradient-based Data Diversification Boosts Generalization in LLM Reasoning
por: Jung, Jaehun, et al.
Publicado: (2025)
por: Jung, Jaehun, et al.
Publicado: (2025)
RewardBench: Evaluating Reward Models for Language Modeling
por: Lambert, Nathan, et al.
Publicado: (2024)
por: Lambert, Nathan, et al.
Publicado: (2024)
What Makes it Ok to Set a Fire? Iterative Self-distillation of Contexts and Rationales for Disambiguating Defeasible Social and Moral Situations
por: Rao, Kavel, et al.
Publicado: (2023)
por: Rao, Kavel, et al.
Publicado: (2023)
Certainly Uncertain: A Benchmark and Metric for Multimodal Epistemic and Aleatoric Awareness
por: Chandu, Khyathi Raghavi, et al.
Publicado: (2024)
por: Chandu, Khyathi Raghavi, et al.
Publicado: (2024)
RESTOR: Knowledge Recovery in Machine Unlearning
por: Rezaei, Keivan, et al.
Publicado: (2024)
por: Rezaei, Keivan, et al.
Publicado: (2024)
L3GO: Language Agents with Chain-of-3D-Thoughts for Generating Unconventional Objects
por: Yamada, Yutaro, et al.
Publicado: (2024)
por: Yamada, Yutaro, et al.
Publicado: (2024)
The Art of Saying No: Contextual Noncompliance in Language Models
por: Brahman, Faeze, et al.
Publicado: (2024)
por: Brahman, Faeze, et al.
Publicado: (2024)
Position: Privacy Is Not Just Memorization!
por: Mireshghallah, Niloofar, et al.
Publicado: (2025)
por: Mireshghallah, Niloofar, et al.
Publicado: (2025)
Trust No Bot: Discovering Personal Disclosures in Human-LLM Conversations in the Wild
por: Mireshghallah, Niloofar, et al.
Publicado: (2024)
por: Mireshghallah, Niloofar, et al.
Publicado: (2024)
Smaller Language Models are Better Black-box Machine-Generated Text Detectors
por: Mireshghallah, Niloofar, et al.
Publicado: (2023)
por: Mireshghallah, Niloofar, et al.
Publicado: (2023)
To Err is AI : A Case Study Informing LLM Flaw Reporting Practices
por: McGregor, Sean, et al.
Publicado: (2024)
por: McGregor, Sean, et al.
Publicado: (2024)
CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation
por: Chen, Tong, et al.
Publicado: (2024)
por: Chen, Tong, et al.
Publicado: (2024)
Spectrum Tuning: Post-Training for Distributional Coverage and In-Context Steerability
por: Sorensen, Taylor, et al.
Publicado: (2025)
por: Sorensen, Taylor, et al.
Publicado: (2025)
Alpaca against Vicuna: Using LLMs to Uncover Memorization of LLMs
por: Kassem, Aly M., et al.
Publicado: (2024)
por: Kassem, Aly M., et al.
Publicado: (2024)
Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
por: Ravichander, Abhilasha, et al.
Publicado: (2025)
por: Ravichander, Abhilasha, et al.
Publicado: (2025)
Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
por: Sorensen, Taylor, et al.
Publicado: (2023)
por: Sorensen, Taylor, et al.
Publicado: (2023)
Surfacing Semantic Orthogonality Across Model Safety Benchmarks: A Multi-Dimensional Analysis
por: Bennion, Jonathan, et al.
Publicado: (2025)
por: Bennion, Jonathan, et al.
Publicado: (2025)
Tailoring Self-Rationalizers with Multi-Reward Distillation
por: Ramnath, Sahana, et al.
Publicado: (2023)
por: Ramnath, Sahana, et al.
Publicado: (2023)
Anonymous Authentication using Attribute-based Encryption
por: Oualha, Nouha
Publicado: (2025)
por: Oualha, Nouha
Publicado: (2025)
Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)
por: Jiang, Liwei, et al.
Publicado: (2025)
por: Jiang, Liwei, et al.
Publicado: (2025)
Reinforcement Learning Improves Traversal of Hierarchical Knowledge in LLMs
por: Zhang, Renfei, et al.
Publicado: (2025)
por: Zhang, Renfei, et al.
Publicado: (2025)
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
por: Zhou, Jijie, et al.
Publicado: (2025)
por: Zhou, Jijie, et al.
Publicado: (2025)
HAICOSYSTEM: An Ecosystem for Sandboxing Safety Risks in Human-AI Interactions
por: Zhou, Xuhui, et al.
Publicado: (2024)
por: Zhou, Xuhui, et al.
Publicado: (2024)
Agent Lumos: Unified and Modular Training for Open-Source Language Agents
por: Yin, Da, et al.
Publicado: (2023)
por: Yin, Da, et al.
Publicado: (2023)
Selective "Selective Prediction": Reducing Unnecessary Abstention in Vision-Language Reasoning
por: Srinivasan, Tejas, et al.
Publicado: (2024)
por: Srinivasan, Tejas, et al.
Publicado: (2024)
CULTURE-GEN: Revealing Global Cultural Perception in Language Models through Natural Language Prompting
por: Li, Huihan, et al.
Publicado: (2024)
por: Li, Huihan, et al.
Publicado: (2024)
Finding Flawed Fictions: Evaluating Complex Reasoning in Language Models via Plot Hole Detection
por: Ahuja, Kabir, et al.
Publicado: (2025)
por: Ahuja, Kabir, et al.
Publicado: (2025)
Experimental Contexts Can Facilitate Robust Semantic Property Inference in Language Models, but Inconsistently
por: Misra, Kanishka, et al.
Publicado: (2024)
por: Misra, Kanishka, et al.
Publicado: (2024)
When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models
por: Li, Yanhong, et al.
Publicado: (2024)
por: Li, Yanhong, et al.
Publicado: (2024)
Audio-Based Linguistic Feature Extraction for Enhancing Multi-lingual and Low-Resource Text-to-Speech
por: Kim, Youngjae, et al.
Publicado: (2024)
por: Kim, Youngjae, et al.
Publicado: (2024)
Ejemplares similares
-
WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
por: Jiang, Liwei, et al.
Publicado: (2024) -
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
por: Han, Seungju, et al.
Publicado: (2024) -
The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage
por: Hallinan, Skyler, et al.
Publicado: (2025) -
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
por: Qiu, Linlu, et al.
Publicado: (2023) -
A Roadmap to Pluralistic Alignment
por: Sorensen, Taylor, et al.
Publicado: (2024)