Who Wrote the Book? Detecting and Attributing LLM Ghostwriters
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Shetty, Anudeex, Xu, Qiongkai, Ohrimenko, Olga, Lau, Jey Han |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
WET: Overcoming Paraphrasing Vulnerabilities in Embeddings-as-a-Service with Linear Transformation Watermarks
par: Shetty, Anudeex, et autres
Publié: (2024)
par: Shetty, Anudeex, et autres
Publié: (2024)
WARDEN: Multi-Directional Backdoor Watermarks for Embedding-as-a-Service Copyright Protection
par: Shetty, Anudeex, et autres
Publié: (2024)
par: Shetty, Anudeex, et autres
Publié: (2024)
Watermarks for Embeddings-as-a-Service Large Language Models
par: Shetty, Anudeex
Publié: (2025)
par: Shetty, Anudeex
Publié: (2025)
Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry
par: Li, Jiang, et autres
Publié: (2026)
par: Li, Jiang, et autres
Publié: (2026)
Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore
par: Wu, Junchao, et autres
Publié: (2024)
par: Wu, Junchao, et autres
Publié: (2024)
Evaluating Evidence Attribution in Generated Fact Checking Explanations
par: Xing, Rui, et autres
Publié: (2024)
par: Xing, Rui, et autres
Publié: (2024)
Who Wrote this Code? Watermarking for Code Generation
par: Lee, Taehyun, et autres
Publié: (2023)
par: Lee, Taehyun, et autres
Publié: (2023)
In Vino Veritas and Vulnerabilities: Examining LLM Safety via Drunk Language Inducement
par: Shetty, Anudeex, et autres
Publié: (2026)
par: Shetty, Anudeex, et autres
Publié: (2026)
Who Wrote This? Identifying Machine vs Human-Generated Text in Hausa
par: Sani, Babangida, et autres
Publié: (2025)
par: Sani, Babangida, et autres
Publié: (2025)
CMA-R:Causal Mediation Analysis for Explaining Rumour Detection
par: Tian, Lin, et autres
Publié: (2024)
par: Tian, Lin, et autres
Publié: (2024)
MoDEM: Mixture of Domain Expert Models
par: Simonds, Toby, et autres
Publié: (2024)
par: Simonds, Toby, et autres
Publié: (2024)
VITAL: A New Dataset for Benchmarking Pluralistic Alignment in Healthcare
par: Shetty, Anudeex, et autres
Publié: (2025)
par: Shetty, Anudeex, et autres
Publié: (2025)
Hidding the Ghostwriters: An Adversarial Evaluation of AI-Generated Student Essay Detection
par: Peng, Xinlin, et autres
Publié: (2024)
par: Peng, Xinlin, et autres
Publié: (2024)
Beyond Seen Data: Improving KBQA Generalization Through Schema-Guided Logical Form Generation
par: Gao, Shengxiang, et autres
Publié: (2025)
par: Gao, Shengxiang, et autres
Publié: (2025)
A Sentiment Consolidation Framework for Meta-Review Generation
par: Li, Miao, et autres
Publié: (2024)
par: Li, Miao, et autres
Publié: (2024)
Interaction Matters: An Evaluation Framework for Interactive Dialogue Assessment on English Second Language Conversations
par: Gao, Rena, et autres
Publié: (2024)
par: Gao, Rena, et autres
Publié: (2024)
On the Interplay between Human Label Variation and Model Fairness
par: Kurniawan, Kemal, et autres
Publié: (2025)
par: Kurniawan, Kemal, et autres
Publié: (2025)
To Aggregate or Not to Aggregate. That is the Question: A Case Study on Annotation Subjectivity in Span Prediction
par: Kurniawan, Kemal, et autres
Publié: (2024)
par: Kurniawan, Kemal, et autres
Publié: (2024)
PatentScore: Multi-dimensional Evaluation of LLM-Generated Patent Claims
par: Yoo, Yongmin, et autres
Publié: (2025)
par: Yoo, Yongmin, et autres
Publié: (2025)
Context Volume Drives Performance: Tackling Domain Shift in Extremely Low-Resource Translation via RAG
par: Setiawan, David Samuel, et autres
Publié: (2026)
par: Setiawan, David Samuel, et autres
Publié: (2026)
Pluralistic Alignment for Healthcare: A Role-Driven Framework
par: Zhong, Jiayou, et autres
Publié: (2025)
par: Zhong, Jiayou, et autres
Publié: (2025)
WHoW: A Cross-domain Approach for Analysing Conversation Moderation
par: Chen, Ming-Bin, et autres
Publié: (2024)
par: Chen, Ming-Bin, et autres
Publié: (2024)
Moderation Matters:Measuring Conversational Moderation Impact in English as a Second Language Group Discussion
par: Gao, Rena, et autres
Publié: (2025)
par: Gao, Rena, et autres
Publié: (2025)
Factual Dialogue Summarization via Learning from Large Language Models
par: Zhu, Rongxin, et autres
Publié: (2024)
par: Zhu, Rongxin, et autres
Publié: (2024)
Decomposed Opinion Summarization with Verified Aspect-Aware Modules
par: Li, Miao, et autres
Publié: (2025)
par: Li, Miao, et autres
Publié: (2025)
COMMUNITYNOTES: A Dataset for Exploring the Helpfulness of Fact-Checking Explanations
par: Xing, Rui, et autres
Publié: (2025)
par: Xing, Rui, et autres
Publié: (2025)
Training and Evaluating with Human Label Variation: An Empirical Study
par: Kurniawan, Kemal, et autres
Publié: (2025)
par: Kurniawan, Kemal, et autres
Publié: (2025)
Controlling Distributional Bias in Multi-Round LLM Generation via KL-Optimized Fine-Tuning
par: Jiang, Yanbei, et autres
Publié: (2026)
par: Jiang, Yanbei, et autres
Publié: (2026)
Overview of the 2024 ALTA Shared Task: Detect Automatic AI-Generated Sentences for Human-AI Hybrid Articles
par: Mollá, Diego, et autres
Publié: (2024)
par: Mollá, Diego, et autres
Publié: (2024)
VISPA: Pluralistic Alignment via Automatic Value Selection and Activation
par: Zheng, Shenyan, et autres
Publié: (2026)
par: Zheng, Shenyan, et autres
Publié: (2026)
Adaptive Cost-Efficient Evaluation for Reliable Patent Claim Generation
par: Yoo, Yongmin, et autres
Publié: (2026)
par: Yoo, Yongmin, et autres
Publié: (2026)
Inclusion-of-Thoughts: Mitigating Preference Instability via Purifying the Decision Space
par: Madani, Mohammad Reza Ghasemi, et autres
Publié: (2026)
par: Madani, Mohammad Reza Ghasemi, et autres
Publié: (2026)
Predicting Sentence Acceptability Judgments in Multimodal Contexts
par: Jang, Hyewon, et autres
Publié: (2026)
par: Jang, Hyewon, et autres
Publié: (2026)
CIG: Measuring Conversational Information Gain in Deliberative Dialogues with Semantic Memory Dynamics
par: Chen, Ming-Bin, et autres
Publié: (2026)
par: Chen, Ming-Bin, et autres
Publié: (2026)
Heterogeneous Dependency Graph-Guided Attentionfor Patent Representation Learning
par: Yoo, Yongmin, et autres
Publié: (2026)
par: Yoo, Yongmin, et autres
Publié: (2026)
An Interpretable and Crosslingual Method for Evaluating Second-Language Dialogues
par: Gao, Rena, et autres
Publié: (2024)
par: Gao, Rena, et autres
Publié: (2024)
WhoSaidIt: Human-LLM Collaborative Annotation for Text-Based Multilingual Speaker-Attribute Classification
par: Gao, Lingyu, et autres
Publié: (2026)
par: Gao, Lingyu, et autres
Publié: (2026)
"I Wrote, I Paused, I Rewrote" Teaching LLMs to Read Between the Lines of Student Writing
par: Zafar, Samra, et autres
Publié: (2025)
par: Zafar, Samra, et autres
Publié: (2025)
Detecting Non-Membership in LLM Training Data via Rank Correlations
par: Shetty, Pranav, et autres
Publié: (2026)
par: Shetty, Pranav, et autres
Publié: (2026)
CERT-ED: Certifiably Robust Text Classification for Edit Distance
par: Huang, Zhuoqun, et autres
Publié: (2024)
par: Huang, Zhuoqun, et autres
Publié: (2024)
Documents similaires
-
WET: Overcoming Paraphrasing Vulnerabilities in Embeddings-as-a-Service with Linear Transformation Watermarks
par: Shetty, Anudeex, et autres
Publié: (2024) -
WARDEN: Multi-Directional Backdoor Watermarks for Embedding-as-a-Service Copyright Protection
par: Shetty, Anudeex, et autres
Publié: (2024) -
Watermarks for Embeddings-as-a-Service Large Language Models
par: Shetty, Anudeex
Publié: (2025) -
Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry
par: Li, Jiang, et autres
Publié: (2026) -
Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore
par: Wu, Junchao, et autres
Publié: (2024)