Smaller Language Models are Better Black-box Machine-Generated Text Detectors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mireshghallah, Niloofar, Mattern, Justus, Gao, Sicun, Shokri, Reza, Berg-Kirkpatrick, Taylor |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
von: Ngong, Ivoline C., et al.
Veröffentlicht: (2024)
von: Ngong, Ivoline C., et al.
Veröffentlicht: (2024)
Readability $\ne$ Learnability: Rethinking the Role of Simplicity in Training Small Language Models
von: Lee, Ivan, et al.
Veröffentlicht: (2025)
von: Lee, Ivan, et al.
Veröffentlicht: (2025)
Position: Privacy Is Not Just Memorization!
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2025)
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2025)
CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation
von: Chen, Tong, et al.
Veröffentlicht: (2024)
von: Chen, Tong, et al.
Veröffentlicht: (2024)
An Evaluation of Explanation Methods for Black-Box Detectors of Machine-Generated Text
von: Schoenegger, Loris, et al.
Veröffentlicht: (2024)
von: Schoenegger, Loris, et al.
Veröffentlicht: (2024)
Alt-Text with Context: Improving Accessibility for Images on Twitter
von: Srivatsan, Nikita, et al.
Veröffentlicht: (2023)
von: Srivatsan, Nikita, et al.
Veröffentlicht: (2023)
Translation or Recitation? Calibrating Evaluation Scores for Machine Translation of Extremely Low-Resource Languages
von: Chen, Danlu, et al.
Veröffentlicht: (2026)
von: Chen, Danlu, et al.
Veröffentlicht: (2026)
Smaller But Better: Unifying Layout Generation with Smaller Large Language Models
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
Continuous Diffusion Models Can Obey Formal Syntax
von: Kim, Jinwoo, et al.
Veröffentlicht: (2026)
von: Kim, Jinwoo, et al.
Veröffentlicht: (2026)
Generating with Confidence: Uncertainty Quantification for Black-box Large Language Models
von: Lin, Zhen, et al.
Veröffentlicht: (2023)
von: Lin, Zhen, et al.
Veröffentlicht: (2023)
The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text
von: Meeus, Matthieu, et al.
Veröffentlicht: (2025)
von: Meeus, Matthieu, et al.
Veröffentlicht: (2025)
Bob's Confetti: Phonetic Memorization Attacks in Music and Video Generation
von: Roh, Jaechul, et al.
Veröffentlicht: (2025)
von: Roh, Jaechul, et al.
Veröffentlicht: (2025)
Boundary-targeted Membership Inference Attacks on Safety Classifiers
von: Hughes, Anthony, et al.
Veröffentlicht: (2026)
von: Hughes, Anthony, et al.
Veröffentlicht: (2026)
DALD: Improving Logits-based Detector without Logits from Black-box LLMs
von: Zeng, Cong, et al.
Veröffentlicht: (2024)
von: Zeng, Cong, et al.
Veröffentlicht: (2024)
ParaPO: Aligning Language Models to Reduce Verbatim Reproduction of Pre-training Data
von: Chen, Tong, et al.
Veröffentlicht: (2025)
von: Chen, Tong, et al.
Veröffentlicht: (2025)
Causally Testing Gender Bias in LLMs: A Case Study on Occupational Bias
von: Chen, Yuen, et al.
Veröffentlicht: (2022)
von: Chen, Yuen, et al.
Veröffentlicht: (2022)
Studying the Soupability of Documents in State Space Models
von: Jafari, Yasaman, et al.
Veröffentlicht: (2025)
von: Jafari, Yasaman, et al.
Veröffentlicht: (2025)
Constrained Sampling for Language Models Should Be Easy: An MCMC Perspective
von: Gonzalez, Emmanuel Anaya, et al.
Veröffentlicht: (2025)
von: Gonzalez, Emmanuel Anaya, et al.
Veröffentlicht: (2025)
Knowledge Editing on Black-box Large Language Models
von: Song, Xiaoshuai, et al.
Veröffentlicht: (2024)
von: Song, Xiaoshuai, et al.
Veröffentlicht: (2024)
Optical Context Compression Is Just (Bad) Autoencoding
von: Lee, Ivan Yee, et al.
Veröffentlicht: (2025)
von: Lee, Ivan Yee, et al.
Veröffentlicht: (2025)
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2023)
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2023)
On Importance of Layer Pruning for Smaller BERT Models and Low Resource Languages
von: Shirke, Mayur, et al.
Veröffentlicht: (2025)
von: Shirke, Mayur, et al.
Veröffentlicht: (2025)
Constrained Adaptive Rejection Sampling
von: Parys, Paweł, et al.
Veröffentlicht: (2025)
von: Parys, Paweł, et al.
Veröffentlicht: (2025)
MIA-Tuner: Adapting Large Language Models as Pre-training Text Detector
von: Fu, Wenjie, et al.
Veröffentlicht: (2024)
von: Fu, Wenjie, et al.
Veröffentlicht: (2024)
Machine Learning from Explanations
von: Tao, Jiashu, et al.
Veröffentlicht: (2025)
von: Tao, Jiashu, et al.
Veröffentlicht: (2025)
Enhancing Generalization in Chain of Thought Reasoning for Smaller Models
von: Yin, Maxwell J., et al.
Veröffentlicht: (2025)
von: Yin, Maxwell J., et al.
Veröffentlicht: (2025)
Mitigating Paraphrase Attacks on Machine-Text Detectors via Paraphrase Inversion
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2024)
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2024)
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2023)
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2023)
Quantifying the Effect of Test Set Contamination on Generative Evaluations
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2026)
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2026)
Merlin's Whisper: Enabling Efficient Reasoning in Large Language Models via Black-box Persuasive Prompting
von: Xia, Heming, et al.
Veröffentlicht: (2025)
von: Xia, Heming, et al.
Veröffentlicht: (2025)
Automated Black-box Prompt Engineering for Personalized Text-to-Image Generation
von: He, Yutong, et al.
Veröffentlicht: (2024)
von: He, Yutong, et al.
Veröffentlicht: (2024)
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
von: Zhou, Jijie, et al.
Veröffentlicht: (2025)
von: Zhou, Jijie, et al.
Veröffentlicht: (2025)
Context-Aware Membership Inference Attacks against Pre-trained Large Language Models
von: Chang, Hongyan, et al.
Veröffentlicht: (2024)
von: Chang, Hongyan, et al.
Veröffentlicht: (2024)
Grammar-Aligned Decoding
von: Park, Kanghee, et al.
Veröffentlicht: (2024)
von: Park, Kanghee, et al.
Veröffentlicht: (2024)
Harnessing Optimization Dynamics for Curvature-Informed Model Merging
von: Mahdavinia, Pouria, et al.
Veröffentlicht: (2025)
von: Mahdavinia, Pouria, et al.
Veröffentlicht: (2025)
Hierarchical Text Classification Using Black Box Large Language Models
von: Yoshimura, Kosuke, et al.
Veröffentlicht: (2025)
von: Yoshimura, Kosuke, et al.
Veröffentlicht: (2025)
What Happens When Small Is Made Smaller? Exploring the Impact of Compression on Small Data Pretrained Language Models
von: Awobade, Busayo, et al.
Veröffentlicht: (2024)
von: Awobade, Busayo, et al.
Veröffentlicht: (2024)
Watermark Smoothing Attacks against Language Models
von: Chang, Hongyan, et al.
Veröffentlicht: (2024)
von: Chang, Hongyan, et al.
Veröffentlicht: (2024)
Discrete Prompt Tuning via Recursive Utilization of Black-box Multimodal Large Language Model for Personalized Visual Emotion Recognition
von: Takahashi, Ryo, et al.
Veröffentlicht: (2025)
von: Takahashi, Ryo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
von: Naseh, Ali, et al.
Veröffentlicht: (2025) -
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
von: Ngong, Ivoline C., et al.
Veröffentlicht: (2024) -
Readability $\ne$ Learnability: Rethinking the Role of Simplicity in Training Small Language Models
von: Lee, Ivan, et al.
Veröffentlicht: (2025) -
Position: Privacy Is Not Just Memorization!
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2025) -
CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation
von: Chen, Tong, et al.
Veröffentlicht: (2024)