Smaller Language Models are Better Black-box Machine-Generated Text Detectors
Fuente:
arXiv
Salvato in:
| Autori principali: | Mireshghallah, Niloofar, Mattern, Justus, Gao, Sicun, Shokri, Reza, Berg-Kirkpatrick, Taylor |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
di: Naseh, Ali, et al.
Pubblicazione: (2025)
di: Naseh, Ali, et al.
Pubblicazione: (2025)
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
di: Ngong, Ivoline C., et al.
Pubblicazione: (2024)
di: Ngong, Ivoline C., et al.
Pubblicazione: (2024)
Readability $\ne$ Learnability: Rethinking the Role of Simplicity in Training Small Language Models
di: Lee, Ivan, et al.
Pubblicazione: (2025)
di: Lee, Ivan, et al.
Pubblicazione: (2025)
Position: Privacy Is Not Just Memorization!
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025)
CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation
di: Chen, Tong, et al.
Pubblicazione: (2024)
di: Chen, Tong, et al.
Pubblicazione: (2024)
An Evaluation of Explanation Methods for Black-Box Detectors of Machine-Generated Text
di: Schoenegger, Loris, et al.
Pubblicazione: (2024)
di: Schoenegger, Loris, et al.
Pubblicazione: (2024)
Alt-Text with Context: Improving Accessibility for Images on Twitter
di: Srivatsan, Nikita, et al.
Pubblicazione: (2023)
di: Srivatsan, Nikita, et al.
Pubblicazione: (2023)
Translation or Recitation? Calibrating Evaluation Scores for Machine Translation of Extremely Low-Resource Languages
di: Chen, Danlu, et al.
Pubblicazione: (2026)
di: Chen, Danlu, et al.
Pubblicazione: (2026)
Smaller But Better: Unifying Layout Generation with Smaller Large Language Models
di: Zhang, Peirong, et al.
Pubblicazione: (2025)
di: Zhang, Peirong, et al.
Pubblicazione: (2025)
Continuous Diffusion Models Can Obey Formal Syntax
di: Kim, Jinwoo, et al.
Pubblicazione: (2026)
di: Kim, Jinwoo, et al.
Pubblicazione: (2026)
Generating with Confidence: Uncertainty Quantification for Black-box Large Language Models
di: Lin, Zhen, et al.
Pubblicazione: (2023)
di: Lin, Zhen, et al.
Pubblicazione: (2023)
The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text
di: Meeus, Matthieu, et al.
Pubblicazione: (2025)
di: Meeus, Matthieu, et al.
Pubblicazione: (2025)
Bob's Confetti: Phonetic Memorization Attacks in Music and Video Generation
di: Roh, Jaechul, et al.
Pubblicazione: (2025)
di: Roh, Jaechul, et al.
Pubblicazione: (2025)
Boundary-targeted Membership Inference Attacks on Safety Classifiers
di: Hughes, Anthony, et al.
Pubblicazione: (2026)
di: Hughes, Anthony, et al.
Pubblicazione: (2026)
DALD: Improving Logits-based Detector without Logits from Black-box LLMs
di: Zeng, Cong, et al.
Pubblicazione: (2024)
di: Zeng, Cong, et al.
Pubblicazione: (2024)
ParaPO: Aligning Language Models to Reduce Verbatim Reproduction of Pre-training Data
di: Chen, Tong, et al.
Pubblicazione: (2025)
di: Chen, Tong, et al.
Pubblicazione: (2025)
Causally Testing Gender Bias in LLMs: A Case Study on Occupational Bias
di: Chen, Yuen, et al.
Pubblicazione: (2022)
di: Chen, Yuen, et al.
Pubblicazione: (2022)
Studying the Soupability of Documents in State Space Models
di: Jafari, Yasaman, et al.
Pubblicazione: (2025)
di: Jafari, Yasaman, et al.
Pubblicazione: (2025)
Constrained Sampling for Language Models Should Be Easy: An MCMC Perspective
di: Gonzalez, Emmanuel Anaya, et al.
Pubblicazione: (2025)
di: Gonzalez, Emmanuel Anaya, et al.
Pubblicazione: (2025)
Knowledge Editing on Black-box Large Language Models
di: Song, Xiaoshuai, et al.
Pubblicazione: (2024)
di: Song, Xiaoshuai, et al.
Pubblicazione: (2024)
Optical Context Compression Is Just (Bad) Autoencoding
di: Lee, Ivan Yee, et al.
Pubblicazione: (2025)
di: Lee, Ivan Yee, et al.
Pubblicazione: (2025)
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2023)
On Importance of Layer Pruning for Smaller BERT Models and Low Resource Languages
di: Shirke, Mayur, et al.
Pubblicazione: (2025)
di: Shirke, Mayur, et al.
Pubblicazione: (2025)
Constrained Adaptive Rejection Sampling
di: Parys, Paweł, et al.
Pubblicazione: (2025)
di: Parys, Paweł, et al.
Pubblicazione: (2025)
MIA-Tuner: Adapting Large Language Models as Pre-training Text Detector
di: Fu, Wenjie, et al.
Pubblicazione: (2024)
di: Fu, Wenjie, et al.
Pubblicazione: (2024)
Machine Learning from Explanations
di: Tao, Jiashu, et al.
Pubblicazione: (2025)
di: Tao, Jiashu, et al.
Pubblicazione: (2025)
Enhancing Generalization in Chain of Thought Reasoning for Smaller Models
di: Yin, Maxwell J., et al.
Pubblicazione: (2025)
di: Yin, Maxwell J., et al.
Pubblicazione: (2025)
Mitigating Paraphrase Attacks on Machine-Text Detectors via Paraphrase Inversion
di: Soto, Rafael Rivera, et al.
Pubblicazione: (2024)
di: Soto, Rafael Rivera, et al.
Pubblicazione: (2024)
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2023)
di: Bhattacharjee, Amrita, et al.
Pubblicazione: (2023)
Quantifying the Effect of Test Set Contamination on Generative Evaluations
di: Schaeffer, Rylan, et al.
Pubblicazione: (2026)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2026)
Merlin's Whisper: Enabling Efficient Reasoning in Large Language Models via Black-box Persuasive Prompting
di: Xia, Heming, et al.
Pubblicazione: (2025)
di: Xia, Heming, et al.
Pubblicazione: (2025)
Automated Black-box Prompt Engineering for Personalized Text-to-Image Generation
di: He, Yutong, et al.
Pubblicazione: (2024)
di: He, Yutong, et al.
Pubblicazione: (2024)
Operationalizing Data Minimization for Privacy-Preserving LLM Prompting
di: Zhou, Jijie, et al.
Pubblicazione: (2025)
di: Zhou, Jijie, et al.
Pubblicazione: (2025)
Context-Aware Membership Inference Attacks against Pre-trained Large Language Models
di: Chang, Hongyan, et al.
Pubblicazione: (2024)
di: Chang, Hongyan, et al.
Pubblicazione: (2024)
Grammar-Aligned Decoding
di: Park, Kanghee, et al.
Pubblicazione: (2024)
di: Park, Kanghee, et al.
Pubblicazione: (2024)
Harnessing Optimization Dynamics for Curvature-Informed Model Merging
di: Mahdavinia, Pouria, et al.
Pubblicazione: (2025)
di: Mahdavinia, Pouria, et al.
Pubblicazione: (2025)
Hierarchical Text Classification Using Black Box Large Language Models
di: Yoshimura, Kosuke, et al.
Pubblicazione: (2025)
di: Yoshimura, Kosuke, et al.
Pubblicazione: (2025)
What Happens When Small Is Made Smaller? Exploring the Impact of Compression on Small Data Pretrained Language Models
di: Awobade, Busayo, et al.
Pubblicazione: (2024)
di: Awobade, Busayo, et al.
Pubblicazione: (2024)
Watermark Smoothing Attacks against Language Models
di: Chang, Hongyan, et al.
Pubblicazione: (2024)
di: Chang, Hongyan, et al.
Pubblicazione: (2024)
Discrete Prompt Tuning via Recursive Utilization of Black-box Multimodal Large Language Model for Personalized Visual Emotion Recognition
di: Takahashi, Ryo, et al.
Pubblicazione: (2025)
di: Takahashi, Ryo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection
di: Naseh, Ali, et al.
Pubblicazione: (2025) -
Differentially Private Learning Needs Better Model Initialization and Self-Distillation
di: Ngong, Ivoline C., et al.
Pubblicazione: (2024) -
Readability $\ne$ Learnability: Rethinking the Role of Simplicity in Training Small Language Models
di: Lee, Ivan, et al.
Pubblicazione: (2025) -
Position: Privacy Is Not Just Memorization!
di: Mireshghallah, Niloofar, et al.
Pubblicazione: (2025) -
CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation
di: Chen, Tong, et al.
Pubblicazione: (2024)