JAMDEC: Unsupervised Authorship Obfuscation using Constrained Decoding over Small Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fisher, Jillian, Lu, Ximing, Jung, Jaehun, Jiang, Liwei, Harchaoui, Zaid, Choi, Yejin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
von: Fisher, Jillian, et al.
Veröffentlicht: (2024)
von: Fisher, Jillian, et al.
Veröffentlicht: (2024)
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing
von: Jung, Jaehun, et al.
Veröffentlicht: (2023)
von: Jung, Jaehun, et al.
Veröffentlicht: (2023)
Information-Theoretic Distillation for Reference-less Summarization
von: Jung, Jaehun, et al.
Veröffentlicht: (2024)
von: Jung, Jaehun, et al.
Veröffentlicht: (2024)
The Invisible Leash: Why RLVR May or May Not Escape Its Origin
von: Wu, Fang, et al.
Veröffentlicht: (2025)
von: Wu, Fang, et al.
Veröffentlicht: (2025)
Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions
von: Acuna, David, et al.
Veröffentlicht: (2025)
von: Acuna, David, et al.
Veröffentlicht: (2025)
Mitigating Self-Preference by Authorship Obfuscation
von: Mahbub, Taslim, et al.
Veröffentlicht: (2025)
von: Mahbub, Taslim, et al.
Veröffentlicht: (2025)
Spectrum Tuning: Post-Training for Distributional Coverage and In-Context Steerability
von: Sorensen, Taylor, et al.
Veröffentlicht: (2025)
von: Sorensen, Taylor, et al.
Veröffentlicht: (2025)
Verifying the Verifiers: Unveiling Pitfalls and Potentials in Fact Verifiers
von: Seo, Wooseok, et al.
Veröffentlicht: (2025)
von: Seo, Wooseok, et al.
Veröffentlicht: (2025)
A Roadmap to Pluralistic Alignment
von: Sorensen, Taylor, et al.
Veröffentlicht: (2024)
von: Sorensen, Taylor, et al.
Veröffentlicht: (2024)
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
von: Qiu, Linlu, et al.
Veröffentlicht: (2023)
von: Qiu, Linlu, et al.
Veröffentlicht: (2023)
ProfBench: Multi-Domain Rubrics requiring Professional Knowledge to Answer and Judge
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
Prismatic Synthesis: Gradient-based Data Diversification Boosts Generalization in LLM Reasoning
von: Jung, Jaehun, et al.
Veröffentlicht: (2025)
von: Jung, Jaehun, et al.
Veröffentlicht: (2025)
Trust or Escalate: LLM Judges with Provable Guarantees for Human Agreement
von: Jung, Jaehun, et al.
Veröffentlicht: (2024)
von: Jung, Jaehun, et al.
Veröffentlicht: (2024)
DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life
von: Chiu, Yu Ying, et al.
Veröffentlicht: (2024)
von: Chiu, Yu Ying, et al.
Veröffentlicht: (2024)
Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
von: Ravichander, Abhilasha, et al.
Veröffentlicht: (2025)
von: Ravichander, Abhilasha, et al.
Veröffentlicht: (2025)
Retro-Search: Exploring Untaken Paths for Deeper and Efficient Reasoning
von: Lu, Ximing, et al.
Veröffentlicht: (2025)
von: Lu, Ximing, et al.
Veröffentlicht: (2025)
ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models
von: Liu, Mingjie, et al.
Veröffentlicht: (2025)
von: Liu, Mingjie, et al.
Veröffentlicht: (2025)
Long Grounded Thoughts: Synthesizing Visual Problems and Reasoning Chains at Scale
von: Acuna, David, et al.
Veröffentlicht: (2025)
von: Acuna, David, et al.
Veröffentlicht: (2025)
ALISON: Fast and Effective Stylometric Authorship Obfuscation
von: Xing, Eric, et al.
Veröffentlicht: (2024)
von: Xing, Eric, et al.
Veröffentlicht: (2024)
Can Language Models Reason about Individualistic Human Values and Preferences?
von: Jiang, Liwei, et al.
Veröffentlicht: (2024)
von: Jiang, Liwei, et al.
Veröffentlicht: (2024)
Every Step Counts: Decoding Trajectories as Authorship Fingerprints of dLLMs
von: Li, Qi, et al.
Veröffentlicht: (2025)
von: Li, Qi, et al.
Veröffentlicht: (2025)
BroRL: Scaling Reinforcement Learning via Broadened Exploration
von: Hu, Jian, et al.
Veröffentlicht: (2025)
von: Hu, Jian, et al.
Veröffentlicht: (2025)
Language Model Meets Prototypes: Towards Interpretable Text Classification Models through Prototypical Networks
von: Wen, Ximing
Veröffentlicht: (2024)
von: Wen, Ximing
Veröffentlicht: (2024)
FDARxBench: Benchmarking Regulatory and Clinical Reasoning on FDA Generic Drug Assessment
von: Xiong, Betty, et al.
Veröffentlicht: (2026)
von: Xiong, Betty, et al.
Veröffentlicht: (2026)
DeepSearch: Overcome the Bottleneck of Reinforcement Learning with Verifiable Rewards via Monte Carlo Tree Search
von: Wu, Fang, et al.
Veröffentlicht: (2025)
von: Wu, Fang, et al.
Veröffentlicht: (2025)
CULTURE-GEN: Revealing Global Cultural Perception in Language Models through Natural Language Prompting
von: Li, Huihan, et al.
Veröffentlicht: (2024)
von: Li, Huihan, et al.
Veröffentlicht: (2024)
Personalized Author Obfuscation with Large Language Models
von: Shokri, Mohammad, et al.
Veröffentlicht: (2025)
von: Shokri, Mohammad, et al.
Veröffentlicht: (2025)
Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
von: Sorensen, Taylor, et al.
Veröffentlicht: (2023)
von: Sorensen, Taylor, et al.
Veröffentlicht: (2023)
From Decoding to Meta-Generation: Inference-time Algorithms for Large Language Models
von: Welleck, Sean, et al.
Veröffentlicht: (2024)
von: Welleck, Sean, et al.
Veröffentlicht: (2024)
Unraveling Interwoven Roles of Large Language Models in Authorship Privacy: Obfuscation, Mimicking, and Verification
von: Nguyen, Tuc, et al.
Veröffentlicht: (2025)
von: Nguyen, Tuc, et al.
Veröffentlicht: (2025)
WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
von: Jiang, Liwei, et al.
Veröffentlicht: (2024)
von: Jiang, Liwei, et al.
Veröffentlicht: (2024)
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
von: Macko, Dominik, et al.
Veröffentlicht: (2024)
von: Macko, Dominik, et al.
Veröffentlicht: (2024)
Obfuscation Rules for Detecting and Detoxifying Korean Toxicity
von: Lee, Yejin, et al.
Veröffentlicht: (2025)
von: Lee, Yejin, et al.
Veröffentlicht: (2025)
The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage
von: Hallinan, Skyler, et al.
Veröffentlicht: (2025)
von: Hallinan, Skyler, et al.
Veröffentlicht: (2025)
Can LLMs Obfuscate Code? A Systematic Analysis of Large Language Models into Assembly Code Obfuscation
von: Mohseni, Seyedreza, et al.
Veröffentlicht: (2024)
von: Mohseni, Seyedreza, et al.
Veröffentlicht: (2024)
Opt-ICL at LeWiDi-2025: Maximizing In-Context Signal from Rater Examples via Meta-Learning
von: Sorensen, Taylor, et al.
Veröffentlicht: (2025)
von: Sorensen, Taylor, et al.
Veröffentlicht: (2025)
Retrieval-Constrained Decoding Reveals Underestimated Parametric Knowledge in Language Models
von: Hamdani, Rajaa El, et al.
Veröffentlicht: (2025)
von: Hamdani, Rajaa El, et al.
Veröffentlicht: (2025)
AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
von: Lu, Ximing, et al.
Veröffentlicht: (2024)
von: Lu, Ximing, et al.
Veröffentlicht: (2024)
Masks and Mimicry: Strategic Obfuscation and Impersonation Attacks on Authorship Verification
von: Alperin, Kenneth, et al.
Veröffentlicht: (2025)
von: Alperin, Kenneth, et al.
Veröffentlicht: (2025)
WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences
von: Lu, Yujie, et al.
Veröffentlicht: (2024)
von: Lu, Yujie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
von: Fisher, Jillian, et al.
Veröffentlicht: (2024) -
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing
von: Jung, Jaehun, et al.
Veröffentlicht: (2023) -
Information-Theoretic Distillation for Reference-less Summarization
von: Jung, Jaehun, et al.
Veröffentlicht: (2024) -
The Invisible Leash: Why RLVR May or May Not Escape Its Origin
von: Wu, Fang, et al.
Veröffentlicht: (2025) -
Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions
von: Acuna, David, et al.
Veröffentlicht: (2025)