Amplifying, Not Learning: Fine-Tuned AI Text Detectors Amplify a Pretrained Direction
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Smirnov, Alexander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Janus Interface: How Fine-Tuning in Large Language Models Amplifies the Privacy Risks
von: Chen, Xiaoyi, et al.
Veröffentlicht: (2023)
von: Chen, Xiaoyi, et al.
Veröffentlicht: (2023)
ChemAmp: Amplified Chemistry Tools via Composable Agents
von: Li, Zhucong, et al.
Veröffentlicht: (2025)
von: Li, Zhucong, et al.
Veröffentlicht: (2025)
Superscopes: Amplifying Internal Feature Representations for Language Model Interpretation
von: Jacobi, Jonathan, et al.
Veröffentlicht: (2025)
von: Jacobi, Jonathan, et al.
Veröffentlicht: (2025)
Amplifying human performance in combinatorial competitive programming
von: Veličković, Petar, et al.
Veröffentlicht: (2024)
von: Veličković, Petar, et al.
Veröffentlicht: (2024)
RIPPLECOT: Amplifying Ripple Effect of Knowledge Editing in Language Models via Chain-of-Thought In-Context Learning
von: Zhao, Zihao, et al.
Veröffentlicht: (2024)
von: Zhao, Zihao, et al.
Veröffentlicht: (2024)
Scalpel vs. Hammer: GRPO Amplifies Existing Capabilities, SFT Replaces Them
von: Rajani, Neel, et al.
Veröffentlicht: (2025)
von: Rajani, Neel, et al.
Veröffentlicht: (2025)
WRAP++: Web discoveRy Amplified Pretraining
von: Zhou, Jiang, et al.
Veröffentlicht: (2026)
von: Zhou, Jiang, et al.
Veröffentlicht: (2026)
BIPEFT: Budget-Guided Iterative Search for Parameter Efficient Fine-Tuning of Large Pretrained Language Models
von: Chang, Aofei, et al.
Veröffentlicht: (2024)
von: Chang, Aofei, et al.
Veröffentlicht: (2024)
Fine-Tuning and Evaluating Conversational AI for Agricultural Advisory
von: Singh, Sanyam, et al.
Veröffentlicht: (2026)
von: Singh, Sanyam, et al.
Veröffentlicht: (2026)
Direct Alignment of Draft Model for Speculative Decoding with Chat-Fine-Tuned LLMs
von: Goel, Raghavv, et al.
Veröffentlicht: (2024)
von: Goel, Raghavv, et al.
Veröffentlicht: (2024)
Robust and Fine-Grained Detection of AI Generated Texts
von: Kadiyala, Ram Mohan Rao, et al.
Veröffentlicht: (2025)
von: Kadiyala, Ram Mohan Rao, et al.
Veröffentlicht: (2025)
Supervised Fine-Tuning as Inverse Reinforcement Learning
von: Sun, Hao
Veröffentlicht: (2024)
von: Sun, Hao
Veröffentlicht: (2024)
Joint Localization and Activation Editing for Low-Resource Fine-Tuning
von: Lai, Wen, et al.
Veröffentlicht: (2025)
von: Lai, Wen, et al.
Veröffentlicht: (2025)
Fine-Tuning Language Models with Reward Learning on Policy
von: Lang, Hao, et al.
Veröffentlicht: (2024)
von: Lang, Hao, et al.
Veröffentlicht: (2024)
Teaching LLMs How to Learn with Contextual Fine-Tuning
von: Choi, Younwoo, et al.
Veröffentlicht: (2025)
von: Choi, Younwoo, et al.
Veröffentlicht: (2025)
IPAD: Inverse Prompt for AI Detection - A Robust and Interpretable LLM-Generated Text Detector
von: Chen, Zheng, et al.
Veröffentlicht: (2025)
von: Chen, Zheng, et al.
Veröffentlicht: (2025)
Proximal Supervised Fine-Tuning
von: Zhu, Wenhong, et al.
Veröffentlicht: (2025)
von: Zhu, Wenhong, et al.
Veröffentlicht: (2025)
Fine-Tuning is Subgraph Search: A New Lens on Learning Dynamics
von: Li, Yueyan, et al.
Veröffentlicht: (2025)
von: Li, Yueyan, et al.
Veröffentlicht: (2025)
Efficient Differentially Private Fine-Tuning of LLMs via Reinforcement Learning
von: Khadangi, Afshin, et al.
Veröffentlicht: (2025)
von: Khadangi, Afshin, et al.
Veröffentlicht: (2025)
Order-Independence Without Fine Tuning
von: McIlroy-Young, Reid, et al.
Veröffentlicht: (2024)
von: McIlroy-Young, Reid, et al.
Veröffentlicht: (2024)
Model Editing by Standard Fine-Tuning
von: Gangadhar, Govind, et al.
Veröffentlicht: (2024)
von: Gangadhar, Govind, et al.
Veröffentlicht: (2024)
ARES: Alternating Reinforcement Learning and Supervised Fine-Tuning for Enhanced Multi-Modal Chain-of-Thought Reasoning Through Diverse AI Feedback
von: Byun, Ju-Seung, et al.
Veröffentlicht: (2024)
von: Byun, Ju-Seung, et al.
Veröffentlicht: (2024)
Instruct-Tuning Pretrained Causal Language Models for Ancient Greek Papyrology and Epigraphy
von: Cullhed, Eric
Veröffentlicht: (2024)
von: Cullhed, Eric
Veröffentlicht: (2024)
Q-SFT: Q-Learning for Language Models via Supervised Fine-Tuning
von: Hong, Joey, et al.
Veröffentlicht: (2024)
von: Hong, Joey, et al.
Veröffentlicht: (2024)
Text to Trust: Evaluating Fine-Tuning and LoRA Trade-offs in Language Models for Unfair Terms of Service Detection
von: Juttu, Noshitha Padma Pratyusha, et al.
Veröffentlicht: (2025)
von: Juttu, Noshitha Padma Pratyusha, et al.
Veröffentlicht: (2025)
Base Models Look Human To AI Detectors
von: Xu, Yixuan Even, et al.
Veröffentlicht: (2026)
von: Xu, Yixuan Even, et al.
Veröffentlicht: (2026)
Generalizable and Stable Finetuning of Pretrained Language Models on Low-Resource Texts
von: Somayajula, Sai Ashish, et al.
Veröffentlicht: (2024)
von: Somayajula, Sai Ashish, et al.
Veröffentlicht: (2024)
Parameter-Efficient Fine-Tuning for Foundation Models
von: Zhang, Dan, et al.
Veröffentlicht: (2025)
von: Zhang, Dan, et al.
Veröffentlicht: (2025)
Automated Text Scoring in the Age of Generative AI for the GPU-poor
von: Ormerod, Christopher Michael, et al.
Veröffentlicht: (2024)
von: Ormerod, Christopher Michael, et al.
Veröffentlicht: (2024)
Fine-tuning can Help Detect Pretraining Data from Large Language Models
von: Zhang, Hengxiang, et al.
Veröffentlicht: (2024)
von: Zhang, Hengxiang, et al.
Veröffentlicht: (2024)
MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning
von: Rüdiger, Sten, et al.
Veröffentlicht: (2026)
von: Rüdiger, Sten, et al.
Veröffentlicht: (2026)
EMORL: Ensemble Multi-Objective Reinforcement Learning for Efficient and Flexible LLM Fine-Tuning
von: Kong, Lingxiao, et al.
Veröffentlicht: (2025)
von: Kong, Lingxiao, et al.
Veröffentlicht: (2025)
Directional Gradient Projection for Robust Fine-Tuning of Foundation Models
von: Huang, Chengyue, et al.
Veröffentlicht: (2025)
von: Huang, Chengyue, et al.
Veröffentlicht: (2025)
Fine-Tuning Improves Information Conveyance in Language Models
von: Cheng, Yuwei, et al.
Veröffentlicht: (2026)
von: Cheng, Yuwei, et al.
Veröffentlicht: (2026)
Boosting Large Language Models with Mask Fine-Tuning
von: Zhang, Mingyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Mingyuan, et al.
Veröffentlicht: (2025)
Parameter-Efficient Fine-Tuning with Discrete Fourier Transform
von: Gao, Ziqi, et al.
Veröffentlicht: (2024)
von: Gao, Ziqi, et al.
Veröffentlicht: (2024)
ROSA: Random Subspace Adaptation for Efficient Fine-Tuning
von: Hameed, Marawan Gamal Abdel, et al.
Veröffentlicht: (2024)
von: Hameed, Marawan Gamal Abdel, et al.
Veröffentlicht: (2024)
Fine-Tuning or Retrieval? Comparing Knowledge Injection in LLMs
von: Ovadia, Oded, et al.
Veröffentlicht: (2023)
von: Ovadia, Oded, et al.
Veröffentlicht: (2023)
Blending Supervised and Reinforcement Fine-Tuning with Prefix Sampling
von: Huang, Zeyu, et al.
Veröffentlicht: (2025)
von: Huang, Zeyu, et al.
Veröffentlicht: (2025)
Understanding the Performance and Estimating the Cost of LLM Fine-Tuning
von: Xia, Yuchen, et al.
Veröffentlicht: (2024)
von: Xia, Yuchen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The Janus Interface: How Fine-Tuning in Large Language Models Amplifies the Privacy Risks
von: Chen, Xiaoyi, et al.
Veröffentlicht: (2023) -
ChemAmp: Amplified Chemistry Tools via Composable Agents
von: Li, Zhucong, et al.
Veröffentlicht: (2025) -
Superscopes: Amplifying Internal Feature Representations for Language Model Interpretation
von: Jacobi, Jonathan, et al.
Veröffentlicht: (2025) -
Amplifying human performance in combinatorial competitive programming
von: Veličković, Petar, et al.
Veröffentlicht: (2024) -
RIPPLECOT: Amplifying Ripple Effect of Knowledge Editing in Language Models via Chain-of-Thought In-Context Learning
von: Zhao, Zihao, et al.
Veröffentlicht: (2024)