DP-RFT: Learning to Generate Synthetic Text via Differentially Private Reinforcement Fine-Tuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Fangyuan, Chen, Sihao, Lin, Zinan, Shi, Taiwei, Graham, Sydney, Zhou, Pei, Wan, Mengting, Stein, Alex, Estellers, Virginia, Chen, Charles, Sharp, Morris, Speyer, Richard, Baltrusaitis, Tadas, Neville, Jennifer, Choi, Eunsol, Yang, Longqi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Differentially Private Synthetic Data via APIs 3: Using Simulators Instead of Foundation Model
von: Lin, Zinan, et al.
Veröffentlicht: (2025)
von: Lin, Zinan, et al.
Veröffentlicht: (2025)
GASP: Gaussian Avatars with Synthetic Priors
von: Saunders, Jack, et al.
Veröffentlicht: (2024)
von: Saunders, Jack, et al.
Veröffentlicht: (2024)
Experiential Reinforcement Learning
von: Shi, Taiwei, et al.
Veröffentlicht: (2026)
von: Shi, Taiwei, et al.
Veröffentlicht: (2026)
Struct-Bench: A Benchmark for Differentially Private Structured Text Generation
von: Wang, Shuaiqi, et al.
Veröffentlicht: (2025)
von: Wang, Shuaiqi, et al.
Veröffentlicht: (2025)
Teaching Language Models To Gather Information Proactively
von: Huang, Tenghao, et al.
Veröffentlicht: (2025)
von: Huang, Tenghao, et al.
Veröffentlicht: (2025)
One Model, All Roles: Multi-Turn, Multi-Agent Self-Play Reinforcement Learning for Conversational Social Intelligence
von: Jiang, Bowen, et al.
Veröffentlicht: (2026)
von: Jiang, Bowen, et al.
Veröffentlicht: (2026)
Eyelid Fold Consistency in Facial Modeling
von: Petikam, Lohit, et al.
Veröffentlicht: (2024)
von: Petikam, Lohit, et al.
Veröffentlicht: (2024)
DP-SAPF: Saliency-Aware Parameter Fine-tuning of Public Models for Differentially Private Image Synthesis
von: Gong, Chen, et al.
Veröffentlicht: (2026)
von: Gong, Chen, et al.
Veröffentlicht: (2026)
Emergent symmetries at criticality in multi field RFT/DP
von: Bartels, Jochen, et al.
Veröffentlicht: (2024)
von: Bartels, Jochen, et al.
Veröffentlicht: (2024)
SimpleEgo: Predicting Probabilistic Body Pose from Egocentric Cameras
von: Cuevas-Velasquez, Hanz, et al.
Veröffentlicht: (2024)
von: Cuevas-Velasquez, Hanz, et al.
Veröffentlicht: (2024)
WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback
von: Shi, Taiwei, et al.
Veröffentlicht: (2024)
von: Shi, Taiwei, et al.
Veröffentlicht: (2024)
DAViD: Data-efficient and Accurate Vision Models from Synthetic Data
von: Saleh, Fatemeh, et al.
Veröffentlicht: (2025)
von: Saleh, Fatemeh, et al.
Veröffentlicht: (2025)
Corporate Communication Companion (CCC): An LLM-empowered Writing Assistant for Workplace Social Media
von: Lu, Zhuoran, et al.
Veröffentlicht: (2024)
von: Lu, Zhuoran, et al.
Veröffentlicht: (2024)
RefreshKV: Updating Small KV Cache During Long-form Generation
von: Xu, Fangyuan, et al.
Veröffentlicht: (2024)
von: Xu, Fangyuan, et al.
Veröffentlicht: (2024)
Visual-RFT: Visual Reinforcement Fine-Tuning
von: Liu, Ziyu, et al.
Veröffentlicht: (2025)
von: Liu, Ziyu, et al.
Veröffentlicht: (2025)
DP-SelFT: Differentially Private Selective Fine-Tuning for Large Language Models
von: Sha, Haichao, et al.
Veröffentlicht: (2026)
von: Sha, Haichao, et al.
Veröffentlicht: (2026)
Understanding Retrieval Augmentation for Long-Form Question Answering
von: Chen, Hung-Ting, et al.
Veröffentlicht: (2023)
von: Chen, Hung-Ting, et al.
Veröffentlicht: (2023)
SympCam: Remote Optical Measurement of Sympathetic Arousal
von: Braun, Björn, et al.
Veröffentlicht: (2024)
von: Braun, Björn, et al.
Veröffentlicht: (2024)
Reason-RFT: Reinforcement Fine-Tuning for Visual Reasoning of Vision Language Models
von: Tan, Huajie, et al.
Veröffentlicht: (2025)
von: Tan, Huajie, et al.
Veröffentlicht: (2025)
ConsistentRFT: Reducing Visual Hallucinations in Flow-based Reinforcement Fine-Tuning
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
Evaluating LLM-Simulated Conversations in Modeling Inconsistent and Uncollaborative Behaviors in Human Social Interaction
von: Kamoi, Ryo, et al.
Veröffentlicht: (2026)
von: Kamoi, Ryo, et al.
Veröffentlicht: (2026)
Differentially Private Synthetic Data via Foundation Model APIs 1: Images
von: Lin, Zinan, et al.
Veröffentlicht: (2023)
von: Lin, Zinan, et al.
Veröffentlicht: (2023)
GenTool: Enhancing Tool Generalization in Language Models through Zero-to-One and Weak-to-Strong Simulation
von: He, Jie, et al.
Veröffentlicht: (2025)
von: He, Jie, et al.
Veröffentlicht: (2025)
DP-RDM: Adapting Diffusion Models to Private Domains Without Fine-Tuning
von: Lebensold, Jonathan, et al.
Veröffentlicht: (2024)
von: Lebensold, Jonathan, et al.
Veröffentlicht: (2024)
Contrastive Learning to Improve Retrieval for Real-world Fact Checking
von: Sriram, Aniruddh, et al.
Veröffentlicht: (2024)
von: Sriram, Aniruddh, et al.
Veröffentlicht: (2024)
Trinity-RFT: A General-Purpose and Unified Framework for Reinforcement Fine-Tuning of Large Language Models
von: Pan, Xuchen, et al.
Veröffentlicht: (2025)
von: Pan, Xuchen, et al.
Veröffentlicht: (2025)
Mesh-RFT: Enhancing Mesh Generation via Fine-grained Reinforcement Fine-Tuning
von: Liu, Jian, et al.
Veröffentlicht: (2025)
von: Liu, Jian, et al.
Veröffentlicht: (2025)
SoRFT: Issue Resolving with Subtask-oriented Reinforced Fine-Tuning
von: Ma, Zexiong, et al.
Veröffentlicht: (2025)
von: Ma, Zexiong, et al.
Veröffentlicht: (2025)
MotionRFT: Unified Reinforcement Fine-Tuning for Text-to-Motion Generation
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
Europäisch - amerikanische Verwandtschaften
von: Speyer, Adolf
Veröffentlicht: (1875)
von: Speyer, Adolf
Veröffentlicht: (1875)
The minimal counterexample to James's conjecture
von: Speyer, Liron
Veröffentlicht: (2026)
von: Speyer, Liron
Veröffentlicht: (2026)
Wild blocks of type $A$ Hecke algebras are strictly wild
von: Speyer, Liron
Veröffentlicht: (2024)
von: Speyer, Liron
Veröffentlicht: (2024)
What can complex systems theory reveal about social inclusion
von: Helene Speyer
Veröffentlicht: (2026)
von: Helene Speyer
Veröffentlicht: (2026)
Wild blocks of type A$A$ Hecke algebras are strictly wild
von: Liron Speyer
Veröffentlicht: (2025)
von: Liron Speyer
Veröffentlicht: (2025)
PE-means: Improved Differentially Private $k$-means Clustering through Private Evolution
von: Humphries, Thomas, et al.
Veröffentlicht: (2026)
von: Humphries, Thomas, et al.
Veröffentlicht: (2026)
DPImageBench: A Unified Benchmark for Differentially Private Image Synthesis
von: Gong, Chen, et al.
Veröffentlicht: (2025)
von: Gong, Chen, et al.
Veröffentlicht: (2025)
How Private are DP-SGD Implementations?
von: Chua, Lynn, et al.
Veröffentlicht: (2024)
von: Chua, Lynn, et al.
Veröffentlicht: (2024)
DP-TLDM: Differentially Private Tabular Latent Diffusion Model
von: Zhu, Chaoyi, et al.
Veröffentlicht: (2024)
von: Zhu, Chaoyi, et al.
Veröffentlicht: (2024)
Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
LMO-DP: Optimizing the Randomization Mechanism for Differentially Private Fine-Tuning (Large) Language Models
von: Yang, Qin, et al.
Veröffentlicht: (2024)
von: Yang, Qin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Differentially Private Synthetic Data via APIs 3: Using Simulators Instead of Foundation Model
von: Lin, Zinan, et al.
Veröffentlicht: (2025) -
GASP: Gaussian Avatars with Synthetic Priors
von: Saunders, Jack, et al.
Veröffentlicht: (2024) -
Experiential Reinforcement Learning
von: Shi, Taiwei, et al.
Veröffentlicht: (2026) -
Struct-Bench: A Benchmark for Differentially Private Structured Text Generation
von: Wang, Shuaiqi, et al.
Veröffentlicht: (2025) -
Teaching Language Models To Gather Information Proactively
von: Huang, Tenghao, et al.
Veröffentlicht: (2025)