Self-Sovereign Agent
Fuente:
arXiv
Saved in:
| Main Authors: | Qu, Wenjie, Zhao, Xuandong, Zhang, Jiaheng, Song, Dawn |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Are You Getting What You Pay For? Auditing Model Substitution in LLM APIs
by: Cai, Will, et al.
Published: (2025)
by: Cai, Will, et al.
Published: (2025)
The Landscape of Memorization in LLMs: Mechanisms, Measurement, and Mitigation
by: Xiong, Alexander, et al.
Published: (2025)
by: Xiong, Alexander, et al.
Published: (2025)
Improving LLM Safety Alignment with Dual-Objective Optimization
by: Zhao, Xuandong, et al.
Published: (2025)
by: Zhao, Xuandong, et al.
Published: (2025)
An Undetectable Watermark for Generative Image Models
by: Gunn, Sam, et al.
Published: (2024)
by: Gunn, Sam, et al.
Published: (2024)
LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage
by: Nie, Yuzhou, et al.
Published: (2024)
by: Nie, Yuzhou, et al.
Published: (2024)
Echoes within the Reasoning: Stealthy and Effective Watermarking via Chain of Thought
by: Lu, Jiacheng, et al.
Published: (2026)
by: Lu, Jiacheng, et al.
Published: (2026)
Differentially Private Post-Processing for Fair Regression
by: Xian, Ruicheng, et al.
Published: (2024)
by: Xian, Ruicheng, et al.
Published: (2024)
Permute-and-Flip: An optimally stable and watermarkable decoder for LLMs
by: Zhao, Xuandong, et al.
Published: (2024)
by: Zhao, Xuandong, et al.
Published: (2024)
Securing LLM Agents Need Intent-to-Execution Integrity
by: Qu, Wenjie, et al.
Published: (2026)
by: Qu, Wenjie, et al.
Published: (2026)
Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications
by: Wu, Yixin, et al.
Published: (2025)
by: Wu, Yixin, et al.
Published: (2025)
Machine Unlearning: Taxonomy, Metrics, Applications, Challenges, and Prospects
by: Li, Na, et al.
Published: (2024)
by: Li, Na, et al.
Published: (2024)
Towards a Systematic Risk Assessment of Deep Neural Network Limitations in Autonomous Driving Perception
by: Pavlitska, Svetlana, et al.
Published: (2026)
by: Pavlitska, Svetlana, et al.
Published: (2026)
Know Your Scientist: KYC as Biosecurity Infrastructure
by: Feldman, Jonathan, et al.
Published: (2026)
by: Feldman, Jonathan, et al.
Published: (2026)
Gaming the Metric, Not the Harm: Certifying Safety Audits against Strategic Platform Manipulation
by: Burnat, Florian A. D., et al.
Published: (2026)
by: Burnat, Florian A. D., et al.
Published: (2026)
Rethinking Anonymity Claims in Synthetic Data Generation: A Model-Centric Privacy Attack Perspective
by: Ganev, Georgi, et al.
Published: (2026)
by: Ganev, Georgi, et al.
Published: (2026)
"bot lane noob" Towards Deployment of NLP-based Toxicity Detectors in Video Games
by: Ave, Jonas, et al.
Published: (2026)
by: Ave, Jonas, et al.
Published: (2026)
FNDaaS: Content-agnostic Detection of Fake News sites
by: Papadopoulos, Panagiotis, et al.
Published: (2022)
by: Papadopoulos, Panagiotis, et al.
Published: (2022)
On the Impact of Multi-dimensional Local Differential Privacy on Fairness
by: Makhlouf, Karima, et al.
Published: (2023)
by: Makhlouf, Karima, et al.
Published: (2023)
LDPKiT: Superimposing Remote Queries for Privacy-Preserving Local Model Training
by: Li, Kexin, et al.
Published: (2024)
by: Li, Kexin, et al.
Published: (2024)
Position Paper: Assessing Robustness, Privacy, and Fairness in Federated Learning Integrated with Foundation Models
by: Wang, Jiaqi, et al.
Published: (2024)
by: Wang, Jiaqi, et al.
Published: (2024)
Sharing is CAIRing: Characterizing Principles and Assessing Properties of Universal Privacy Evaluation for Synthetic Tabular Data
by: Hyrup, Tobias, et al.
Published: (2023)
by: Hyrup, Tobias, et al.
Published: (2023)
Personalized Differential Privacy for Ridge Regression
by: Acharya, Krishna, et al.
Published: (2024)
by: Acharya, Krishna, et al.
Published: (2024)
Privacy-Preserving Data Linkage Across Private and Public Datasets for Collaborative Agriculture Research
by: Zafar, Osama, et al.
Published: (2024)
by: Zafar, Osama, et al.
Published: (2024)
Disparate Impact on Group Accuracy of Linearization for Private Inference
by: Das, Saswat, et al.
Published: (2024)
by: Das, Saswat, et al.
Published: (2024)
Minerva: A File-Based Ransomware Detector
by: Hitaj, Dorjan, et al.
Published: (2023)
by: Hitaj, Dorjan, et al.
Published: (2023)
Blameless Users in a Clean Room: Defining Copyright Protection for Generative Models
by: Cohen, Aloni
Published: (2025)
by: Cohen, Aloni
Published: (2025)
Privacy-Preserving Dataset Combination
by: Fuentes, Keren, et al.
Published: (2025)
by: Fuentes, Keren, et al.
Published: (2025)
Towards eco friendly cybersecurity: machine learning based anomaly detection with carbon and energy metrics
by: Aashish, KC, et al.
Published: (2025)
by: Aashish, KC, et al.
Published: (2025)
I'm Sorry Dave: How the old world of personnel security can inform the new world of AI insider risk
by: Martin, Paul, et al.
Published: (2025)
by: Martin, Paul, et al.
Published: (2025)
RiM: Record, Improve and Maintain Physical Well-being using Federated Learning
by: Mishra, Aditya, et al.
Published: (2025)
by: Mishra, Aditya, et al.
Published: (2025)
Learning Fair Robustness via Domain Mixup
by: Zhong, Meiyu, et al.
Published: (2024)
by: Zhong, Meiyu, et al.
Published: (2024)
Privacy Constrained Fairness Estimation for Decision Trees
by: van der Steen, Florian, et al.
Published: (2023)
by: van der Steen, Florian, et al.
Published: (2023)
Conscious Data Contribution via Community-Driven Chain-of-Thought Distillation
by: Libon, Lena, et al.
Published: (2025)
by: Libon, Lena, et al.
Published: (2025)
Fragments to Facts: Partial-Information Fragment Inference from LLMs
by: Rosenblatt, Lucas, et al.
Published: (2025)
by: Rosenblatt, Lucas, et al.
Published: (2025)
FairDP: Certified Fairness with Differential Privacy
by: Tran, Khang, et al.
Published: (2023)
by: Tran, Khang, et al.
Published: (2023)
Systematically Assessing the Security Risks of AI/ML-enabled Connected Healthcare Systems
by: Elnawawy, Mohammed, et al.
Published: (2024)
by: Elnawawy, Mohammed, et al.
Published: (2024)
Guarantees of confidentiality via Hammersley-Chapman-Robbins bounds
by: Chaudhuri, Kamalika, et al.
Published: (2024)
by: Chaudhuri, Kamalika, et al.
Published: (2024)
Sovereign 2.0: Control-Plane Sovereignty for Cloud Systems Under Disruption
by: Stark, Justin, et al.
Published: (2026)
by: Stark, Justin, et al.
Published: (2026)
Password-Activated Shutdown Protocols for Misaligned Frontier Agents
by: Williams, Kai, et al.
Published: (2025)
by: Williams, Kai, et al.
Published: (2025)
Privacy at a Price: Exploring its Dual Impact on AI Fairness
by: Yang, Mengmeng, et al.
Published: (2024)
by: Yang, Mengmeng, et al.
Published: (2024)
Similar Items
-
Are You Getting What You Pay For? Auditing Model Substitution in LLM APIs
by: Cai, Will, et al.
Published: (2025) -
The Landscape of Memorization in LLMs: Mechanisms, Measurement, and Mitigation
by: Xiong, Alexander, et al.
Published: (2025) -
Improving LLM Safety Alignment with Dual-Objective Optimization
by: Zhao, Xuandong, et al.
Published: (2025) -
An Undetectable Watermark for Generative Image Models
by: Gunn, Sam, et al.
Published: (2024) -
LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage
by: Nie, Yuzhou, et al.
Published: (2024)