SWAN: Semantic Watermarking with Abstract Meaning Representation
Fuente:
arXiv
Saved in:
| Main Authors: | Ye, Ziping, Dey, Gourab, Christodoulopoulos, Christos, Peris, Charith, Ramakrishna, Anil, Ruan, Weitong, Galstyan, Aram, Chang, Kai-Wei, Gupta, Rahul, Mehrabi, Ninareh |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Attribute Controlled Fine-tuning for Large Language Models: A Case Study on Detoxification
by: Meng, Tao, et al.
Published: (2024)
by: Meng, Tao, et al.
Published: (2024)
Towards Safety Reasoning in LLMs: AI-agentic Deliberation for Policy-embedded CoT Data Creation
by: Kumarage, Tharindu, et al.
Published: (2025)
by: Kumarage, Tharindu, et al.
Published: (2025)
K-Edit: Language Model Editing with Contextual Knowledge Awareness
by: Markowitz, Elan, et al.
Published: (2025)
by: Markowitz, Elan, et al.
Published: (2025)
On the steerability of large language models toward data-driven personas
by: Li, Junyi, et al.
Published: (2023)
by: Li, Junyi, et al.
Published: (2023)
Tree-of-Traversals: A Zero-Shot Reasoning Algorithm for Augmenting Black-box Language Models with Knowledge Graphs
by: Markowitz, Elan, et al.
Published: (2024)
by: Markowitz, Elan, et al.
Published: (2024)
Data Advisor: Dynamic Data Curation for Safety Alignment of Large Language Models
by: Wang, Fei, et al.
Published: (2024)
by: Wang, Fei, et al.
Published: (2024)
Customize Multi-modal RAI Guardrails with Precedent-based predictions
by: Yang, Cheng-Fu, et al.
Published: (2025)
by: Yang, Cheng-Fu, et al.
Published: (2025)
Kaleidoscopic Teaming in Multi Agent Simulations
by: Mehrabi, Ninareh, et al.
Published: (2025)
by: Mehrabi, Ninareh, et al.
Published: (2025)
Tokenization Matters: Navigating Data-Scarce Tokenization for Gender Inclusive Language Technologies
by: Ovalle, Anaelia, et al.
Published: (2023)
by: Ovalle, Anaelia, et al.
Published: (2023)
Prompt Perturbation Consistency Learning for Robust Language Models
by: Qiang, Yao, et al.
Published: (2024)
by: Qiang, Yao, et al.
Published: (2024)
Partial Federated Learning
by: Feng, Tiantian, et al.
Published: (2024)
by: Feng, Tiantian, et al.
Published: (2024)
Robust Persona-Aware Toxicity Detection with Prompt Optimization and Learned Ensembling
by: Atil, Berk, et al.
Published: (2026)
by: Atil, Berk, et al.
Published: (2026)
Making Sense Of Distributed Representations With Activation Spectroscopy
by: Reing, Kyle, et al.
Published: (2025)
by: Reing, Kyle, et al.
Published: (2025)
Not Every Token Needs Forgetting: Selective Unlearning to Limit Change in Utility in Large Language Model Unlearning
by: Wan, Yixin, et al.
Published: (2025)
by: Wan, Yixin, et al.
Published: (2025)
Diagnosing Memorization in Chain-of-Thought Reasoning, One Token at a Time
by: Li, Huihan, et al.
Published: (2025)
by: Li, Huihan, et al.
Published: (2025)
Evaluating the Critical Risks of Amazon's Nova Premier under the Frontier Model Safety Framework
by: Krishna, Satyapriya, et al.
Published: (2025)
by: Krishna, Satyapriya, et al.
Published: (2025)
ARES: Adaptive Red-Teaming and End-to-End Repair of Policy-Reward System
by: Liang, Jiacheng, et al.
Published: (2026)
by: Liang, Jiacheng, et al.
Published: (2026)
FLIRT: Feedback Loop In-context Red Teaming
by: Mehrabi, Ninareh, et al.
Published: (2023)
by: Mehrabi, Ninareh, et al.
Published: (2023)
FERRET: Framework for Expansion Reliant Red Teaming
by: Mehrabi, Ninareh, et al.
Published: (2026)
by: Mehrabi, Ninareh, et al.
Published: (2026)
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
by: Parekh, Tanmay, et al.
Published: (2025)
by: Parekh, Tanmay, et al.
Published: (2025)
Asymmetric Phase Coding Audio Watermarking
by: Yang, Guang, et al.
Published: (2026)
by: Yang, Guang, et al.
Published: (2026)
Assessing Visual Privacy Risks in Multimodal AI: A Novel Taxonomy-Grounded Evaluation of Vision-Language Models
by: Tsaprazlis, Efthymios, et al.
Published: (2025)
by: Tsaprazlis, Efthymios, et al.
Published: (2025)
Learning Morphisms with Gauss-Newton Approximation for Growing Networks
by: Lawton, Neal, et al.
Published: (2024)
by: Lawton, Neal, et al.
Published: (2024)
Survey of Abstract Meaning Representation: Then, Now, Future
by: Mansouri, Behrooz
Published: (2025)
by: Mansouri, Behrooz
Published: (2025)
Abstract Meaning Representation for Hospital Discharge Summarization
by: Landes, Paul, et al.
Published: (2025)
by: Landes, Paul, et al.
Published: (2025)
Rethinking Visual Privacy: A Compositional Privacy Risk Framework for Severity Assessment with VLMs
by: Tsaprazlis, Efthymios, et al.
Published: (2026)
by: Tsaprazlis, Efthymios, et al.
Published: (2026)
Emergent Strategic Reasoning Risks in AI: A Taxonomy-Driven Evaluation Framework
by: Kumarage, Tharindu, et al.
Published: (2026)
by: Kumarage, Tharindu, et al.
Published: (2026)
KG-LLM-Bench: A Scalable Benchmark for Evaluating LLM Reasoning on Textualized Knowledge Graphs
by: Markowitz, Elan, et al.
Published: (2025)
by: Markowitz, Elan, et al.
Published: (2025)
RoWSFormer: A Robust Watermarking Framework with Swin Transformer for Enhanced Geometric Attack Resilience
by: Chen, Weitong, et al.
Published: (2024)
by: Chen, Weitong, et al.
Published: (2024)
When Actions Go Off-Task: Detecting and Correcting Misaligned Actions in Computer-Use Agents
by: Ning, Yuting, et al.
Published: (2026)
by: Ning, Yuting, et al.
Published: (2026)
Breaking the Benchmark: Revealing LLM Bias via Minimal Contextual Augmentation
by: Miandoab, Kaveh Eskandari, et al.
Published: (2025)
by: Miandoab, Kaveh Eskandari, et al.
Published: (2025)
Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning
by: Jeoung, Sullam, et al.
Published: (2024)
by: Jeoung, Sullam, et al.
Published: (2024)
SemEval-2025 Task 4: Unlearning sensitive content from Large Language Models
by: Ramakrishna, Anil, et al.
Published: (2025)
by: Ramakrishna, Anil, et al.
Published: (2025)
LUME: LLM Unlearning with Multitask Evaluations
by: Ramakrishna, Anil, et al.
Published: (2025)
by: Ramakrishna, Anil, et al.
Published: (2025)
Contrastive Learning with Enhanced Abstract Representations using Grouped Loss of Abstract Semantic Supervision
by: Suissa, Omri, et al.
Published: (2025)
by: Suissa, Omri, et al.
Published: (2025)
The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans
by: Chlapanis, Odysseas S., et al.
Published: (2026)
by: Chlapanis, Odysseas S., et al.
Published: (2026)
Evaluating Differentially Private Synthetic Data Generation in High-Stakes Domains
by: Ramesh, Krithika, et al.
Published: (2024)
by: Ramesh, Krithika, et al.
Published: (2024)
Persian Abstract Meaning Representation: Annotation Guidelines and Gold Standard Dataset
by: Takhshid, Reza, et al.
Published: (2022)
by: Takhshid, Reza, et al.
Published: (2022)
Lost in Translationese? Reducing Translation Effect Using Abstract Meaning Representation
by: Wein, Shira, et al.
Published: (2023)
by: Wein, Shira, et al.
Published: (2023)
Asking Back: Interaction-Layer Antidistillation Watermarks
by: Yang, Guang, et al.
Published: (2026)
by: Yang, Guang, et al.
Published: (2026)
Similar Items
-
Attribute Controlled Fine-tuning for Large Language Models: A Case Study on Detoxification
by: Meng, Tao, et al.
Published: (2024) -
Towards Safety Reasoning in LLMs: AI-agentic Deliberation for Policy-embedded CoT Data Creation
by: Kumarage, Tharindu, et al.
Published: (2025) -
K-Edit: Language Model Editing with Contextual Knowledge Awareness
by: Markowitz, Elan, et al.
Published: (2025) -
On the steerability of large language models toward data-driven personas
by: Li, Junyi, et al.
Published: (2023) -
Tree-of-Traversals: A Zero-Shot Reasoning Algorithm for Augmenting Black-box Language Models with Knowledge Graphs
by: Markowitz, Elan, et al.
Published: (2024)