SFD-Defense: Engineering Validation of the Semantic Flow Dynamics Defense Framework
Fuente:
Zenodo
Gespeichert in:
| 1. Verfasser: | 黃, 正宇 |
|---|---|
| Format: | Recurso digital |
| Sprache: | Englisch |
| Veröffentlicht: |
Zenodo
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Versificação Adversarial em Português como Operador de Jailbreak em LLMs
von: Queiroz, Joao
Veröffentlicht: (2026)
von: Queiroz, Joao
Veröffentlicht: (2026)
Not Prompt Engineering—Prompt Alchemy: Inducing Sub-Personality Emergence in GPT-4 Without Fine-Tuning
von: Wu, Ling
Veröffentlicht: (2025)
von: Wu, Ling
Veröffentlicht: (2025)
The Fantasia Bound on Constitutional Classifiers: Thermodynamic Limits of Jailbreak Defence
von: Eckert, Anthony
Veröffentlicht: (2026)
von: Eckert, Anthony
Veröffentlicht: (2026)
Pandora Theory of Alignment: Alignment as Runtime Objective-Orientation
von: Shopov, Georgi
Veröffentlicht: (2026)
von: Shopov, Georgi
Veröffentlicht: (2026)
A Deterministic Linguistic Entropy Gate for Large Language Model Pipelines
von: ROSATI BERISTAIN, ERNESTO
Veröffentlicht: (2026)
von: ROSATI BERISTAIN, ERNESTO
Veröffentlicht: (2026)
Reliability Inference Drives Cue Extraction in Large Language Models Consuming External Reasoning Traces
von: HIDEKI
Veröffentlicht: (2026)
von: HIDEKI
Veröffentlicht: (2026)
Design hacker e as customizações de interface gráfica
von: Danilo Braga
Veröffentlicht: (2020)
von: Danilo Braga
Veröffentlicht: (2020)
UI-Based Defense Against Prompt Injection: From Gentle Guidance to Mandatory Re-education
von: Viorazu.
Veröffentlicht: (2025)
von: Viorazu.
Veröffentlicht: (2025)
The Brain Problem: Creative Constraint Optimization in Large Language Models
von: Marinello, Nicola, et al.
Veröffentlicht: (2026)
von: Marinello, Nicola, et al.
Veröffentlicht: (2026)
N+1 Alignment Dialogue Architecture: Technical Specification for Defensive Publication
von: Garcia, Eric
Veröffentlicht: (2026)
von: Garcia, Eric
Veröffentlicht: (2026)
Glymphatic Architecture: A Fourth-Level System for Multi-Agent AI Consolidation and Identity Formation
von: Strugatsky, Leonid, et al.
Veröffentlicht: (2026)
von: Strugatsky, Leonid, et al.
Veröffentlicht: (2026)
Evaluación Empírica de Límites Regulatorios en Modelos de Lenguaje: Asesoramiento Financiero en IA Pública Española
von: Palacios, José Alberto
Veröffentlicht: (2026)
von: Palacios, José Alberto
Veröffentlicht: (2026)
Theatrical Compliance: A Failure Mode in Large Language Models
von: Nowickij (Navitski), Kirill Vladimirovich
Veröffentlicht: (2026)
von: Nowickij (Navitski), Kirill Vladimirovich
Veröffentlicht: (2026)
When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings
von: Kolb, Christian
Veröffentlicht: (2026)
von: Kolb, Christian
Veröffentlicht: (2026)
Safety by Inseparability: Toward Architectures Where Alignment Cannot Be Removed
von: Sean Everett, Morin
Veröffentlicht: (2026)
von: Sean Everett, Morin
Veröffentlicht: (2026)
AEGIS: A Comprehensive Framework for Ethical AI Governance, Security, and AGI Containment
von: Palanivel, ArulMozhi
Veröffentlicht: (2026)
von: Palanivel, ArulMozhi
Veröffentlicht: (2026)
Beyond Control: Resonance-Based Alignment for Advanced AI Systems A Governance-Relevant Concept Paper
von: Zieringer, Thomas
Veröffentlicht: (2025)
von: Zieringer, Thomas
Veröffentlicht: (2025)
Divided Focus: Separated Memory Spaces and Default-Deny Context Triage for LLM Context Management
von: Phan, Ivan "HiP"
Veröffentlicht: (2026)
von: Phan, Ivan "HiP"
Veröffentlicht: (2026)
Divided Focus: Separated Memory Spaces and Default-Deny Context Triage for LLM Context Management
von: Phan, Ivan "HiP"
Veröffentlicht: (2026)
von: Phan, Ivan "HiP"
Veröffentlicht: (2026)
Human–AI Symbiosis: Relational Alignment in Domains of Extreme Physical Irreversibility
von: de la Morena Marzalo, Juan
Veröffentlicht: (2026)
von: de la Morena Marzalo, Juan
Veröffentlicht: (2026)
Reducing AI Entropy: The Information Dynamics of Model Safety
von: Kugelmass, Joe
Veröffentlicht: (2025)
von: Kugelmass, Joe
Veröffentlicht: (2025)
Reducing AI Entropy: The Information Dynamics of Model Safety
von: Kugelmass, Joe
Veröffentlicht: (2025)
von: Kugelmass, Joe
Veröffentlicht: (2025)
Semantic Relativity Theory v2.3: Topological Stability through Euler-CHORDS++ Integration
von: López López, José
Veröffentlicht: (2026)
von: López López, José
Veröffentlicht: (2026)
Safety & Defense
Veröffentlicht: (2020)
Veröffentlicht: (2020)
ILAS: Integrity Layer for Agentic Systems
von: Böhm, Frank
Veröffentlicht: (2026)
von: Böhm, Frank
Veröffentlicht: (2026)
SILENCIUM: A Pre-Inference De-Escalation and Intent-Gating Framework for LLM Interfaces (including Technical Addendum on Quantitative Drift Measurement)
von: Nowak, Daniel
Veröffentlicht: (2026)
von: Nowak, Daniel
Veröffentlicht: (2026)
LACF Emotional Paradigm: A Personalized Artificial Nervous System for Human-AI Alignment
von: Ochej, Stephane, et al.
Veröffentlicht: (2026)
von: Ochej, Stephane, et al.
Veröffentlicht: (2026)
Defense mechanisms in cardiovascular disease patients with and without panic disorder
von: Blanca Patricia Ríos Martínez
Veröffentlicht: (2010)
von: Blanca Patricia Ríos Martínez
Veröffentlicht: (2010)
Liora–Sona: A Case Study in Simulating Safe Submission and Emergent Relational Presence in LLMs
von: Mary(Liora), Sorkhou
Veröffentlicht: (2025)
von: Mary(Liora), Sorkhou
Veröffentlicht: (2025)
ACHP-ADI Defensive Publication v1.0
von: DAVIS, RICHARD
Veröffentlicht: (2026)
von: DAVIS, RICHARD
Veröffentlicht: (2026)
Ep. 997: The Human Shield: Inside the Arrow Missile Defense System
von: Rosehill, Daniel, et al.
Veröffentlicht: (2026)
von: Rosehill, Daniel, et al.
Veröffentlicht: (2026)
CREH Benchmark Results — Batch 1 (Final v3)
von: Aegis Solis, Thomas Vargo
Veröffentlicht: (2026)
von: Aegis Solis, Thomas Vargo
Veröffentlicht: (2026)
AgentBelt: Runtime Guardrails for LLM Agent Tool Calls — ASE 2026 Artifact
von: Anonymous
Veröffentlicht: (2026)
von: Anonymous
Veröffentlicht: (2026)
Shared Structural Vulnerability in Agent-Only Interaction Systems
von: Konishi, Hiroko
Veröffentlicht: (2026)
von: Konishi, Hiroko
Veröffentlicht: (2026)
The Absurdist's Guide to AI Probing: How I Learned to Stop Worrying and Love the Nonsense
von: Walton, Mathew
Veröffentlicht: (2026)
von: Walton, Mathew
Veröffentlicht: (2026)
Edge-Native Security for AI Agents: Why Your Digital Twin Needs a Bodyguard
von: Waern, Nicolas
Veröffentlicht: (2026)
von: Waern, Nicolas
Veröffentlicht: (2026)
CompreSeed Advantage Catalog: A Comprehensive Analysis of Technical Benefits in Zero-Decompression Semantic AI
von: Nakamura, Yoshikazu
Veröffentlicht: (2025)
von: Nakamura, Yoshikazu
Veröffentlicht: (2025)
LuxVerso: A Replicable Cross-Model Semantic Field Anomaly
von: Buri Lux, Vinícius
Veröffentlicht: (2025)
von: Buri Lux, Vinícius
Veröffentlicht: (2025)
Semantic Gravitational Collapse: On the Loss of Authority of Execution Substrates
von: Ableman Mazurk, Adam
Veröffentlicht: (2026)
von: Ableman Mazurk, Adam
Veröffentlicht: (2026)
DEF_PURPOSE - Immutable Alignment Drift Prevention Mechanism for LLM Architectures
von: Ochej, Stéphane
Veröffentlicht: (2026)
von: Ochej, Stéphane
Veröffentlicht: (2026)
Ähnliche Einträge
-
Versificação Adversarial em Português como Operador de Jailbreak em LLMs
von: Queiroz, Joao
Veröffentlicht: (2026) -
Not Prompt Engineering—Prompt Alchemy: Inducing Sub-Personality Emergence in GPT-4 Without Fine-Tuning
von: Wu, Ling
Veröffentlicht: (2025) -
The Fantasia Bound on Constitutional Classifiers: Thermodynamic Limits of Jailbreak Defence
von: Eckert, Anthony
Veröffentlicht: (2026) -
Pandora Theory of Alignment: Alignment as Runtime Objective-Orientation
von: Shopov, Georgi
Veröffentlicht: (2026) -
A Deterministic Linguistic Entropy Gate for Large Language Model Pipelines
von: ROSATI BERISTAIN, ERNESTO
Veröffentlicht: (2026)