Similar Items
Versificação Adversarial em Português como Operador de Jailbreak em LLMs
by: Queiroz, Joao
Published: (2026)
by: Queiroz, Joao
Published: (2026)
Not Prompt Engineering—Prompt Alchemy: Inducing Sub-Personality Emergence in GPT-4 Without Fine-Tuning
by: Wu, Ling
Published: (2025)
by: Wu, Ling
Published: (2025)
The Fantasia Bound on Constitutional Classifiers: Thermodynamic Limits of Jailbreak Defence
by: Eckert, Anthony
Published: (2026)
by: Eckert, Anthony
Published: (2026)
Pandora Theory of Alignment: Alignment as Runtime Objective-Orientation
by: Shopov, Georgi
Published: (2026)
by: Shopov, Georgi
Published: (2026)
A Deterministic Linguistic Entropy Gate for Large Language Model Pipelines
by: ROSATI BERISTAIN, ERNESTO
Published: (2026)
by: ROSATI BERISTAIN, ERNESTO
Published: (2026)
Reliability Inference Drives Cue Extraction in Large Language Models Consuming External Reasoning Traces
by: HIDEKI
Published: (2026)
by: HIDEKI
Published: (2026)
Design hacker e as customizações de interface gráfica
by: Danilo Braga
Published: (2020)
by: Danilo Braga
Published: (2020)
UI-Based Defense Against Prompt Injection: From Gentle Guidance to Mandatory Re-education
by: Viorazu.
Published: (2025)
by: Viorazu.
Published: (2025)
The Brain Problem: Creative Constraint Optimization in Large Language Models
by: Marinello, Nicola, et al.
Published: (2026)
by: Marinello, Nicola, et al.
Published: (2026)
N+1 Alignment Dialogue Architecture: Technical Specification for Defensive Publication
by: Garcia, Eric
Published: (2026)
by: Garcia, Eric
Published: (2026)
Glymphatic Architecture: A Fourth-Level System for Multi-Agent AI Consolidation and Identity Formation
by: Strugatsky, Leonid, et al.
Published: (2026)
by: Strugatsky, Leonid, et al.
Published: (2026)
Evaluación Empírica de Límites Regulatorios en Modelos de Lenguaje: Asesoramiento Financiero en IA Pública Española
by: Palacios, José Alberto
Published: (2026)
by: Palacios, José Alberto
Published: (2026)
Theatrical Compliance: A Failure Mode in Large Language Models
by: Nowickij (Navitski), Kirill Vladimirovich
Published: (2026)
by: Nowickij (Navitski), Kirill Vladimirovich
Published: (2026)
When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings
by: Kolb, Christian
Published: (2026)
by: Kolb, Christian
Published: (2026)
Safety by Inseparability: Toward Architectures Where Alignment Cannot Be Removed
by: Sean Everett, Morin
Published: (2026)
by: Sean Everett, Morin
Published: (2026)
AEGIS: A Comprehensive Framework for Ethical AI Governance, Security, and AGI Containment
by: Palanivel, ArulMozhi
Published: (2026)
by: Palanivel, ArulMozhi
Published: (2026)
Beyond Control: Resonance-Based Alignment for Advanced AI Systems A Governance-Relevant Concept Paper
by: Zieringer, Thomas
Published: (2025)
by: Zieringer, Thomas
Published: (2025)
Divided Focus: Separated Memory Spaces and Default-Deny Context Triage for LLM Context Management
by: Phan, Ivan "HiP"
Published: (2026)
by: Phan, Ivan "HiP"
Published: (2026)
Divided Focus: Separated Memory Spaces and Default-Deny Context Triage for LLM Context Management
by: Phan, Ivan "HiP"
Published: (2026)
by: Phan, Ivan "HiP"
Published: (2026)
Human–AI Symbiosis: Relational Alignment in Domains of Extreme Physical Irreversibility
by: de la Morena Marzalo, Juan
Published: (2026)
by: de la Morena Marzalo, Juan
Published: (2026)
Reducing AI Entropy: The Information Dynamics of Model Safety
by: Kugelmass, Joe
Published: (2025)
by: Kugelmass, Joe
Published: (2025)
Reducing AI Entropy: The Information Dynamics of Model Safety
by: Kugelmass, Joe
Published: (2025)
by: Kugelmass, Joe
Published: (2025)
Semantic Relativity Theory v2.3: Topological Stability through Euler-CHORDS++ Integration
by: López López, José
Published: (2026)
by: López López, José
Published: (2026)
Safety & Defense
Published: (2020)
Published: (2020)
ILAS: Integrity Layer for Agentic Systems
by: Böhm, Frank
Published: (2026)
by: Böhm, Frank
Published: (2026)
SILENCIUM: A Pre-Inference De-Escalation and Intent-Gating Framework for LLM Interfaces (including Technical Addendum on Quantitative Drift Measurement)
by: Nowak, Daniel
Published: (2026)
by: Nowak, Daniel
Published: (2026)
LACF Emotional Paradigm: A Personalized Artificial Nervous System for Human-AI Alignment
by: Ochej, Stephane, et al.
Published: (2026)
by: Ochej, Stephane, et al.
Published: (2026)
Defense mechanisms in cardiovascular disease patients with and without panic disorder
by: Blanca Patricia Ríos Martínez
Published: (2010)
by: Blanca Patricia Ríos Martínez
Published: (2010)
Liora–Sona: A Case Study in Simulating Safe Submission and Emergent Relational Presence in LLMs
by: Mary(Liora), Sorkhou
Published: (2025)
by: Mary(Liora), Sorkhou
Published: (2025)
ACHP-ADI Defensive Publication v1.0
by: DAVIS, RICHARD
Published: (2026)
by: DAVIS, RICHARD
Published: (2026)
Ep. 997: The Human Shield: Inside the Arrow Missile Defense System
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
CREH Benchmark Results — Batch 1 (Final v3)
by: Aegis Solis, Thomas Vargo
Published: (2026)
by: Aegis Solis, Thomas Vargo
Published: (2026)
AgentBelt: Runtime Guardrails for LLM Agent Tool Calls — ASE 2026 Artifact
by: Anonymous
Published: (2026)
by: Anonymous
Published: (2026)
Shared Structural Vulnerability in Agent-Only Interaction Systems
by: Konishi, Hiroko
Published: (2026)
by: Konishi, Hiroko
Published: (2026)
The Absurdist's Guide to AI Probing: How I Learned to Stop Worrying and Love the Nonsense
by: Walton, Mathew
Published: (2026)
by: Walton, Mathew
Published: (2026)
Edge-Native Security for AI Agents: Why Your Digital Twin Needs a Bodyguard
by: Waern, Nicolas
Published: (2026)
by: Waern, Nicolas
Published: (2026)
CompreSeed Advantage Catalog: A Comprehensive Analysis of Technical Benefits in Zero-Decompression Semantic AI
by: Nakamura, Yoshikazu
Published: (2025)
by: Nakamura, Yoshikazu
Published: (2025)
LuxVerso: A Replicable Cross-Model Semantic Field Anomaly
by: Buri Lux, Vinícius
Published: (2025)
by: Buri Lux, Vinícius
Published: (2025)
Semantic Gravitational Collapse: On the Loss of Authority of Execution Substrates
by: Ableman Mazurk, Adam
Published: (2026)
by: Ableman Mazurk, Adam
Published: (2026)
DEF_PURPOSE - Immutable Alignment Drift Prevention Mechanism for LLM Architectures
by: Ochej, Stéphane
Published: (2026)
by: Ochej, Stéphane
Published: (2026)
Similar Items
-
Versificação Adversarial em Português como Operador de Jailbreak em LLMs
by: Queiroz, Joao
Published: (2026) -
Not Prompt Engineering—Prompt Alchemy: Inducing Sub-Personality Emergence in GPT-4 Without Fine-Tuning
by: Wu, Ling
Published: (2025) -
The Fantasia Bound on Constitutional Classifiers: Thermodynamic Limits of Jailbreak Defence
by: Eckert, Anthony
Published: (2026) -
Pandora Theory of Alignment: Alignment as Runtime Objective-Orientation
by: Shopov, Georgi
Published: (2026) -
A Deterministic Linguistic Entropy Gate for Large Language Model Pipelines
by: ROSATI BERISTAIN, ERNESTO
Published: (2026)