DynaGuard: A Dynamic Guardian Model With User-Defined Policies
Fuente:
arXiv
Saved in:
| Main Authors: | Hoover, Monte, Baherwani, Vatsal, Jain, Neel, Saifullah, Khalid, Vincent, Joseph, Jain, Chirag, Rad, Melissa Kazemi, Bruss, C. Bayan, Panda, Ashwinee, Goldstein, Tom |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
by: Panda, Ashwinee, et al.
Published: (2025)
by: Panda, Ashwinee, et al.
Published: (2025)
Analysis of Attention in Video Diffusion Transformers
by: Wen, Yuxin, et al.
Published: (2025)
by: Wen, Yuxin, et al.
Published: (2025)
Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models
by: Jain, Neel, et al.
Published: (2024)
by: Jain, Neel, et al.
Published: (2024)
Characterizing Motion Encoding in Video Diffusion Timesteps
by: Baherwani, Vatsal, et al.
Published: (2025)
by: Baherwani, Vatsal, et al.
Published: (2025)
LoRI: Reducing Cross-Task Interference in Multi-Task Low-Rank Adaptation
by: Zhang, Juzheng, et al.
Published: (2025)
by: Zhang, Juzheng, et al.
Published: (2025)
Exploiting Sparsity for Long Context Inference: Million Token Contexts on Commodity GPUs
by: Synk, Ryan, et al.
Published: (2025)
by: Synk, Ryan, et al.
Published: (2025)
Speculating Experts Accelerates Inference for Mixture-of-Experts
by: Madan, Vivan, et al.
Published: (2026)
by: Madan, Vivan, et al.
Published: (2026)
Zero-shot Multivariate Time Series Forecasting Using Tabular Prior Fitted Networks
by: Jayawardhana, Mayuka, et al.
Published: (2026)
by: Jayawardhana, Mayuka, et al.
Published: (2026)
Multi-Token Prediction via Self-Distillation
by: Kirchenbauer, John, et al.
Published: (2026)
by: Kirchenbauer, John, et al.
Published: (2026)
FineGRAIN: Evaluating Failure Modes of Text-to-Image Models with Vision Language Model Judges
by: Hayes, Kevin David, et al.
Published: (2025)
by: Hayes, Kevin David, et al.
Published: (2025)
Coercing LLMs to do and reveal (almost) anything
by: Geiping, Jonas, et al.
Published: (2024)
by: Geiping, Jonas, et al.
Published: (2024)
A Simple Baseline for Predicting Events with Auto-Regressive Tabular Transformers
by: Stein, Alex, et al.
Published: (2024)
by: Stein, Alex, et al.
Published: (2024)
Identifying and Evaluating Inactive Heads in Pretrained LLMs
by: Sandoval-Segura, Pedro, et al.
Published: (2025)
by: Sandoval-Segura, Pedro, et al.
Published: (2025)
WatchGuardian: Enabling User-Defined Personalized Just-in-Time Intervention on Smartwatch
by: Lei, Ying, et al.
Published: (2025)
by: Lei, Ying, et al.
Published: (2025)
GenQA: Generating Millions of Instructions from a Handful of Prompts
by: Chen, Jiuhai, et al.
Published: (2024)
by: Chen, Jiuhai, et al.
Published: (2024)
AI versus AI in Financial Crimes and Detection: GenAI Crime Waves to Co-Evolutionary AI
by: Kurshan, Eren, et al.
Published: (2024)
by: Kurshan, Eren, et al.
Published: (2024)
Integrating Sequential and Relational Modeling for User Events: Datasets and Prediction Tasks
by: Fathony, Rizal, et al.
Published: (2025)
by: Fathony, Rizal, et al.
Published: (2025)
Exotic Lovelock black holes and extended quasitopological electromagnetism
by: Ali, Askar, et al.
Published: (2025)
by: Ali, Askar, et al.
Published: (2025)
Power-Yang-Mills black holes and black branes in quartic quasi-topological gravity
by: Ali, Askar, et al.
Published: (2022)
by: Ali, Askar, et al.
Published: (2022)
Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
by: Hans, Abhimanyu, et al.
Published: (2024)
by: Hans, Abhimanyu, et al.
Published: (2024)
CinePile: A Long Video Question Answering Dataset and Benchmark
by: Rawal, Ruchit, et al.
Published: (2024)
by: Rawal, Ruchit, et al.
Published: (2024)
BEDTime: A Unified Benchmark for Automatically Describing Time Series
by: Sen, Medhasweta, et al.
Published: (2025)
by: Sen, Medhasweta, et al.
Published: (2025)
Who Guards the Guardians? The Challenges of Evaluating Identifiability of Learned Representations
by: Joshi, Shruti, et al.
Published: (2026)
by: Joshi, Shruti, et al.
Published: (2026)
On the Coverage Required for Diploid Genome Assembly
by: Mahajan, Daanish, et al.
Published: (2024)
by: Mahajan, Daanish, et al.
Published: (2024)
Sectoral Stock Market Volatility Under Climate Policy Uncertainty
by: Chirag Jain, et al.
Published: (2026)
by: Chirag Jain, et al.
Published: (2026)
LiveBench: A Challenging, Contamination-Limited LLM Benchmark
by: White, Colin, et al.
Published: (2024)
by: White, Colin, et al.
Published: (2024)
matlab codes for Adaptive Fractional-Order MPC for Boost Converters: Real-Time Implementation via Grey Wolf Optimization paper.
by: Khalid, Saifullah, et al.
Published: (2026)
by: Khalid, Saifullah, et al.
Published: (2026)
Multi-Agent Coordination in Autonomous Vehicle Routing: A Simulation-Based Study of Communication, Memory, and Routing Loops
by: Saifullah, KM Khalid, et al.
Published: (2025)
by: Saifullah, KM Khalid, et al.
Published: (2025)
Gemstones: A Model Suite for Multi-Faceted Scaling Laws
by: McLeish, Sean, et al.
Published: (2025)
by: McLeish, Sean, et al.
Published: (2025)
Exploring Zero-Shot App Review Classification with ChatGPT: Challenges and Potential
by: Chaudhary, Mohit, et al.
Published: (2025)
by: Chaudhary, Mohit, et al.
Published: (2025)
Modeling and Predicting Multi-Turn Answer Instability in Large Language Models
by: He, Jiahang, et al.
Published: (2025)
by: He, Jiahang, et al.
Published: (2025)
A Unified Approach to Memory-Sample Tradeoffs for Detecting Planted Structures
by: Garg, Sumegha, et al.
Published: (2026)
by: Garg, Sumegha, et al.
Published: (2026)
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
by: Geiping, Jonas, et al.
Published: (2025)
by: Geiping, Jonas, et al.
Published: (2025)
Surface Characterization, High‐Temperature Tribological Properties of Cu‐Based Tungsten Developed through Microwave Radiations
by: Khalid Bashir, et al.
Published: (2025)
by: Khalid Bashir, et al.
Published: (2025)
Sentiment Analysis in Software Engineering: Evaluating Generative Pre-trained Transformers
by: Saifullah, KM Khalid, et al.
Published: (2025)
by: Saifullah, KM Khalid, et al.
Published: (2025)
The Privacy Guardian Agent: Towards Trustworthy AI Privacy Agents
by: Freiberger, Vincent
Published: (2026)
by: Freiberger, Vincent
Published: (2026)
Influencia de la presión y la temperatura en la reacción de hidroformilación en medio bifásico del 1-hexeno en régimen continuo con el complejo catalítico [(µ-Pz)(CO)(TFFTS)Rh]2 (I)
by: Arnoldo Bruss
Published: (2006)
by: Arnoldo Bruss
Published: (2006)
Modulation by context of a scene in monkey anterior inferotemporal cortex during a saccadic eye movement task
by: Bruss Lima
Published: (2003)
by: Bruss Lima
Published: (2003)
What do we learn from inverting CLIP models?
by: Kazemi, Hamid, et al.
Published: (2024)
by: Kazemi, Hamid, et al.
Published: (2024)
FAST: Factorizable Attention for Speeding up Transformers
by: Gerami, Armin, et al.
Published: (2024)
by: Gerami, Armin, et al.
Published: (2024)
Similar Items
-
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
by: Panda, Ashwinee, et al.
Published: (2025) -
Analysis of Attention in Video Diffusion Transformers
by: Wen, Yuxin, et al.
Published: (2025) -
Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models
by: Jain, Neel, et al.
Published: (2024) -
Characterizing Motion Encoding in Video Diffusion Timesteps
by: Baherwani, Vatsal, et al.
Published: (2025) -
LoRI: Reducing Cross-Task Interference in Multi-Task Low-Rank Adaptation
by: Zhang, Juzheng, et al.
Published: (2025)