Active Attacks: Red-teaming LLMs via Adaptive Environments
Fuente:
arXiv
Saved in:
| Main Authors: | Yun, Taeyoung, St-Charles, Pierre-Luc, Park, Jinkyoo, Bengio, Yoshua, Kim, Minsu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Local Search GFlowNets
by: Kim, Minsu, et al.
Published: (2023)
by: Kim, Minsu, et al.
Published: (2023)
Guided Trajectory Generation with Diffusion Models for Offline Model-based Optimization
by: Yun, Taeyoung, et al.
Published: (2024)
by: Yun, Taeyoung, et al.
Published: (2024)
GTA: Generative Trajectory Augmentation with Guidance for Offline Reinforcement Learning
by: Lee, Jaewoo, et al.
Published: (2024)
by: Lee, Jaewoo, et al.
Published: (2024)
Adaptive teachers for amortized samplers
by: Kim, Minsu, et al.
Published: (2024)
by: Kim, Minsu, et al.
Published: (2024)
Learning to Scale Logits for Temperature-Conditional GFlowNets
by: Kim, Minsu, et al.
Published: (2023)
by: Kim, Minsu, et al.
Published: (2023)
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
by: Lee, Gyubin, et al.
Published: (2025)
by: Lee, Gyubin, et al.
Published: (2025)
Automated Kernel Discovery Towards Understanding High-dimensional Bayesian Optimization
by: Yun, Taeyoung, et al.
Published: (2026)
by: Yun, Taeyoung, et al.
Published: (2026)
An Offline Meta Black-box Optimization Framework for Adaptive Design of Urban Traffic Light Management Systems
by: Yun, Taeyoung, et al.
Published: (2024)
by: Yun, Taeyoung, et al.
Published: (2024)
Latent Veracity Inference for Identifying Errors in Stepwise Reasoning
by: Kim, Minsu, et al.
Published: (2025)
by: Kim, Minsu, et al.
Published: (2025)
Ant Colony Sampling with GFlowNets for Combinatorial Optimization
by: Kim, Minsu, et al.
Published: (2024)
by: Kim, Minsu, et al.
Published: (2024)
Improved Off-policy Reinforcement Learning in Biological Sequence Design
by: Kim, Hyeonah, et al.
Published: (2024)
by: Kim, Hyeonah, et al.
Published: (2024)
Analysis of Fourier Neural Operators via Effective Field Theory
by: Kim, Taeyoung
Published: (2025)
by: Kim, Taeyoung
Published: (2025)
Aligning Few-Step Generative Models by Amortizing Sample-based Variational Inference
by: Lee, Jaewoo, et al.
Published: (2026)
by: Lee, Jaewoo, et al.
Published: (2026)
Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning
by: Song, Chihyeon, et al.
Published: (2025)
by: Song, Chihyeon, et al.
Published: (2025)
Self-Evolving Curriculum for LLM Reasoning
by: Chen, Xiaoyin, et al.
Published: (2025)
by: Chen, Xiaoyin, et al.
Published: (2025)
ANT: Adaptive Noise Schedule for Time Series Diffusion Models
by: Lee, Seunghan, et al.
Published: (2024)
by: Lee, Seunghan, et al.
Published: (2024)
In-Context Parametric Inference: Point or Distribution Estimators?
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
Loop Corrections to the Training Error and Generalization Gap of Random Feature Models
by: Kim, Taeyoung
Published: (2026)
by: Kim, Taeyoung
Published: (2026)
Red-teaming Activation Probes using Prompted LLMs
by: Blandfort, Phil, et al.
Published: (2025)
by: Blandfort, Phil, et al.
Published: (2025)
Offline Model-Based Optimization: Comprehensive Review
by: Kim, Minsu, et al.
Published: (2025)
by: Kim, Minsu, et al.
Published: (2025)
GFlowNet Foundations
by: Bengio, Yoshua, et al.
Published: (2021)
by: Bengio, Yoshua, et al.
Published: (2021)
To Predict or Not To Predict? Proportionally Masked Autoencoders for Tabular Data Imputation
by: Kim, Jungkyu, et al.
Published: (2024)
by: Kim, Jungkyu, et al.
Published: (2024)
Monte Carlo Tree Diffusion for System 2 Planning
by: Yoon, Jaesik, et al.
Published: (2025)
by: Yoon, Jaesik, et al.
Published: (2025)
Shaping Inductive Bias in Diffusion Models through Frequency-Based Noise Control
by: Jiralerspong, Thomas, et al.
Published: (2025)
by: Jiralerspong, Thomas, et al.
Published: (2025)
In-Context Reinforcement Learning through Bayesian Fusion of Context and Value Prior
by: Berkes, Anaïs, et al.
Published: (2026)
by: Berkes, Anaïs, et al.
Published: (2026)
A Complexity-Based Theory of Compositionality
by: Elmoznino, Eric, et al.
Published: (2024)
by: Elmoznino, Eric, et al.
Published: (2024)
Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?
by: Bengio, Yoshua, et al.
Published: (2025)
by: Bengio, Yoshua, et al.
Published: (2025)
FALCON: Few-step Accurate Likelihoods for Continuous Flows
by: Rehman, Danyal, et al.
Published: (2025)
by: Rehman, Danyal, et al.
Published: (2025)
Were RNNs All We Needed?
by: Feng, Leo, et al.
Published: (2024)
by: Feng, Leo, et al.
Published: (2024)
Why Rectified Power Unit Networks Fail and How to Improve It: An Effective Field Theory Perspective
by: Kim, Taeyoung, et al.
Published: (2024)
by: Kim, Taeyoung, et al.
Published: (2024)
Efficient Causal Graph Discovery Using Large Language Models
by: Jiralerspong, Thomas, et al.
Published: (2024)
by: Jiralerspong, Thomas, et al.
Published: (2024)
Unlearning via Sparse Representations
by: Shah, Vedant, et al.
Published: (2023)
by: Shah, Vedant, et al.
Published: (2023)
Expert-Guided LLM Reasoning for Battery Discovery: From AI-Driven Hypothesis to Synthesis and Characterization
by: Liu, Shengchao, et al.
Published: (2025)
by: Liu, Shengchao, et al.
Published: (2025)
Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
Learning What Matters: Steering Diffusion via Spectrally Anisotropic Forward Noise
by: Scimeca, Luca, et al.
Published: (2025)
by: Scimeca, Luca, et al.
Published: (2025)
Channel Normalization for Time Series Channel Identification
by: Lee, Seunghan, et al.
Published: (2025)
by: Lee, Seunghan, et al.
Published: (2025)
Dataset-Driven Channel Masks in Transformers for Multivariate Time Series
by: Lee, Seunghan, et al.
Published: (2024)
by: Lee, Seunghan, et al.
Published: (2024)
Soft Contrastive Learning for Time Series
by: Lee, Seunghan, et al.
Published: (2023)
by: Lee, Seunghan, et al.
Published: (2023)
Learning to Embed Time Series Patches Independently
by: Lee, Seunghan, et al.
Published: (2023)
by: Lee, Seunghan, et al.
Published: (2023)
Diffusion Fine-Tuning via Reparameterized Policy Gradient of the Soft Q-Function
by: Kang, Hyeongyu, et al.
Published: (2025)
by: Kang, Hyeongyu, et al.
Published: (2025)
Similar Items
-
Local Search GFlowNets
by: Kim, Minsu, et al.
Published: (2023) -
Guided Trajectory Generation with Diffusion Models for Offline Model-based Optimization
by: Yun, Taeyoung, et al.
Published: (2024) -
GTA: Generative Trajectory Augmentation with Guidance for Offline Reinforcement Learning
by: Lee, Jaewoo, et al.
Published: (2024) -
Adaptive teachers for amortized samplers
by: Kim, Minsu, et al.
Published: (2024) -
Learning to Scale Logits for Temperature-Conditional GFlowNets
by: Kim, Minsu, et al.
Published: (2023)