Enregistré dans:
| Auteurs principaux: | Lee, Wonjun, Lee, Doehyeon, Choi, Eugene, Yu, Sangyoon, Yousefpour, Ashkan, Park, Haon, Ham, Bumsub, Kim, Suhyun |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2502.04757 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Jailbreaking on Text-to-Video Models via Scene Splitting Strategy
par: Lee, Wonjun, et autres
Publié: (2025)
par: Lee, Wonjun, et autres
Publié: (2025)
Maximizing the Position Embedding for Vision Transformers with Global Average Pooling
par: Lee, Wonjun, et autres
Publié: (2025)
par: Lee, Wonjun, et autres
Publié: (2025)
M2S: Multi-turn to Single-turn jailbreak in Red Teaming for LLMs
par: Ha, Junwoo, et autres
Publié: (2025)
par: Ha, Junwoo, et autres
Publié: (2025)
sudo rm -rf agentic_security
par: Lee, Sejin, et autres
Publié: (2025)
par: Lee, Sejin, et autres
Publié: (2025)
ObjexMT: Objective Extraction and Metacognitive Calibration for LLM-as-a-Judge under Multi-Turn Jailbreaks
par: Kim, Hyunjun, et autres
Publié: (2025)
par: Kim, Hyunjun, et autres
Publié: (2025)
X-Teaming Evolutionary M2S: Automated Discovery of Multi-turn to Single-turn Jailbreak Templates
par: Kim, Hyunjun, et autres
Publié: (2025)
par: Kim, Hyunjun, et autres
Publié: (2025)
Selective Vision is the Challenge for Visual Reasoning: A Benchmark for Visual Argument Understanding
par: Chung, Jiwan, et autres
Publié: (2024)
par: Chung, Jiwan, et autres
Publié: (2024)
AZ-NAS: Assembling Zero-Cost Proxies for Network Architecture Search
par: Lee, Junghyup, et autres
Publié: (2024)
par: Lee, Junghyup, et autres
Publié: (2024)
Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection
par: Lee, Sanghoon, et autres
Publié: (2026)
par: Lee, Sanghoon, et autres
Publié: (2026)
3DPillars: Pillar-based two-stage 3D object detection
par: Noh, Jongyoun, et autres
Publié: (2025)
par: Noh, Jongyoun, et autres
Publié: (2025)
Large Language Models Still Exhibit Bias in Long Text
par: Jeung, Wonje, et autres
Publié: (2024)
par: Jeung, Wonje, et autres
Publié: (2024)
TAS-LoRA: Transformer Architecture Search with Mixture-of-LoRA Experts
par: Jeon, Jeimin, et autres
Publié: (2026)
par: Jeon, Jeimin, et autres
Publié: (2026)
Efficient Few-Shot Neural Architecture Search by Counting the Number of Nonlinear Functions
par: Oh, Youngmin, et autres
Publié: (2024)
par: Oh, Youngmin, et autres
Publié: (2024)
Scheduling Weight Transitions for Quantization-Aware Training
par: Lee, Junghyup, et autres
Publié: (2024)
par: Lee, Junghyup, et autres
Publié: (2024)
Representation Bending for Large Language Model Safety
par: Yousefpour, Ashkan, et autres
Publié: (2025)
par: Yousefpour, Ashkan, et autres
Publié: (2025)
Disentangled Representations for Short-Term and Long-Term Person Re-Identification
par: Eom, Chanho, et autres
Publié: (2024)
par: Eom, Chanho, et autres
Publié: (2024)
UniSAFE: A Comprehensive Benchmark for Safety Evaluation of Unified Multimodal Models
par: Lee, Segyu, et autres
Publié: (2026)
par: Lee, Segyu, et autres
Publié: (2026)
How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?
par: Lee, Seongyun, et autres
Publié: (2024)
par: Lee, Seongyun, et autres
Publié: (2024)
Aligning Large Language Models by On-Policy Self-Judgment
par: Lee, Sangkyu, et autres
Publié: (2024)
par: Lee, Sangkyu, et autres
Publié: (2024)
Toward INT4 Fixed-Point Training via Exploring Quantization Error for Gradients
par: Kim, Dohyung, et autres
Publié: (2024)
par: Kim, Dohyung, et autres
Publié: (2024)
AccuQuant: Simulating Multiple Denoising Steps for Quantizing Diffusion Models
par: Lee, Seunghoon, et autres
Publié: (2025)
par: Lee, Seunghoon, et autres
Publié: (2025)
FYI: Flip Your Images for Dataset Distillation
par: Son, Byunggwan, et autres
Publié: (2024)
par: Son, Byunggwan, et autres
Publié: (2024)
Relational Feature Caching for Accelerating Diffusion Transformers
par: Son, Byunggwan, et autres
Publié: (2026)
par: Son, Byunggwan, et autres
Publié: (2026)
Instance-Aware Group Quantization for Vision Transformers
par: Moon, Jaehyeon, et autres
Publié: (2024)
par: Moon, Jaehyeon, et autres
Publié: (2024)
MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models
par: Lee, Suhyun, et autres
Publié: (2026)
par: Lee, Suhyun, et autres
Publié: (2026)
Improving Visual Token Reduction via Rectifying Distortions for Efficient Multimodal LLM Inference
par: Cho, Hyeonwoo, et autres
Publié: (2026)
par: Cho, Hyeonwoo, et autres
Publié: (2026)
GrowTAS: Progressive Expansion from Small to Large Subnets for Efficient ViT Architecture Search
par: Lee, Hyunju, et autres
Publié: (2025)
par: Lee, Hyunju, et autres
Publié: (2025)
Can Structural Cues Save LLMs? Evaluating Language Models in Massive Document Streams
par: Lee, Yukyung, et autres
Publié: (2026)
par: Lee, Yukyung, et autres
Publié: (2026)
Probabilistic Precision and Recall Towards Reliable Evaluation of Generative Models
par: Park, Dogyun, et autres
Publié: (2023)
par: Park, Dogyun, et autres
Publié: (2023)
EconCausal: A Context-Aware Economic Reasoning Benchmark for Large Language Models
par: Lee, Donggyu, et autres
Publié: (2025)
par: Lee, Donggyu, et autres
Publié: (2025)
Subnet-Aware Dynamic Supernet Training for Neural Architecture Search
par: Jeon, Jeimin, et autres
Publié: (2025)
par: Jeon, Jeimin, et autres
Publié: (2025)
Cerberus: Attribute-based person re-identification using semantic IDs
par: Eom, Chanho, et autres
Publié: (2024)
par: Eom, Chanho, et autres
Publié: (2024)
Large Language Models can Share Images, Too!
par: Lee, Young-Jun, et autres
Publié: (2023)
par: Lee, Young-Jun, et autres
Publié: (2023)
Safety-Aligned Weights Are Not Enough: Refusal-Teacher-Guided Finetuning Enhances Safety and Downstream Performance under Harmful Finetuning Attacks
par: Ham, Seokil, et autres
Publié: (2025)
par: Ham, Seokil, et autres
Publié: (2025)
Eliciting and Analyzing Emergent Misalignment in State-of-the-Art Large Language Models
par: Panpatil, Siddhant, et autres
Publié: (2025)
par: Panpatil, Siddhant, et autres
Publié: (2025)
FLEUR: An Explainable Reference-Free Evaluation Metric for Image Captioning Using a Large Multimodal Model
par: Lee, Yebin, et autres
Publié: (2024)
par: Lee, Yebin, et autres
Publié: (2024)
Benign-to-Toxic Jailbreaking: Inducing Harmful Responses from Harmless Prompts
par: Kim, Hee-Seon, et autres
Publié: (2025)
par: Kim, Hee-Seon, et autres
Publié: (2025)
ChatEXAONEPath: An Expert-level Multimodal Large Language Model for Histopathology Using Whole Slide Images
par: Kim, Sangwook, et autres
Publié: (2025)
par: Kim, Sangwook, et autres
Publié: (2025)
Enhancing Dialogue Speech Recognition with Robust Contextual Awareness via Noise Representation Learning
par: Lee, Wonjun, et autres
Publié: (2024)
par: Lee, Wonjun, et autres
Publié: (2024)
Read Like a Radiologist: Efficient Vision-Language Model for 3D Medical Imaging Interpretation
par: Lee, Changsun, et autres
Publié: (2024)
par: Lee, Changsun, et autres
Publié: (2024)
Documents similaires
-
Jailbreaking on Text-to-Video Models via Scene Splitting Strategy
par: Lee, Wonjun, et autres
Publié: (2025) -
Maximizing the Position Embedding for Vision Transformers with Global Average Pooling
par: Lee, Wonjun, et autres
Publié: (2025) -
M2S: Multi-turn to Single-turn jailbreak in Red Teaming for LLMs
par: Ha, Junwoo, et autres
Publié: (2025) -
sudo rm -rf agentic_security
par: Lee, Sejin, et autres
Publié: (2025) -
ObjexMT: Objective Extraction and Metacognitive Calibration for LLM-as-a-Judge under Multi-Turn Jailbreaks
par: Kim, Hyunjun, et autres
Publié: (2025)