Salvato in:
| Autori principali: | Kim, Chaeyun, Lim, YongTaek, Kim, Kihyun, Kim, Junghwan, Kim, Minwoo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.20170 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming
di: Jung, MinJae, et al.
Pubblicazione: (2026)
di: Jung, MinJae, et al.
Pubblicazione: (2026)
PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI
di: Deng, Wesley Hanwen, et al.
Pubblicazione: (2026)
di: Deng, Wesley Hanwen, et al.
Pubblicazione: (2026)
EPLKG: Efficient Prompt Learning with Knowledge Graph
di: Lim, YongTaek, et al.
Pubblicazione: (2023)
di: Lim, YongTaek, et al.
Pubblicazione: (2023)
Nunchi-Bench: Benchmarking Language Models on Cultural Reasoning with a Focus on Korean Superstition
di: Kim, Kyuhee, et al.
Pubblicazione: (2025)
di: Kim, Kyuhee, et al.
Pubblicazione: (2025)
The ASIR Courage Model: A Phase-Dynamic Framework for Truth Transitions in Human and AI Systems
di: Kim, Hyo Jin
Pubblicazione: (2026)
di: Kim, Hyo Jin
Pubblicazione: (2026)
Red Teaming AI Red Teaming
di: Majumdar, Subhabrata, et al.
Pubblicazione: (2025)
di: Majumdar, Subhabrata, et al.
Pubblicazione: (2025)
Addressing Selection Bias in Computerized Adaptive Testing: A User-Wise Aggregate Influence Function Approach
di: Kwon, Soonwoo, et al.
Pubblicazione: (2023)
di: Kwon, Soonwoo, et al.
Pubblicazione: (2023)
A Study on the Framework for Evaluating the Ethics and Trustworthiness of Generative AI
di: Jeong, Cheonsu, et al.
Pubblicazione: (2025)
di: Jeong, Cheonsu, et al.
Pubblicazione: (2025)
Designing Ethical Learning for Agentic AI: Toegye Yi Hwang's Ethical Emotion Regulation Framework
di: Kim, Ji Yeon
Pubblicazione: (2026)
di: Kim, Ji Yeon
Pubblicazione: (2026)
SAIF: A Comprehensive Framework for Evaluating the Risks of Generative AI in the Public Sector
di: Lee, Kyeongryul, et al.
Pubblicazione: (2025)
di: Lee, Kyeongryul, et al.
Pubblicazione: (2025)
Accuracy of a Large Language Model in Distinguishing Anti- And Pro-vaccination Messages on Social Media: The Case of Human Papillomavirus Vaccination
di: Kim, Soojong, et al.
Pubblicazione: (2024)
di: Kim, Soojong, et al.
Pubblicazione: (2024)
Geometry-Correct Diffusion Posterior Sampling with Denoiser-Pullback Curvature Guidance and Manifold-Aligned Damping
di: Shin, Seunghyeok, et al.
Pubblicazione: (2026)
di: Shin, Seunghyeok, et al.
Pubblicazione: (2026)
AgentSUMO: An Agentic Framework for Interactive Simulation Scenario Generation in SUMO via Large Language Models
di: Jeong, Minwoo, et al.
Pubblicazione: (2025)
di: Jeong, Minwoo, et al.
Pubblicazione: (2025)
Model Fusion through Bayesian Optimization in Language Model Fine-Tuning
di: Jang, Chaeyun, et al.
Pubblicazione: (2024)
di: Jang, Chaeyun, et al.
Pubblicazione: (2024)
Benign-to-Toxic Jailbreaking: Inducing Harmful Responses from Harmless Prompts
di: Kim, Hee-Seon, et al.
Pubblicazione: (2025)
di: Kim, Hee-Seon, et al.
Pubblicazione: (2025)
KTCF: Actionable Recourse in Knowledge Tracing via Counterfactual Explanations for Education
di: Kim, Woojin, et al.
Pubblicazione: (2026)
di: Kim, Woojin, et al.
Pubblicazione: (2026)
Knowledge Tracing in Programming Education Integrating Students' Questions
di: Kim, Doyoun, et al.
Pubblicazione: (2025)
di: Kim, Doyoun, et al.
Pubblicazione: (2025)
Perceptions of Discriminatory Decisions of Artificial Intelligence: Unpacking the Role of Individual Characteristics
di: Kim, Soojong
Pubblicazione: (2024)
di: Kim, Soojong
Pubblicazione: (2024)
How to Capture and Study Conversations Between Research Participants and ChatGPT: GPT for Researchers (g4r.org)
di: Kim, Jin
Pubblicazione: (2025)
di: Kim, Jin
Pubblicazione: (2025)
Susceptibility of Large Language Models to User-Driven Factors in Medical Queries
di: Lim, Kyung Ho, et al.
Pubblicazione: (2025)
di: Lim, Kyung Ho, et al.
Pubblicazione: (2025)
FAST-DIPS: Adjoint-Free Analytic Steps and Hard-Constrained Likelihood Correction for Diffusion-Prior Inverse Problems
di: Kim, Minwoo, et al.
Pubblicazione: (2026)
di: Kim, Minwoo, et al.
Pubblicazione: (2026)
Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs
di: Oh, Gyutaek, et al.
Pubblicazione: (2025)
di: Oh, Gyutaek, et al.
Pubblicazione: (2025)
Anecdoctoring: Automated Red-Teaming Across Language and Place
di: Cuevas, Alejandro, et al.
Pubblicazione: (2025)
di: Cuevas, Alejandro, et al.
Pubblicazione: (2025)
Benchmark Early and Red Team Often: A Framework for Assessing and Managing Dual-Use Hazards of AI Foundation Models
di: Barrett, Anthony M., et al.
Pubblicazione: (2024)
di: Barrett, Anthony M., et al.
Pubblicazione: (2024)
Red Teaming AI Policy: A Taxonomy of Avoision and the EU AI Act
di: Yew, Rui-Jie, et al.
Pubblicazione: (2025)
di: Yew, Rui-Jie, et al.
Pubblicazione: (2025)
Bridging the Missing-Modality Gap: Improving Text-Only Calibration of Vision Language Models
di: Kim, Mingyeong, et al.
Pubblicazione: (2026)
di: Kim, Mingyeong, et al.
Pubblicazione: (2026)
Red Teaming for Generative AI, Report on a Copyright-Focused Exercise Completed in an Academic Medical Center
di: Wen, James, et al.
Pubblicazione: (2025)
di: Wen, James, et al.
Pubblicazione: (2025)
CharDiff-LP: A Diffusion Model with Character-Level Guidance for License Plate Image Restoration
di: Na, Kihyun, et al.
Pubblicazione: (2025)
di: Na, Kihyun, et al.
Pubblicazione: (2025)
The Precautionary Principle and the Innovation Principle: Incompatible Guides for AI Innovation Governance?
di: Kaivanto, Kim
Pubblicazione: (2025)
di: Kaivanto, Kim
Pubblicazione: (2025)
Students' Perceived Roles, Opportunities, and Challenges of a Generative AI-powered Teachable Agent: A Case of Middle School Math Class
di: Song, Yukyeong, et al.
Pubblicazione: (2024)
di: Song, Yukyeong, et al.
Pubblicazione: (2024)
Can You Share Your Story? Modeling Clients' Metacognition and Openness for LLM Therapist Evaluation
di: Kim, Minju, et al.
Pubblicazione: (2025)
di: Kim, Minju, et al.
Pubblicazione: (2025)
Integrating Generative AI in BIM Education: Insights from Classroom Implementation
di: Sahraoui, Islem, et al.
Pubblicazione: (2025)
di: Sahraoui, Islem, et al.
Pubblicazione: (2025)
Rethinking Garment Conditioning in Diffusion-based Virtual Try-On
di: Na, Kihyun, et al.
Pubblicazione: (2025)
di: Na, Kihyun, et al.
Pubblicazione: (2025)
Taxonomy and Analysis of Sensitive User Queries in Generative AI Search
di: Jo, Hwiyeol, et al.
Pubblicazione: (2024)
di: Jo, Hwiyeol, et al.
Pubblicazione: (2024)
RedTopic: Toward Topic-Diverse Red Teaming of Large Language Models
di: Ding, Jiale, et al.
Pubblicazione: (2025)
di: Ding, Jiale, et al.
Pubblicazione: (2025)
RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity
di: Shin, Jisu, et al.
Pubblicazione: (2025)
di: Shin, Jisu, et al.
Pubblicazione: (2025)
Position: The AI Conference Peer Review Crisis Demands Author Feedback and Reviewer Rewards
di: Kim, Jaeho, et al.
Pubblicazione: (2025)
di: Kim, Jaeho, et al.
Pubblicazione: (2025)
Verbalized Confidence Triggers Self-Verification: Emergent Behavior Without Explicit Reasoning Supervision
di: Jang, Chaeyun, et al.
Pubblicazione: (2025)
di: Jang, Chaeyun, et al.
Pubblicazione: (2025)
Can AI be a moral victim? The role of moral patiency and ownership perceptions in ethical judgments of using AI-generated content
di: Choung, Hyesun, et al.
Pubblicazione: (2026)
di: Choung, Hyesun, et al.
Pubblicazione: (2026)
Risk-Adjusted Harm Scoring for Automated Red Teaming for LLMs in Financial Services
di: Dimino, Fabrizio, et al.
Pubblicazione: (2026)
di: Dimino, Fabrizio, et al.
Pubblicazione: (2026)
Documenti analoghi
-
STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming
di: Jung, MinJae, et al.
Pubblicazione: (2026) -
PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI
di: Deng, Wesley Hanwen, et al.
Pubblicazione: (2026) -
EPLKG: Efficient Prompt Learning with Knowledge Graph
di: Lim, YongTaek, et al.
Pubblicazione: (2023) -
Nunchi-Bench: Benchmarking Language Models on Cultural Reasoning with a Focus on Korean Superstition
di: Kim, Kyuhee, et al.
Pubblicazione: (2025) -
The ASIR Courage Model: A Phase-Dynamic Framework for Truth Transitions in Human and AI Systems
di: Kim, Hyo Jin
Pubblicazione: (2026)