From Seed to Harvest: Augmenting Human Creativity with AI for Red-teaming Text-to-Image Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Quaye, Jessica, Rastogi, Charvi, Parrish, Alicia, Inel, Oana, Kahng, Minsuk, Aroyo, Lora, Reddi, Vijay Janapa |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Adversarial Nibbler: An Open Red-Teaming Method for Identifying Diverse Harms in Text-to-Image Generation
por: Quaye, Jessica, et al.
Publicado: (2024)
por: Quaye, Jessica, et al.
Publicado: (2024)
Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South
por: Rastogi, Charvi, et al.
Publicado: (2026)
por: Rastogi, Charvi, et al.
Publicado: (2026)
TinyTorch: Building Machine Learning Systems from First Principles
por: Reddi, Vijay Janapa
Publicado: (2026)
por: Reddi, Vijay Janapa
Publicado: (2026)
Generative AI Agents in Autonomous Machines: A Safety Perspective
por: Jabbour, Jason, et al.
Publicado: (2024)
por: Jabbour, Jason, et al.
Publicado: (2024)
Whose View of Safety? A Deep DIVE Dataset for Pluralistic Alignment of Text-to-Image Models
por: Rastogi, Charvi, et al.
Publicado: (2025)
por: Rastogi, Charvi, et al.
Publicado: (2025)
"Just a strange pic": Evaluating 'safety' in GenAI Image safety annotation tasks from diverse annotators' perspectives
por: Wang, Ding, et al.
Publicado: (2025)
por: Wang, Ding, et al.
Publicado: (2025)
Decoding Safety Feedback from Diverse Raters: A Data-driven Lens on Responsiveness to Severity
por: Mishra, Pushkar, et al.
Publicado: (2025)
por: Mishra, Pushkar, et al.
Publicado: (2025)
The Magnificent Seven Challenges and Opportunities in Domain-Specific Accelerator Design for Autonomous Systems
por: Neuman, Sabrina M., et al.
Publicado: (2024)
por: Neuman, Sabrina M., et al.
Publicado: (2024)
Data-Prompt Co-Evolution: Growing Test Sets to Refine LLM Behavior
por: Lee, Minjae, et al.
Publicado: (2025)
por: Lee, Minjae, et al.
Publicado: (2025)
Insights on Disagreement Patterns in Multimodal Safety Perception across Diverse Rater Groups
por: Rastogi, Charvi, et al.
Publicado: (2024)
por: Rastogi, Charvi, et al.
Publicado: (2024)
SocratiQ: A Generative AI-Powered Learning Companion for Personalized Education and Broader Accessibility
por: Jabbour, Jason, et al.
Publicado: (2025)
por: Jabbour, Jason, et al.
Publicado: (2025)
FedStaleWeight: Buffered Asynchronous Federated Learning with Fair Aggregation via Staleness Reweighting
por: Ma, Jeffrey, et al.
Publicado: (2024)
por: Ma, Jeffrey, et al.
Publicado: (2024)
Automatic Histograms: Leveraging Language Models for Text Dataset Exploration
por: Reif, Emily, et al.
Publicado: (2024)
por: Reif, Emily, et al.
Publicado: (2024)
When Cars Have Stereotypes: Auditing Demographic Bias in Objects from Text-to-Image Models
por: Choi, Dasol, et al.
Publicado: (2025)
por: Choi, Dasol, et al.
Publicado: (2025)
Understanding the Dataset Practitioners Behind Large Language Model Development
por: Qian, Crystal, et al.
Publicado: (2024)
por: Qian, Crystal, et al.
Publicado: (2024)
VLSlice: Interactive Vision-and-Language Slice Discovery
por: Slyman, Eric, et al.
Publicado: (2023)
por: Slyman, Eric, et al.
Publicado: (2023)
Materiality and Risk in the Age of Pervasive AI Sensors
por: Sloane, Mona, et al.
Publicado: (2024)
por: Sloane, Mona, et al.
Publicado: (2024)
Multi-turn Evaluation of Anthropomorphic Behaviours in Large Language Models
por: Ibrahim, Lujain, et al.
Publicado: (2025)
por: Ibrahim, Lujain, et al.
Publicado: (2025)
Tabula: Efficiently Computing Nonlinear Activation Functions for Secure Neural Network Inference
por: Lam, Maximilian, et al.
Publicado: (2022)
por: Lam, Maximilian, et al.
Publicado: (2022)
SAFEPATH: Preventing Harmful Reasoning in Chain-of-Thought via Early Alignment
por: Jeung, Wonje, et al.
Publicado: (2025)
por: Jeung, Wonje, et al.
Publicado: (2025)
Aligning Object Detector Bounding Boxes with Human Preference
por: Strafforello, Ombretta, et al.
Publicado: (2024)
por: Strafforello, Ombretta, et al.
Publicado: (2024)
From Perception to Decision: Assessing the Role of Chart Types Affordances in High-Level Decision Tasks
por: Li, Yixuan, et al.
Publicado: (2024)
por: Li, Yixuan, et al.
Publicado: (2024)
Whom do Explanations Serve? A Systematic Literature Survey of User Characteristics in Explainable Recommender Systems Evaluation
por: Wardatzky, Kathrin, et al.
Publicado: (2024)
por: Wardatzky, Kathrin, et al.
Publicado: (2024)
D-RDW: Diversity-Driven Random Walks for News Recommender Systems
por: Li, Runze, et al.
Publicado: (2025)
por: Li, Runze, et al.
Publicado: (2025)
Informfully Recommenders -- Reproducibility Framework for Diversity-aware Intra-session Recommendations
por: Heitz, Lucien, et al.
Publicado: (2025)
por: Heitz, Lucien, et al.
Publicado: (2025)
TinyML Security: Exploring Vulnerabilities in Resource-Constrained Machine Learning Systems
por: Huckelberry, Jacob, et al.
Publicado: (2024)
por: Huckelberry, Jacob, et al.
Publicado: (2024)
GRASP: A Disagreement Analysis Framework to Assess Group Associations in Perspectives
por: Prabhakaran, Vinodkumar, et al.
Publicado: (2023)
por: Prabhakaran, Vinodkumar, et al.
Publicado: (2023)
QuArch: A Question-Answering Dataset for AI Agents in Computer Architecture
por: Prakash, Shvetank, et al.
Publicado: (2025)
por: Prakash, Shvetank, et al.
Publicado: (2025)
Multi-Agent Reinforcement Learning for Sample-Efficient Deep Neural Network Mapping
por: Krishnan, Srivatsan, et al.
Publicado: (2025)
por: Krishnan, Srivatsan, et al.
Publicado: (2025)
Adaptive Surrogate Gradients for Sequential Reinforcement Learning in Spiking Neural Networks
por: Berghe, Korneel Van den, et al.
Publicado: (2025)
por: Berghe, Korneel Van den, et al.
Publicado: (2025)
Generative AI in Embodied Systems: System-Level Analysis of Performance, Efficiency and Scalability
por: Wan, Zishen, et al.
Publicado: (2025)
por: Wan, Zishen, et al.
Publicado: (2025)
Who Defines "Best"? Towards Interactive, User-Defined Evaluation of LLM Leaderboards
por: Jung, Minji, et al.
Publicado: (2026)
por: Jung, Minji, et al.
Publicado: (2026)
Slm-mux: Orchestrating small language models for reasoning
por: Wang, Chenyu, et al.
Publicado: (2025)
por: Wang, Chenyu, et al.
Publicado: (2025)
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
por: Tschand, Arya, et al.
Publicado: (2025)
por: Tschand, Arya, et al.
Publicado: (2025)
From Seed to Harvest: Growing the Macdonald Campus Seed Library
por: Ingalls, Dana
Publicado: (2021)
por: Ingalls, Dana
Publicado: (2021)
Paradoxical Vitiligo Induced by Adalimumab in Ankylosing Spondylitis With Uveitis: Clinical Management and Therapeutic Considerations
por: Tuba Yuce Inel
Publicado: (2026)
por: Tuba Yuce Inel
Publicado: (2026)
SWE-fficiency: Can Language Models Optimize Real-World Repositories on Real Workloads?
por: Ma, Jeffrey Jian, et al.
Publicado: (2025)
por: Ma, Jeffrey Jian, et al.
Publicado: (2025)
Interactive Prompt Debugging with Sequence Salience
por: Tenney, Ian, et al.
Publicado: (2024)
por: Tenney, Ian, et al.
Publicado: (2024)
DWARF: Disease-weighted network for attention map refinement
por: Luo, Haozhe, et al.
Publicado: (2024)
por: Luo, Haozhe, et al.
Publicado: (2024)
COSMIC: Enabling Full-Stack Co-Design and Optimization of Distributed Machine Learning Systems
por: Raju, Aditi, et al.
Publicado: (2025)
por: Raju, Aditi, et al.
Publicado: (2025)
Ejemplares similares
-
Adversarial Nibbler: An Open Red-Teaming Method for Identifying Diverse Harms in Text-to-Image Generation
por: Quaye, Jessica, et al.
Publicado: (2024) -
Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South
por: Rastogi, Charvi, et al.
Publicado: (2026) -
TinyTorch: Building Machine Learning Systems from First Principles
por: Reddi, Vijay Janapa
Publicado: (2026) -
Generative AI Agents in Autonomous Machines: A Safety Perspective
por: Jabbour, Jason, et al.
Publicado: (2024) -
Whose View of Safety? A Deep DIVE Dataset for Pluralistic Alignment of Text-to-Image Models
por: Rastogi, Charvi, et al.
Publicado: (2025)