FIGURA: A Modular Prompt Engineering Method for Artistic Figure Photography in Safety-Filtered Text-to-Image Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Cazzaniga, Luca |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Fast Text-Driven Approach for Generating Artistic Content
von: Lupascu, Marian, et al.
Veröffentlicht: (2022)
von: Lupascu, Marian, et al.
Veröffentlicht: (2022)
An Effective Image Copy-Move Forgery Detection Using Entropy Information
von: Jiang, Li, et al.
Veröffentlicht: (2023)
von: Jiang, Li, et al.
Veröffentlicht: (2023)
TraceRouter: Robust Safety for Large Foundation Models via Path-Level Intervention
von: Shi, Chuancheng, et al.
Veröffentlicht: (2026)
von: Shi, Chuancheng, et al.
Veröffentlicht: (2026)
Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
Transcending Fusion: A Multi-Scale Alignment Method for Remote Sensing Image-Text Retrieval
von: Yang, Rui, et al.
Veröffentlicht: (2024)
von: Yang, Rui, et al.
Veröffentlicht: (2024)
A Framework for Critical Evaluation of Text-to-Image Models: Integrating Art Historical Analysis, Artistic Exploration, and Critical Prompt Engineering
von: Foka, Amalia
Veröffentlicht: (2024)
von: Foka, Amalia
Veröffentlicht: (2024)
TextRefiner: Internal Visual Feature as Efficient Refiner for Vision-Language Models Prompt Tuning
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
FakingRecipe: Detecting Fake News on Short Video Platforms from the Perspective of Creative Process
von: Bu, Yuyan, et al.
Veröffentlicht: (2024)
von: Bu, Yuyan, et al.
Veröffentlicht: (2024)
Prompt-aware of Frame Sampling for Efficient Text-Video Retrieval
von: Zhang, Deyu, et al.
Veröffentlicht: (2025)
von: Zhang, Deyu, et al.
Veröffentlicht: (2025)
A User-Friendly Framework for Generating Model-Preferred Prompts in Text-to-Image Synthesis
von: Hei, Nailei, et al.
Veröffentlicht: (2024)
von: Hei, Nailei, et al.
Veröffentlicht: (2024)
WordArt Designer API: User-Driven Artistic Typography Synthesis with Large Language Models on ModelScope
von: He, Jun-Yan, et al.
Veröffentlicht: (2024)
von: He, Jun-Yan, et al.
Veröffentlicht: (2024)
Exploring Phrase-Level Grounding with Text-to-Image Diffusion Model
von: Yang, Danni, et al.
Veröffentlicht: (2024)
von: Yang, Danni, et al.
Veröffentlicht: (2024)
Modularized Zero-shot VQA with Pre-trained Models
von: Cao, Rui, et al.
Veröffentlicht: (2023)
von: Cao, Rui, et al.
Veröffentlicht: (2023)
From Data Deluge to Data Curation: A Filtering-WoRA Paradigm for Efficient Text-based Person Search
von: Sun, Jintao, et al.
Veröffentlicht: (2024)
von: Sun, Jintao, et al.
Veröffentlicht: (2024)
VG-TVP: Multimodal Procedural Planning via Visually Grounded Text-Video Prompting
von: Ilaslan, Muhammet Furkan, et al.
Veröffentlicht: (2024)
von: Ilaslan, Muhammet Furkan, et al.
Veröffentlicht: (2024)
TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree Sequencing
von: Lionar, Stefan, et al.
Veröffentlicht: (2025)
von: Lionar, Stefan, et al.
Veröffentlicht: (2025)
Embedding an Ethical Mind: Aligning Text-to-Image Synthesis via Lightweight Value Optimization
von: Wang, Xingqi, et al.
Veröffentlicht: (2024)
von: Wang, Xingqi, et al.
Veröffentlicht: (2024)
Scaling Prompt Instructed Zero Shot Composed Image Retrieval with Image-Only Data
von: Duan, Yiqun, et al.
Veröffentlicht: (2025)
von: Duan, Yiqun, et al.
Veröffentlicht: (2025)
Beyond Walking: A Large-Scale Image-Text Benchmark for Text-based Person Anomaly Search
von: Yang, Shuyu, et al.
Veröffentlicht: (2024)
von: Yang, Shuyu, et al.
Veröffentlicht: (2024)
Bringing Textual Prompt to AI-Generated Image Quality Assessment
von: Qu, Bowen, et al.
Veröffentlicht: (2024)
von: Qu, Bowen, et al.
Veröffentlicht: (2024)
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation
von: Wu, Xun, et al.
Veröffentlicht: (2024)
von: Wu, Xun, et al.
Veröffentlicht: (2024)
G-Refine: A General Quality Refiner for Text-to-Image Generation
von: Li, Chunyi, et al.
Veröffentlicht: (2024)
von: Li, Chunyi, et al.
Veröffentlicht: (2024)
PDA: Text-Augmented Defense Framework for Robust Vision-Language Models against Adversarial Image Attacks
von: Xu, Jingning, et al.
Veröffentlicht: (2026)
von: Xu, Jingning, et al.
Veröffentlicht: (2026)
Visual Semantic Description Generation with MLLMs for Image-Text Matching
von: Chen, Junyu, et al.
Veröffentlicht: (2025)
von: Chen, Junyu, et al.
Veröffentlicht: (2025)
Noisy-Correspondence Learning for Text-to-Image Person Re-identification
von: Qin, Yang, et al.
Veröffentlicht: (2023)
von: Qin, Yang, et al.
Veröffentlicht: (2023)
GSCodec Studio: A Modular Framework for Gaussian Splat Compression
von: Li, Sicheng, et al.
Veröffentlicht: (2025)
von: Li, Sicheng, et al.
Veröffentlicht: (2025)
Deep Boosting Learning: A Brand-new Cooperative Approach for Image-Text Matching
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
von: Diao, Haiwen, et al.
Veröffentlicht: (2024)
CustomContrast: A Multilevel Contrastive Perspective For Subject-Driven Text-to-Image Customization
von: Chen, Nan, et al.
Veröffentlicht: (2024)
von: Chen, Nan, et al.
Veröffentlicht: (2024)
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection
von: Cheung, Tsun-Hin, et al.
Veröffentlicht: (2024)
von: Cheung, Tsun-Hin, et al.
Veröffentlicht: (2024)
OS-HGAdapter: Open Semantic Hypergraph Adapter for Large Language Models Assisted Entropy-Enhanced Image-Text Alignment
von: Chen, Rongjun, et al.
Veröffentlicht: (2025)
von: Chen, Rongjun, et al.
Veröffentlicht: (2025)
InteractDiffusion: Interaction Control in Text-to-Image Diffusion Models
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2023)
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2023)
MAO: Efficient Model-Agnostic Optimization of Prompt Tuning for Vision-Language Models
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
DPC: Dual-Prompt Collaboration for Tuning Vision-Language Models
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
Cross-Modal and Uni-Modal Soft-Label Alignment for Image-Text Retrieval
von: Huang, Hailang, et al.
Veröffentlicht: (2024)
von: Huang, Hailang, et al.
Veröffentlicht: (2024)
Mitigating GenAI-powered Evidence Pollution for Out-of-Context Multimodal Misinformation Detection
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
Prompt-A-Video: Prompt Your Video Diffusion Model via Preference-Aligned LLM
von: Ji, Yatai, et al.
Veröffentlicht: (2024)
von: Ji, Yatai, et al.
Veröffentlicht: (2024)
Magic3DSketch: Create Colorful 3D Models From Sketch-Based 3D Modeling Guided by Text and Language-Image Pre-Training
von: Zang, Ying, et al.
Veröffentlicht: (2024)
von: Zang, Ying, et al.
Veröffentlicht: (2024)
CBVS: A Large-Scale Chinese Image-Text Benchmark for Real-World Short Video Search Scenarios
von: Qiao, Xiangshuo, et al.
Veröffentlicht: (2024)
von: Qiao, Xiangshuo, et al.
Veröffentlicht: (2024)
Hand1000: Generating Realistic Hands from Text with Only 1,000 Images
von: Zhang, Haozhuo, et al.
Veröffentlicht: (2024)
von: Zhang, Haozhuo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Fast Text-Driven Approach for Generating Artistic Content
von: Lupascu, Marian, et al.
Veröffentlicht: (2022) -
An Effective Image Copy-Move Forgery Detection Using Entropy Information
von: Jiang, Li, et al.
Veröffentlicht: (2023) -
TraceRouter: Robust Safety for Large Foundation Models via Path-Level Intervention
von: Shi, Chuancheng, et al.
Veröffentlicht: (2026) -
Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement
von: Gao, Jiayi, et al.
Veröffentlicht: (2025) -
Transcending Fusion: A Multi-Scale Alignment Method for Remote Sensing Image-Text Retrieval
von: Yang, Rui, et al.
Veröffentlicht: (2024)