Surrealistic-like Image Generation with Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Ayten, Elif, Wang, Shuai, Snoep, Hjalmar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
di: Cao, Bryan Bo, et al.
Pubblicazione: (2024)
di: Cao, Bryan Bo, et al.
Pubblicazione: (2024)
Towards Interpretable Visual Decoding with Attention to Brain Representations
di: Feng, Pinyuan, et al.
Pubblicazione: (2025)
di: Feng, Pinyuan, et al.
Pubblicazione: (2025)
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
di: Merugu, Ranjith, et al.
Pubblicazione: (2025)
di: Merugu, Ranjith, et al.
Pubblicazione: (2025)
Time Step Generating: A Universal Synthesized Deepfake Image Detector
di: Zeng, Ziyue, et al.
Pubblicazione: (2024)
di: Zeng, Ziyue, et al.
Pubblicazione: (2024)
Transfer learning with generative models for object detection on limited datasets
di: Paiano, Matteo, et al.
Pubblicazione: (2024)
di: Paiano, Matteo, et al.
Pubblicazione: (2024)
Eye-gaze Guided Multi-modal Alignment for Medical Representation Learning
di: Ma, Chong, et al.
Pubblicazione: (2024)
di: Ma, Chong, et al.
Pubblicazione: (2024)
The Impact of Image Resolution on Face Detection: A Comparative Analysis of MTCNN, YOLOv XI and YOLOv XII models
di: Ömercikoğlu, Ahmet Can, et al.
Pubblicazione: (2025)
di: Ömercikoğlu, Ahmet Can, et al.
Pubblicazione: (2025)
Deep Learning for Generating Computational PIN-4 Immunohistochemistry Staining from Prostate Biopsy H&E Images
di: Tran, Vietbao, et al.
Pubblicazione: (2026)
di: Tran, Vietbao, et al.
Pubblicazione: (2026)
Person detection and re-identification in open-world settings of retail stores and public spaces
di: Brkljač, Branko, et al.
Pubblicazione: (2025)
di: Brkljač, Branko, et al.
Pubblicazione: (2025)
Parameter-efficient fine-tuning (PEFT) of Vision Foundation Models for Atypical Mitotic Figure Classification
di: Ramchandani, Lavish, et al.
Pubblicazione: (2025)
di: Ramchandani, Lavish, et al.
Pubblicazione: (2025)
A Landmark-Aware Visual Navigation Dataset
di: Johnson, Faith, et al.
Pubblicazione: (2024)
di: Johnson, Faith, et al.
Pubblicazione: (2024)
HieraEdgeNet: A Multi-Scale Edge-Enhanced Framework for Automated Pollen Recognition
di: Long, Yuchong, et al.
Pubblicazione: (2025)
di: Long, Yuchong, et al.
Pubblicazione: (2025)
Multi-Agent Object Detection Framework Based on Raspberry Pi YOLO Detector and Slack-Ollama Natural Language Interface
di: Kalušev, Vladimir, et al.
Pubblicazione: (2026)
di: Kalušev, Vladimir, et al.
Pubblicazione: (2026)
Balanced conic rectified flow
di: Kim, Shin Seong, et al.
Pubblicazione: (2025)
di: Kim, Shin Seong, et al.
Pubblicazione: (2025)
Can Agentic AI Match the Performance of Human Data Scientists?
di: Luo, An, et al.
Pubblicazione: (2025)
di: Luo, An, et al.
Pubblicazione: (2025)
AgentDS Technical Report: Benchmarking the Future of Human-AI Collaboration in Domain-Specific Data Science
di: Luo, An, et al.
Pubblicazione: (2026)
di: Luo, An, et al.
Pubblicazione: (2026)
Beyond RGB: Leveraging Vision Transformers for Thermal Weapon Segmentation
di: Kambhatla, Akhila, et al.
Pubblicazione: (2025)
di: Kambhatla, Akhila, et al.
Pubblicazione: (2025)
MVTamperBench: Evaluating Robustness of Vision-Language Models
di: Agarwal, Amit, et al.
Pubblicazione: (2024)
di: Agarwal, Amit, et al.
Pubblicazione: (2024)
Fast TILs -- A Pipeline for Efficient TILs Estimation in Non-Small Cell Lung Cancer
di: Shvetsov, Nikita, et al.
Pubblicazione: (2024)
di: Shvetsov, Nikita, et al.
Pubblicazione: (2024)
Motion-Based Sign Language Video Summarization using Curvature and Torsion
di: Sartinas, Evangelos G., et al.
Pubblicazione: (2023)
di: Sartinas, Evangelos G., et al.
Pubblicazione: (2023)
Ensemble YOLO Framework for Multi-Domain Mitotic Figure Detection in Histopathology Images
di: Kelam, Navya Sri, et al.
Pubblicazione: (2025)
di: Kelam, Navya Sri, et al.
Pubblicazione: (2025)
Fixed-Threshold Evaluation of a Hybrid CNN-ViT for AI-Generated Image Detection Across Photos and Art
di: Khan, Md Ashik, et al.
Pubblicazione: (2025)
di: Khan, Md Ashik, et al.
Pubblicazione: (2025)
Proceedings of the 20th International Conference on Knowledge, Information and Creativity Support Systems (KICSS 2025)
di: Hayama, Edited by Tessai, et al.
Pubblicazione: (2025)
di: Hayama, Edited by Tessai, et al.
Pubblicazione: (2025)
Sequential PatchCore: Anomaly Detection for Surface Inspection using Synthetic Impurities
di: Mao, Runzhou, et al.
Pubblicazione: (2025)
di: Mao, Runzhou, et al.
Pubblicazione: (2025)
Customizing Graph Neural Networks using Path Reweighting
di: Chen, Jianpeng, et al.
Pubblicazione: (2021)
di: Chen, Jianpeng, et al.
Pubblicazione: (2021)
Banana Ripeness Level Classification using a Simple CNN Model Trained with Real and Synthetic Datasets
di: Chuquimarca, Luis, et al.
Pubblicazione: (2025)
di: Chuquimarca, Luis, et al.
Pubblicazione: (2025)
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
di: Rychkovskiy, Denis
Pubblicazione: (2025)
di: Rychkovskiy, Denis
Pubblicazione: (2025)
AssistedDS: Benchmarking How External Domain Knowledge Assists LLMs in Automated Data Science
di: Luo, An, et al.
Pubblicazione: (2025)
di: Luo, An, et al.
Pubblicazione: (2025)
A Machine Learning Approach for the Efficient Estimation of Ground-Level Air Temperature in Urban Areas
di: Delgado-Enales, Iñigo, et al.
Pubblicazione: (2024)
di: Delgado-Enales, Iñigo, et al.
Pubblicazione: (2024)
Intersymbolic AI: Interlinking Symbolic AI and Subsymbolic AI
di: Platzer, André
Pubblicazione: (2024)
di: Platzer, André
Pubblicazione: (2024)
Vision Transformer-based Model for Severity Quantification of Lung Pneumonia Using Chest X-ray Images
di: Slika, Bouthaina, et al.
Pubblicazione: (2023)
di: Slika, Bouthaina, et al.
Pubblicazione: (2023)
Non-destructive Identification of Oyster Species is possible from Hyperspectral Images with Machine Learning
di: Waters, Ethan Kane, et al.
Pubblicazione: (2026)
di: Waters, Ethan Kane, et al.
Pubblicazione: (2026)
Event-ECC: Asynchronous Tracking of Events with Continuous Optimization
di: Zafeiri, Maria, et al.
Pubblicazione: (2024)
di: Zafeiri, Maria, et al.
Pubblicazione: (2024)
Knowledge Distillation: Enhancing Neural Network Compression with Integrated Gradients
di: Hernandez, David E., et al.
Pubblicazione: (2025)
di: Hernandez, David E., et al.
Pubblicazione: (2025)
Approximating Discrimination Within Models When Faced With Several Non-Binary Sensitive Attributes
di: Bian, Yijun, et al.
Pubblicazione: (2024)
di: Bian, Yijun, et al.
Pubblicazione: (2024)
Does Machine Bring in Extra Bias in Learning? Approximating Fairness in Models Promptly
di: Bian, Yijun, et al.
Pubblicazione: (2024)
di: Bian, Yijun, et al.
Pubblicazione: (2024)
Reference-Guided Verdict: LLMs-as-Judges in Automatic Evaluation of Free-Form QA
di: Badshah, Sher, et al.
Pubblicazione: (2024)
di: Badshah, Sher, et al.
Pubblicazione: (2024)
Generating Natural-Language Surgical Feedback: From Structured Representation to Domain-Grounded Evaluation
di: Nasriddinov, Firdavs, et al.
Pubblicazione: (2025)
di: Nasriddinov, Firdavs, et al.
Pubblicazione: (2025)
A Multimodal Pipeline for Clinical Data Extraction: Applying Vision-Language Models to Scans of Transfusion Reaction Reports
di: Schäfer, Henning, et al.
Pubblicazione: (2025)
di: Schäfer, Henning, et al.
Pubblicazione: (2025)
Fixed-Budget Parameter-Efficient Training with Frozen Encoders Improves Multimodal Chest X-Ray Classification
di: Khan, Md Ashik, et al.
Pubblicazione: (2025)
di: Khan, Md Ashik, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
di: Cao, Bryan Bo, et al.
Pubblicazione: (2024) -
Towards Interpretable Visual Decoding with Attention to Brain Representations
di: Feng, Pinyuan, et al.
Pubblicazione: (2025) -
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
di: Merugu, Ranjith, et al.
Pubblicazione: (2025) -
Time Step Generating: A Universal Synthesized Deepfake Image Detector
di: Zeng, Ziyue, et al.
Pubblicazione: (2024) -
Transfer learning with generative models for object detection on limited datasets
di: Paiano, Matteo, et al.
Pubblicazione: (2024)