Salvato in:
| Autori principali: | Jamil, Sofia, Reddy, Bollampalli Areen, Kumar, Raghvendra, Saha, Sriparna, Joseph, K J, Goswami, Koustava |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2501.05839 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PoemTale Diffusion: Minimising Information Loss in Poem to Image Generation with Multi-Stage Prompt Refinement
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
Do It Yourself (DIY): Modifying Images for Poems in a Zero-Shot Setting Using Weighted Prompt Manipulation
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
Crossing Borders: A Multimodal Challenge for Indian Poetry Translation and Image Generation
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
GASCADE: Grouped Summarization of Adverse Drug Event for Enhanced Cancer Pharmacovigilance
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs
di: Wang, Shanshan, et al.
Pubblicazione: (2026)
di: Wang, Shanshan, et al.
Pubblicazione: (2026)
Step-by-step Layered Design Generation
di: Khan, Faizan Farooq, et al.
Pubblicazione: (2025)
di: Khan, Faizan Farooq, et al.
Pubblicazione: (2025)
SafaRi:Adaptive Sequence Transformer for Weakly Supervised Referring Expression Segmentation
di: Nag, Sayan, et al.
Pubblicazione: (2024)
di: Nag, Sayan, et al.
Pubblicazione: (2024)
PixelDiT: Pixel Diffusion Transformers for Image Generation
di: Yu, Yongsheng, et al.
Pubblicazione: (2025)
di: Yu, Yongsheng, et al.
Pubblicazione: (2025)
When Background Matters: Breaking Medical Vision Language Models by Transferable Attack
di: Ghosh, Akash, et al.
Pubblicazione: (2026)
di: Ghosh, Akash, et al.
Pubblicazione: (2026)
Agentic Design Review System
di: Nag, Sayan, et al.
Pubblicazione: (2025)
di: Nag, Sayan, et al.
Pubblicazione: (2025)
Ask Me Again Differently: GRAS for Measuring Bias in Vision Language Models on Gender, Race, Age, and Skin Tone
di: Malik, Shaivi, et al.
Pubblicazione: (2025)
di: Malik, Shaivi, et al.
Pubblicazione: (2025)
BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources
di: Kumar, Raghvendra, et al.
Pubblicazione: (2026)
di: Kumar, Raghvendra, et al.
Pubblicazione: (2026)
Enhancing Adverse Drug Event Detection with Multimodal Dataset: Corpus Creation and Model Development
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
Pixel-Perfect Depth with Semantics-Prompted Diffusion Transformers
di: Xu, Gangwei, et al.
Pubblicazione: (2025)
di: Xu, Gangwei, et al.
Pubblicazione: (2025)
FedMRL: Data Heterogeneity Aware Federated Multi-agent Deep Reinforcement Learning for Medical Imaging
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and Generation
di: Jiang, Liyao, et al.
Pubblicazione: (2024)
di: Jiang, Liyao, et al.
Pubblicazione: (2024)
Pixels, Patterns, but No Poetry: To See The World like Humans
di: Gao, Hongcheng, et al.
Pubblicazione: (2025)
di: Gao, Hongcheng, et al.
Pubblicazione: (2025)
Can Better Text Semantics in Prompt Tuning Improve VLM Generalization?
di: Kuchibhotla, Hari Chandana, et al.
Pubblicazione: (2024)
di: Kuchibhotla, Hari Chandana, et al.
Pubblicazione: (2024)
Poetry2Image: An Iterative Correction Framework for Images Generated from Chinese Classical Poetry
di: Jiang, Jing, et al.
Pubblicazione: (2024)
di: Jiang, Jing, et al.
Pubblicazione: (2024)
PromptSafe: Gated Prompt Tuning for Safe Text-to-Image Generation
di: Jing, Zonglei, et al.
Pubblicazione: (2025)
di: Jing, Zonglei, et al.
Pubblicazione: (2025)
Edify Image: High-Quality Image Generation with Pixel Space Laplacian Diffusion Models
di: NVIDIA, et al.
Pubblicazione: (2024)
di: NVIDIA, et al.
Pubblicazione: (2024)
PromptRR: Diffusion Models as Prompt Generators for Single Image Reflection Removal
di: Wang, Tao, et al.
Pubblicazione: (2024)
di: Wang, Tao, et al.
Pubblicazione: (2024)
ToxVidLM: A Multimodal Framework for Toxicity Detection in Code-Mixed Videos
di: Maity, Krishanu, et al.
Pubblicazione: (2024)
di: Maity, Krishanu, et al.
Pubblicazione: (2024)
Exploring the Frontier of Vision-Language Models: A Survey of Current Methodologies and Future Directions
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
Concept-to-Pixel: Prompt-Free Universal Medical Image Segmentation
di: Chen, Haoyun, et al.
Pubblicazione: (2026)
di: Chen, Haoyun, et al.
Pubblicazione: (2026)
From Pampas to Pixels: Fine-Tuning Diffusion Models for Gaúcho Heritage
di: Amadeus, Marcellus, et al.
Pubblicazione: (2024)
di: Amadeus, Marcellus, et al.
Pubblicazione: (2024)
CarePilot: A Multi-Agent Framework for Long-Horizon Computer Task Automation in Healthcare
di: Ghosh, Akash, et al.
Pubblicazione: (2026)
di: Ghosh, Akash, et al.
Pubblicazione: (2026)
AntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image Detectors
di: Chang, You-Ming, et al.
Pubblicazione: (2023)
di: Chang, You-Ming, et al.
Pubblicazione: (2023)
PixelRush: Ultra-Fast, Training-Free High-Resolution Image Generation via One-step Diffusion
di: Lai, Hong-Phuc, et al.
Pubblicazione: (2026)
di: Lai, Hong-Phuc, et al.
Pubblicazione: (2026)
Simulating Post-Neoadjuvant Chemotherapy Breast Cancer MRI via Diffusion Model with Prompt Tuning
di: Kim, Jonghun, et al.
Pubblicazione: (2025)
di: Kim, Jonghun, et al.
Pubblicazione: (2025)
PixIE: Prompted Pixel-Space Low-Light Image Enhancement
di: Lin, Ruirui, et al.
Pubblicazione: (2026)
di: Lin, Ruirui, et al.
Pubblicazione: (2026)
Latent Forcing: Reordering the Diffusion Trajectory for Pixel-Space Image Generation
di: Baade, Alan, et al.
Pubblicazione: (2026)
di: Baade, Alan, et al.
Pubblicazione: (2026)
FreeMorph: Tuning-Free Generalized Image Morphing with Diffusion Model
di: Cao, Yukang, et al.
Pubblicazione: (2025)
di: Cao, Yukang, et al.
Pubblicazione: (2025)
A Comprehensive Survey of Hallucination in Large Language, Image, Video and Audio Foundation Models
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
Semi-supervised Chinese Poem-to-Painting Generation via Cycle-consistent Adversarial Networks
di: Lu, Zhengyang, et al.
Pubblicazione: (2024)
di: Lu, Zhengyang, et al.
Pubblicazione: (2024)
PixelFlow: Pixel-Space Generative Models with Flow
di: Chen, Shoufa, et al.
Pubblicazione: (2025)
di: Chen, Shoufa, et al.
Pubblicazione: (2025)
LesionGen: A Concept-Guided Diffusion Model for Dermatology Image Synthesis
di: Fayyad, Jamil, et al.
Pubblicazione: (2025)
di: Fayyad, Jamil, et al.
Pubblicazione: (2025)
SANSKRITI: A Comprehensive Benchmark for Evaluating Language Models' Knowledge of Indian Culture
di: Maji, Arijit, et al.
Pubblicazione: (2025)
di: Maji, Arijit, et al.
Pubblicazione: (2025)
Pixel Is Not a Barrier: An Effective Evasion Attack for Pixel-Domain Diffusion Models
di: Shih, Chun-Yen, et al.
Pubblicazione: (2024)
di: Shih, Chun-Yen, et al.
Pubblicazione: (2024)
URSimulator: Human-Perception-Driven Prompt Tuning for Enhanced Virtual Urban Renewal via Diffusion Models
di: Hu, Chuanbo, et al.
Pubblicazione: (2024)
di: Hu, Chuanbo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
PoemTale Diffusion: Minimising Information Loss in Poem to Image Generation with Multi-Stage Prompt Refinement
di: Jamil, Sofia, et al.
Pubblicazione: (2025) -
Do It Yourself (DIY): Modifying Images for Poems in a Zero-Shot Setting Using Weighted Prompt Manipulation
di: Jamil, Sofia, et al.
Pubblicazione: (2025) -
Crossing Borders: A Multimodal Challenge for Indian Poetry Translation and Image Generation
di: Jamil, Sofia, et al.
Pubblicazione: (2025) -
GASCADE: Grouped Summarization of Adverse Drug Event for Enhanced Cancer Pharmacovigilance
di: Jamil, Sofia, et al.
Pubblicazione: (2025) -
Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs
di: Wang, Shanshan, et al.
Pubblicazione: (2026)