AttriStory: Fine-grained Attribute Realization for Visual Storytelling with Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Sreenivas, Manogna, Kumar, Rohit, Biswas, Soma |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Open Set Single Image Test Time Adaptation of Vision Language Models
by: Sreenivas, Manogna, et al.
Published: (2024)
by: Sreenivas, Manogna, et al.
Published: (2024)
Segmentation Assisted Incremental Test Time Adaptation in an Open World
by: Sreenivas, Manogna, et al.
Published: (2025)
by: Sreenivas, Manogna, et al.
Published: (2025)
AttriCtrl: Fine-Grained Control of Aesthetic Attribute Intensity in Diffusion Models
by: Chen, Die, et al.
Published: (2025)
by: Chen, Die, et al.
Published: (2025)
TACLE: Task and Class-aware Exemplar-free Semi-supervised Class Incremental Learning
by: Kalla, Jayateja, et al.
Published: (2024)
by: Kalla, Jayateja, et al.
Published: (2024)
AttriBE: Quantifying Attribute Expressivity in Body Embeddings for Recognition and Identification
by: Pal, Basudha, et al.
Published: (2026)
by: Pal, Basudha, et al.
Published: (2026)
FiVA: Fine-grained Visual Attribute Dataset for Text-to-Image Diffusion Models
by: Wu, Tong, et al.
Published: (2024)
by: Wu, Tong, et al.
Published: (2024)
AttriPrompter: Auto-Prompting with Attribute Semantics for Zero-shot Nuclei Detection via Visual-Language Pre-trained Models
by: Wu, Yongjian, et al.
Published: (2024)
by: Wu, Yongjian, et al.
Published: (2024)
Beyond Binary Preference: Aligning Diffusion Models to Fine-grained Criteria by Decoupling Attributes
by: Meng, Chenye, et al.
Published: (2026)
by: Meng, Chenye, et al.
Published: (2026)
Fully Unsupervised Self-debiasing of Text-to-Image Diffusion Models
by: Vardhana, Korada Sri, et al.
Published: (2025)
by: Vardhana, Korada Sri, et al.
Published: (2025)
Let Storytelling Tell Vivid Stories: An Expressive and Fluent Multimodal Storyteller
by: Zang, Chuanqi, et al.
Published: (2024)
by: Zang, Chuanqi, et al.
Published: (2024)
Story3D-Agent: Exploring 3D Storytelling Visualization with Large Language Models
by: Huang, Yuzhou, et al.
Published: (2024)
by: Huang, Yuzhou, et al.
Published: (2024)
AttriHuman-3D: Editable 3D Human Avatar Generation with Attribute Decomposition and Indexing
by: Yang, Fan, et al.
Published: (2023)
by: Yang, Fan, et al.
Published: (2023)
InstructAttribute: Fine-grained Object Attributes editing with Instruction
by: Yin, Xingxi, et al.
Published: (2025)
by: Yin, Xingxi, et al.
Published: (2025)
Towards Generative Class Prompt Learning for Fine-grained Visual Recognition
by: Chattopadhyay, Soumitri, et al.
Published: (2024)
by: Chattopadhyay, Soumitri, et al.
Published: (2024)
AttriPrompt: Dynamic Prompt Composition Learning for CLIP
by: Zhan, Qiqi, et al.
Published: (2025)
by: Zhan, Qiqi, et al.
Published: (2025)
Intelligent Grimm -- Open-ended Visual Storytelling via Latent Diffusion Models
by: Liu, Chang, et al.
Published: (2023)
by: Liu, Chang, et al.
Published: (2023)
CoVLM: Leveraging Consensus from Vision-Language Models for Semi-supervised Multi-modal Fake News Detection
by: Devank, et al.
Published: (2024)
by: Devank, et al.
Published: (2024)
FineFACE: Fair Facial Attribute Classification Leveraging Fine-grained Features
by: Manzoor, Ayesha, et al.
Published: (2024)
by: Manzoor, Ayesha, et al.
Published: (2024)
AggSS: An Aggregated Self-Supervised Approach for Class-Incremental Learning
by: Kalla, Jayateja, et al.
Published: (2024)
by: Kalla, Jayateja, et al.
Published: (2024)
Multiple Stochastic Prompt Tuning for Few-shot Adaptation under Extreme Domain Shift
by: Brahma, Debarshi, et al.
Published: (2025)
by: Brahma, Debarshi, et al.
Published: (2025)
StoryMem: Multi-shot Long Video Storytelling with Memory
by: Zhang, Kaiwen, et al.
Published: (2025)
by: Zhang, Kaiwen, et al.
Published: (2025)
ContextualStory: Consistent Visual Storytelling with Spatially-Enhanced and Storyline Context
by: Zheng, Sixiao, et al.
Published: (2024)
by: Zheng, Sixiao, et al.
Published: (2024)
AttriCLIP: A Non-Incremental Learner for Incremental Knowledge Learning
by: Wang, Runqi, et al.
Published: (2023)
by: Wang, Runqi, et al.
Published: (2023)
DiVE-k: Differential Visual Reasoning for Fine-grained Image Recognition
by: Kumar, Raja, et al.
Published: (2025)
by: Kumar, Raja, et al.
Published: (2025)
ViSTA: Visual Storytelling using Multi-modal Adapters for Text-to-Image Diffusion Models
by: Dong, Sibo, et al.
Published: (2025)
by: Dong, Sibo, et al.
Published: (2025)
Democratizing Fine-grained Visual Recognition with Large Language Models
by: Liu, Mingxuan, et al.
Published: (2024)
by: Liu, Mingxuan, et al.
Published: (2024)
FineDiffusion: Scaling up Diffusion Models for Fine-grained Image Generation with 10,000 Classes
by: Pan, Ziying, et al.
Published: (2024)
by: Pan, Ziying, et al.
Published: (2024)
ReDiStory: Region-Disentangled Diffusion for Consistent Visual Story Generation
by: Sarkar, Ayushman, et al.
Published: (2026)
by: Sarkar, Ayushman, et al.
Published: (2026)
Improving Visual Storytelling with Multimodal Large Language Models
by: Lin, Xiaochuan, et al.
Published: (2024)
by: Lin, Xiaochuan, et al.
Published: (2024)
Boosting Consistency in Story Visualization with Rich-Contextual Conditional Diffusion Models
by: Shen, Fei, et al.
Published: (2024)
by: Shen, Fei, et al.
Published: (2024)
DeCorStory: Gram-Schmidt Prompt Embedding Decorrelation for Consistent Storytelling
by: Sarkar, Ayushman, et al.
Published: (2026)
by: Sarkar, Ayushman, et al.
Published: (2026)
DSAA: Dual-Stage Attribute Activation for Fine-grained Open Vocabulary Detection
by: Jiang, Donghong, et al.
Published: (2026)
by: Jiang, Donghong, et al.
Published: (2026)
FOCUS: Fine-grained Optimization with Semantic Guided Understanding for Pedestrian Attributes Recognition
by: An, Hongyan, et al.
Published: (2025)
by: An, Hongyan, et al.
Published: (2025)
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training
by: Huang, Jiale, et al.
Published: (2024)
by: Huang, Jiale, et al.
Published: (2024)
Fine-grained Image-to-LiDAR Contrastive Distillation with Visual Foundation Models
by: Zhang, Yifan, et al.
Published: (2024)
by: Zhang, Yifan, et al.
Published: (2024)
HiGFA: Hierarchical Guidance for Fine-grained Data Augmentation with Diffusion Models
by: Lu, Zhiguang, et al.
Published: (2025)
by: Lu, Zhiguang, et al.
Published: (2025)
EFDiT: Efficient Fine-grained Image Generation Using Diffusion Transformer Models
by: Wang, Kun, et al.
Published: (2025)
by: Wang, Kun, et al.
Published: (2025)
Customized Visual Storytelling with Unified Multimodal LLMs
by: Li, Wei-Hua, et al.
Published: (2026)
by: Li, Wei-Hua, et al.
Published: (2026)
Audit & Repair: An Agentic Framework for Consistent Story Visualization in Text-to-Image Diffusion Models
by: Akdemir, Kiymet, et al.
Published: (2025)
by: Akdemir, Kiymet, et al.
Published: (2025)
Face Adapter for Pre-Trained Diffusion Models with Fine-Grained ID and Attribute Control
by: Han, Yue, et al.
Published: (2024)
by: Han, Yue, et al.
Published: (2024)
Similar Items
-
Efficient Open Set Single Image Test Time Adaptation of Vision Language Models
by: Sreenivas, Manogna, et al.
Published: (2024) -
Segmentation Assisted Incremental Test Time Adaptation in an Open World
by: Sreenivas, Manogna, et al.
Published: (2025) -
AttriCtrl: Fine-Grained Control of Aesthetic Attribute Intensity in Diffusion Models
by: Chen, Die, et al.
Published: (2025) -
TACLE: Task and Class-aware Exemplar-free Semi-supervised Class Incremental Learning
by: Kalla, Jayateja, et al.
Published: (2024) -
AttriBE: Quantifying Attribute Expressivity in Body Embeddings for Recognition and Identification
by: Pal, Basudha, et al.
Published: (2026)