StoryTailor:A Zero-Shot Pipeline for Action-Rich Multi-Subject Visual Narratives
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Jinghao, Zhang, Yuhe, Geng, GuoHua, Li, Kang, Zhang, Han |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Color and Lines: Zero-Shot Style-Specific Image Variations with Coordinated Semantics
von: Hu, Jinghao, et al.
Veröffentlicht: (2024)
von: Hu, Jinghao, et al.
Veröffentlicht: (2024)
Telling Stories for Common Sense Zero-Shot Action Recognition
von: Gowda, Shreyank N, et al.
Veröffentlicht: (2023)
von: Gowda, Shreyank N, et al.
Veröffentlicht: (2023)
Boosting Consistency in Story Visualization with Rich-Contextual Conditional Diffusion Models
von: Shen, Fei, et al.
Veröffentlicht: (2024)
von: Shen, Fei, et al.
Veröffentlicht: (2024)
Zero-Shot Skeleton-based Action Recognition with Dual Visual-Text Alignment
von: Kuang, Jidong, et al.
Veröffentlicht: (2024)
von: Kuang, Jidong, et al.
Veröffentlicht: (2024)
SubZero: Composing Subject, Style, and Action via Zero-Shot Personalization
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
DreamingComics: A Story Visualization Pipeline via Subject and Layout Customized Generation using Video Models
von: Kwon, Patrick, et al.
Veröffentlicht: (2025)
von: Kwon, Patrick, et al.
Veröffentlicht: (2025)
Expanding Zero-Shot Object Counting with Rich Prompts
von: Zhu, Huilin, et al.
Veröffentlicht: (2025)
von: Zhu, Huilin, et al.
Veröffentlicht: (2025)
Multi-Stage VLM Pipeline for Zero-Shot Traffic Accident Understanding
von: Tatematsu, Fumiya, et al.
Veröffentlicht: (2026)
von: Tatematsu, Fumiya, et al.
Veröffentlicht: (2026)
One Shot Learning for Edge Detection on Point Clouds
von: Tu, Zhikun, et al.
Veröffentlicht: (2026)
von: Tu, Zhikun, et al.
Veröffentlicht: (2026)
Zero-Shot Action Generalization with Limited Observations
von: Alchihabi, Abdullah, et al.
Veröffentlicht: (2025)
von: Alchihabi, Abdullah, et al.
Veröffentlicht: (2025)
Training-Free Zero-Shot Temporal Action Detection with Vision-Language Models
von: Han, Chaolei, et al.
Veröffentlicht: (2025)
von: Han, Chaolei, et al.
Veröffentlicht: (2025)
DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion
von: He, Huiguo, et al.
Veröffentlicht: (2024)
von: He, Huiguo, et al.
Veröffentlicht: (2024)
Epsilon: Exploring Comprehensive Visual-Semantic Projection for Multi-Label Zero-Shot Learning
von: Liu, Ziming, et al.
Veröffentlicht: (2024)
von: Liu, Ziming, et al.
Veröffentlicht: (2024)
MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation
von: Zhang, Haojie, et al.
Veröffentlicht: (2026)
von: Zhang, Haojie, et al.
Veröffentlicht: (2026)
Zero-Shot Skeleton-Based Action Recognition With Prototype-Guided Feature Alignment
von: Zhou, Kai, et al.
Veröffentlicht: (2025)
von: Zhou, Kai, et al.
Veröffentlicht: (2025)
SUGAR: Subject-Driven Video Customization in a Zero-Shot Manner
von: Zhou, Yufan, et al.
Veröffentlicht: (2024)
von: Zhou, Yufan, et al.
Veröffentlicht: (2024)
Subject-Aware Multi-Granularity Alignment for Zero-Shot EEG-to-Image Retrieval
von: Jiang, Lin, et al.
Veröffentlicht: (2026)
von: Jiang, Lin, et al.
Veröffentlicht: (2026)
OZ-TAL: Online Zero-Shot Temporal Action Localization
von: Han, Chaolei, et al.
Veröffentlicht: (2026)
von: Han, Chaolei, et al.
Veröffentlicht: (2026)
Multi-Granularity Mutual Refinement Network for Zero-Shot Learning
von: Wang, Ning, et al.
Veröffentlicht: (2025)
von: Wang, Ning, et al.
Veröffentlicht: (2025)
Neuron: Learning Context-Aware Evolving Representations for Zero-Shot Skeleton Action Recognition
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
Visual-Semantic Graph Matching Net for Zero-Shot Learning
von: Duan, Bowen, et al.
Veröffentlicht: (2024)
von: Duan, Bowen, et al.
Veröffentlicht: (2024)
Learning Visual Proxy for Compositional Zero-Shot Learning
von: Zhang, Shiyu, et al.
Veröffentlicht: (2025)
von: Zhang, Shiyu, et al.
Veröffentlicht: (2025)
VLCounter: Text-aware Visual Representation for Zero-Shot Object Counting
von: Kang, Seunggu, et al.
Veröffentlicht: (2023)
von: Kang, Seunggu, et al.
Veröffentlicht: (2023)
Diverse and Tailored Image Generation for Zero-shot Multi-label Classification
von: Zhang, Kaixin, et al.
Veröffentlicht: (2024)
von: Zhang, Kaixin, et al.
Veröffentlicht: (2024)
ZEBRA: Towards Zero-Shot Cross-Subject Generalization for Universal Brain Visual Decoding
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
von: Wei, Yujie, et al.
Veröffentlicht: (2024)
von: Wei, Yujie, et al.
Veröffentlicht: (2024)
SVIP: Semantically Contextualized Visual Patches for Zero-Shot Learning
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
FALIP: Visual Prompt as Foveal Attention Boosts CLIP Zero-Shot Performance
von: Zhuang, Jiedong, et al.
Veröffentlicht: (2024)
von: Zhuang, Jiedong, et al.
Veröffentlicht: (2024)
VisAgent: Narrative-Preserving Story Visualization Framework
von: Kim, Seungkwon, et al.
Veröffentlicht: (2025)
von: Kim, Seungkwon, et al.
Veröffentlicht: (2025)
ViStoryBench: Comprehensive Benchmark Suite for Story Visualization
von: Zhuang, Cailin, et al.
Veröffentlicht: (2025)
von: Zhuang, Cailin, et al.
Veröffentlicht: (2025)
Text-Enhanced Zero-Shot Action Recognition: A training-free approach
von: Bosetti, Massimo, et al.
Veröffentlicht: (2024)
von: Bosetti, Massimo, et al.
Veröffentlicht: (2024)
LAGO: Language-Guided Adaptive Object-Region Focus for Zero-Shot Visual-Text Alignment
von: Hu, Junyi, et al.
Veröffentlicht: (2026)
von: Hu, Junyi, et al.
Veröffentlicht: (2026)
Continual Learning Improves Zero-Shot Action Recognition
von: Gowda, Shreyank N, et al.
Veröffentlicht: (2024)
von: Gowda, Shreyank N, et al.
Veröffentlicht: (2024)
Novel Semantic Prompting for Zero-Shot Action Recognition
von: Iqbal, Salman, et al.
Veröffentlicht: (2026)
von: Iqbal, Salman, et al.
Veröffentlicht: (2026)
Test-Time Zero-Shot Temporal Action Localization
von: Liberatori, Benedetta, et al.
Veröffentlicht: (2024)
von: Liberatori, Benedetta, et al.
Veröffentlicht: (2024)
AnyLogo: Symbiotic Subject-Driven Diffusion System with Gemini Status
von: Zhang, Jinghao, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghao, et al.
Veröffentlicht: (2024)
Dual Expert Distillation Network for Generalized Zero-Shot Learning
von: Rao, Zhijie, et al.
Veröffentlicht: (2024)
von: Rao, Zhijie, et al.
Veröffentlicht: (2024)
Visual and Semantic Prompt Collaboration for Generalized Zero-Shot Learning
von: Jiang, Huajie, et al.
Veröffentlicht: (2025)
von: Jiang, Huajie, et al.
Veröffentlicht: (2025)
Attend and Enrich: Enhanced Visual Prompt for Zero-Shot Learning
von: Liu, Man, et al.
Veröffentlicht: (2024)
von: Liu, Man, et al.
Veröffentlicht: (2024)
Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos
von: Han, Mingfei, et al.
Veröffentlicht: (2023)
von: Han, Mingfei, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Beyond Color and Lines: Zero-Shot Style-Specific Image Variations with Coordinated Semantics
von: Hu, Jinghao, et al.
Veröffentlicht: (2024) -
Telling Stories for Common Sense Zero-Shot Action Recognition
von: Gowda, Shreyank N, et al.
Veröffentlicht: (2023) -
Boosting Consistency in Story Visualization with Rich-Contextual Conditional Diffusion Models
von: Shen, Fei, et al.
Veröffentlicht: (2024) -
Zero-Shot Skeleton-based Action Recognition with Dual Visual-Text Alignment
von: Kuang, Jidong, et al.
Veröffentlicht: (2024) -
SubZero: Composing Subject, Style, and Action via Zero-Shot Personalization
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)