yley123/MatSynth-Captions-Text-Prompts-for-Material-Videos: MatSynth-Captions-Text-Prompts-for-Material-Videos
Fuente:
Zenodo
Guardado en:
| Autor principal: | BOWEN |
|---|---|
| Formato: | Recurso digital |
| Publicado: |
Zenodo
2025
|
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MatSynth: A Modern PBR Materials Dataset
por: Vecchio, Giuseppe, et al.
Publicado: (2024)
por: Vecchio, Giuseppe, et al.
Publicado: (2024)
VideoMat: Extracting PBR Materials from Video Diffusion Models
por: Munkberg, Jacob, et al.
Publicado: (2025)
por: Munkberg, Jacob, et al.
Publicado: (2025)
VideoMat: Extracting PBR Materials from Video Diffusion Models
por: J. Munkberg, et al.
Publicado: (2025)
por: J. Munkberg, et al.
Publicado: (2025)
Image Captions are Natural Prompts for Text-to-Image Models
por: Lei, Shiye, et al.
Publicado: (2023)
por: Lei, Shiye, et al.
Publicado: (2023)
Text Data-Centric Image Captioning with Interactive Prompts
por: Wang, Yiyu, et al.
Publicado: (2024)
por: Wang, Yiyu, et al.
Publicado: (2024)
VRMDiff: Text-Guided Video Referring Matting Generation of Diffusion
por: Yang, Lehan, et al.
Publicado: (2025)
por: Yang, Lehan, et al.
Publicado: (2025)
VideoNeuMat: Neural Material Extraction from Generative Video Models
por: Xue, Bowen, et al.
Publicado: (2026)
por: Xue, Bowen, et al.
Publicado: (2026)
MatAtlas: Text-driven Consistent Geometry Texturing and Material Assignment
por: Ceylan, Duygu, et al.
Publicado: (2024)
por: Ceylan, Duygu, et al.
Publicado: (2024)
Synth$^2$: Boosting Visual-Language Models with Synthetic Captions and Image Embeddings
por: Sharifzadeh, Sahand, et al.
Publicado: (2024)
por: Sharifzadeh, Sahand, et al.
Publicado: (2024)
Generative Video Matting
por: Ge, Yongtao, et al.
Publicado: (2025)
por: Ge, Yongtao, et al.
Publicado: (2025)
VideoMatGen: PBR Materials through Joint Generative Modeling
por: Hasselgren, Jon, et al.
Publicado: (2026)
por: Hasselgren, Jon, et al.
Publicado: (2026)
Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting
por: Tang, Yunlong, et al.
Publicado: (2025)
por: Tang, Yunlong, et al.
Publicado: (2025)
Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M)
por: Merchant, Nicholas, et al.
Publicado: (2025)
por: Merchant, Nicholas, et al.
Publicado: (2025)
MatNexus: A Comprehensive Text Mining and Analysis Suite for Materials Discover
por: Zhang, Lei, et al.
Publicado: (2023)
por: Zhang, Lei, et al.
Publicado: (2023)
It's Just Another Day: Unique Video Captioning by Discriminative Prompting
por: Perrett, Toby, et al.
Publicado: (2024)
por: Perrett, Toby, et al.
Publicado: (2024)
HowToCaption: Prompting LLMs to Transform Video Annotations at Scale
por: Shvetsova, Nina, et al.
Publicado: (2023)
por: Shvetsova, Nina, et al.
Publicado: (2023)
MatAnyone: Stable Video Matting with Consistent Memory Propagation
por: Yang, Peiqing, et al.
Publicado: (2025)
por: Yang, Peiqing, et al.
Publicado: (2025)
Expertized Caption Auto-Enhancement for Video-Text Retrieval
por: Yang, Baoyao, et al.
Publicado: (2025)
por: Yang, Baoyao, et al.
Publicado: (2025)
Pretrained Image-Text Models are Secretly Video Captioners
por: Zhang, Chunhui, et al.
Publicado: (2025)
por: Zhang, Chunhui, et al.
Publicado: (2025)
Learning to Rank Caption Chains for Video-Text Alignment
por: Blume, Ansel, et al.
Publicado: (2026)
por: Blume, Ansel, et al.
Publicado: (2026)
VC4VG: Optimizing Video Captions for Text-to-Video Generation
por: Du, Yang, et al.
Publicado: (2025)
por: Du, Yang, et al.
Publicado: (2025)
VoCap: Video Object Captioning and Segmentation from Any Prompt
por: Uijlings, Jasper, et al.
Publicado: (2025)
por: Uijlings, Jasper, et al.
Publicado: (2025)
ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning
por: Kim, Taewhan, et al.
Publicado: (2024)
por: Kim, Taewhan, et al.
Publicado: (2024)
LeMat-Synth: a multi-modal toolbox to curate broad synthesis procedure databases from scientific literature
por: Lederbauer, Magdalena, et al.
Publicado: (2025)
por: Lederbauer, Magdalena, et al.
Publicado: (2025)
Post-Training Quantization for Video Matting
por: Zhu, Tianrui, et al.
Publicado: (2025)
por: Zhu, Tianrui, et al.
Publicado: (2025)
SynthVLM: Towards High-Quality and Efficient Synthesis of Image-Caption Datasets for Vision-Language Models
por: Liu, Zheng, et al.
Publicado: (2024)
por: Liu, Zheng, et al.
Publicado: (2024)
TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation
por: Bansal, Hritik, et al.
Publicado: (2024)
por: Bansal, Hritik, et al.
Publicado: (2024)
MatAnyone 2: Scaling Video Matting via a Learned Quality Evaluator
por: Yang, Peiqing, et al.
Publicado: (2025)
por: Yang, Peiqing, et al.
Publicado: (2025)
MatPhys: Learning Material-Aware Physics Parameters for Deformable Object Simulation from Videos
por: Yang, Yang, et al.
Publicado: (2026)
por: Yang, Yang, et al.
Publicado: (2026)
TA-Prompting: Enhancing Video Large Language Models for Dense Video Captioning via Temporal Anchors
por: Cheng, Wei-Yuan, et al.
Publicado: (2026)
por: Cheng, Wei-Yuan, et al.
Publicado: (2026)
VidCapBench: A Comprehensive Benchmark of Video Captioning for Controllable Text-to-Video Generation
por: Chen, Xinlong, et al.
Publicado: (2025)
por: Chen, Xinlong, et al.
Publicado: (2025)
Narrating the Video: Boosting Text-Video Retrieval via Comprehensive Utilization of Frame-Level Captions
por: Hur, Chan, et al.
Publicado: (2025)
por: Hur, Chan, et al.
Publicado: (2025)
Live Video Captioning
por: Blanco-Fernández, Eduardo, et al.
Publicado: (2024)
por: Blanco-Fernández, Eduardo, et al.
Publicado: (2024)
Captioning for Text-Video Retrieval via Dual-Group Direct Preference Optimization
por: Lee, Ji Soo, et al.
Publicado: (2025)
por: Lee, Ji Soo, et al.
Publicado: (2025)
OPCap:Object-aware Prompting Captioning
por: Huang, Feiyang
Publicado: (2024)
por: Huang, Feiyang
Publicado: (2024)
Robustness Assessment and Enhancement of Text Watermarking for Google's SynthID
por: Han, Xia, et al.
Publicado: (2025)
por: Han, Xia, et al.
Publicado: (2025)
RxnCaption: Reformulating Reaction Diagram Parsing as Visual Prompt Guided Captioning
por: Song, Jiahe, et al.
Publicado: (2025)
por: Song, Jiahe, et al.
Publicado: (2025)
RealMat: Realistic Materials with Diffusion and Reinforcement Learning
por: Zhou, Xilong, et al.
Publicado: (2025)
por: Zhou, Xilong, et al.
Publicado: (2025)
MatFuse: Controllable Material Generation with Diffusion Models
por: Vecchio, Giuseppe, et al.
Publicado: (2023)
por: Vecchio, Giuseppe, et al.
Publicado: (2023)
SynthTextEval: Synthetic Text Data Generation and Evaluation for High-Stakes Domains
por: Ramesh, Krithika, et al.
Publicado: (2025)
por: Ramesh, Krithika, et al.
Publicado: (2025)
Ejemplares similares
-
MatSynth: A Modern PBR Materials Dataset
por: Vecchio, Giuseppe, et al.
Publicado: (2024) -
VideoMat: Extracting PBR Materials from Video Diffusion Models
por: Munkberg, Jacob, et al.
Publicado: (2025) -
VideoMat: Extracting PBR Materials from Video Diffusion Models
por: J. Munkberg, et al.
Publicado: (2025) -
Image Captions are Natural Prompts for Text-to-Image Models
por: Lei, Shiye, et al.
Publicado: (2023) -
Text Data-Centric Image Captioning with Interactive Prompts
por: Wang, Yiyu, et al.
Publicado: (2024)