Dense Video Captioning Using Unsupervised Semantic Information
Fuente:
arXiv
Saved in:
| Main Authors: | Estevam, Valter, Laroca, Rayson, Pedrini, Helio, Menotti, David |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CEZSAR: A Contrastive Embedding Method for Zero-Shot Action Recognition
by: Estevam, Valter, et al.
Published: (2026)
by: Estevam, Valter, et al.
Published: (2026)
Advancing Multinational License Plate Recognition Through Synthetic and Real Data Fusion: A Comprehensive Evaluation
by: Laroca, Rayson, et al.
Published: (2026)
by: Laroca, Rayson, et al.
Published: (2026)
Improving Small Drone Detection Through Multi-Scale Processing and Data Augmentation
by: Laroca, Rayson, et al.
Published: (2025)
by: Laroca, Rayson, et al.
Published: (2025)
Enhancing License Plate Super-Resolution: A Layout-Aware and Character-Driven Approach
by: Nascimento, Valfride, et al.
Published: (2024)
by: Nascimento, Valfride, et al.
Published: (2024)
Toward Enhancing Vehicle Color Recognition in Adverse Conditions: A Dataset and Benchmark
by: Lima, Gabriel E., et al.
Published: (2024)
by: Lima, Gabriel E., et al.
Published: (2024)
Less is more: concatenating videos for Sign Language Translation from a small set of signs
by: da Silva, David Vinicius, et al.
Published: (2024)
by: da Silva, David Vinicius, et al.
Published: (2024)
LPLCv2: An Expanded Dataset for Fine-Grained License Plate Legibility Classification
by: Wojcik, Lucas, et al.
Published: (2026)
by: Wojcik, Lucas, et al.
Published: (2026)
Multi-Feature Aggregation in Diffusion Models for Enhanced Face Super-Resolution
by: Santos, Marcelo dos, et al.
Published: (2024)
by: Santos, Marcelo dos, et al.
Published: (2024)
Toward Unified Fine-Grained Vehicle Classification and Automatic License Plate Recognition
by: Lima, Gabriel E., et al.
Published: (2026)
by: Lima, Gabriel E., et al.
Published: (2026)
LPLC: A Dataset for License Plate Legibility Classification
by: Wojcik, Lucas, et al.
Published: (2025)
by: Wojcik, Lucas, et al.
Published: (2025)
Toward Advancing License Plate Super-Resolution in Real-World Scenarios: A Dataset and Benchmark
by: Nascimento, Valfride, et al.
Published: (2025)
by: Nascimento, Valfride, et al.
Published: (2025)
Além do Desempenho: Um Estudo da Confiabilidade de Detectores de Deepfakes
by: Lopes, Lucas, et al.
Published: (2026)
by: Lopes, Lucas, et al.
Published: (2026)
Identification of Deforestation Areas in the Amazon Rainforest Using Change Detection Models
by: Konishi, Christian Massao, et al.
Published: (2025)
by: Konishi, Christian Massao, et al.
Published: (2025)
Streaming Dense Video Captioning
by: Zhou, Xingyi, et al.
Published: (2024)
by: Zhou, Xingyi, et al.
Published: (2024)
SoccerNet-Caption: Dense Video Captioning for Soccer Broadcasts Commentaries
by: Mkhallati, Hassan, et al.
Published: (2023)
by: Mkhallati, Hassan, et al.
Published: (2023)
CCNeXt: An Effective Self-Supervised Stereo Depth Estimation Approach
by: Lopes, Alexandre, et al.
Published: (2025)
by: Lopes, Alexandre, et al.
Published: (2025)
P-NOC: adversarial training of CAM generating networks for robust weakly supervised semantic segmentation priors
by: David, Lucas, et al.
Published: (2023)
by: David, Lucas, et al.
Published: (2023)
Sali4Vid: Saliency-Aware Video Reweighting and Adaptive Caption Retrieval for Dense Video Captioning
by: Jeon, MinJu, et al.
Published: (2025)
by: Jeon, MinJu, et al.
Published: (2025)
Dense Video Object Captioning from Disjoint Supervision
by: Zhou, Xingyi, et al.
Published: (2023)
by: Zhou, Xingyi, et al.
Published: (2023)
Technical Report for Soccernet 2023 -- Dense Video Captioning
by: Ruan, Zheng, et al.
Published: (2024)
by: Ruan, Zheng, et al.
Published: (2024)
Explicit Temporal-Semantic Modeling for Dense Video Captioning via Context-Aware Cross-Modal Interaction
by: Jia, Mingda, et al.
Published: (2025)
by: Jia, Mingda, et al.
Published: (2025)
Time-Scaling State-Space Models for Dense Video Captioning
by: Piergiovanni, AJ, et al.
Published: (2025)
by: Piergiovanni, AJ, et al.
Published: (2025)
Follow the Saliency: Supervised Saliency for Retrieval-augmented Dense Video Captioning
by: Choi, Seung hee, et al.
Published: (2026)
by: Choi, Seung hee, et al.
Published: (2026)
PR-DETR: Injecting Position and Relation Prior for Dense Video Captioning
by: Li, Yizhe, et al.
Published: (2025)
by: Li, Yizhe, et al.
Published: (2025)
Weakly Supervised Attention-based Models Using Activation Maps for Citrus Mite and Insect Pest Classification
by: Bollis, Edson, et al.
Published: (2021)
by: Bollis, Edson, et al.
Published: (2021)
Exo2EgoDVC: Dense Video Captioning of Egocentric Procedural Activities Using Web Instructional Videos
by: Ohkawa, Takehiko, et al.
Published: (2023)
by: Ohkawa, Takehiko, et al.
Published: (2023)
Exploring Temporal Event Cues for Dense Video Captioning in Cyclic Co-learning
by: Xie, Zhuyang, et al.
Published: (2024)
by: Xie, Zhuyang, et al.
Published: (2024)
HiCM$^2$: Hierarchical Compact Memory Modeling for Dense Video Captioning
by: Kim, Minkuk, et al.
Published: (2024)
by: Kim, Minkuk, et al.
Published: (2024)
Do You Remember? Dense Video Captioning with Cross-Modal Memory Retrieval
by: Kim, Minkuk, et al.
Published: (2024)
by: Kim, Minkuk, et al.
Published: (2024)
PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning
by: Xu, Lin, et al.
Published: (2024)
by: Xu, Lin, et al.
Published: (2024)
Enhancing Traffic Safety with Parallel Dense Video Captioning for End-to-End Event Analysis
by: Shoman, Maged, et al.
Published: (2024)
by: Shoman, Maged, et al.
Published: (2024)
CodecCap: High-Fidelity Codec-Inspired Residual Modeling for Dense Video Captioning
by: Lin, Zihan, et al.
Published: (2026)
by: Lin, Zihan, et al.
Published: (2026)
Dense Video Captioning using Graph-based Sentence Summarization
by: Zhang, Zhiwang, et al.
Published: (2025)
by: Zhang, Zhiwang, et al.
Published: (2025)
Self-Organizing Visual Prototypes for Non-Parametric Representation Learning
by: Silva, Thalles, et al.
Published: (2025)
by: Silva, Thalles, et al.
Published: (2025)
Learning from Memory: Non-Parametric Memory Augmented Self-Supervised Learning of Visual Features
by: Silva, Thalles, et al.
Published: (2024)
by: Silva, Thalles, et al.
Published: (2024)
SGCap: Decoding Semantic Group for Zero-shot Video Captioning
by: Pan, Zeyu, et al.
Published: (2025)
by: Pan, Zeyu, et al.
Published: (2025)
Set Prediction Guided by Semantic Concepts for Diverse Video Captioning
by: Lu, Yifan, et al.
Published: (2023)
by: Lu, Yifan, et al.
Published: (2023)
Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization
by: Zhang, Zhiwang, et al.
Published: (2025)
by: Zhang, Zhiwang, et al.
Published: (2025)
VidChain: Chain-of-Tasks with Metric-based Direct Preference Optimization for Dense Video Captioning
by: Lee, Ji Soo, et al.
Published: (2025)
by: Lee, Ji Soo, et al.
Published: (2025)
Stay in your Lane: Role Specific Queries with Overlap Suppression Loss for Dense Video Captioning
by: Baek, Seung Hyup, et al.
Published: (2026)
by: Baek, Seung Hyup, et al.
Published: (2026)
Similar Items
-
CEZSAR: A Contrastive Embedding Method for Zero-Shot Action Recognition
by: Estevam, Valter, et al.
Published: (2026) -
Advancing Multinational License Plate Recognition Through Synthetic and Real Data Fusion: A Comprehensive Evaluation
by: Laroca, Rayson, et al.
Published: (2026) -
Improving Small Drone Detection Through Multi-Scale Processing and Data Augmentation
by: Laroca, Rayson, et al.
Published: (2025) -
Enhancing License Plate Super-Resolution: A Layout-Aware and Character-Driven Approach
by: Nascimento, Valfride, et al.
Published: (2024) -
Toward Enhancing Vehicle Color Recognition in Adverse Conditions: A Dataset and Benchmark
by: Lima, Gabriel E., et al.
Published: (2024)