HighlightMe: Detecting Highlights from Human-Centric Videos
Fuente:
arXiv
Guardado en:
| Autores principales: | Bhattacharya, Uttaran, Wu, Gang, Petrangeli, Stefano, Swaminathan, Viswanathan, Manocha, Dinesh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2021
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Show Me What I Like: Detecting User-Specific Video Highlights Using Content-Based Multi-Head Attention
por: Bhattacharya, Uttaran, et al.
Publicado: (2022)
por: Bhattacharya, Uttaran, et al.
Publicado: (2022)
Speech2UnifiedExpressions: Synchronous Synthesis of Co-Speech Affective Face and Body Expressions from Affordable Inputs
por: Bhattacharya, Uttaran, et al.
Publicado: (2024)
por: Bhattacharya, Uttaran, et al.
Publicado: (2024)
Proc3D: Procedural 3D Generation and Parametric Editing of 3D Shapes with Large Language Models
por: Raji, Fadlullah, et al.
Publicado: (2026)
por: Raji, Fadlullah, et al.
Publicado: (2026)
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
por: Bhattacharya, Uttaran, et al.
Publicado: (2019)
por: Bhattacharya, Uttaran, et al.
Publicado: (2019)
ShotVL: Human-Centric Highlight Frame Retrieval via Language Queries
por: Xue, Wangyu, et al.
Publicado: (2024)
por: Xue, Wangyu, et al.
Publicado: (2024)
Unleash the Potential of CLIP for Video Highlight Detection
por: Han, Donghoon, et al.
Publicado: (2024)
por: Han, Donghoon, et al.
Publicado: (2024)
Unsupervised Video Highlight Detection by Learning from Audio and Visual Recurrence
por: Islam, Zahidul, et al.
Publicado: (2024)
por: Islam, Zahidul, et al.
Publicado: (2024)
KFFocus: Highlighting Keyframes for Enhanced Video Understanding
por: Nie, Ming, et al.
Publicado: (2025)
por: Nie, Ming, et al.
Publicado: (2025)
Inst4DGS: Instance-Decomposed 4D Gaussian Splatting with Multi-Video Label Permutation Learning
por: Lee, Yonghan, et al.
Publicado: (2026)
por: Lee, Yonghan, et al.
Publicado: (2026)
Unsupervised Modality-Transferable Video Highlight Detection with Representation Activation Sequence Learning
por: Li, Tingtian, et al.
Publicado: (2024)
por: Li, Tingtian, et al.
Publicado: (2024)
DiffusionVMR: Diffusion Model for Joint Video Moment Retrieval and Highlight Detection
por: Zhao, Henghao, et al.
Publicado: (2023)
por: Zhao, Henghao, et al.
Publicado: (2023)
Unsupervised Transcript-assisted Video Summarization and Highlight Detection
por: Barbakos, Spyros, et al.
Publicado: (2025)
por: Barbakos, Spyros, et al.
Publicado: (2025)
Take an Emotion Walk: Perceiving Emotions from Gaits Using Hierarchical Attention Pooling and Affective Mapping
por: Bhattacharya, Uttaran, et al.
Publicado: (2019)
por: Bhattacharya, Uttaran, et al.
Publicado: (2019)
Gameplay Highlights Generation
por: Edithal, Vignesh, et al.
Publicado: (2025)
por: Edithal, Vignesh, et al.
Publicado: (2025)
Multi-modal Fusion and Query Refinement Network for Video Moment Retrieval and Highlight Detection
por: Xu, Yifang, et al.
Publicado: (2025)
por: Xu, Yifang, et al.
Publicado: (2025)
Differentiable Frequency-based Disentanglement for Aerial Video Action Recognition
por: Kothandaraman, Divya, et al.
Publicado: (2022)
por: Kothandaraman, Divya, et al.
Publicado: (2022)
Efficient and Robust Registration on the 3D Special Euclidean Group
por: Bhattacharya, Uttaran, et al.
Publicado: (2019)
por: Bhattacharya, Uttaran, et al.
Publicado: (2019)
Saliency-Guided DETR for Moment Retrieval and Highlight Detection
por: Gordeev, Aleksandr, et al.
Publicado: (2024)
por: Gordeev, Aleksandr, et al.
Publicado: (2024)
Moment and Highlight Detection via MLLM Frame Segmentation
por: Jiwanta, I Putu Andika Bagas, et al.
Publicado: (2025)
por: Jiwanta, I Putu Andika Bagas, et al.
Publicado: (2025)
Automated Detection of Sport Highlights from Audio and Video Sources
por: Della Santa, Francesco, et al.
Publicado: (2025)
por: Della Santa, Francesco, et al.
Publicado: (2025)
Sounding Highlights: Dual-Pathway Audio Encoders for Audio-Visual Video Highlight Detection
por: Joo, Seohyun, et al.
Publicado: (2026)
por: Joo, Seohyun, et al.
Publicado: (2026)
UAV4D: Dynamic Neural Rendering of Human-Centric UAV Imagery using Gaussian Splatting
por: Choi, Jaehoon, et al.
Publicado: (2025)
por: Choi, Jaehoon, et al.
Publicado: (2025)
MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic Learning
por: Ma, Hongxu, et al.
Publicado: (2025)
por: Ma, Hongxu, et al.
Publicado: (2025)
Test-Time Adaptation for Video Highlight Detection Using Meta-Auxiliary Learning and Cross-Modality Hallucinations
por: Islam, Zahidul, et al.
Publicado: (2025)
por: Islam, Zahidul, et al.
Publicado: (2025)
Generating Narrated Lecture Videos from Slides with Synchronized Highlights
por: Holmberg, Alexander
Publicado: (2025)
por: Holmberg, Alexander
Publicado: (2025)
Lighthouse: A User-Friendly Library for Reproducible Video Moment Retrieval and Highlight Detection
por: Nishimura, Taichi, et al.
Publicado: (2024)
por: Nishimura, Taichi, et al.
Publicado: (2024)
GT-SVJ: Generative-Transformer-Based Self-Supervised Video Judge For Efficient Video Reward Modeling
por: Shekhar, Shivanshu, et al.
Publicado: (2026)
por: Shekhar, Shivanshu, et al.
Publicado: (2026)
PACE: Data-Driven Virtual Agent Interaction in Dense and Cluttered Environments
por: Mullen, James, et al.
Publicado: (2023)
por: Mullen, James, et al.
Publicado: (2023)
Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight Detection
por: Um, Sung Jin, et al.
Publicado: (2025)
por: Um, Sung Jin, et al.
Publicado: (2025)
VideoLights: Feature Refinement and Cross-Task Alignment Transformer for Joint Video Highlight Detection and Moment Retrieval
por: Paul, Dhiman, et al.
Publicado: (2024)
por: Paul, Dhiman, et al.
Publicado: (2024)
Commentary Generation for Soccer Highlights
por: Ravuru, Chidaksh
Publicado: (2025)
por: Ravuru, Chidaksh
Publicado: (2025)
Large Content And Behavior Models To Understand, Simulate, And Optimize Content And Behavior
por: Khandelwal, Ashmit, et al.
Publicado: (2023)
por: Khandelwal, Ashmit, et al.
Publicado: (2023)
COACH: Collaborative Agents for Contextual Highlighting -- A Multi-Agent Framework for Sports Video Analysis
por: Wong, Tsz-To, et al.
Publicado: (2025)
por: Wong, Tsz-To, et al.
Publicado: (2025)
HUMOTO: A 4D Dataset of Mocap Human Object Interactions
por: Lu, Jiaxin, et al.
Publicado: (2025)
por: Lu, Jiaxin, et al.
Publicado: (2025)
CoSTL: Comprehensive Spatial-Temporal Representation Learning for Moment Retrieval and Highlight Detection
por: Dong, Xin, et al.
Publicado: (2026)
por: Dong, Xin, et al.
Publicado: (2026)
A Highlight Removal Method for Capsule Endoscopy Images
por: Zhang, Shaojie, et al.
Publicado: (2024)
por: Zhang, Shaojie, et al.
Publicado: (2024)
Dual-Hybrid Attention Network for Specular Highlight Removal
por: Guo, Xiaojiao, et al.
Publicado: (2024)
por: Guo, Xiaojiao, et al.
Publicado: (2024)
HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting
por: Lee, Jeongeun, et al.
Publicado: (2025)
por: Lee, Jeongeun, et al.
Publicado: (2025)
SimMotionEdit: Text-Based Human Motion Editing with Motion Similarity Prediction
por: Li, Zhengyuan, et al.
Publicado: (2025)
por: Li, Zhengyuan, et al.
Publicado: (2025)
SCP: Soft Conditional Prompt Learning for Aerial Video Action Recognition
por: Wang, Xijun, et al.
Publicado: (2023)
por: Wang, Xijun, et al.
Publicado: (2023)
Ejemplares similares
-
Show Me What I Like: Detecting User-Specific Video Highlights Using Content-Based Multi-Head Attention
por: Bhattacharya, Uttaran, et al.
Publicado: (2022) -
Speech2UnifiedExpressions: Synchronous Synthesis of Co-Speech Affective Face and Body Expressions from Affordable Inputs
por: Bhattacharya, Uttaran, et al.
Publicado: (2024) -
Proc3D: Procedural 3D Generation and Parametric Editing of 3D Shapes with Large Language Models
por: Raji, Fadlullah, et al.
Publicado: (2026) -
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
por: Bhattacharya, Uttaran, et al.
Publicado: (2019) -
ShotVL: Human-Centric Highlight Frame Retrieval via Language Queries
por: Xue, Wangyu, et al.
Publicado: (2024)