TKN: Transformer-based Keypoint Prediction Network For Real-time Video Prediction
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Haoran, Li, XiaoLu, Lin, Yihang, Hao, Yanbin, Xie, Haiyong, Zhou, Pengyuan, Liao, Yong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Noise-NeRF: Hide Information in Neural Radiance Fields using Trainable Noise
di: Huang, Qinglong, et al.
Pubblicazione: (2024)
di: Huang, Qinglong, et al.
Pubblicazione: (2024)
3D-GOI: 3D GAN Omni-Inversion for Multifaceted and Multi-object Editing
di: Li, Haoran, et al.
Pubblicazione: (2023)
di: Li, Haoran, et al.
Pubblicazione: (2023)
Viewport Prediction for Volumetric Video Streaming by Exploring Video Saliency and Trajectory Information
di: Li, Jie, et al.
Pubblicazione: (2023)
di: Li, Jie, et al.
Pubblicazione: (2023)
A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming
di: Zhou, Pengyuan, et al.
Pubblicazione: (2024)
di: Zhou, Pengyuan, et al.
Pubblicazione: (2024)
Real-time Video Prediction With Fast Video Interpolation Model and Prediction Training
di: Hirose, Shota, et al.
Pubblicazione: (2025)
di: Hirose, Shota, et al.
Pubblicazione: (2025)
DreamScene: 3D Gaussian-based Text-to-3D Scene Generation via Formation Pattern Sampling
di: Li, Haoran, et al.
Pubblicazione: (2024)
di: Li, Haoran, et al.
Pubblicazione: (2024)
Sparse2Dense: A Keypoint-driven Generative Framework for Human Video Compression and Vertex Prediction
di: Chen, Bolin, et al.
Pubblicazione: (2025)
di: Chen, Bolin, et al.
Pubblicazione: (2025)
Video Prediction Transformers without Recurrence or Convolution
di: Tang, Yujin, et al.
Pubblicazione: (2024)
di: Tang, Yujin, et al.
Pubblicazione: (2024)
Recent development on nanomaterial‐based biosensors for identifying thyroid tumor biomarkers
di: Kun Xu, et al.
Pubblicazione: (2024)
di: Kun Xu, et al.
Pubblicazione: (2024)
Detecting AI-Generated Video via Frame Consistency
di: Ma, Long, et al.
Pubblicazione: (2024)
di: Ma, Long, et al.
Pubblicazione: (2024)
Steering Video Diffusion Transformers with Massive Activations
di: Cheng, Xianhang, et al.
Pubblicazione: (2026)
di: Cheng, Xianhang, et al.
Pubblicazione: (2026)
MacFormer: Map-Agent Coupled Transformer for Real-time and Robust Trajectory Prediction
di: Feng, Chen, et al.
Pubblicazione: (2023)
di: Feng, Chen, et al.
Pubblicazione: (2023)
FAKER: Full-body Anonymization with Human Keypoint Extraction for Real-time Video Deidentification
di: Ban, Byunghyun, et al.
Pubblicazione: (2024)
di: Ban, Byunghyun, et al.
Pubblicazione: (2024)
GMatch: A Lightweight, Geometry-Constrained Keypoint Matcher for Zero-Shot 6DoF Pose Estimation in Robotic Grasp Tasks
di: Yang, Ming, et al.
Pubblicazione: (2025)
di: Yang, Ming, et al.
Pubblicazione: (2025)
How does spatial structure affect psychological restoration? A method based on Graph Neural Networks and Street View Imagery
di: Ma, Haoran, et al.
Pubblicazione: (2023)
di: Ma, Haoran, et al.
Pubblicazione: (2023)
Enhancing Bandwidth Efficiency for Video Motion Transfer Applications using Deep Learning Based Keypoint Prediction
di: Bai, Xue, et al.
Pubblicazione: (2024)
di: Bai, Xue, et al.
Pubblicazione: (2024)
Keypoint Detection and Description for Raw Bayer Images
di: Lin, Jiakai, et al.
Pubblicazione: (2025)
di: Lin, Jiakai, et al.
Pubblicazione: (2025)
High-Fidelity Document Stain Removal via A Large-Scale Real-World Dataset and A Memory-Augmented Transformer
di: Li, Mingxian, et al.
Pubblicazione: (2024)
di: Li, Mingxian, et al.
Pubblicazione: (2024)
Empowering Vector Graphics with Consistently Arbitrary Viewing and View-dependent Visibility
di: Li, Yidi, et al.
Pubblicazione: (2025)
di: Li, Yidi, et al.
Pubblicazione: (2025)
Unified Dense Prediction of Video Diffusion
di: Yang, Lehan, et al.
Pubblicazione: (2025)
di: Yang, Lehan, et al.
Pubblicazione: (2025)
Accelerating Diffusion Transformer via Gradient-Optimized Cache
di: Qiu, Junxiang, et al.
Pubblicazione: (2025)
di: Qiu, Junxiang, et al.
Pubblicazione: (2025)
SD-Net: Symmetric-Aware Keypoint Prediction and Domain Adaptation for 6D Pose Estimation In Bin-picking Scenarios
di: Huang, Ding-Tao, et al.
Pubblicazione: (2024)
di: Huang, Ding-Tao, et al.
Pubblicazione: (2024)
LongLive: Real-time Interactive Long Video Generation
di: Yang, Shuai, et al.
Pubblicazione: (2025)
di: Yang, Shuai, et al.
Pubblicazione: (2025)
Proact-VL: A Proactive VideoLLM for Real-Time AI Companions
di: Yan, Weicai, et al.
Pubblicazione: (2026)
di: Yan, Weicai, et al.
Pubblicazione: (2026)
GIFStream: 4D Gaussian-based Immersive Video with Feature Stream
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
Expressive Keypoints for Skeleton-based Action Recognition via Skeleton Transformation
di: Yang, Yijie, et al.
Pubblicazione: (2024)
di: Yang, Yijie, et al.
Pubblicazione: (2024)
GKNet: Graph-based Keypoints Network for Monocular Pose Estimation of Non-cooperative Spacecraft
di: Ma, Weizhao, et al.
Pubblicazione: (2025)
di: Ma, Weizhao, et al.
Pubblicazione: (2025)
Comparative and Interpretative Analysis of CNN and Transformer Models in Predicting Wildfire Spread Using Remote Sensing Data
di: Zhou, Yihang, et al.
Pubblicazione: (2025)
di: Zhou, Yihang, et al.
Pubblicazione: (2025)
Towards Online Real-Time Memory-based Video Inpainting Transformers
di: Thiry, Guillaume, et al.
Pubblicazione: (2024)
di: Thiry, Guillaume, et al.
Pubblicazione: (2024)
Photovoltaic Defect Image Generator with Boundary Alignment Smoothing Constraint for Domain Shift Mitigation
di: Li, Dongying, et al.
Pubblicazione: (2025)
di: Li, Dongying, et al.
Pubblicazione: (2025)
A Detector-oblivious Multi-arm Network for Keypoint Matching
di: Shen, Xuelun, et al.
Pubblicazione: (2021)
di: Shen, Xuelun, et al.
Pubblicazione: (2021)
PosMLP-Video: Spatial and Temporal Relative Position Encoding for Efficient Video Recognition
di: Hao, Yanbin, et al.
Pubblicazione: (2024)
di: Hao, Yanbin, et al.
Pubblicazione: (2024)
A Real-to-Sim-to-Real Approach to Robotic Manipulation with VLM-Generated Iterative Keypoint Rewards
di: Patel, Shivansh, et al.
Pubblicazione: (2025)
di: Patel, Shivansh, et al.
Pubblicazione: (2025)
DreamScene: 3D Gaussian-based End-to-end Text-to-3D Scene Generation
di: Li, Haoran, et al.
Pubblicazione: (2025)
di: Li, Haoran, et al.
Pubblicazione: (2025)
Transformer-based Video Saliency Prediction with High Temporal Dimension Decoding
di: Moradi, Morteza, et al.
Pubblicazione: (2024)
di: Moradi, Morteza, et al.
Pubblicazione: (2024)
OccupancyDETR: Using DETR for Mixed Dense-sparse 3D Occupancy Prediction
di: Jia, Yupeng, et al.
Pubblicazione: (2023)
di: Jia, Yupeng, et al.
Pubblicazione: (2023)
Accelerating Diffusion Transformer via Error-Optimized Cache
di: Qiu, Junxiang, et al.
Pubblicazione: (2025)
di: Qiu, Junxiang, et al.
Pubblicazione: (2025)
Satellite Streaming Video QoE Prediction: A Real-World Subjective Database and Network-Level Prediction Models
di: Chen, Bowen, et al.
Pubblicazione: (2024)
di: Chen, Bowen, et al.
Pubblicazione: (2024)
SCP: Spatial Causal Prediction in Video
di: Zhao, Yanguang, et al.
Pubblicazione: (2026)
di: Zhao, Yanguang, et al.
Pubblicazione: (2026)
Keypoint-based Dynamic Object 6-DoF Pose Tracking via Event Camera
di: Wang, Zhe, et al.
Pubblicazione: (2026)
di: Wang, Zhe, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Noise-NeRF: Hide Information in Neural Radiance Fields using Trainable Noise
di: Huang, Qinglong, et al.
Pubblicazione: (2024) -
3D-GOI: 3D GAN Omni-Inversion for Multifaceted and Multi-object Editing
di: Li, Haoran, et al.
Pubblicazione: (2023) -
Viewport Prediction for Volumetric Video Streaming by Exploring Video Saliency and Trajectory Information
di: Li, Jie, et al.
Pubblicazione: (2023) -
A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming
di: Zhou, Pengyuan, et al.
Pubblicazione: (2024) -
Real-time Video Prediction With Fast Video Interpolation Model and Prediction Training
di: Hirose, Shota, et al.
Pubblicazione: (2025)