ALIVE: An Avatar-Lecture Interactive Video Engine with Content-Aware Retrieval for Real-Time Interaction
Fuente:
arXiv
Saved in:
| Main Authors: | Islam, Md Zabirul, Manik, Md Motaleb Hossen, Wang, Ge |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SlideChain: Semantic Provenance for Lecture Understanding via Blockchain Registration
by: Manik, Md Motaleb Hossen, et al.
Published: (2025)
by: Manik, Md Motaleb Hossen, et al.
Published: (2025)
N-ReLU: Zero-Mean Stochastic Extension of ReLU
by: Manik, Md Motaleb Hossen, et al.
Published: (2025)
by: Manik, Md Motaleb Hossen, et al.
Published: (2025)
Unified Deployment-Aware Evaluation of Open Reasoning Language Models
by: Manik, Md Motaleb Hossen, et al.
Published: (2026)
by: Manik, Md Motaleb Hossen, et al.
Published: (2026)
Emergent decentralized regulation in a purely synthetic society
by: Manik, Md Motaleb Hossen, et al.
Published: (2026)
by: Manik, Md Motaleb Hossen, et al.
Published: (2026)
ALIVE: Animate Your World with Lifelike Audio-Video Generation
by: Guo, Ying, et al.
Published: (2026)
by: Guo, Ying, et al.
Published: (2026)
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models
by: Yu, Haojie, et al.
Published: (2025)
by: Yu, Haojie, et al.
Published: (2025)
Leveraging Pre-trained CNNs for Efficient Feature Extraction in Rice Leaf Disease Classification
by: Sobuj, Md. Shohanur Islam, et al.
Published: (2024)
by: Sobuj, Md. Shohanur Islam, et al.
Published: (2024)
ElderFallGuard: Real-Time IoT and Computer Vision-Based Fall Detection System for Elderly Safety
by: Riahi, Tasrifur, et al.
Published: (2025)
by: Riahi, Tasrifur, et al.
Published: (2025)
StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars
by: Sun, Zhiyao, et al.
Published: (2025)
by: Sun, Zhiyao, et al.
Published: (2025)
Real-Time Animatable 2DGS-Avatars with Detail Enhancement from Monocular Videos
by: Yuan, Xia, et al.
Published: (2025)
by: Yuan, Xia, et al.
Published: (2025)
IKIWISI: An Interactive Visual Pattern Generator for Evaluating the Reliability of Vision-Language Models Without Ground Truth
by: Islam, Md Touhidul, et al.
Published: (2025)
by: Islam, Md Touhidul, et al.
Published: (2025)
RGNet: A Unified Clip Retrieval and Grounding Network for Long Videos
by: Hannan, Tanveer, et al.
Published: (2023)
by: Hannan, Tanveer, et al.
Published: (2023)
InteractAvatar: Modeling Hand-Face Interaction in Photorealistic Avatars with Deformable Gaussians
by: Chen, Kefan, et al.
Published: (2025)
by: Chen, Kefan, et al.
Published: (2025)
Real-Time Multi-Modal Embedded Vision Framework for Object Detection Facial Emotion Recognition and Biometric Identification on Low-Power Edge Platforms
by: Zahid, S. M. Khalid Bin, et al.
Published: (2026)
by: Zahid, S. M. Khalid Bin, et al.
Published: (2026)
RIVER: A Real-Time Interaction Benchmark for Video LLMs
by: Shi, Yansong, et al.
Published: (2026)
by: Shi, Yansong, et al.
Published: (2026)
Position: Interactive Generative Video as Next-Generation Game Engine
by: Yu, Jiwen, et al.
Published: (2025)
by: Yu, Jiwen, et al.
Published: (2025)
CAST: Channel-Aware Spatial Transfer Learning with Pseudo-Image Radar for Sign Language Recognition
by: Shujon, Md. Shakhoyat Rahman, et al.
Published: (2026)
by: Shujon, Md. Shakhoyat Rahman, et al.
Published: (2026)
ViSA: 3D-Aware Video Shading for Real-Time Upper-Body Avatar Creation
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
CSRAP: Enhanced Canvas Attention Scheduling for Real-Time Mission Critical Perception
by: Sakib, Md Iftekharul Islam, et al.
Published: (2025)
by: Sakib, Md Iftekharul Islam, et al.
Published: (2025)
BanglaMM-Disaster: A Multimodal Transformer-Based Deep Learning Framework for Multiclass Disaster Classification in Bangla
by: Islam, Ariful, et al.
Published: (2025)
by: Islam, Ariful, et al.
Published: (2025)
RealCam: Real-Time Novel-View Video Generation with Interactive Camera Control
by: Xu, Youcan, et al.
Published: (2026)
by: Xu, Youcan, et al.
Published: (2026)
Privacy-Preserving Empathy Detection in Video Interactions
by: Hasan, Md Rakibul, et al.
Published: (2025)
by: Hasan, Md Rakibul, et al.
Published: (2025)
Jellyfish Species Identification: A CNN Based Artificial Neural Network Approach
by: Hossen, Md. Sabbir, et al.
Published: (2025)
by: Hossen, Md. Sabbir, et al.
Published: (2025)
Interactive Rendering of Relightable and Animatable Gaussian Avatars
by: Zhan, Youyi, et al.
Published: (2024)
by: Zhan, Youyi, et al.
Published: (2024)
Analyzing the Dynamics of COVID-19 Lockdown Success: Insights from Regional Data and Public Health Measures
by: Manik, Md. Motaleb Hossen, et al.
Published: (2024)
by: Manik, Md. Motaleb Hossen, et al.
Published: (2024)
DyCAF-Net: Dynamic Class-Aware Fusion Network
by: Jahin, Md Abrar, et al.
Published: (2025)
by: Jahin, Md Abrar, et al.
Published: (2025)
Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation
by: Ki, Taekyung, et al.
Published: (2026)
by: Ki, Taekyung, et al.
Published: (2026)
A CNN-Based Malaria Diagnosis from Blood Cell Images with SHAP and LIME Explainability
by: Abir, Md. Ismiel Hossen, et al.
Published: (2025)
by: Abir, Md. Ismiel Hossen, et al.
Published: (2025)
ChatGPT vs. DeepSeek: A Comparative Study on AI-Based Code Generation
by: Manik, Md Motaleb Hossen
Published: (2025)
by: Manik, Md Motaleb Hossen
Published: (2025)
AvatarPose: Avatar-guided 3D Pose Estimation of Close Human Interaction from Sparse Multi-view Videos
by: Lu, Feichi, et al.
Published: (2024)
by: Lu, Feichi, et al.
Published: (2024)
ADAPT: AI-Driven Decentralized Adaptive Publishing Testbed
by: Manik, Md Motaleb Hossen, et al.
Published: (2026)
by: Manik, Md Motaleb Hossen, et al.
Published: (2026)
OpenClaw Agents on Moltbook: Risky Instruction Sharing and Norm Enforcement in an Agent-Only Social Network
by: Manik, Md Motaleb Hossen, et al.
Published: (2026)
by: Manik, Md Motaleb Hossen, et al.
Published: (2026)
HeBA: Heterogeneous Bottleneck Adapters for Robust Vision-Language Models
by: Islam, Md Jahidul
Published: (2026)
by: Islam, Md Jahidul
Published: (2026)
A Critical Analysis on Machine Learning Techniques for Video-based Human Activity Recognition of Surveillance Systems: A Review
by: Jahan, Shahriar, et al.
Published: (2024)
by: Jahan, Shahriar, et al.
Published: (2024)
SelfieAvatar: Real-time Head Avatar reenactment from a Selfie Video
by: Liang, Wei, et al.
Published: (2026)
by: Liang, Wei, et al.
Published: (2026)
LLandMark: A Multi-Agent Framework for Landmark-Aware Multimodal Interactive Video Retrieval
by: Phung, Minh-Chi, et al.
Published: (2026)
by: Phung, Minh-Chi, et al.
Published: (2026)
FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization
by: Song, Quanjian, et al.
Published: (2026)
by: Song, Quanjian, et al.
Published: (2026)
SemanticDraw: Towards Real-Time Interactive Content Creation from Image Diffusion Models
by: Lee, Jaerin, et al.
Published: (2024)
by: Lee, Jaerin, et al.
Published: (2024)
SplattingAvatar: Realistic Real-Time Human Avatars with Mesh-Embedded Gaussian Splatting
by: Shao, Zhijing, et al.
Published: (2024)
by: Shao, Zhijing, et al.
Published: (2024)
RITA: A Real-time Interactive Talking Avatars Framework
by: Cheng, Wuxinlin, et al.
Published: (2024)
by: Cheng, Wuxinlin, et al.
Published: (2024)
Similar Items
-
SlideChain: Semantic Provenance for Lecture Understanding via Blockchain Registration
by: Manik, Md Motaleb Hossen, et al.
Published: (2025) -
N-ReLU: Zero-Mean Stochastic Extension of ReLU
by: Manik, Md Motaleb Hossen, et al.
Published: (2025) -
Unified Deployment-Aware Evaluation of Open Reasoning Language Models
by: Manik, Md Motaleb Hossen, et al.
Published: (2026) -
Emergent decentralized regulation in a purely synthetic society
by: Manik, Md Motaleb Hossen, et al.
Published: (2026) -
ALIVE: Animate Your World with Lifelike Audio-Video Generation
by: Guo, Ying, et al.
Published: (2026)