Gespeichert in:
| Hauptverfasser: | Wang, Wenqing, Fu, Yun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2508.12163 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Talking Tennis: Language Feedback from 3D Biomechanical Action Recognition
von: Dashore, Arushi, et al.
Veröffentlicht: (2025)
von: Dashore, Arushi, et al.
Veröffentlicht: (2025)
DexAvatar: 3D Sign Language Reconstruction with Hand and Body Pose Priors
von: Kundu, Kaustubh, et al.
Veröffentlicht: (2025)
von: Kundu, Kaustubh, et al.
Veröffentlicht: (2025)
A Medical Low-Back Pain Physical Rehabilitation Dataset for Human Body Movement Analysis
von: Nguyen, Sao Mai, et al.
Veröffentlicht: (2024)
von: Nguyen, Sao Mai, et al.
Veröffentlicht: (2024)
Seeing in the Dark: A Teacher-Student Framework for Dark Video Action Recognition via Knowledge Distillation and Contrastive Learning
von: Dass, Sharana Dharshikgan Suresh, et al.
Veröffentlicht: (2025)
von: Dass, Sharana Dharshikgan Suresh, et al.
Veröffentlicht: (2025)
ActNetFormer: Transformer-ResNet Hybrid Method for Semi-Supervised Action Recognition in Videos
von: Dass, Sharana Dharshikgan Suresh, et al.
Veröffentlicht: (2024)
von: Dass, Sharana Dharshikgan Suresh, et al.
Veröffentlicht: (2024)
EmoGene: Audio-Driven Emotional 3D Talking-Head Generation
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
RClicks: Realistic Click Simulation for Benchmarking Interactive Segmentation
von: Antonov, Anton, et al.
Veröffentlicht: (2024)
von: Antonov, Anton, et al.
Veröffentlicht: (2024)
Towards a GENEA Leaderboard -- an Extended, Living Benchmark for Evaluating and Advancing Conversational Motion Synthesis
von: Nagy, Rajmund, et al.
Veröffentlicht: (2024)
von: Nagy, Rajmund, et al.
Veröffentlicht: (2024)
EgoPoser: Robust Real-Time Egocentric Pose Estimation from Sparse and Intermittent Observations Everywhere
von: Jiang, Jiaxi, et al.
Veröffentlicht: (2023)
von: Jiang, Jiaxi, et al.
Veröffentlicht: (2023)
Visual Aesthetic Benchmark: Can Frontier Models Judge Beauty?
von: Feng, Yichen, et al.
Veröffentlicht: (2026)
von: Feng, Yichen, et al.
Veröffentlicht: (2026)
SPEAK: Speech-Driven Pose and Emotion-Adjustable Talking Head Generation
von: Cai, Changpeng, et al.
Veröffentlicht: (2024)
von: Cai, Changpeng, et al.
Veröffentlicht: (2024)
Human Motion Capture from Loose and Sparse Inertial Sensors with Garment-aware Diffusion Models
von: Ilic, Andela, et al.
Veröffentlicht: (2025)
von: Ilic, Andela, et al.
Veröffentlicht: (2025)
Group Inertial Poser: Multi-Person Pose and Global Translation from Sparse Inertial Sensors and Ultra-Wideband Ranging
von: Xue, Ying, et al.
Veröffentlicht: (2025)
von: Xue, Ying, et al.
Veröffentlicht: (2025)
AI-Based Screening for Depression and Social Anxiety Through Eye Tracking: An Exploratory Study
von: Chlasta, Karol, et al.
Veröffentlicht: (2025)
von: Chlasta, Karol, et al.
Veröffentlicht: (2025)
OUGS: Active View Selection via Object-aware Uncertainty Estimation in 3DGS
von: Li, Haiyi, et al.
Veröffentlicht: (2025)
von: Li, Haiyi, et al.
Veröffentlicht: (2025)
Adaptive Prompt Elicitation for Text-to-Image Generation
von: Wen, Xinyi, et al.
Veröffentlicht: (2026)
von: Wen, Xinyi, et al.
Veröffentlicht: (2026)
Neural Harmonic Textures for High-Quality Primitive Based Neural Reconstruction
von: Condor, Jorge, et al.
Veröffentlicht: (2026)
von: Condor, Jorge, et al.
Veröffentlicht: (2026)
Real Time Captioning of Sign Language Gestures in Video Meetings
von: Mukherjee, Sharanya, et al.
Veröffentlicht: (2025)
von: Mukherjee, Sharanya, et al.
Veröffentlicht: (2025)
StyleID: A Perception-Aware Dataset and Metric for Stylization-Agnostic Facial Identity Recognition
von: Yun, Kwan, et al.
Veröffentlicht: (2026)
von: Yun, Kwan, et al.
Veröffentlicht: (2026)
VLSlice: Interactive Vision-and-Language Slice Discovery
von: Slyman, Eric, et al.
Veröffentlicht: (2023)
von: Slyman, Eric, et al.
Veröffentlicht: (2023)
Learning Hierarchical Image Segmentation For Recognition and By Recognition
von: Ke, Tsung-Wei, et al.
Veröffentlicht: (2022)
von: Ke, Tsung-Wei, et al.
Veröffentlicht: (2022)
Enhancing Cross-Modal Contextual Congruence for Crowdfunding Success using Knowledge-infused Learning
von: Padhi, Trilok, et al.
Veröffentlicht: (2024)
von: Padhi, Trilok, et al.
Veröffentlicht: (2024)
Improving the Performance of Unimodal Dynamic Hand-Gesture Recognition with Multimodal Training
von: Abavisani, Mahdi, et al.
Veröffentlicht: (2018)
von: Abavisani, Mahdi, et al.
Veröffentlicht: (2018)
Claycode: Stylable and Deformable 2D Scannable Codes
von: Maida, Marco, et al.
Veröffentlicht: (2025)
von: Maida, Marco, et al.
Veröffentlicht: (2025)
Multimodal Fusion of EMG and Vision for Human Grasp Intent Inference in Prosthetic Hand Control
von: Zandigohar, Mehrshad, et al.
Veröffentlicht: (2021)
von: Zandigohar, Mehrshad, et al.
Veröffentlicht: (2021)
LLM-Guided Agentic Floor Plan Parsing for Accessible Indoor Navigation of Blind and Low-Vision People
von: Ayanzadeh, Aydin, et al.
Veröffentlicht: (2026)
von: Ayanzadeh, Aydin, et al.
Veröffentlicht: (2026)
Explaining Explainability: Recommendations for Effective Use of Concept Activation Vectors
von: Nicolson, Angus, et al.
Veröffentlicht: (2024)
von: Nicolson, Angus, et al.
Veröffentlicht: (2024)
Vision-Language Cross-Attention for Real-Time Autonomous Driving
von: Patapati, Santosh, et al.
Veröffentlicht: (2025)
von: Patapati, Santosh, et al.
Veröffentlicht: (2025)
Mahalanobis PatchCore: Covariance-Aware and Streaming-Compatible Industrial Anomaly Detection
von: Ferrari, Niccolò, et al.
Veröffentlicht: (2026)
von: Ferrari, Niccolò, et al.
Veröffentlicht: (2026)
Underwater SONAR Image Classification and Analysis using LIME-based Explainable Artificial Intelligence
von: Natarajan, Purushothaman, et al.
Veröffentlicht: (2024)
von: Natarajan, Purushothaman, et al.
Veröffentlicht: (2024)
Topological Structure Description for Artcode Detection Using the Shape of Orientation Histogram
von: Xu, Liming, et al.
Veröffentlicht: (2025)
von: Xu, Liming, et al.
Veröffentlicht: (2025)
Approximate UMAP allows for high-rate online visualization of high-dimensional data streams
von: Wassenaar, Peter, et al.
Veröffentlicht: (2024)
von: Wassenaar, Peter, et al.
Veröffentlicht: (2024)
Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark
von: Nagy, Rajmund, et al.
Veröffentlicht: (2025)
von: Nagy, Rajmund, et al.
Veröffentlicht: (2025)
OpenFake: An Open Dataset and Platform Toward Real-World Deepfake Detection
von: Livernoche, Victor, et al.
Veröffentlicht: (2025)
von: Livernoche, Victor, et al.
Veröffentlicht: (2025)
TouchInsight: Uncertainty-aware Rapid Touch and Text Input for Mixed Reality from Egocentric Vision
von: Streli, Paul, et al.
Veröffentlicht: (2024)
von: Streli, Paul, et al.
Veröffentlicht: (2024)
Multimodal 3D Fusion and In-Situ Learning for Spatially Aware AI
von: Xu, Chengyuan, et al.
Veröffentlicht: (2024)
von: Xu, Chengyuan, et al.
Veröffentlicht: (2024)
Phase-Aware Wavelet-Based-Scattering Encoder-Decoder for Dense Predictions
von: Marrakchi, Ghassen, et al.
Veröffentlicht: (2026)
von: Marrakchi, Ghassen, et al.
Veröffentlicht: (2026)
NTIRE 2026 Challenge on Video Saliency Prediction: Methods and Results
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2026)
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2026)
BREPS: Bounding-Box Robustness Evaluation of Promptable Segmentation
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2026)
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2026)
Hierarchical Pre-Training of Vision Encoders with Large Language Models
von: Lee, Eugene, et al.
Veröffentlicht: (2026)
von: Lee, Eugene, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Talking Tennis: Language Feedback from 3D Biomechanical Action Recognition
von: Dashore, Arushi, et al.
Veröffentlicht: (2025) -
DexAvatar: 3D Sign Language Reconstruction with Hand and Body Pose Priors
von: Kundu, Kaustubh, et al.
Veröffentlicht: (2025) -
A Medical Low-Back Pain Physical Rehabilitation Dataset for Human Body Movement Analysis
von: Nguyen, Sao Mai, et al.
Veröffentlicht: (2024) -
Seeing in the Dark: A Teacher-Student Framework for Dark Video Action Recognition via Knowledge Distillation and Contrastive Learning
von: Dass, Sharana Dharshikgan Suresh, et al.
Veröffentlicht: (2025) -
ActNetFormer: Transformer-ResNet Hybrid Method for Semi-Supervised Action Recognition in Videos
von: Dass, Sharana Dharshikgan Suresh, et al.
Veröffentlicht: (2024)