Zero-Shot Open-Vocabulary Human Motion Grounding with Test-Time Training
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Yunjiao, Chen, Xinyan, Qian, Junlang, Xie, Lihua, Yang, Jianfei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AdaPose: Towards Cross-Site Device-Free Human Pose Estimation with Commodity WiFi
by: Zhou, Yunjiao, et al.
Published: (2023)
by: Zhou, Yunjiao, et al.
Published: (2023)
T3DNet: Compressing Point Cloud Models for Lightweight 3D Recognition
by: Yang, Zhiyuan, et al.
Published: (2024)
by: Yang, Zhiyuan, et al.
Published: (2024)
MaskFi: Unsupervised Learning of WiFi and Vision Representations for Multimodal Human Activity Recognition
by: Yang, Jianfei, et al.
Published: (2024)
by: Yang, Jianfei, et al.
Published: (2024)
FACTOR: Counterfactual Training-Free Test-Time Adaptation for Open-Vocabulary Object Detection
by: Zhao, Kaixiang, et al.
Published: (2026)
by: Zhao, Kaixiang, et al.
Published: (2026)
SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
by: Li, Rong, et al.
Published: (2024)
by: Li, Rong, et al.
Published: (2024)
mmPred: Radar-based Human Motion Prediction in the Dark
by: Fan, Junqiao, et al.
Published: (2025)
by: Fan, Junqiao, et al.
Published: (2025)
Zero-Shot Open-Vocabulary Tracking with Large Pre-Trained Models
by: Chu, Wen-Hsuan, et al.
Published: (2023)
by: Chu, Wen-Hsuan, et al.
Published: (2023)
M4Human: A Large-Scale Multimodal mmWave Radar Benchmark for Human Mesh Reconstruction
by: Fan, Junqiao, et al.
Published: (2025)
by: Fan, Junqiao, et al.
Published: (2025)
SkeFi: Cross-Modal Knowledge Transfer for Wireless Skeleton-Based Action Recognition
by: Huang, Shunyu, et al.
Published: (2026)
by: Huang, Shunyu, et al.
Published: (2026)
X-Fi: A Modality-Invariant Foundation Model for Multimodal Human Sensing
by: Chen, Xinyan, et al.
Published: (2024)
by: Chen, Xinyan, et al.
Published: (2024)
GHOST: Grounded Human Motion Generation with Open Vocabulary Scene-and-Text Contexts
by: Milacski, Zoltán Á., et al.
Published: (2024)
by: Milacski, Zoltán Á., et al.
Published: (2024)
Synthetic Captions for Open-Vocabulary Zero-Shot Segmentation
by: Lebailly, Tim, et al.
Published: (2025)
by: Lebailly, Tim, et al.
Published: (2025)
RSVG-ZeroOV: Exploring a Training-Free Framework for Zero-Shot Open-Vocabulary Visual Grounding in Remote Sensing Images
by: Li, Ke, et al.
Published: (2025)
by: Li, Ke, et al.
Published: (2025)
A Simple Framework for Open-Vocabulary Zero-Shot Segmentation
by: Stegmüller, Thomas, et al.
Published: (2024)
by: Stegmüller, Thomas, et al.
Published: (2024)
Generative Dataset Distillation using Min-Max Diffusion Model
by: Fan, Junqiao, et al.
Published: (2025)
by: Fan, Junqiao, et al.
Published: (2025)
Training-Free Class Purification for Open-Vocabulary Semantic Segmentation
by: Chen, Qi, et al.
Published: (2025)
by: Chen, Qi, et al.
Published: (2025)
Exploring Vision-Language Models for Open-Vocabulary Zero-Shot Action Segmentation
by: Unmesh, Asim, et al.
Published: (2026)
by: Unmesh, Asim, et al.
Published: (2026)
Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding
by: Yuan, Zhihao, et al.
Published: (2023)
by: Yuan, Zhihao, et al.
Published: (2023)
Interactive Test-Time Adaptation with Reliable Spatial-Temporal Voxels for Multi-Modal Segmentation
by: Cao, Haozhi, et al.
Published: (2024)
by: Cao, Haozhi, et al.
Published: (2024)
Prototype-Aware Multimodal Alignment for Open-Vocabulary Visual Grounding
by: Xie, Jiangnan, et al.
Published: (2025)
by: Xie, Jiangnan, et al.
Published: (2025)
DOZE: A Dataset for Open-Vocabulary Zero-Shot Object Navigation in Dynamic Environments
by: Ma, Ji, et al.
Published: (2024)
by: Ma, Ji, et al.
Published: (2024)
ForgeryTTT: Zero-Shot Image Manipulation Localization with Test-Time Training
by: Liu, Weihuang, et al.
Published: (2024)
by: Liu, Weihuang, et al.
Published: (2024)
Can We Evaluate Domain Adaptation Models Without Target-Domain Labels?
by: Yang, Jianfei, et al.
Published: (2023)
by: Yang, Jianfei, et al.
Published: (2023)
ViQAgent: Zero-Shot Video Question Answering via Agent with Open-Vocabulary Grounding Validation
by: Montes, Tony, et al.
Published: (2025)
by: Montes, Tony, et al.
Published: (2025)
Zero-Shot Dual-Path Integration Framework for Open-Vocabulary 3D Instance Segmentation
by: Ton, Tri, et al.
Published: (2024)
by: Ton, Tri, et al.
Published: (2024)
Test-Time Optimization for Domain Adaptive Open Vocabulary Segmentation
by: De Silva, Ulindu, et al.
Published: (2025)
by: De Silva, Ulindu, et al.
Published: (2025)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
by: Tai, Hanchen, et al.
Published: (2024)
by: Tai, Hanchen, et al.
Published: (2024)
Video-GroundingDINO: Towards Open-Vocabulary Spatio-Temporal Video Grounding
by: Wasim, Syed Talal, et al.
Published: (2023)
by: Wasim, Syed Talal, et al.
Published: (2023)
Diffusion Model is a Good Pose Estimator from 3D RF-Vision
by: Fan, Junqiao, et al.
Published: (2024)
by: Fan, Junqiao, et al.
Published: (2024)
RADSeg: Unleashing Parameter and Compute Efficient Zero-Shot Open-Vocabulary Segmentation Using Agglomerative Models
by: Alama, Omar, et al.
Published: (2025)
by: Alama, Omar, et al.
Published: (2025)
Structure-aware Prompt Adaptation from Seen to Unseen for Open-Vocabulary Compositional Zero-Shot Learning
by: Duan, Yihang, et al.
Published: (2026)
by: Duan, Yihang, et al.
Published: (2026)
Boundary-Aware Test-Time Adaptation for Zero-Shot Medical Image Segmentation
by: Xu, Chenlin, et al.
Published: (2025)
by: Xu, Chenlin, et al.
Published: (2025)
SegTTA: Training-Free Test-Time Augmentation for Zero-Shot Medical Imaging Segmentation
by: Yao, Yihong, et al.
Published: (2026)
by: Yao, Yihong, et al.
Published: (2026)
Test-Time Degradation Adaptation for Open-Set Image Restoration
by: Gou, Yuanbiao, et al.
Published: (2023)
by: Gou, Yuanbiao, et al.
Published: (2023)
Reconstructing In-the-Wild Open-Vocabulary Human-Object Interactions
by: Wen, Boran, et al.
Published: (2025)
by: Wen, Boran, et al.
Published: (2025)
Diffusion Model is Secretly a Training-free Open Vocabulary Semantic Segmenter
by: Wang, Jinglong, et al.
Published: (2023)
by: Wang, Jinglong, et al.
Published: (2023)
Test-Time Zero-Shot Temporal Action Localization
by: Liberatori, Benedetta, et al.
Published: (2024)
by: Liberatori, Benedetta, et al.
Published: (2024)
ZS-VCOS: Zero-Shot Video Camouflaged Object Segmentation By Optical Flow and Open Vocabulary Object Detection
by: Guo, Wenqi, et al.
Published: (2025)
by: Guo, Wenqi, et al.
Published: (2025)
OpenHuman4D: Open-Vocabulary 4D Human Parsing
by: Suzuki, Keito, et al.
Published: (2025)
by: Suzuki, Keito, et al.
Published: (2025)
Direct Segmentation without Logits Optimization for Training-Free Open-Vocabulary Semantic Segmentation
by: Li, Jiahao, et al.
Published: (2026)
by: Li, Jiahao, et al.
Published: (2026)
Similar Items
-
AdaPose: Towards Cross-Site Device-Free Human Pose Estimation with Commodity WiFi
by: Zhou, Yunjiao, et al.
Published: (2023) -
T3DNet: Compressing Point Cloud Models for Lightweight 3D Recognition
by: Yang, Zhiyuan, et al.
Published: (2024) -
MaskFi: Unsupervised Learning of WiFi and Vision Representations for Multimodal Human Activity Recognition
by: Yang, Jianfei, et al.
Published: (2024) -
FACTOR: Counterfactual Training-Free Test-Time Adaptation for Open-Vocabulary Object Detection
by: Zhao, Kaixiang, et al.
Published: (2026) -
SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
by: Li, Rong, et al.
Published: (2024)