TSTMotion: Training-free Scene-aware Text-to-motion Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guo, Ziyan, Qu, Haoxuan, Rahmani, Hossein, Soh, Dewen, Hu, Ping, Ke, Qiuhong, Liu, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GPT-Connect: Interaction between Text-Driven Human Motion Generator and 3D Scenes in a Training-free Manner
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024)
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024)
LongDiff: Training-Free Long Video Generation in One Go
von: Li, Zhuoling, et al.
Veröffentlicht: (2025)
von: Li, Zhuoling, et al.
Veröffentlicht: (2025)
Translating Signals to Languages for sEMG-Based Activity Recognition
von: Wang, Ming, et al.
Veröffentlicht: (2026)
von: Wang, Ming, et al.
Veröffentlicht: (2026)
DisC-GS: Discontinuity-aware Gaussian Splatting
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024)
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024)
An Image-like Diffusion Method for Human-Object Interaction Detection
von: Hui, Xiaofei, et al.
Veröffentlicht: (2025)
von: Hui, Xiaofei, et al.
Veröffentlicht: (2025)
Answering from Sure to Uncertain: Uncertainty-Aware Curriculum Learning for Video Question Answering
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
Unified Prompt Attack Against Text-to-Image Generation Models
von: Peng, Duo, et al.
Veröffentlicht: (2025)
von: Peng, Duo, et al.
Veröffentlicht: (2025)
Boosting Skeleton-based Zero-Shot Action Recognition with Training-Free Test-Time Adaptation
von: Zhu, Jingmin, et al.
Veröffentlicht: (2025)
von: Zhu, Jingmin, et al.
Veröffentlicht: (2025)
Recent Advances of Continual Learning in Computer Vision: An Overview
von: Qu, Haoxuan, et al.
Veröffentlicht: (2021)
von: Qu, Haoxuan, et al.
Veröffentlicht: (2021)
GaussianBlock: Building Part-Aware Compositional and Editable 3D Scene by Primitives and Gaussians
von: Jiang, Shuyi, et al.
Veröffentlicht: (2024)
von: Jiang, Shuyi, et al.
Veröffentlicht: (2024)
Learning to Generate Cross-Task Unexploitable Examples
von: Qu, Haoxuan, et al.
Veröffentlicht: (2025)
von: Qu, Haoxuan, et al.
Veröffentlicht: (2025)
When Visual Privacy Protection Meets Multimodal Large Language Models
von: Hui, Xiaofei, et al.
Veröffentlicht: (2026)
von: Hui, Xiaofei, et al.
Veröffentlicht: (2026)
A Mixed-Primitive-based Gaussian Splatting Method for Surface Reconstruction
von: Qu, Haoxuan, et al.
Veröffentlicht: (2025)
von: Qu, Haoxuan, et al.
Veröffentlicht: (2025)
TSkel-Mamba: Temporal Dynamic Modeling via State Space Model for Human Skeleton-based Action Recognition
von: Liu, Yanan, et al.
Veröffentlicht: (2025)
von: Liu, Yanan, et al.
Veröffentlicht: (2025)
Sports-QA: A Large-Scale Video Question Answering Benchmark for Complex and Professional Sports
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
UPAM: Unified Prompt Attack in Text-to-Image Generation Models Against Both Textual Filters and Visual Checkers
von: Peng, Duo, et al.
Veröffentlicht: (2024)
von: Peng, Duo, et al.
Veröffentlicht: (2024)
ToolFG: Towards Well-Grounded Fine-Grained Image Classification
von: Xue, Yu, et al.
Veröffentlicht: (2026)
von: Xue, Yu, et al.
Veröffentlicht: (2026)
Leveraging Text Localization for Scene Text Removal via Text-aware Masked Image Modeling
von: Wang, Zixiao, et al.
Veröffentlicht: (2024)
von: Wang, Zixiao, et al.
Veröffentlicht: (2024)
MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm
von: Guo, Ziyan, et al.
Veröffentlicht: (2025)
von: Guo, Ziyan, et al.
Veröffentlicht: (2025)
MicroscopyMatching: Towards a Ready-to-use Framework for Microscopy Image Analysis in Diverse Conditions
von: Hui, Xiaofei, et al.
Veröffentlicht: (2026)
von: Hui, Xiaofei, et al.
Veröffentlicht: (2026)
AI-Generated Content (AIGC) for Various Data Modalities: A Survey
von: Foo, Lin Geng, et al.
Veröffentlicht: (2023)
von: Foo, Lin Geng, et al.
Veröffentlicht: (2023)
Diff-Tracker: Text-to-Image Diffusion Models are Unsupervised Trackers
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2024)
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2024)
DynaPURLS: Dynamic Refinement of Part-Aware Representations for Skeleton-Based Zero-Shot Action Recognition
von: Zhu, Jingmin, et al.
Veröffentlicht: (2025)
von: Zhu, Jingmin, et al.
Veröffentlicht: (2025)
Text2Traffic: A Text-to-Image Generation and Editing Method for Traffic Scenes
von: Lv, Feng, et al.
Veröffentlicht: (2025)
von: Lv, Feng, et al.
Veröffentlicht: (2025)
SceneTracker: Long-term Scene Flow Estimation Network
von: Wang, Bo, et al.
Veröffentlicht: (2024)
von: Wang, Bo, et al.
Veröffentlicht: (2024)
Revisiting Tampered Scene Text Detection in the Era of Generative AI
von: Qu, Chenfan, et al.
Veröffentlicht: (2024)
von: Qu, Chenfan, et al.
Veröffentlicht: (2024)
Part-aware Unified Representation of Language and Skeleton for Zero-shot Action Recognition
von: Zhu, Anqi, et al.
Veröffentlicht: (2024)
von: Zhu, Anqi, et al.
Veröffentlicht: (2024)
Enhancing Human-Centered Dynamic Scene Understanding via Multiple LLMs Collaborated Reasoning
von: Zhang, Hang, et al.
Veröffentlicht: (2024)
von: Zhang, Hang, et al.
Veröffentlicht: (2024)
Future Forcing: Future-aware Training-free KV Cache Policy for Autoregressive Video Generation
von: Luo, Jiayi, et al.
Veröffentlicht: (2026)
von: Luo, Jiayi, et al.
Veröffentlicht: (2026)
LoVoRA: Text-guided and Mask-free Video Object Removal and Addition with Learnable Object-aware Localization
von: Xiao, Zhihan, et al.
Veröffentlicht: (2025)
von: Xiao, Zhihan, et al.
Veröffentlicht: (2025)
CMMLoc: Advancing Text-to-PointCloud Localization with Cauchy-Mixture-Model Based Framework
von: Xu, Yanlong, et al.
Veröffentlicht: (2025)
von: Xu, Yanlong, et al.
Veröffentlicht: (2025)
LLMs are Good Action Recognizers
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024)
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024)
Omni2Sound: Towards Unified Video-Text-to-Audio Generation
von: Dai, Yusheng, et al.
Veröffentlicht: (2026)
von: Dai, Yusheng, et al.
Veröffentlicht: (2026)
Off-the-shelf ChatGPT is a Good Few-shot Human Motion Predictor
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024)
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024)
TextMastero: Mastering High-Quality Scene Text Editing in Diverse Languages and Styles
von: Wang, Tong, et al.
Veröffentlicht: (2024)
von: Wang, Tong, et al.
Veröffentlicht: (2024)
Training-free Composite Scene Generation for Layout-to-Image Synthesis
von: Liu, Jiaqi, et al.
Veröffentlicht: (2024)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2024)
GlyphMastero: A Glyph Encoder for High-Fidelity Scene Text Editing
von: Wang, Tong, et al.
Veröffentlicht: (2025)
von: Wang, Tong, et al.
Veröffentlicht: (2025)
Action Detection via an Image Diffusion Process
von: Foo, Lin Geng, et al.
Veröffentlicht: (2024)
von: Foo, Lin Geng, et al.
Veröffentlicht: (2024)
TATTOO: Training-free AesTheTic-aware Outfit recOmmendation
von: Wu, Yuntian, et al.
Veröffentlicht: (2025)
von: Wu, Yuntian, et al.
Veröffentlicht: (2025)
6D-Diff: A Keypoint Diffusion Framework for 6D Object Pose Estimation
von: Xu, Li, et al.
Veröffentlicht: (2023)
von: Xu, Li, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
GPT-Connect: Interaction between Text-Driven Human Motion Generator and 3D Scenes in a Training-free Manner
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024) -
LongDiff: Training-Free Long Video Generation in One Go
von: Li, Zhuoling, et al.
Veröffentlicht: (2025) -
Translating Signals to Languages for sEMG-Based Activity Recognition
von: Wang, Ming, et al.
Veröffentlicht: (2026) -
DisC-GS: Discontinuity-aware Gaussian Splatting
von: Qu, Haoxuan, et al.
Veröffentlicht: (2024) -
An Image-like Diffusion Method for Human-Object Interaction Detection
von: Hui, Xiaofei, et al.
Veröffentlicht: (2025)