Salvato in:
| Autori principali: | Wen, Boran, Huang, Dingbang, Zhang, Zichen, Zhou, Jiahong, Deng, Jianbin, Gong, Jingyu, Chen, Yulong, Ma, Lizhuang, Li, Yong-Lu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2503.15898 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Interacted Object Grounding in Spatio-Temporal Human-Object Interactions
di: Liu, Xiaoyang, et al.
Pubblicazione: (2024)
di: Liu, Xiaoyang, et al.
Pubblicazione: (2024)
Efficient and Scalable Monocular Human-Object Interaction Motion Reconstruction
di: Wen, Boran, et al.
Pubblicazione: (2025)
di: Wen, Boran, et al.
Pubblicazione: (2025)
Textual Decomposition Then Sub-motion-space Scattering for Open-Vocabulary Motion Generation
di: Fan, Ke, et al.
Pubblicazione: (2024)
di: Fan, Ke, et al.
Pubblicazione: (2024)
WildOS: Open-Vocabulary Object Search in the Wild
di: Shah, Hardik, et al.
Pubblicazione: (2026)
di: Shah, Hardik, et al.
Pubblicazione: (2026)
UniForward: Unified 3D Scene and Semantic Field Reconstruction via Feed-Forward Gaussian Splatting from Only Sparse-View Images
di: Tian, Qijian, et al.
Pubblicazione: (2025)
di: Tian, Qijian, et al.
Pubblicazione: (2025)
Streamlined Open-Vocabulary Human-Object Interaction Detection
di: Sun, Chang, et al.
Pubblicazione: (2026)
di: Sun, Chang, et al.
Pubblicazione: (2026)
OpenObj: Open-Vocabulary Object-Level Neural Radiance Fields with Fine-Grained Understanding
di: Deng, Yinan, et al.
Pubblicazione: (2024)
di: Deng, Yinan, et al.
Pubblicazione: (2024)
S2GS: Streaming Semantic Gaussian Splatting for Online Scene Understanding and Reconstruction
di: Zhang, Renhe, et al.
Pubblicazione: (2026)
di: Zhang, Renhe, et al.
Pubblicazione: (2026)
Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection
di: Hu, Yupeng, et al.
Pubblicazione: (2025)
di: Hu, Yupeng, et al.
Pubblicazione: (2025)
Monocular Human-Object Reconstruction in the Wild
di: Huo, Chaofan, et al.
Pubblicazione: (2024)
di: Huo, Chaofan, et al.
Pubblicazione: (2024)
CAGS: Open-Vocabulary 3D Scene Understanding with Context-Aware Gaussian Splatting
di: Sun, Wei, et al.
Pubblicazione: (2025)
di: Sun, Wei, et al.
Pubblicazione: (2025)
Emphasizing Semantic Consistency of Salient Posture for Speech-Driven Gesture Generation
di: Liu, Fengqi, et al.
Pubblicazione: (2024)
di: Liu, Fengqi, et al.
Pubblicazione: (2024)
Open-Vocabulary Camouflaged Object Segmentation
di: Pang, Youwei, et al.
Pubblicazione: (2023)
di: Pang, Youwei, et al.
Pubblicazione: (2023)
AnyHome: Open-Vocabulary Generation of Structured and Textured 3D Homes
di: Fu, Rao, et al.
Pubblicazione: (2023)
di: Fu, Rao, et al.
Pubblicazione: (2023)
ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection
di: Nguyen, Minh Anh, et al.
Pubblicazione: (2026)
di: Nguyen, Minh Anh, et al.
Pubblicazione: (2026)
GSCompleter: A Distillation-Free Plugin for Metric-Aware 3D Gaussian Splatting Completion in Seconds
di: Gao, Ao, et al.
Pubblicazione: (2026)
di: Gao, Ao, et al.
Pubblicazione: (2026)
LOVON: Legged Open-Vocabulary Object Navigator
di: Peng, Daojie, et al.
Pubblicazione: (2025)
di: Peng, Daojie, et al.
Pubblicazione: (2025)
Open-Vocabulary Object Detection via Language Hierarchy
di: Huang, Jiaxing, et al.
Pubblicazione: (2024)
di: Huang, Jiaxing, et al.
Pubblicazione: (2024)
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
di: Zhang, Yupeng, et al.
Pubblicazione: (2025)
di: Zhang, Yupeng, et al.
Pubblicazione: (2025)
LLMs Meet VLMs: Boost Open Vocabulary Object Detection with Fine-grained Descriptors
di: Jin, Sheng, et al.
Pubblicazione: (2024)
di: Jin, Sheng, et al.
Pubblicazione: (2024)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
di: Liu, Yumeng, et al.
Pubblicazione: (2024)
di: Liu, Yumeng, et al.
Pubblicazione: (2024)
SuperMat: Physically Consistent PBR Material Estimation at Interactive Rates
di: Hong, Yijia, et al.
Pubblicazione: (2024)
di: Hong, Yijia, et al.
Pubblicazione: (2024)
TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation
di: Li, Dingbang, et al.
Pubblicazione: (2024)
di: Li, Dingbang, et al.
Pubblicazione: (2024)
Boosting Open-Vocabulary Object Detection by Handling Background Samples
di: Zeng, Ruizhe, et al.
Pubblicazione: (2024)
di: Zeng, Ruizhe, et al.
Pubblicazione: (2024)
Scaling Open-Vocabulary Object Detection
di: Minderer, Matthias, et al.
Pubblicazione: (2023)
di: Minderer, Matthias, et al.
Pubblicazione: (2023)
DEMOS: Dynamic Environment Motion Synthesis in 3D Scenes via Local Spherical-BEV Perception
di: Gong, Jingyu, et al.
Pubblicazione: (2024)
di: Gong, Jingyu, et al.
Pubblicazione: (2024)
RHOBIN Challenge: Reconstruction of Human Object Interaction
di: Xie, Xianghui, et al.
Pubblicazione: (2024)
di: Xie, Xianghui, et al.
Pubblicazione: (2024)
Open-Vocabulary Camouflaged Object Segmentation with Cascaded Vision Language Models
di: Zhao, Kai, et al.
Pubblicazione: (2025)
di: Zhao, Kai, et al.
Pubblicazione: (2025)
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
di: Zhou, Donghao, et al.
Pubblicazione: (2026)
di: Zhou, Donghao, et al.
Pubblicazione: (2026)
Classifier-Centric Adaptive Framework for Open-Vocabulary Camouflaged Object Segmentation
di: Zhang, Hanyu, et al.
Pubblicazione: (2025)
di: Zhang, Hanyu, et al.
Pubblicazione: (2025)
ObjectFinder: An Open-Vocabulary Assistive System for Interactive Object Search by Blind People
di: Liu, Ruiping, et al.
Pubblicazione: (2024)
di: Liu, Ruiping, et al.
Pubblicazione: (2024)
Taking A Closer Look at Interacting Objects: Interaction-Aware Open Vocabulary Scene Graph Generation
di: Li, Lin, et al.
Pubblicazione: (2025)
di: Li, Lin, et al.
Pubblicazione: (2025)
Self-Calibrated CLIP for Training-Free Open-Vocabulary Segmentation
di: Bai, Sule, et al.
Pubblicazione: (2024)
di: Bai, Sule, et al.
Pubblicazione: (2024)
Unsupervised Open-Vocabulary Object Localization in Videos
di: Fan, Ke, et al.
Pubblicazione: (2023)
di: Fan, Ke, et al.
Pubblicazione: (2023)
Retrieval-Augmented Open-Vocabulary Object Detection
di: Kim, Jooyeon, et al.
Pubblicazione: (2024)
di: Kim, Jooyeon, et al.
Pubblicazione: (2024)
OpenGraph: Open-Vocabulary Hierarchical 3D Graph Representation in Large-Scale Outdoor Environments
di: Deng, Yinan, et al.
Pubblicazione: (2024)
di: Deng, Yinan, et al.
Pubblicazione: (2024)
LED: LLM Enhanced Open-Vocabulary Object Detection without Human Curated Data Generation
di: Zhou, Yang, et al.
Pubblicazione: (2025)
di: Zhou, Yang, et al.
Pubblicazione: (2025)
State and Scene Enhanced Prototypes for Weakly Supervised Open-Vocabulary Object Detection
di: Zhou, Jiaying, et al.
Pubblicazione: (2025)
di: Zhou, Jiaying, et al.
Pubblicazione: (2025)
Locate Anything on Earth: Advancing Open-Vocabulary Object Detection for Remote Sensing Community
di: Pan, Jiancheng, et al.
Pubblicazione: (2024)
di: Pan, Jiancheng, et al.
Pubblicazione: (2024)
Taming SAM3 in the Wild: A Concept Bank for Open-Vocabulary Segmentation
di: Pei, Gensheng, et al.
Pubblicazione: (2026)
di: Pei, Gensheng, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Interacted Object Grounding in Spatio-Temporal Human-Object Interactions
di: Liu, Xiaoyang, et al.
Pubblicazione: (2024) -
Efficient and Scalable Monocular Human-Object Interaction Motion Reconstruction
di: Wen, Boran, et al.
Pubblicazione: (2025) -
Textual Decomposition Then Sub-motion-space Scattering for Open-Vocabulary Motion Generation
di: Fan, Ke, et al.
Pubblicazione: (2024) -
WildOS: Open-Vocabulary Object Search in the Wild
di: Shah, Hardik, et al.
Pubblicazione: (2026) -
UniForward: Unified 3D Scene and Semantic Field Reconstruction via Feed-Forward Gaussian Splatting from Only Sparse-View Images
di: Tian, Qijian, et al.
Pubblicazione: (2025)