Salvato in:
| Autori principali: | Milacski, Zoltán Á., Niinuma, Koichiro, Kawamura, Ryosuke, de la Torre, Fernando, Jeni, László A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2405.18438 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CoGS: Controllable Gaussian Splatting
di: Yu, Heng, et al.
Pubblicazione: (2023)
di: Yu, Heng, et al.
Pubblicazione: (2023)
Gaussian Splatting Lucas-Kanade
di: Xie, Liuyue, et al.
Pubblicazione: (2024)
di: Xie, Liuyue, et al.
Pubblicazione: (2024)
Don't Look Twice: Faster Video Transformers with Run-Length Tokenization
di: Choudhury, Rohan, et al.
Pubblicazione: (2024)
di: Choudhury, Rohan, et al.
Pubblicazione: (2024)
SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion
di: Cho, Jungbin, et al.
Pubblicazione: (2025)
di: Cho, Jungbin, et al.
Pubblicazione: (2025)
Occlusion Sensitivity Analysis with Augmentation Subspace Perturbation in Deep Feature Space
di: Valois, Pedro, et al.
Pubblicazione: (2023)
di: Valois, Pedro, et al.
Pubblicazione: (2023)
OnlineHMR: Video-based Online World-Grounded Human Mesh Recovery
di: Zhao, Yiwen, et al.
Pubblicazione: (2026)
di: Zhao, Yiwen, et al.
Pubblicazione: (2026)
DiSRT-In-Bed: Diffusion-Based Sim-to-Real Transfer Framework for In-Bed Human Mesh Recovery
di: Gao, Jing, et al.
Pubblicazione: (2025)
di: Gao, Jing, et al.
Pubblicazione: (2025)
Zero-Shot Open-Vocabulary Human Motion Grounding with Test-Time Training
di: Zhou, Yunjiao, et al.
Pubblicazione: (2025)
di: Zhou, Yunjiao, et al.
Pubblicazione: (2025)
Scalable Dynamic Origin-Destination Demand Estimation Enhanced by High-Resolution Satellite Imagery Data
di: Liu, Jiachao, et al.
Pubblicazione: (2025)
di: Liu, Jiachao, et al.
Pubblicazione: (2025)
Self-Prompting Diffusion Transformer for Open-Vocabulary Scene Text Editing via In-Context Learning
di: Li, Hongxi, et al.
Pubblicazione: (2026)
di: Li, Hongxi, et al.
Pubblicazione: (2026)
Through the Curved Cover: Synthesizing Cover Aberrated Scenes with Refractive Field
di: Xie, Liuyue, et al.
Pubblicazione: (2024)
di: Xie, Liuyue, et al.
Pubblicazione: (2024)
AlignDiff: Learning Physically-Grounded Camera Alignment via Diffusion
di: Xie, Liuyue, et al.
Pubblicazione: (2025)
di: Xie, Liuyue, et al.
Pubblicazione: (2025)
GHOST: Ground-projected Hypotheses from Observed Structure-from-Motion Trajectories
di: Frelek, Tomasz, et al.
Pubblicazione: (2026)
di: Frelek, Tomasz, et al.
Pubblicazione: (2026)
3D-LFM: Lifting Foundation Model
di: Dabhi, Mosam, et al.
Pubblicazione: (2023)
di: Dabhi, Mosam, et al.
Pubblicazione: (2023)
Open-Vocabulary Functional 3D Human-Scene Interaction Generation
di: Liu, Jie, et al.
Pubblicazione: (2026)
di: Liu, Jie, et al.
Pubblicazione: (2026)
Generating Human Interaction Motions in Scenes with Text Control
di: Yi, Hongwei, et al.
Pubblicazione: (2024)
di: Yi, Hongwei, et al.
Pubblicazione: (2024)
4Real: Towards Photorealistic 4D Scene Generation via Video Diffusion Models
di: Yu, Heng, et al.
Pubblicazione: (2024)
di: Yu, Heng, et al.
Pubblicazione: (2024)
Open-Vocabulary Indoor Object Grounding with 3D Hierarchical Scene Graph
di: Linok, Sergey, et al.
Pubblicazione: (2025)
di: Linok, Sergey, et al.
Pubblicazione: (2025)
Open-Vocabulary Domain Generalization in Urban-Scene Segmentation
di: Zhao, Dong, et al.
Pubblicazione: (2026)
di: Zhao, Dong, et al.
Pubblicazione: (2026)
Multi-Camera Self-Calibration in Sports Motion Capture: Leveraging Human and Stick Poses
di: Yang, Fan, et al.
Pubblicazione: (2026)
di: Yang, Fan, et al.
Pubblicazione: (2026)
SCORE: Scene Context Matters in Open-Vocabulary Remote Sensing Instance Segmentation
di: Huang, Shiqi, et al.
Pubblicazione: (2025)
di: Huang, Shiqi, et al.
Pubblicazione: (2025)
Generating Human Motion in 3D Scenes from Text Descriptions
di: Cen, Zhi, et al.
Pubblicazione: (2024)
di: Cen, Zhi, et al.
Pubblicazione: (2024)
Classifying the Unknown: In-Context Learning for Open-Vocabulary Text and Symbol Recognition
di: Simon, Tom, et al.
Pubblicazione: (2025)
di: Simon, Tom, et al.
Pubblicazione: (2025)
Unified Spherical Frontend: Learning Rotation-Equivariant Representations of Spherical Images from Any Camera
di: Yu, Mukai, et al.
Pubblicazione: (2025)
di: Yu, Mukai, et al.
Pubblicazione: (2025)
CAGS: Open-Vocabulary 3D Scene Understanding with Context-Aware Gaussian Splatting
di: Sun, Wei, et al.
Pubblicazione: (2025)
di: Sun, Wei, et al.
Pubblicazione: (2025)
Context-based Motion Retrieval using Open Vocabulary Methods for Autonomous Driving
di: Englmeier, Stefan, et al.
Pubblicazione: (2025)
di: Englmeier, Stefan, et al.
Pubblicazione: (2025)
Open-Vocabulary Scene Text Recognition via Pseudo-Image Labeling and Margin Loss
di: Ren, Xuhua, et al.
Pubblicazione: (2024)
di: Ren, Xuhua, et al.
Pubblicazione: (2024)
Beyond Bare Queries: Open-Vocabulary Object Grounding with 3D Scene Graph
di: Linok, Sergey, et al.
Pubblicazione: (2024)
di: Linok, Sergey, et al.
Pubblicazione: (2024)
Open Vocabulary Semantic Scene Sketch Understanding
di: Bourouis, Ahmed, et al.
Pubblicazione: (2023)
di: Bourouis, Ahmed, et al.
Pubblicazione: (2023)
Contrastive Prompts Improve Disentanglement in Text-to-Image Diffusion Models
di: Wu, Chen, et al.
Pubblicazione: (2024)
di: Wu, Chen, et al.
Pubblicazione: (2024)
Text-controlled Motion Mamba: Text-Instructed Temporal Grounding of Human Motion
di: Wang, Xinghan, et al.
Pubblicazione: (2024)
di: Wang, Xinghan, et al.
Pubblicazione: (2024)
OpenScan: A Benchmark for Generalized Open-Vocabulary 3D Scene Understanding
di: Zhao, Youjun, et al.
Pubblicazione: (2024)
di: Zhao, Youjun, et al.
Pubblicazione: (2024)
SceneAssistant: A Visual Feedback Agent for Open-Vocabulary 3D Scene Generation
di: Luo, Jun, et al.
Pubblicazione: (2026)
di: Luo, Jun, et al.
Pubblicazione: (2026)
ROVI: A VLM-LLM Re-Captioned Dataset for Open-Vocabulary Instance-Grounded Text-to-Image Generation
di: Peng, Cihang, et al.
Pubblicazione: (2025)
di: Peng, Cihang, et al.
Pubblicazione: (2025)
GHOST: Gaussian Hypothesis Open-Set Technique
di: Rabinowitz, Ryan, et al.
Pubblicazione: (2025)
di: Rabinowitz, Ryan, et al.
Pubblicazione: (2025)
Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes
di: Zhou, Changqing, et al.
Pubblicazione: (2026)
di: Zhou, Changqing, et al.
Pubblicazione: (2026)
Interaction-Centric Knowledge Infusion and Transfer for Open-Vocabulary Scene Graph Generation
di: Li, Lin, et al.
Pubblicazione: (2025)
di: Li, Lin, et al.
Pubblicazione: (2025)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
di: Tai, Hanchen, et al.
Pubblicazione: (2024)
di: Tai, Hanchen, et al.
Pubblicazione: (2024)
Affogato: Learning Open-Vocabulary Affordance Grounding with Automated Data Generation at Scale
di: Lee, Junha, et al.
Pubblicazione: (2025)
di: Lee, Junha, et al.
Pubblicazione: (2025)
Textual Decomposition Then Sub-motion-space Scattering for Open-Vocabulary Motion Generation
di: Fan, Ke, et al.
Pubblicazione: (2024)
di: Fan, Ke, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CoGS: Controllable Gaussian Splatting
di: Yu, Heng, et al.
Pubblicazione: (2023) -
Gaussian Splatting Lucas-Kanade
di: Xie, Liuyue, et al.
Pubblicazione: (2024) -
Don't Look Twice: Faster Video Transformers with Run-Length Tokenization
di: Choudhury, Rohan, et al.
Pubblicazione: (2024) -
SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion
di: Cho, Jungbin, et al.
Pubblicazione: (2025) -
Occlusion Sensitivity Analysis with Augmentation Subspace Perturbation in Deep Feature Space
di: Valois, Pedro, et al.
Pubblicazione: (2023)