Toward Human Understanding with Controllable Synthesis
Fuente:
arXiv
Guardado en:
| Autores principales: | Cuevas-Velasquez, Hanz, Patel, Priyanka, Feng, Haiwen, Black, Michael |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CameraHMR: Aligning People with Perspective
por: Patel, Priyanka, et al.
Publicado: (2024)
por: Patel, Priyanka, et al.
Publicado: (2024)
TokenHMR: Advancing Human Mesh Recovery with a Tokenized Pose Representation
por: Dwivedi, Sai Kumar, et al.
Publicado: (2024)
por: Dwivedi, Sai Kumar, et al.
Publicado: (2024)
SimpleEgo: Predicting Probabilistic Body Pose from Egocentric Cameras
por: Cuevas-Velasquez, Hanz, et al.
Publicado: (2024)
por: Cuevas-Velasquez, Hanz, et al.
Publicado: (2024)
ChatPose: Chatting about 3D Human Pose
por: Feng, Yao, et al.
Publicado: (2023)
por: Feng, Yao, et al.
Publicado: (2023)
InterDyn: Controllable Interactive Dynamics with Video Diffusion Models
por: Akkerman, Rick, et al.
Publicado: (2024)
por: Akkerman, Rick, et al.
Publicado: (2024)
PromptHMR: Promptable Human Mesh Recovery
por: Wang, Yufu, et al.
Publicado: (2025)
por: Wang, Yufu, et al.
Publicado: (2025)
Generative Zoo
por: Niewiadomski, Tomasz, et al.
Publicado: (2024)
por: Niewiadomski, Tomasz, et al.
Publicado: (2024)
BEDLAM2.0: Synthetic Humans and Cameras in Motion
por: Tesch, Joachim, et al.
Publicado: (2025)
por: Tesch, Joachim, et al.
Publicado: (2025)
ETCH: Generalizing Body Fitting to Clothed Humans via Equivariant Tightness
por: Li, Boqian, et al.
Publicado: (2025)
por: Li, Boqian, et al.
Publicado: (2025)
Two Heads are Better than One: Geometric-Latent Attention for Point Cloud Classification and Segmentation
por: Cuevas-Velasquez, Hanz, et al.
Publicado: (2021)
por: Cuevas-Velasquez, Hanz, et al.
Publicado: (2021)
MAMMA: Markerless & Automatic Multi-Person Motion Action Capture
por: Cuevas-Velasquez, Hanz, et al.
Publicado: (2025)
por: Cuevas-Velasquez, Hanz, et al.
Publicado: (2025)
Predicting 4D Hand Trajectory from Monocular Videos
por: Ye, Yufei, et al.
Publicado: (2025)
por: Ye, Yufei, et al.
Publicado: (2025)
Re-Thinking Inverse Graphics With Large Language Models
por: Kulits, Peter, et al.
Publicado: (2024)
por: Kulits, Peter, et al.
Publicado: (2024)
Global Point Cloud Registration Network for Large Transformations
por: Cuevas-Velasquez, Hanz, et al.
Publicado: (2024)
por: Cuevas-Velasquez, Hanz, et al.
Publicado: (2024)
GenLit: Reformulating Single-Image Relighting as Video Generation
por: Bharadwaj, Shrisha, et al.
Publicado: (2024)
por: Bharadwaj, Shrisha, et al.
Publicado: (2024)
Explorative Inbetweening of Time and Space
por: Feng, Haiwen, et al.
Publicado: (2024)
por: Feng, Haiwen, et al.
Publicado: (2024)
From Detection to Anticipation: Online Understanding of Struggles across Various Tasks and Activities
por: Feng, Shijia, et al.
Publicado: (2025)
por: Feng, Shijia, et al.
Publicado: (2025)
E-React: Towards Emotionally Controlled Synthesis of Human Reactions
por: Zhu, Chen, et al.
Publicado: (2025)
por: Zhu, Chen, et al.
Publicado: (2025)
St4RTrack: Simultaneous 4D Reconstruction and Tracking in the World
por: Feng, Haiwen, et al.
Publicado: (2025)
por: Feng, Haiwen, et al.
Publicado: (2025)
ChatHuman: Chatting about 3D Humans with Tools
por: Lin, Jing, et al.
Publicado: (2024)
por: Lin, Jing, et al.
Publicado: (2024)
Generating Human Interaction Motions in Scenes with Text Control
por: Yi, Hongwei, et al.
Publicado: (2024)
por: Yi, Hongwei, et al.
Publicado: (2024)
Half-Physics: Enabling Kinematic 3D Human Model with Physical Interactions
por: Siyao, Li, et al.
Publicado: (2025)
por: Siyao, Li, et al.
Publicado: (2025)
SynthForge: Synthesizing High-Quality Face Dataset with Controllable 3D Generative Models
por: Rawat, Abhay, et al.
Publicado: (2024)
por: Rawat, Abhay, et al.
Publicado: (2024)
Automatic Synthesis of High-Quality Triplet Data for Composed Image Retrieval
por: Li, Haiwen, et al.
Publicado: (2025)
por: Li, Haiwen, et al.
Publicado: (2025)
Towards Automated Initial Probe Placement in Transthoracic Teleultrasound Using Human Mesh and Skeleton Recovery
por: Lee, Yu Chung, et al.
Publicado: (2026)
por: Lee, Yu Chung, et al.
Publicado: (2026)
Controllable Human-Object Interaction Synthesis
por: Li, Jiaman, et al.
Publicado: (2023)
por: Li, Jiaman, et al.
Publicado: (2023)
WHAM: Reconstructing World-grounded Humans with Accurate 3D Motion
por: Shin, Soyong, et al.
Publicado: (2023)
por: Shin, Soyong, et al.
Publicado: (2023)
Unimotion: Unifying 3D Human Motion Synthesis and Understanding
por: Li, Chuqiao, et al.
Publicado: (2024)
por: Li, Chuqiao, et al.
Publicado: (2024)
SINC: Spatial Composition of 3D Human Motions for Simultaneous Action Generation
por: Athanasiou, Nikos, et al.
Publicado: (2023)
por: Athanasiou, Nikos, et al.
Publicado: (2023)
From Skin to Skeleton: Towards Biomechanically Accurate 3D Digital Humans
por: Keller, Marilyn, et al.
Publicado: (2025)
por: Keller, Marilyn, et al.
Publicado: (2025)
CacheFlow: Compressive Streaming Memory for Efficient Long-Form Video Understanding
por: Patel, Shrenik, et al.
Publicado: (2025)
por: Patel, Shrenik, et al.
Publicado: (2025)
VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?
por: Liu, Qing'an, et al.
Publicado: (2026)
por: Liu, Qing'an, et al.
Publicado: (2026)
HIS-GPT: Towards 3D Human-In-Scene Multimodal Understanding
por: Zhao, Jiahe, et al.
Publicado: (2025)
por: Zhao, Jiahe, et al.
Publicado: (2025)
Can Large Language Models Understand Symbolic Graphics Programs?
por: Qiu, Zeju, et al.
Publicado: (2024)
por: Qiu, Zeju, et al.
Publicado: (2024)
Supervising 3D Talking Head Avatars with Analysis-by-Audio-Synthesis
por: Daněček, Radek, et al.
Publicado: (2025)
por: Daněček, Radek, et al.
Publicado: (2025)
Towards Consistent and Controllable Image Synthesis for Face Editing
por: Wei, Mengting, et al.
Publicado: (2025)
por: Wei, Mengting, et al.
Publicado: (2025)
AWOL: Analysis WithOut synthesis using Language
por: Zuffi, Silvia, et al.
Publicado: (2024)
por: Zuffi, Silvia, et al.
Publicado: (2024)
EvoStruggle: A Dataset Capturing the Evolution of Struggle across Activities and Skill Levels
por: Feng, Shijia, et al.
Publicado: (2025)
por: Feng, Shijia, et al.
Publicado: (2025)
Im2Haircut: Single-view Strand-based Hair Reconstruction for Human Avatars
por: Sklyarova, Vanessa, et al.
Publicado: (2025)
por: Sklyarova, Vanessa, et al.
Publicado: (2025)
Moving by Looking: Towards Vision-Driven Avatar Motion Generation
por: Diomataris, Markos, et al.
Publicado: (2025)
por: Diomataris, Markos, et al.
Publicado: (2025)
Ejemplares similares
-
CameraHMR: Aligning People with Perspective
por: Patel, Priyanka, et al.
Publicado: (2024) -
TokenHMR: Advancing Human Mesh Recovery with a Tokenized Pose Representation
por: Dwivedi, Sai Kumar, et al.
Publicado: (2024) -
SimpleEgo: Predicting Probabilistic Body Pose from Egocentric Cameras
por: Cuevas-Velasquez, Hanz, et al.
Publicado: (2024) -
ChatPose: Chatting about 3D Human Pose
por: Feng, Yao, et al.
Publicado: (2023) -
InterDyn: Controllable Interactive Dynamics with Video Diffusion Models
por: Akkerman, Rick, et al.
Publicado: (2024)