Generation of Complex 3D Human Motion by Temporal and Spatial Composition of Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Mandelli, Lorenzo, Berretti, Stefano |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
High-Entropy Tokens as Multimodal Failure Points in Vision-Language Models
by: He, Mengqi, et al.
Published: (2025)
by: He, Mengqi, et al.
Published: (2025)
Measuring Diversity in Co-creative Image Generation
by: Ibarrola, Francisco, et al.
Published: (2024)
by: Ibarrola, Francisco, et al.
Published: (2024)
Sparse vs Contiguous Adversarial Pixel Perturbations in Multimodal Models: An Empirical Analysis
by: Botocan, Cristian-Alexandru, et al.
Published: (2024)
by: Botocan, Cristian-Alexandru, et al.
Published: (2024)
A Review of Pseudo-Labeling for Computer Vision
by: Kage, Patrick, et al.
Published: (2024)
by: Kage, Patrick, et al.
Published: (2024)
Efficient Diffusion Training through Parallelization with Truncated Karhunen-Loève Expansion
by: Ren, Yumeng, et al.
Published: (2025)
by: Ren, Yumeng, et al.
Published: (2025)
Visual Language Models show widespread visual deficits on neuropsychological tests
by: Tangtartharakul, Gene, et al.
Published: (2025)
by: Tangtartharakul, Gene, et al.
Published: (2025)
Towards Infusing Auxiliary Knowledge for Distracted Driver Detection
by: Balappanawar, Ishwar B, et al.
Published: (2024)
by: Balappanawar, Ishwar B, et al.
Published: (2024)
Convolutional Neural Networks Can (Meta-)Learn the Same-Different Relation
by: Gupta, Max, et al.
Published: (2025)
by: Gupta, Max, et al.
Published: (2025)
Context-Aware Full Body Anonymization using Text-to-Image Diffusion Models
by: Zwick, Pascal, et al.
Published: (2024)
by: Zwick, Pascal, et al.
Published: (2024)
FUTURE-AI: International consensus guideline for trustworthy and deployable artificial intelligence in healthcare
by: Lekadir, Karim, et al.
Published: (2023)
by: Lekadir, Karim, et al.
Published: (2023)
MAR-MAER: Metric-Aware and Ambiguity-Adaptive Autoregressive Image Generation
by: Dong, Kai, et al.
Published: (2026)
by: Dong, Kai, et al.
Published: (2026)
Beyond Specialization: Assessing the Capabilities of MLLMs in Age and Gender Estimation
by: Kuprashevich, Maksim, et al.
Published: (2024)
by: Kuprashevich, Maksim, et al.
Published: (2024)
EncQA: Benchmarking Vision-Language Models on Visual Encodings for Charts
by: Mukherjee, Kushin, et al.
Published: (2025)
by: Mukherjee, Kushin, et al.
Published: (2025)
Domain Generalized Stereo Matching with Uncertainty-guided Data Augmentation
by: Du, Shuangli, et al.
Published: (2025)
by: Du, Shuangli, et al.
Published: (2025)
Fine-Grained Open-Vocabulary Object Detection with Fined-Grained Prompts: Task, Dataset and Benchmark
by: Liu, Ying, et al.
Published: (2025)
by: Liu, Ying, et al.
Published: (2025)
Bridging SFT and DPO for Diffusion Model Alignment with Self-Sampling Preference Optimization
by: Zhang, Daoan, et al.
Published: (2024)
by: Zhang, Daoan, et al.
Published: (2024)
Vision Transformer-based Model for Severity Quantification of Lung Pneumonia Using Chest X-ray Images
by: Slika, Bouthaina, et al.
Published: (2023)
by: Slika, Bouthaina, et al.
Published: (2023)
Pose Matters: Evaluating Vision Transformers and CNNs for Human Action Recognition on Small COCO Subsets
by: Tang, MingZe, et al.
Published: (2025)
by: Tang, MingZe, et al.
Published: (2025)
Statistical Analysis of the Impact of Quaternion Components in Convolutional Neural Networks
by: Altamirano-Gómez, Gerardo, et al.
Published: (2024)
by: Altamirano-Gómez, Gerardo, et al.
Published: (2024)
Massively Multi-Person 3D Human Motion Forecasting with Scene Context
by: Mueller, Felix B, et al.
Published: (2024)
by: Mueller, Felix B, et al.
Published: (2024)
StippleDiffusion: Capacity-Constrained Stippling using Controlled Diffusion
by: Gilad, Ofir, et al.
Published: (2026)
by: Gilad, Ofir, et al.
Published: (2026)
Generating Image Adversarial Examples by Embedding Digital Watermarks
by: Xiang, Yuexin, et al.
Published: (2020)
by: Xiang, Yuexin, et al.
Published: (2020)
CerberusDet: Unified Multi-Dataset Object Detection
by: Tolstykh, Irina, et al.
Published: (2024)
by: Tolstykh, Irina, et al.
Published: (2024)
Object-centric proto-symbolic behavioural reasoning from pixels
by: van Bergen, Ruben, et al.
Published: (2024)
by: van Bergen, Ruben, et al.
Published: (2024)
Generative inpainting of incomplete Euclidean distance matrices of trajectories generated by a fractional Brownian motion
by: Lobashev, Alexander, et al.
Published: (2024)
by: Lobashev, Alexander, et al.
Published: (2024)
Efficient Estimation of Regularized Tyler's M-Estimator Using Approximate LOOCV
by: Abou-Moustafa, Karim
Published: (2025)
by: Abou-Moustafa, Karim
Published: (2025)
Visual Categorization Across Minds and Models: Cognitive Analysis of Human Labeling and Neuro-Symbolic Integration
by: Kabgere, Chethana Prasad
Published: (2025)
by: Kabgere, Chethana Prasad
Published: (2025)
Measuring proximity to standard planes during fetal brain ultrasound scanning
by: Di Vece, Chiara, et al.
Published: (2024)
by: Di Vece, Chiara, et al.
Published: (2024)
Deep EM with Hierarchical Latent Label Modelling for Multi-Site Prostate Lesion Segmentation
by: Yan, Wen, et al.
Published: (2026)
by: Yan, Wen, et al.
Published: (2026)
Sample as You Infer: Predictive Coding With Langevin Dynamics
by: Zahid, Umais, et al.
Published: (2023)
by: Zahid, Umais, et al.
Published: (2023)
Mechanistically Interpretable Neural Encoding Reveals Fine-Grained Functional Selectivity in Human Visual Cortex
by: Grosbard, Idan Daniel, et al.
Published: (2026)
by: Grosbard, Idan Daniel, et al.
Published: (2026)
DIsoN: Decentralized Isolation Networks for Out-of-Distribution Detection in Medical Imaging
by: Wagner, Felix, et al.
Published: (2025)
by: Wagner, Felix, et al.
Published: (2025)
Quaternion Convolutional Neural Networks: Current Advances and Future Directions
by: Altamirano-Gomez, Gerardo, et al.
Published: (2023)
by: Altamirano-Gomez, Gerardo, et al.
Published: (2023)
Efficient Neural Network Encoding for 3D Color Lookup Tables
by: Zehtab, Vahid, et al.
Published: (2024)
by: Zehtab, Vahid, et al.
Published: (2024)
Synthetic Photography Detection: A Visual Guidance for Identifying Synthetic Images Created by AI
by: Mathys, Melanie, et al.
Published: (2024)
by: Mathys, Melanie, et al.
Published: (2024)
Synthetic Image Generation in Cyber Influence Operations: An Emergent Threat?
by: Mathys, Melanie, et al.
Published: (2024)
by: Mathys, Melanie, et al.
Published: (2024)
Exploring Transfer Learning for Deep Learning Polyp Detection in Colonoscopy Images Using YOLOv8
by: Vazquez, Fabian, et al.
Published: (2025)
by: Vazquez, Fabian, et al.
Published: (2025)
CoMA: Complementary Masking and Hierarchical Dynamic Multi-Window Self-Attention in a Unified Pre-training Framework
by: Li, Jiaxuan, et al.
Published: (2025)
by: Li, Jiaxuan, et al.
Published: (2025)
SeNeDiF-OOD: Semantic Nested Dichotomy Fusion for Out-of-Distribution Detection Methodology in Open-World Classification. A Case Study on Monument Style Classification
by: Antequera-Sánchez, Ignacio, et al.
Published: (2026)
by: Antequera-Sánchez, Ignacio, et al.
Published: (2026)
Skeleton-based sign language recognition using a dual-stream spatio-temporal dynamic graph convolutional network
by: Liu, Liangjin, et al.
Published: (2025)
by: Liu, Liangjin, et al.
Published: (2025)
Similar Items
-
High-Entropy Tokens as Multimodal Failure Points in Vision-Language Models
by: He, Mengqi, et al.
Published: (2025) -
Measuring Diversity in Co-creative Image Generation
by: Ibarrola, Francisco, et al.
Published: (2024) -
Sparse vs Contiguous Adversarial Pixel Perturbations in Multimodal Models: An Empirical Analysis
by: Botocan, Cristian-Alexandru, et al.
Published: (2024) -
A Review of Pseudo-Labeling for Computer Vision
by: Kage, Patrick, et al.
Published: (2024) -
Efficient Diffusion Training through Parallelization with Truncated Karhunen-Loève Expansion
by: Ren, Yumeng, et al.
Published: (2025)