BodyShapeGPT: SMPL Body Shape Manipulation with LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Árbol, Baldomero R., Casas, Dan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
What Shape Is Optimal for Masks in Text Removal?
von: Nakada, Hyakka, et al.
Veröffentlicht: (2025)
von: Nakada, Hyakka, et al.
Veröffentlicht: (2025)
Large Body Language Models
von: Punjwani, Saif, et al.
Veröffentlicht: (2024)
von: Punjwani, Saif, et al.
Veröffentlicht: (2024)
Using Motion Cues to Supervise Single-Frame Body Pose and Shape Estimation in Low Data Regimes
von: Davydov, Andrey, et al.
Veröffentlicht: (2024)
von: Davydov, Andrey, et al.
Veröffentlicht: (2024)
DL-EWF: Deep Learning Empowering Women's Fashion with Grounded-Segment-Anything Segmentation for Body Shape Classification
von: Asghari, Fatemeh, et al.
Veröffentlicht: (2024)
von: Asghari, Fatemeh, et al.
Veröffentlicht: (2024)
SimpleGPT: Improving GPT via A Simple Normalization Strategy
von: Chen, Marco, et al.
Veröffentlicht: (2026)
von: Chen, Marco, et al.
Veröffentlicht: (2026)
Visual Whole-Body Control for Legged Loco-Manipulation
von: Liu, Minghuan, et al.
Veröffentlicht: (2024)
von: Liu, Minghuan, et al.
Veröffentlicht: (2024)
VIST-GPT: Ushering in the Era of Visual Storytelling with LLMs?
von: Gado, Mohamed, et al.
Veröffentlicht: (2025)
von: Gado, Mohamed, et al.
Veröffentlicht: (2025)
Examining the Robustness of Homogeneity Bias to Hyperparameter Adjustments in GPT-4
von: Lee, Messi H. J.
Veröffentlicht: (2025)
von: Lee, Messi H. J.
Veröffentlicht: (2025)
HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale
von: Chen, Junying, et al.
Veröffentlicht: (2024)
von: Chen, Junying, et al.
Veröffentlicht: (2024)
Optimizing GPT for Video Understanding: Zero-Shot Performance and Prompt Engineering
von: Beliaev, Mark, et al.
Veröffentlicht: (2025)
von: Beliaev, Mark, et al.
Veröffentlicht: (2025)
ShizhenGPT: Towards Multimodal LLMs for Traditional Chinese Medicine
von: Chen, Junying, et al.
Veröffentlicht: (2025)
von: Chen, Junying, et al.
Veröffentlicht: (2025)
VolumetricSMPL: A Neural Volumetric Body Model for Efficient Interactions, Contacts, and Collisions
von: Mihajlovic, Marko, et al.
Veröffentlicht: (2025)
von: Mihajlovic, Marko, et al.
Veröffentlicht: (2025)
BodyGPS: Anatomical Positioning System
von: Yerebakan, Halid Ziya, et al.
Veröffentlicht: (2025)
von: Yerebakan, Halid Ziya, et al.
Veröffentlicht: (2025)
EMMA: Efficient Visual Alignment in Multi-Modal LLMs
von: Ghazanfari, Sara, et al.
Veröffentlicht: (2024)
von: Ghazanfari, Sara, et al.
Veröffentlicht: (2024)
VLLaVO: Mitigating Visual Gap through LLMs
von: Chen, Shuhao, et al.
Veröffentlicht: (2024)
von: Chen, Shuhao, et al.
Veröffentlicht: (2024)
Lumos : Empowering Multimodal LLMs with Scene Text Recognition
von: Shenoy, Ashish, et al.
Veröffentlicht: (2024)
von: Shenoy, Ashish, et al.
Veröffentlicht: (2024)
ShapeR: Robust Conditional 3D Shape Generation from Casual Captures
von: Siddiqui, Yawar, et al.
Veröffentlicht: (2026)
von: Siddiqui, Yawar, et al.
Veröffentlicht: (2026)
Millimeter-wave Imaging for Anthropometric Body Measurement
von: Senne, Miriam, et al.
Veröffentlicht: (2026)
von: Senne, Miriam, et al.
Veröffentlicht: (2026)
JPEG-LM: LLMs as Image Generators with Canonical Codec Representations
von: Han, Xiaochuang, et al.
Veröffentlicht: (2024)
von: Han, Xiaochuang, et al.
Veröffentlicht: (2024)
Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
SpurLens: Automatic Detection of Spurious Cues in Multimodal LLMs
von: Hosseini, Parsa, et al.
Veröffentlicht: (2025)
von: Hosseini, Parsa, et al.
Veröffentlicht: (2025)
PyTorchGeoNodes: Enabling Differentiable Shape Programs for 3D Shape Reconstruction
von: Stekovic, Sinisa, et al.
Veröffentlicht: (2024)
von: Stekovic, Sinisa, et al.
Veröffentlicht: (2024)
QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs
von: Daxberger, Erik, et al.
Veröffentlicht: (2025)
von: Daxberger, Erik, et al.
Veröffentlicht: (2025)
LLMs as Visual Explainers: Advancing Image Classification with Evolving Visual Descriptions
von: Han, Songhao, et al.
Veröffentlicht: (2023)
von: Han, Songhao, et al.
Veröffentlicht: (2023)
GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
Shallow Signed Distance Functions for Kinematic Collision Bodies
von: Akar, Osman, et al.
Veröffentlicht: (2024)
von: Akar, Osman, et al.
Veröffentlicht: (2024)
Fast Neural Inverse Kinematics on Human Body Motions
von: Tolpin, David, et al.
Veröffentlicht: (2025)
von: Tolpin, David, et al.
Veröffentlicht: (2025)
Significance-Gain Pair Encoding for LLMs: A Statistical Alternative to Frequency-Based Subword Merging
von: Nouri, Azam
Veröffentlicht: (2026)
von: Nouri, Azam
Veröffentlicht: (2026)
Gloss2Text: Sign Language Gloss translation using LLMs and Semantically Aware Label Smoothing
von: Fayyazsanavi, Pooya, et al.
Veröffentlicht: (2024)
von: Fayyazsanavi, Pooya, et al.
Veröffentlicht: (2024)
ElectroVizQA: How well do Multi-modal LLMs perform in Electronics Visual Question Answering?
von: Meshram, Pragati Shuddhodhan, et al.
Veröffentlicht: (2024)
von: Meshram, Pragati Shuddhodhan, et al.
Veröffentlicht: (2024)
MoD-DPO: Towards Mitigating Cross-modal Hallucinations in Omni LLMs using Modality Decoupled Preference Optimization
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2026)
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2026)
Weakly Supervised Segmentation of Vertebral Bodies with Iterative Slice-propagation
von: Peng, Shiqi, et al.
Veröffentlicht: (2024)
von: Peng, Shiqi, et al.
Veröffentlicht: (2024)
AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling
von: Zhan, Jun, et al.
Veröffentlicht: (2024)
von: Zhan, Jun, et al.
Veröffentlicht: (2024)
A Study of Shape Modeling Against Noise
von: Long, Cheng, et al.
Veröffentlicht: (2024)
von: Long, Cheng, et al.
Veröffentlicht: (2024)
Shape Generation via Weight Space Learning
von: Plattner, Maximilian, et al.
Veröffentlicht: (2025)
von: Plattner, Maximilian, et al.
Veröffentlicht: (2025)
S-GRPO: Unified Post-Training for Large Vision-Language Models
von: Yan, Yuming, et al.
Veröffentlicht: (2026)
von: Yan, Yuming, et al.
Veröffentlicht: (2026)
Harnessing GPT-4V(ision) for Insurance: A Preliminary Exploration
von: Lin, Chenwei, et al.
Veröffentlicht: (2024)
von: Lin, Chenwei, et al.
Veröffentlicht: (2024)
TWIST: Teleoperated Whole-Body Imitation System
von: Ze, Yanjie, et al.
Veröffentlicht: (2025)
von: Ze, Yanjie, et al.
Veröffentlicht: (2025)
Total Selfie: Generating Full-Body Selfies
von: Chen, Bowei, et al.
Veröffentlicht: (2023)
von: Chen, Bowei, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
What Shape Is Optimal for Masks in Text Removal?
von: Nakada, Hyakka, et al.
Veröffentlicht: (2025) -
Large Body Language Models
von: Punjwani, Saif, et al.
Veröffentlicht: (2024) -
Using Motion Cues to Supervise Single-Frame Body Pose and Shape Estimation in Low Data Regimes
von: Davydov, Andrey, et al.
Veröffentlicht: (2024) -
DL-EWF: Deep Learning Empowering Women's Fashion with Grounded-Segment-Anything Segmentation for Body Shape Classification
von: Asghari, Fatemeh, et al.
Veröffentlicht: (2024) -
SimpleGPT: Improving GPT via A Simple Normalization Strategy
von: Chen, Marco, et al.
Veröffentlicht: (2026)