Bones Can't Be Triangles: Accurate and Efficient Vertebrae Keypoint Estimation through Collaborative Error Revision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Jinhee, Kim, Taesung, Choo, Jaegul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Morphology-Aware Interactive Keypoint Estimation
von: Kim, Jinhee, et al.
Veröffentlicht: (2022)
von: Kim, Jinhee, et al.
Veröffentlicht: (2022)
EPIC: Effective Prompting for Imbalanced-Class Data Synthesis in Tabular Data Classification via Large Language Models
von: Kim, Jinhee, et al.
Veröffentlicht: (2024)
von: Kim, Jinhee, et al.
Veröffentlicht: (2024)
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
Attend-and-Refine: Interactive keypoint estimation and quantitative cervical vertebrae analysis for bone age assessment
von: Kim, Jinhee, et al.
Veröffentlicht: (2025)
von: Kim, Jinhee, et al.
Veröffentlicht: (2025)
SurFhead: Affine Rig Blending for Geometrically Accurate 2D Gaussian Surfel Head Avatars
von: Lee, Jaeseong, et al.
Veröffentlicht: (2024)
von: Lee, Jaeseong, et al.
Veröffentlicht: (2024)
InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion
von: Jin, Hoiyeong, et al.
Veröffentlicht: (2025)
von: Jin, Hoiyeong, et al.
Veröffentlicht: (2025)
Good Noise Makes Good Edits: A Training-Free Diffusion-Based Video Editing with Image and Text Prompts
von: Choi, Saemee, et al.
Veröffentlicht: (2025)
von: Choi, Saemee, et al.
Veröffentlicht: (2025)
SelfSwapper: Self-Supervised Face Swapping via Shape Agnostic Masked AutoEncoder
von: Lee, Jaeseong, et al.
Veröffentlicht: (2024)
von: Lee, Jaeseong, et al.
Veröffentlicht: (2024)
ConceptScope: Characterizing Dataset Bias via Disentangled Visual Concepts
von: Choi, Jinho, et al.
Veröffentlicht: (2025)
von: Choi, Jinho, et al.
Veröffentlicht: (2025)
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
von: Kim, Donghu, et al.
Veröffentlicht: (2024)
von: Kim, Donghu, et al.
Veröffentlicht: (2024)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
von: Kim, Kangyeol, et al.
Veröffentlicht: (2024)
von: Kim, Kangyeol, et al.
Veröffentlicht: (2024)
Memory-Efficient Fine-Tuning Diffusion Transformers via Dynamic Patch Sampling and Block Skipping
von: Park, Sunghyun, et al.
Veröffentlicht: (2026)
von: Park, Sunghyun, et al.
Veröffentlicht: (2026)
TexAvatars : Hybrid Texel-3D Representations for Stable Rigging of Photorealistic Gaussian Head Avatars
von: Lee, Jaeseong, et al.
Veröffentlicht: (2025)
von: Lee, Jaeseong, et al.
Veröffentlicht: (2025)
SpineCLUE: Automatic Vertebrae Identification Using Contrastive Learning and Uncertainty Estimation
von: Zhang, Sheng, et al.
Veröffentlicht: (2024)
von: Zhang, Sheng, et al.
Veröffentlicht: (2024)
DesignLab: Designing Slides Through Iterative Detection and Correction
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
Why Can't I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition
von: Ahn, Geo, et al.
Veröffentlicht: (2026)
von: Ahn, Geo, et al.
Veröffentlicht: (2026)
Deformable Dynamic Convolution for Accurate yet Efficient Spatio-Temporal Traffic Prediction
von: Jin, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Jin, Hyeonseok, et al.
Veröffentlicht: (2025)
Training Spatial-Frequency Visual Prompts and Probabilistic Clusters for Accurate Black-Box Transfer Learning
von: Cho, Wonwoo, et al.
Veröffentlicht: (2024)
von: Cho, Wonwoo, et al.
Veröffentlicht: (2024)
Time Blindness: Why Video-Language Models Can't See What Humans Can?
von: Upadhyay, Ujjwal, et al.
Veröffentlicht: (2025)
von: Upadhyay, Ujjwal, et al.
Veröffentlicht: (2025)
Fair Generation without Unfair Distortions: Debiasing Text-to-Image Generation with Entanglement-Free Attention
von: Park, Jeonghoon, et al.
Veröffentlicht: (2025)
von: Park, Jeonghoon, et al.
Veröffentlicht: (2025)
Towards Calibrated Robust Fine-Tuning of Vision-Language Models
von: Oh, Changdae, et al.
Veröffentlicht: (2023)
von: Oh, Changdae, et al.
Veröffentlicht: (2023)
Enhancing Interpretability of Vertebrae Fracture Grading using Human-interpretable Prototypes
von: Sinhamahapatra, Poulami, et al.
Veröffentlicht: (2024)
von: Sinhamahapatra, Poulami, et al.
Veröffentlicht: (2024)
KeyRe-ID: Keypoint-Guided Person Re-Identification using Part-Aware Representation in Videos
von: Kim, Jinseong, et al.
Veröffentlicht: (2025)
von: Kim, Jinseong, et al.
Veröffentlicht: (2025)
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
von: Kim, Kinam, et al.
Veröffentlicht: (2025)
von: Kim, Kinam, et al.
Veröffentlicht: (2025)
Panoptic Segmentation and Labelling of Lumbar Spine Vertebrae using Modified Attention Unet
von: Pal, Rikathi, et al.
Veröffentlicht: (2024)
von: Pal, Rikathi, et al.
Veröffentlicht: (2024)
Inscanner: Dual-Phase Detection and Classification of Auxiliary Insulation Using YOLOv8 Models
von: Kim, Youngtae, et al.
Veröffentlicht: (2025)
von: Kim, Youngtae, et al.
Veröffentlicht: (2025)
Your Vision-Language Model Can't Even Count to 20: Exposing the Failures of VLMs in Compositional Counting
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
When Words Can't Capture It All: Towards Video-Based User Complaint Text Generation with Multimodal Video Complaint Dataset
von: Das, Sarmistha, et al.
Veröffentlicht: (2025)
von: Das, Sarmistha, et al.
Veröffentlicht: (2025)
Can't See the Forest for the Trees: Benchmarking Multimodal Safety Awareness for Multimodal LLMs
von: Wang, Wenxuan, et al.
Veröffentlicht: (2025)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2025)
Collaborative Edge-to-Server Inference for Vision-Language Models
von: Song, Soochang, et al.
Veröffentlicht: (2025)
von: Song, Soochang, et al.
Veröffentlicht: (2025)
TV-LiVE: Training-Free, Text-Guided Video Editing via Layer Informed Vitality Exploitation
von: Kim, Min-Jung, et al.
Veröffentlicht: (2025)
von: Kim, Min-Jung, et al.
Veröffentlicht: (2025)
Learning to Make Keypoints Sub-Pixel Accurate
von: Kim, Shinjeong, et al.
Veröffentlicht: (2024)
von: Kim, Shinjeong, et al.
Veröffentlicht: (2024)
SD-Net: Symmetric-Aware Keypoint Prediction and Domain Adaptation for 6D Pose Estimation In Bin-picking Scenarios
von: Huang, Ding-Tao, et al.
Veröffentlicht: (2024)
von: Huang, Ding-Tao, et al.
Veröffentlicht: (2024)
Keypoints as Dynamic Centroids for Unified Human Pose and Segmentation
von: Ahmad, Niaz, et al.
Veröffentlicht: (2025)
von: Ahmad, Niaz, et al.
Veröffentlicht: (2025)
Solving Video Inverse Problems Using Image Diffusion Models
von: Kwon, Taesung, et al.
Veröffentlicht: (2024)
von: Kwon, Taesung, et al.
Veröffentlicht: (2024)
VISION-XL: High Definition Video Inverse Problem Solver using Latent Image Diffusion Models
von: Kwon, Taesung, et al.
Veröffentlicht: (2024)
von: Kwon, Taesung, et al.
Veröffentlicht: (2024)
Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects
von: Jo, Kyungmin, et al.
Veröffentlicht: (2024)
von: Jo, Kyungmin, et al.
Veröffentlicht: (2024)
Scaling Up Personalized Image Aesthetic Assessment via Task Vector Customization
von: Yun, Jooyeol, et al.
Veröffentlicht: (2024)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2024)
Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Morphology-Aware Interactive Keypoint Estimation
von: Kim, Jinhee, et al.
Veröffentlicht: (2022) -
EPIC: Effective Prompting for Imbalanced-Class Data Synthesis in Tabular Data Classification via Large Language Models
von: Kim, Jinhee, et al.
Veröffentlicht: (2024) -
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models
von: Kim, Jeongho, et al.
Veröffentlicht: (2024) -
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
von: Kim, Jeongho, et al.
Veröffentlicht: (2024) -
Attend-and-Refine: Interactive keypoint estimation and quantitative cervical vertebrae analysis for bone age assessment
von: Kim, Jinhee, et al.
Veröffentlicht: (2025)