Physically Guided Visual Mass Estimation from a Single RGB Image
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Sungjae, Jeong, Junhan, Hong, Yeonjoo, Kim, Kwang In |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GraspCorrect: Robotic Grasp Correction via Vision-Language Model-Guided Feedback
von: Lee, Sungjae, et al.
Veröffentlicht: (2025)
von: Lee, Sungjae, et al.
Veröffentlicht: (2025)
Local Representative Token Guided Merging for Text-to-Image Generation
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
Predicting Depth Maps from Single RGB Images and Addressing Missing Information in Depth Estimation
von: Chaar, Mohamad Mofeed, et al.
Veröffentlicht: (2025)
von: Chaar, Mohamad Mofeed, et al.
Veröffentlicht: (2025)
VLM6D: VLM based 6Dof Pose Estimation based on RGB-D Images
von: Sarowar, Md Selim, et al.
Veröffentlicht: (2025)
von: Sarowar, Md Selim, et al.
Veröffentlicht: (2025)
Test-Time Mixup Augmentation for Data and Class-Specific Uncertainty Estimation in Deep Learning Image Classification
von: Lee, Hansang, et al.
Veröffentlicht: (2022)
von: Lee, Hansang, et al.
Veröffentlicht: (2022)
GuidNoise: Single-Pair Guided Diffusion for Generalized Noise Synthesis
von: Kim, Changjin, et al.
Veröffentlicht: (2025)
von: Kim, Changjin, et al.
Veröffentlicht: (2025)
Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model
von: Lee, Donghwna, et al.
Veröffentlicht: (2024)
von: Lee, Donghwna, et al.
Veröffentlicht: (2024)
RDPN6D: Residual-based Dense Point-wise Network for 6Dof Object Pose Estimation Based on RGB-D Images
von: Hong, Zong-Wei, et al.
Veröffentlicht: (2024)
von: Hong, Zong-Wei, et al.
Veröffentlicht: (2024)
Fourier-Guided Attention Upsampling for Image Super-Resolution
von: Choi, Daejune, et al.
Veröffentlicht: (2025)
von: Choi, Daejune, et al.
Veröffentlicht: (2025)
CoBELa: Steering Transparent Generation via Concept Bottlenecks on Energy Landscapes
von: Kim, Sangwon, et al.
Veröffentlicht: (2025)
von: Kim, Sangwon, et al.
Veröffentlicht: (2025)
DeClotH: Decomposable 3D Cloth and Human Body Reconstruction from a Single Image
von: Nam, Hyeongjin, et al.
Veröffentlicht: (2025)
von: Nam, Hyeongjin, et al.
Veröffentlicht: (2025)
Adversarial Attack for RGB-Event based Visual Object Tracking
von: Chen, Qiang, et al.
Veröffentlicht: (2025)
von: Chen, Qiang, et al.
Veröffentlicht: (2025)
QueensCAMP: an RGB-D dataset for robust Visual SLAM
von: Bruno, Hudson M. S., et al.
Veröffentlicht: (2024)
von: Bruno, Hudson M. S., et al.
Veröffentlicht: (2024)
Pix2Next: Leveraging Vision Foundation Models for RGB to NIR Image Translation
von: Jin, Youngwan, et al.
Veröffentlicht: (2024)
von: Jin, Youngwan, et al.
Veröffentlicht: (2024)
Real-Time Person Image Synthesis Using a Flow Matching Model
von: Jeong, Jiwoo, et al.
Veröffentlicht: (2025)
von: Jeong, Jiwoo, et al.
Veröffentlicht: (2025)
Image-Guided Semantic Pseudo-LiDAR Point Generation for 3D Object Detection
von: Lee, Minseung, et al.
Veröffentlicht: (2024)
von: Lee, Minseung, et al.
Veröffentlicht: (2024)
VisDoT : Enhancing Visual Reasoning through Human-Like Interpretation Grounding and Decomposition of Thought
von: Lee, Eunsoo, et al.
Veröffentlicht: (2026)
von: Lee, Eunsoo, et al.
Veröffentlicht: (2026)
Human Interaction-Aware 3D Reconstruction from a Single Image
von: Kim, Gwanghyun, et al.
Veröffentlicht: (2026)
von: Kim, Gwanghyun, et al.
Veröffentlicht: (2026)
Cycle-Correspondence Loss: Learning Dense View-Invariant Visual Features from Unlabeled and Unordered RGB Images
von: Adrian, David B., et al.
Veröffentlicht: (2024)
von: Adrian, David B., et al.
Veröffentlicht: (2024)
Reflexive Guidance: Improving OoDD in Vision-Language Models via Self-Guided Image-Adaptive Concept Generation
von: Kim, Jihyo, et al.
Veröffentlicht: (2024)
von: Kim, Jihyo, et al.
Veröffentlicht: (2024)
CaddieSet: A Golf Swing Dataset with Human Joint Features and Ball Information
von: Jung, Seunghyeon, et al.
Veröffentlicht: (2025)
von: Jung, Seunghyeon, et al.
Veröffentlicht: (2025)
VM-BHINet:Vision Mamba Bimanual Hand Interaction Network for 3D Interacting Hand Mesh Recovery From a Single RGB Image
von: Bi, Han, et al.
Veröffentlicht: (2025)
von: Bi, Han, et al.
Veröffentlicht: (2025)
Training-Free Text-Guided Image Editing with Visual Autoregressive Model
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
EBDM: Exemplar-guided Image Translation with Brownian-bridge Diffusion Models
von: Lee, Eungbean, et al.
Veröffentlicht: (2024)
von: Lee, Eungbean, et al.
Veröffentlicht: (2024)
WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting
von: Wang, Lezhong, et al.
Veröffentlicht: (2026)
von: Wang, Lezhong, et al.
Veröffentlicht: (2026)
D3T: Distinctive Dual-Domain Teacher Zigzagging Across RGB-Thermal Gap for Domain-Adaptive Object Detection
von: Do, Dinh Phat, et al.
Veröffentlicht: (2024)
von: Do, Dinh Phat, et al.
Veröffentlicht: (2024)
A 3DGS-Diffusion Self-Supervised Framework for Normal Estimation from a Single Image
von: Liang, Yanxing, et al.
Veröffentlicht: (2025)
von: Liang, Yanxing, et al.
Veröffentlicht: (2025)
Learning to Predict Aboveground Biomass from RGB Images with 3D Synthetic Scenes
von: Zuffi, Silvia
Veröffentlicht: (2025)
von: Zuffi, Silvia
Veröffentlicht: (2025)
Sparse Imagination for Efficient Visual World Model Planning
von: Chun, Junha, et al.
Veröffentlicht: (2025)
von: Chun, Junha, et al.
Veröffentlicht: (2025)
SIA: Enhancing Safety via Intent Awareness for Vision-Language Models
von: Na, Youngjin, et al.
Veröffentlicht: (2025)
von: Na, Youngjin, et al.
Veröffentlicht: (2025)
Preserve or Modify? Context-Aware Evaluation for Balancing Preservation and Modification in Text-Guided Image Editing
von: Kim, Yoonjeon, et al.
Veröffentlicht: (2024)
von: Kim, Yoonjeon, et al.
Veröffentlicht: (2024)
APT: Improving Diffusion Models for High Resolution Image Generation with Adaptive Path Tracing
von: Han, Sangmin, et al.
Veröffentlicht: (2025)
von: Han, Sangmin, et al.
Veröffentlicht: (2025)
From Illusion to Intention: Visual Rationale Learning for Vision-Language Reasoning
von: Wang, Changpeng, et al.
Veröffentlicht: (2025)
von: Wang, Changpeng, et al.
Veröffentlicht: (2025)
An Object-Based Deep Learning Approach for Building Height Estimation from Single SAR Images
von: Memar, Babak, et al.
Veröffentlicht: (2025)
von: Memar, Babak, et al.
Veröffentlicht: (2025)
mEOL: Training-Free Instruction-Guided Multimodal Embedder for Vector Graphics and Image Retrieval
von: Kim, Kyeong Seon, et al.
Veröffentlicht: (2026)
von: Kim, Kyeong Seon, et al.
Veröffentlicht: (2026)
Conditional Diffusion Model for Longitudinal Medical Image Generation
von: Dao, Duy-Phuong, et al.
Veröffentlicht: (2024)
von: Dao, Duy-Phuong, et al.
Veröffentlicht: (2024)
Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models
von: Jung, Mingi, et al.
Veröffentlicht: (2025)
von: Jung, Mingi, et al.
Veröffentlicht: (2025)
Semi-Supervised Audio-Visual Video Action Recognition with Audio Source Localization Guided Mixup
von: Kang, Seokun, et al.
Veröffentlicht: (2025)
von: Kang, Seokun, et al.
Veröffentlicht: (2025)
MAST: Mask-Guided Attention Mass Allocation for Training-Free Multi-Style Transfer
von: Kang, Dongkyung, et al.
Veröffentlicht: (2026)
von: Kang, Dongkyung, et al.
Veröffentlicht: (2026)
Text-Guided Variational Image Generation for Industrial Anomaly Detection and Segmentation
von: Lee, Mingyu, et al.
Veröffentlicht: (2024)
von: Lee, Mingyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GraspCorrect: Robotic Grasp Correction via Vision-Language Model-Guided Feedback
von: Lee, Sungjae, et al.
Veröffentlicht: (2025) -
Local Representative Token Guided Merging for Text-to-Image Generation
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025) -
Predicting Depth Maps from Single RGB Images and Addressing Missing Information in Depth Estimation
von: Chaar, Mohamad Mofeed, et al.
Veröffentlicht: (2025) -
VLM6D: VLM based 6Dof Pose Estimation based on RGB-D Images
von: Sarowar, Md Selim, et al.
Veröffentlicht: (2025) -
Test-Time Mixup Augmentation for Data and Class-Specific Uncertainty Estimation in Deep Learning Image Classification
von: Lee, Hansang, et al.
Veröffentlicht: (2022)