Distribution-Aware Calibration for Object Detection with Noisy Bounding Boxes
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Donghao, Li, Jialin, Li, Jinpeng, Huang, Jiancheng, Nie, Qiang, Liu, Yong, Gao, Bin-Bin, Wang, Qiong, Heng, Pheng-Ann, Chen, Guangyong |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MagicTailor: Component-Controllable Personalization in Text-to-Image Diffusion Models
by: Zhou, Donghao, et al.
Published: (2024)
by: Zhou, Donghao, et al.
Published: (2024)
SignVTCL: Multi-Modal Continuous Sign Language Recognition Enhanced by Visual-Textual Contrastive Learning
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
Perceive and Calibrate: Analyzing and Enhancing Robustness of Medical Multi-Modal Large Language Models
by: XU, Dunyuan, et al.
Published: (2025)
by: XU, Dunyuan, et al.
Published: (2025)
Boosting In-Silicon Directed Evolution with Fine-Tuned Protein Language Model and Tree Search
by: Yang, Yaodong, et al.
Published: (2025)
by: Yang, Yaodong, et al.
Published: (2025)
Adaptive Image Zoom-in with Bounding Box Transformation for UAV Object Detection
by: Wang, Tao, et al.
Published: (2026)
by: Wang, Tao, et al.
Published: (2026)
DisCo-Layout: Disentangling and Coordinating Semantic and Physical Refinement in a Multi-Agent Framework for 3D Indoor Layout Synthesis
by: Gao, Jialin, et al.
Published: (2025)
by: Gao, Jialin, et al.
Published: (2025)
Point Cloud Understanding via Attention-Driven Contrastive Learning
by: Wang, Yi, et al.
Published: (2024)
by: Wang, Yi, et al.
Published: (2024)
MM-Mixing: Multi-Modal Mixing Alignment for 3D Understanding
by: Wang, Jiaze, et al.
Published: (2024)
by: Wang, Jiaze, et al.
Published: (2024)
SoftPatch: Unsupervised Anomaly Detection with Noisy Data
by: Jiang, Xi, et al.
Published: (2024)
by: Jiang, Xi, et al.
Published: (2024)
NBBOX: Noisy Bounding Box Improves Remote Sensing Object Detection
by: Kim, Yechan, et al.
Published: (2024)
by: Kim, Yechan, et al.
Published: (2024)
From Learning to Unlearning: Biomedical Security Protection in Multimodal Large Language Models
by: Xu, Dunyuan, et al.
Published: (2025)
by: Xu, Dunyuan, et al.
Published: (2025)
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis
by: Yu, Yang, et al.
Published: (2026)
by: Yu, Yang, et al.
Published: (2026)
Med-Evo: Test-time Self-evolution for Medical Multimodal Large Language Models
by: Xu, Dunyuan, et al.
Published: (2026)
by: Xu, Dunyuan, et al.
Published: (2026)
SFANet: Spatial-Frequency Attention Network for Weather Forecasting
by: Wang, Jiaze, et al.
Published: (2024)
by: Wang, Jiaze, et al.
Published: (2024)
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
by: Zhou, Donghao, et al.
Published: (2026)
by: Zhou, Donghao, et al.
Published: (2026)
Dual Ensembled Multiagent Q-Learning with Hypernet Regularizer
by: Yang, Yaodong, et al.
Published: (2025)
by: Yang, Yaodong, et al.
Published: (2025)
Multi‐Task Mixture Density Graph Neural Networks for Predicting Catalyst Performance
by: Chen Liang, et al.
Published: (2024)
by: Chen Liang, et al.
Published: (2024)
Temporal‐multimodal consistency alignment for Alzheimer's cognitive assessment prediction
by: Xikai Yang, et al.
Published: (2025)
by: Xikai Yang, et al.
Published: (2025)
BoxSeg: Quality-Aware and Peer-Assisted Learning for Box-supervised Instance Segmentation
by: Lai, Jinxiang, et al.
Published: (2025)
by: Lai, Jinxiang, et al.
Published: (2025)
Unlocking Potential Binders: Multimodal Pretraining DEL-Fusion for Denoising DNA-Encoded Libraries
by: Gu, Chunbin, et al.
Published: (2024)
by: Gu, Chunbin, et al.
Published: (2024)
MEJO: MLLM-Engaged Surgical Triplet Recognition via Inter- and Intra-Task Joint Optimization
by: Zhang, Yiyi, et al.
Published: (2025)
by: Zhang, Yiyi, et al.
Published: (2025)
Revisiting Shadow Detection: A New Benchmark Dataset for Complex World
by: Hu, Xiaowei, et al.
Published: (2019)
by: Hu, Xiaowei, et al.
Published: (2019)
Protein Inverse Folding From Structure Feedback
by: Xu, Junde, et al.
Published: (2025)
by: Xu, Junde, et al.
Published: (2025)
Adaptive Negative Evidential Deep Learning for Open-set Semi-supervised Learning
by: Yu, Yang, et al.
Published: (2023)
by: Yu, Yang, et al.
Published: (2023)
IdentityStory: Taming Your Identity-Preserving Generator for Human-Centric Story Generation
by: Zhou, Donghao, et al.
Published: (2025)
by: Zhou, Donghao, et al.
Published: (2025)
SeaDAG: Semi-autoregressive Diffusion for Conditional Directed Acyclic Graph Generation
by: Zhou, Xinyi, et al.
Published: (2024)
by: Zhou, Xinyi, et al.
Published: (2024)
Medical Large Vision Language Models with Multi-Image Visual Ability
by: Yang, Xikai, et al.
Published: (2025)
by: Yang, Xikai, et al.
Published: (2025)
UniHOPE: A Unified Approach for Hand-Only and Hand-Object Pose Estimation
by: Wang, Yinqiao, et al.
Published: (2025)
by: Wang, Yinqiao, et al.
Published: (2025)
Unlocking Positive Transfer in Incrementally Learning Surgical Instruments: A Self-reflection Hierarchical Prompt Framework
by: Zhu, Yu, et al.
Published: (2026)
by: Zhu, Yu, et al.
Published: (2026)
MMAD: A Comprehensive Benchmark for Multimodal Large Language Models in Industrial Anomaly Detection
by: Jiang, Xi, et al.
Published: (2024)
by: Jiang, Xi, et al.
Published: (2024)
Surgical Workflow Recognition and Blocking Effectiveness Detection in Laparoscopic Liver Resections with Pringle Maneuver
by: Guo, Diandian, et al.
Published: (2024)
by: Guo, Diandian, et al.
Published: (2024)
OBMO: One Bounding Box Multiple Objects for Monocular 3D Object Detection
by: Huang, Chenxi, et al.
Published: (2022)
by: Huang, Chenxi, et al.
Published: (2022)
OPA-Pack: Object-Property-Aware Robotic Bin Packing
by: Pan, Jia-Hui, et al.
Published: (2025)
by: Pan, Jia-Hui, et al.
Published: (2025)
Distortion-Aware Adversarial Attacks on Bounding Boxes of Object Detectors
by: Phuc, Pham, et al.
Published: (2024)
by: Phuc, Pham, et al.
Published: (2024)
SurgPub-Video: A Comprehensive Surgical Video Dataset for Enhanced Surgical Intelligence in Vision-Language Model
by: Li, Yaoqian, et al.
Published: (2025)
by: Li, Yaoqian, et al.
Published: (2025)
FPDIoU Loss: A Loss Function for Efficient Bounding Box Regression of Rotated Object Detection
by: Ma, Siliang, et al.
Published: (2024)
by: Ma, Siliang, et al.
Published: (2024)
HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images
by: Liu, Yichen, et al.
Published: (2026)
by: Liu, Yichen, et al.
Published: (2026)
SceneDecorator: Towards Scene-Oriented Story Generation with Scene Planning and Scene Consistency
by: Song, Quanjian, et al.
Published: (2025)
by: Song, Quanjian, et al.
Published: (2025)
SciVerse: Unveiling the Knowledge Comprehension and Visual Reasoning of LMMs on Multi-modal Scientific Problems
by: Guo, Ziyu, et al.
Published: (2025)
by: Guo, Ziyu, et al.
Published: (2025)
GEAR: GEometry-motion Alternating Refinement for Articulated Object Modeling with Gaussian Splatting
by: Li, Jialin, et al.
Published: (2026)
by: Li, Jialin, et al.
Published: (2026)
Similar Items
-
MagicTailor: Component-Controllable Personalization in Text-to-Image Diffusion Models
by: Zhou, Donghao, et al.
Published: (2024) -
SignVTCL: Multi-Modal Continuous Sign Language Recognition Enhanced by Visual-Textual Contrastive Learning
by: Chen, Hao, et al.
Published: (2024) -
Perceive and Calibrate: Analyzing and Enhancing Robustness of Medical Multi-Modal Large Language Models
by: XU, Dunyuan, et al.
Published: (2025) -
Boosting In-Silicon Directed Evolution with Fine-Tuned Protein Language Model and Tree Search
by: Yang, Yaodong, et al.
Published: (2025) -
Adaptive Image Zoom-in with Bounding Box Transformation for UAV Object Detection
by: Wang, Tao, et al.
Published: (2026)