Adapted Center and Scale Prediction: More Stable and More Accurate
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wang, Wenhao, Zhang, Jusheng |
|---|---|
| Format: | Preprint |
| Publié: |
2020
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Shape-IoU: More Accurate Metric considering Bounding Box Shape and Scale
par: Zhang, Hao, et autres
Publié: (2023)
par: Zhang, Hao, et autres
Publié: (2023)
Towards More Accurate Diffusion Model Acceleration with A Timestep Tuner
par: Xia, Mengfei, et autres
Publié: (2023)
par: Xia, Mengfei, et autres
Publié: (2023)
Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors
par: Wang, Xiangchen, et autres
Publié: (2025)
par: Wang, Xiangchen, et autres
Publié: (2025)
Reduce the Artifacts Bias for More Generalizable AI-Generated Image Detection
par: Li, Yiheng, et autres
Publié: (2026)
par: Li, Yiheng, et autres
Publié: (2026)
Towards More Accurate Personalized Image Generation: Addressing Overfitting and Evaluation Bias
par: Li, Mingxiao, et autres
Publié: (2025)
par: Li, Mingxiao, et autres
Publié: (2025)
MeMix: Writing Less, Remembering More for Streaming 3D Reconstruction
par: Dong, Jiacheng, et autres
Publié: (2026)
par: Dong, Jiacheng, et autres
Publié: (2026)
More Pictures Say More: Visual Intersection Network for Open Set Object Detection
par: Dong, Bingcheng, et autres
Publié: (2024)
par: Dong, Bingcheng, et autres
Publié: (2024)
More Clear, More Flexible, More Precise: A Comprehensive Oriented Object Detection benchmark for UAV
par: Ye, Kai, et autres
Publié: (2025)
par: Ye, Kai, et autres
Publié: (2025)
Scaling Laws in Patchification: An Image Is Worth 50,176 Tokens And More
par: Wang, Feng, et autres
Publié: (2025)
par: Wang, Feng, et autres
Publié: (2025)
Discriminative Consensus Mining with A Thousand Groups for More Accurate Co-Salient Object Detection
par: Zheng, Peng
Publié: (2024)
par: Zheng, Peng
Publié: (2024)
Alignment and Adversarial Robustness: Are More Human-Like Models More Secure?
par: Hoak, Blaine, et autres
Publié: (2025)
par: Hoak, Blaine, et autres
Publié: (2025)
3DAlign-DAER: Dynamic Attention Policy and Efficient Retrieval Strategy for Fine-grained 3D-Text Alignment at Scale
par: Fan, Yijia, et autres
Publié: (2025)
par: Fan, Yijia, et autres
Publié: (2025)
SEED: Towards More Accurate Semantic Evaluation for Visual Brain Decoding
par: Park, Juhyeon, et autres
Publié: (2025)
par: Park, Juhyeon, et autres
Publié: (2025)
Samba+: General and Accurate Salient Object Detection via A More Unified Mamba-based Framework
par: Zhao, Wenzhuo, et autres
Publié: (2026)
par: Zhao, Wenzhuo, et autres
Publié: (2026)
HMAFlow: Learning More Accurate Optical Flow via Hierarchical Motion Field Alignment
par: Ma, Dianbo, et autres
Publié: (2024)
par: Ma, Dianbo, et autres
Publié: (2024)
Hypergraph Vision Transformers: Images are More than Nodes, More than Edges
par: Fixelle, Joshua
Publié: (2025)
par: Fixelle, Joshua
Publié: (2025)
The More You See in 2D, the More You Perceive in 3D
par: Han, Xinyang, et autres
Publié: (2024)
par: Han, Xinyang, et autres
Publié: (2024)
More Images, More Problems? A Controlled Analysis of VLM Failure Modes
par: Das, Anurag, et autres
Publié: (2026)
par: Das, Anurag, et autres
Publié: (2026)
Proper Body Landmark Subset Enables More Accurate and 5X Faster Recognition of Isolated Signs in LIBRAS
par: Santos, Daniele L. V. dos, et autres
Publié: (2025)
par: Santos, Daniele L. V. dos, et autres
Publié: (2025)
Does Seeing More Mean Knowing More? Mono-Anchored Advantage Normalization for Multi-Source Visual Reasoning
par: Zeng, Fanhu, et autres
Publié: (2026)
par: Zeng, Fanhu, et autres
Publié: (2026)
GloSplat: Joint Pose-Appearance Optimization for Faster and More Accurate 3D Reconstruction
par: Xiong, Tianyu, et autres
Publié: (2026)
par: Xiong, Tianyu, et autres
Publié: (2026)
SAM2-Adapter: Evaluating & Adapting Segment Anything 2 in Downstream Tasks: Camouflage, Shadow, Medical Image Segmentation, and More
par: Chen, Tianrun, et autres
Publié: (2024)
par: Chen, Tianrun, et autres
Publié: (2024)
Towards More Transparent and Accurate Cancer Diagnosis with an Unsupervised CAE Approach
par: Tabatabaei, Zahra, et autres
Publié: (2023)
par: Tabatabaei, Zahra, et autres
Publié: (2023)
RESBev: Making BEV Perception More Robust
par: Zhuo, Lifeng, et autres
Publié: (2026)
par: Zhuo, Lifeng, et autres
Publié: (2026)
AccDiffusion v2: Towards More Accurate Higher-Resolution Diffusion Extrapolation
par: Lin, Zhihang, et autres
Publié: (2024)
par: Lin, Zhihang, et autres
Publié: (2024)
AffordanceSAM: Segment Anything Once More in Affordance Grounding
par: Jiang, Dengyang, et autres
Publié: (2025)
par: Jiang, Dengyang, et autres
Publié: (2025)
LIME: Less Is More for MLLM Evaluation
par: Zhu, King, et autres
Publié: (2024)
par: Zhu, King, et autres
Publié: (2024)
Exploring Richer and More Accurate Information via Frequency Selection for Image Restoration
par: Gao, Hu, et autres
Publié: (2024)
par: Gao, Hu, et autres
Publié: (2024)
Do More Details Always Introduce More Hallucinations in LVLM-based Image Captioning?
par: Feng, Mingqian, et autres
Publié: (2024)
par: Feng, Mingqian, et autres
Publié: (2024)
MetroGS: Efficient and Stable Reconstruction of Geometrically Accurate High-Fidelity Large-Scale Scenes
par: Chen, Kehua, et autres
Publié: (2025)
par: Chen, Kehua, et autres
Publié: (2025)
Less-to-More Generalization: Unlocking More Controllability by In-Context Generation
par: Wu, Shaojin, et autres
Publié: (2025)
par: Wu, Shaojin, et autres
Publié: (2025)
One-for-More: Continual Diffusion Model for Anomaly Detection
par: Li, Xiaofan, et autres
Publié: (2025)
par: Li, Xiaofan, et autres
Publié: (2025)
Focaler-IoU: More Focused Intersection over Union Loss
par: Zhang, Hao, et autres
Publié: (2024)
par: Zhang, Hao, et autres
Publié: (2024)
Floating No More: Object-Ground Reconstruction from a Single Image
par: Man, Yunze, et autres
Publié: (2024)
par: Man, Yunze, et autres
Publié: (2024)
No More Sibling Rivalry: Debiasing Human-Object Interaction Detection
par: Yang, Bin, et autres
Publié: (2025)
par: Yang, Bin, et autres
Publié: (2025)
Vision Transformers Need More Than Registers
par: Shi, Cheng, et autres
Publié: (2026)
par: Shi, Cheng, et autres
Publié: (2026)
Less is More: Discovering Concise Network Explanations
par: Kondapaneni, Neehar, et autres
Publié: (2024)
par: Kondapaneni, Neehar, et autres
Publié: (2024)
Towards More Unified In-context Visual Understanding
par: Sheng, Dianmo, et autres
Publié: (2023)
par: Sheng, Dianmo, et autres
Publié: (2023)
Towards More Accurate Fake Detection on Images Generated from Advanced Generative and Neural Rendering Models
par: Dong, Chengdong, et autres
Publié: (2024)
par: Dong, Chengdong, et autres
Publié: (2024)
From One to More: Contextual Part Latents for 3D Generation
par: Dong, Shaocong, et autres
Publié: (2025)
par: Dong, Shaocong, et autres
Publié: (2025)
Documents similaires
-
Shape-IoU: More Accurate Metric considering Bounding Box Shape and Scale
par: Zhang, Hao, et autres
Publié: (2023) -
Towards More Accurate Diffusion Model Acceleration with A Timestep Tuner
par: Xia, Mengfei, et autres
Publié: (2023) -
Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors
par: Wang, Xiangchen, et autres
Publié: (2025) -
Reduce the Artifacts Bias for More Generalizable AI-Generated Image Detection
par: Li, Yiheng, et autres
Publié: (2026) -
Towards More Accurate Personalized Image Generation: Addressing Overfitting and Evaluation Bias
par: Li, Mingxiao, et autres
Publié: (2025)