D-FINE: Redefine Regression Task in DETRs as Fine-grained Distribution Refinement
Fuente:
arXiv
Saved in:
| Main Authors: | Peng, Yansong, Li, Hebei, Wu, Peixi, Zhang, Yueyi, Sun, Xiaoyan, Wu, Feng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scene Adaptive Sparse Transformer for Event-based Object Detection
by: Peng, Yansong, et al.
Published: (2024)
by: Peng, Yansong, et al.
Published: (2024)
Efficient Event-Based Semantic Segmentation via Exploiting Frame-Event Fusion: A Hybrid Neural Network Approach
by: Li, Hebei, et al.
Published: (2025)
by: Li, Hebei, et al.
Published: (2025)
Create Anything Anywhere: Layout-Controllable Personalized Diffusion Model for Multiple Subjects
by: Li, Wei, et al.
Published: (2025)
by: Li, Wei, et al.
Published: (2025)
Dome-DETR: DETR with Density-Oriented Feature-Query Manipulation for Efficient Tiny Object Detection
by: Hu, Zhangchi, et al.
Published: (2025)
by: Hu, Zhangchi, et al.
Published: (2025)
DASH: 4D Hash Encoding with Self-Supervised Decomposition for Real-Time Dynamic Scene Rendering
by: Chen, Jie, et al.
Published: (2025)
by: Chen, Jie, et al.
Published: (2025)
FACM: Flow-Anchored Consistency Models
by: Peng, Yansong, et al.
Published: (2025)
by: Peng, Yansong, et al.
Published: (2025)
Deep Multi-Threshold Spiking-UNet for Image Processing
by: Li, Hebei, et al.
Published: (2023)
by: Li, Hebei, et al.
Published: (2023)
Event-assisted Low-Light Video Object Segmentation
by: Li, Hebei, et al.
Published: (2024)
by: Li, Hebei, et al.
Published: (2024)
Unbiased Regression Loss for DETRs
by: Edric, et al.
Published: (2024)
by: Edric, et al.
Published: (2024)
RiO-DETR: DETR for Real-time Oriented Object Detection
by: Hu, Zhangchi, et al.
Published: (2026)
by: Hu, Zhangchi, et al.
Published: (2026)
Enhanced Object Detection: A Study on Vast Vocabulary Object Detection Track for V3Det Challenge 2024
by: Wu, Peixi, et al.
Published: (2024)
by: Wu, Peixi, et al.
Published: (2024)
Efficient Spiking Point Mamba for Point Cloud Analysis
by: Wu, Peixi, et al.
Published: (2025)
by: Wu, Peixi, et al.
Published: (2025)
EE-MLLM: A Data-Efficient and Compute-Efficient Multimodal Large Language Model
by: Ma, Feipeng, et al.
Published: (2024)
by: Ma, Feipeng, et al.
Published: (2024)
Dual DETRs for Multi-Label Temporal Action Detection
by: Zhu, Yuhan, et al.
Published: (2024)
by: Zhu, Yuhan, et al.
Published: (2024)
Motion Generation from Fine-grained Textual Descriptions
by: Li, Kunhang, et al.
Published: (2024)
by: Li, Kunhang, et al.
Published: (2024)
Integrating Diverse Assignment Strategies into DETRs
by: Zhang, Yiwei, et al.
Published: (2026)
by: Zhang, Yiwei, et al.
Published: (2026)
Ranking-based Adaptive Query Generation for DETRs in Crowded Pedestrian Detection
by: Gao, Feng, et al.
Published: (2023)
by: Gao, Feng, et al.
Published: (2023)
LongDWM: Cross-Granularity Distillation for Building a Long-Term Driving World Model
by: Wang, Xiaodong, et al.
Published: (2025)
by: Wang, Xiaodong, et al.
Published: (2025)
Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings
by: Wu, Peixi, et al.
Published: (2026)
by: Wu, Peixi, et al.
Published: (2026)
DETRs Beat YOLOs on Real-time Object Detection
by: Zhao, Yian, et al.
Published: (2023)
by: Zhao, Yian, et al.
Published: (2023)
Enhancing Object Discovery for Unsupervised Instance Segmentation and Object Detection
by: Feng, Xingyu, et al.
Published: (2025)
by: Feng, Xingyu, et al.
Published: (2025)
ImplantFormer: Vision Transformer based Implant Position Regression Using Dental CBCT Data
by: Yang, Xinquan, et al.
Published: (2022)
by: Yang, Xinquan, et al.
Published: (2022)
EvTexture: Event-driven Texture Enhancement for Video Super-Resolution
by: Kai, Dachun, et al.
Published: (2024)
by: Kai, Dachun, et al.
Published: (2024)
Graph Relation Distillation for Efficient Biomedical Instance Segmentation
by: Liu, Xiaoyu, et al.
Published: (2024)
by: Liu, Xiaoyu, et al.
Published: (2024)
ProphetDWM: A Driving World Model for Rolling Out Future Actions and Videos
by: Wang, Xiaodong, et al.
Published: (2025)
by: Wang, Xiaodong, et al.
Published: (2025)
FreeGen: Feed-Forward Reconstruction-Generation Co-Training for Free-Viewpoint Driving Scene Synthesis
by: Chen, Shijie, et al.
Published: (2025)
by: Chen, Shijie, et al.
Published: (2025)
CLoCKDistill: Consistent Location-and-Context-aware Knowledge Distillation for DETRs
by: Lan, Qizhen, et al.
Published: (2025)
by: Lan, Qizhen, et al.
Published: (2025)
Textual Inversion and Self-supervised Refinement for Radiology Report Generation
by: Luo, Yuanjiang, et al.
Published: (2024)
by: Luo, Yuanjiang, et al.
Published: (2024)
Task Adaptive Feature Distribution Based Network for Few-shot Fine-grained Target Classification
by: Li, Ping, et al.
Published: (2024)
by: Li, Ping, et al.
Published: (2024)
LLaDA-VLA: Vision Language Diffusion Action Models
by: Wen, Yuqing, et al.
Published: (2025)
by: Wen, Yuqing, et al.
Published: (2025)
On the Suitability of Reinforcement Fine-Tuning to Visual Tasks
by: Chen, Xiaxu, et al.
Published: (2025)
by: Chen, Xiaxu, et al.
Published: (2025)
Enhancing DETRs Variants through Improved Content Query and Similar Query Aggregation
by: Zhang, Yingying, et al.
Published: (2024)
by: Zhang, Yingying, et al.
Published: (2024)
Unsupervised Cross-Domain Regression for Fine-grained 3D Game Character Reconstruction
by: Wen, Qi, et al.
Published: (2024)
by: Wen, Qi, et al.
Published: (2024)
MMHead: Towards Fine-grained Multi-modal 3D Facial Animation
by: Wu, Sijing, et al.
Published: (2024)
by: Wu, Sijing, et al.
Published: (2024)
MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis
by: He, Wanggui, et al.
Published: (2024)
by: He, Wanggui, et al.
Published: (2024)
Refining CNN-based Heatmap Regression with Gradient-based Corner Points for Electrode Localization
by: Wu, Lin
Published: (2024)
by: Wu, Lin
Published: (2024)
Refining Segmentation On-the-Fly: An Interactive Framework for Point Cloud Semantic Segmentation
by: Zhang, Peng, et al.
Published: (2024)
by: Zhang, Peng, et al.
Published: (2024)
AU-Blendshape for Fine-grained Stylized 3D Facial Expression Manipulation
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
FINE: Factorizing Knowledge for Initialization of Variable-sized Diffusion Models
by: Xie, Yucheng, et al.
Published: (2024)
by: Xie, Yucheng, et al.
Published: (2024)
DASK: Distribution Rehearsing via Adaptive Style Kernel Learning for Exemplar-Free Lifelong Person Re-Identification
by: Xu, Kunlun, et al.
Published: (2024)
by: Xu, Kunlun, et al.
Published: (2024)
Similar Items
-
Scene Adaptive Sparse Transformer for Event-based Object Detection
by: Peng, Yansong, et al.
Published: (2024) -
Efficient Event-Based Semantic Segmentation via Exploiting Frame-Event Fusion: A Hybrid Neural Network Approach
by: Li, Hebei, et al.
Published: (2025) -
Create Anything Anywhere: Layout-Controllable Personalized Diffusion Model for Multiple Subjects
by: Li, Wei, et al.
Published: (2025) -
Dome-DETR: DETR with Density-Oriented Feature-Query Manipulation for Efficient Tiny Object Detection
by: Hu, Zhangchi, et al.
Published: (2025) -
DASH: 4D Hash Encoding with Self-Supervised Decomposition for Real-Time Dynamic Scene Rendering
by: Chen, Jie, et al.
Published: (2025)