Determinantal Point Process as an alternative to NMS
Fuente:
arXiv
Saved in:
| Main Authors: | Some, Samik, Gupta, Mithun Das, Namboodiri, Vinay P. |
|---|---|
| Format: | Preprint |
| Published: |
2020
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Trusting Semantic Segmentation Networks
by: Some, Samik, et al.
Published: (2024)
by: Some, Samik, et al.
Published: (2024)
Can Unsupervised Segmentation Reduce Annotation Costs for Video Semantic Segmentation?
by: Some, Samik, et al.
Published: (2026)
by: Some, Samik, et al.
Published: (2026)
StyleYourSmile: Cross-Domain Face Retargeting Without Paired Multi-Style Data
by: Dey, Avirup, et al.
Published: (2025)
by: Dey, Avirup, et al.
Published: (2025)
CLIP Adaptation by Intra-modal Overlap Reduction
by: Kravets, Alexey, et al.
Published: (2024)
by: Kravets, Alexey, et al.
Published: (2024)
TalkLoRA: Low-Rank Adaptation for Speech-Driven Animation
by: Saunders, Jack, et al.
Published: (2024)
by: Saunders, Jack, et al.
Published: (2024)
Dubbing for Everyone: Data-Efficient Visual Dubbing using Neural Rendering Priors
by: Saunders, Jack, et al.
Published: (2024)
by: Saunders, Jack, et al.
Published: (2024)
Rethinking Few Shot CLIP Benchmarks: A Critical Analysis in the Inductive Setting
by: Kravets, Alexey, et al.
Published: (2025)
by: Kravets, Alexey, et al.
Published: (2025)
MedFocusCLIP : Improving few shot classification in medical datasets using pixel wise attention
by: Arora, Aadya, et al.
Published: (2025)
by: Arora, Aadya, et al.
Published: (2025)
Self-supervised Representation Learning for Cell Event Recognition through Time Arrow Prediction
by: Chen, Cangxiong, et al.
Published: (2024)
by: Chen, Cangxiong, et al.
Published: (2024)
EIDT-V: Exploiting Intersections in Diffusion Trajectories for Model-Agnostic, Zero-Shot, Training-Free Text-to-Video Generation
by: Jagpal, Diljeet, et al.
Published: (2025)
by: Jagpal, Diljeet, et al.
Published: (2025)
RAW: Robust Avatar Watermarking -- Benchmarking and Baseline
by: Parry, Jack, et al.
Published: (2026)
by: Parry, Jack, et al.
Published: (2026)
Interpretability Transfer from Language to Vision via Sparse Autoencoders
by: Kravets, Alexey, et al.
Published: (2026)
by: Kravets, Alexey, et al.
Published: (2026)
VERSE: Virtual-Gradient Aware Streaming Lifelong Learning with Anytime Inference
by: Banerjee, Soumya, et al.
Published: (2023)
by: Banerjee, Soumya, et al.
Published: (2023)
RISSOLE: Parameter-efficient Diffusion Models via Block-wise Generation and Retrieval-Guidance
by: Mukherjee, Avideep, et al.
Published: (2024)
by: Mukherjee, Avideep, et al.
Published: (2024)
Differentiable NMS via Sinkhorn Matching for End-to-End Fabric Defect Detection
by: Lu, Zhengyang, et al.
Published: (2025)
by: Lu, Zhengyang, et al.
Published: (2025)
Diverse Video Generation with Determinantal Point Process-Guided Policy Optimization
by: Kazimi, Tahira, et al.
Published: (2025)
by: Kazimi, Tahira, et al.
Published: (2025)
Fusion2Print: Deep Flash-Non-Flash Fusion for Contactless Fingerprint Matching
by: Sahoo, Roja, et al.
Published: (2026)
by: Sahoo, Roja, et al.
Published: (2026)
Zero-Shot Class Unlearning in CLIP with Synthetic Samples
by: Kravets, A., et al.
Published: (2024)
by: Kravets, A., et al.
Published: (2024)
Illumination-Aware Contactless Fingerprint Spoof Detection via Paired Flash-Non-Flash Imaging
by: Sahoo, Roja, et al.
Published: (2026)
by: Sahoo, Roja, et al.
Published: (2026)
Heracles: A Hybrid SSM-Transformer Model for High-Resolution Image and Time-Series Analysis
by: Patro, Badri N., et al.
Published: (2024)
by: Patro, Badri N., et al.
Published: (2024)
YOLO26: An Analysis of NMS-Free End to End Framework for Real-Time Object Detection
by: Chakrabarty, Sudip
Published: (2026)
by: Chakrabarty, Sudip
Published: (2026)
Enhancement-Driven Pretraining for Robust Fingerprint Representation Learning
by: Gavas, Ekta, et al.
Published: (2024)
by: Gavas, Ekta, et al.
Published: (2024)
Towards Accurate Lip-to-Speech Synthesis in-the-Wild
by: Hegde, Sindhu, et al.
Published: (2024)
by: Hegde, Sindhu, et al.
Published: (2024)
Multi-Object Advertisement Creative Generation
by: Gao, Jialu, et al.
Published: (2026)
by: Gao, Jialu, et al.
Published: (2026)
A Novel Cloud-Based Diffusion-Guided Hybrid Model for High-Accuracy Accident Detection in Intelligent Transportation Systems
by: Sai, Siva, et al.
Published: (2025)
by: Sai, Siva, et al.
Published: (2025)
PS-StyleGAN: Illustrative Portrait Sketching using Attention-Based Style Adaptation
by: Jain, Kushal Kumar, et al.
Published: (2024)
by: Jain, Kushal Kumar, et al.
Published: (2024)
Multiscale Real-Time Object Detection in the NMS-Free Era: A Comparative Performance Evaluation of YOLOv8 and YOLO26
by: Oguine, Chidera G., et al.
Published: (2026)
by: Oguine, Chidera G., et al.
Published: (2026)
Convolutional Prompting meets Language Models for Continual Learning
by: Roy, Anurag, et al.
Published: (2024)
by: Roy, Anurag, et al.
Published: (2024)
CLIP4Sketch: Enhancing Sketch to Mugshot Matching through Dataset Augmentation using Diffusion Models
by: Jain, Kushal Kumar, et al.
Published: (2024)
by: Jain, Kushal Kumar, et al.
Published: (2024)
NOVO: Unlearning-Compliant Vision Transformers
by: Roy, Soumya, et al.
Published: (2025)
by: Roy, Soumya, et al.
Published: (2025)
Efficient Point Transformer with Dynamic Token Aggregating for LiDAR Point Cloud Processing
by: Lu, Dening, et al.
Published: (2024)
by: Lu, Dening, et al.
Published: (2024)
Large Language Model-Guided Semantic Alignment for Human Activity Recognition
by: Yan, Hua, et al.
Published: (2024)
by: Yan, Hua, et al.
Published: (2024)
Spectral Informed Mamba for Robust Point Cloud Processing
by: Bahri, Ali, et al.
Published: (2025)
by: Bahri, Ali, et al.
Published: (2025)
Efficient Text-Guided Convolutional Adapter for the Diffusion Model
by: Das, Aryan, et al.
Published: (2026)
by: Das, Aryan, et al.
Published: (2026)
NaVIP: An Image-Centric Indoor Navigation Solution for Visually Impaired People
by: Yu, Jun, et al.
Published: (2024)
by: Yu, Jun, et al.
Published: (2024)
PointTransformerX: Portable and Efficient 3D Point Cloud Processing without Sparse Algorithms
by: Reichardt, Laurenz, et al.
Published: (2026)
by: Reichardt, Laurenz, et al.
Published: (2026)
ObitoNet: Multimodal High-Resolution Point Cloud Reconstruction
by: Thapliyal, Apoorv, et al.
Published: (2024)
by: Thapliyal, Apoorv, et al.
Published: (2024)
PIT-QMM: A Large Multimodal Model For No-Reference Point Cloud Quality Assessment
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
A Strong Baseline for Point Cloud Registration via Direct Superpoints Matching
by: Gupta, Aniket, et al.
Published: (2023)
by: Gupta, Aniket, et al.
Published: (2023)
CurveCloudNet: Processing Point Clouds with 1D Structure
by: Stearns, Colton, et al.
Published: (2023)
by: Stearns, Colton, et al.
Published: (2023)
Similar Items
-
Trusting Semantic Segmentation Networks
by: Some, Samik, et al.
Published: (2024) -
Can Unsupervised Segmentation Reduce Annotation Costs for Video Semantic Segmentation?
by: Some, Samik, et al.
Published: (2026) -
StyleYourSmile: Cross-Domain Face Retargeting Without Paired Multi-Style Data
by: Dey, Avirup, et al.
Published: (2025) -
CLIP Adaptation by Intra-modal Overlap Reduction
by: Kravets, Alexey, et al.
Published: (2024) -
TalkLoRA: Low-Rank Adaptation for Speech-Driven Animation
by: Saunders, Jack, et al.
Published: (2024)