OmniCount: Multi-label Object Counting with Semantic-Geometric Priors
Fuente:
arXiv
Guardado en:
| Autores principales: | Mondal, Anindya, Nag, Sauradip, Zhu, Xiatian, Dutta, Anjan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Actor-agnostic Multi-label Action Recognition with Multi-modal Query
por: Mondal, Anindya, et al.
Publicado: (2023)
por: Mondal, Anindya, et al.
Publicado: (2023)
A Single-Parameter Factor-Graph Image Prior
por: Wang, Tianyang, et al.
Publicado: (2026)
por: Wang, Tianyang, et al.
Publicado: (2026)
Gabor is Enough: Interpretable Deep Denoising with a Gabor Synthesis Dictionary Prior
por: Janjušević, Nikola, et al.
Publicado: (2022)
por: Janjušević, Nikola, et al.
Publicado: (2022)
SPAT: A Semantic Port-Aware Adaptive-Rate Transmission Protocol for Semantic Communication
por: Wang, Yunhao, et al.
Publicado: (2026)
por: Wang, Yunhao, et al.
Publicado: (2026)
Sparse Signal Reconstruction for Overdispersed Low-photon Count Biomedical Imaging Using $\ell_p$ Total Variation
por: Lu, Yu, et al.
Publicado: (2024)
por: Lu, Yu, et al.
Publicado: (2024)
Solving a Nonlinear Blind Inverse Problem for Tagged MRI with Physics and Deep Generative Priors
por: Bian, Zhangxing, et al.
Publicado: (2026)
por: Bian, Zhangxing, et al.
Publicado: (2026)
Preprocessing Algorithm Leveraging Geometric Modeling for Scale Correction in Hyperspectral Images for Improved Unmixing Performance
por: Sumanasekara, Praveen, et al.
Publicado: (2025)
por: Sumanasekara, Praveen, et al.
Publicado: (2025)
Semantic Communications with Computer Vision Sensing for Edge Video Transmission
por: Peng, Yubo, et al.
Publicado: (2025)
por: Peng, Yubo, et al.
Publicado: (2025)
Vision Transformer Based Semantic Communications for Next Generation Wireless Networks
por: Mohsin, Muhammad Ahmed, et al.
Publicado: (2025)
por: Mohsin, Muhammad Ahmed, et al.
Publicado: (2025)
WaveletGaussian: Wavelet-domain Diffusion for Sparse-view 3D Gaussian Object Reconstruction
por: Nguyen, Hung, et al.
Publicado: (2025)
por: Nguyen, Hung, et al.
Publicado: (2025)
Mamba-FCS: Joint Spatio- Frequency Feature Fusion, Change-Guided Attention, and SeK Loss for Enhanced Semantic Change Detection in Remote Sensing
por: Wijenayake, Buddhi, et al.
Publicado: (2025)
por: Wijenayake, Buddhi, et al.
Publicado: (2025)
EFCNet: Every Feature Counts for Small Medical Object Segmentation
por: Kong, Lingjie, et al.
Publicado: (2024)
por: Kong, Lingjie, et al.
Publicado: (2024)
CountLoop: Training-Free High-Instance Image Generation via Iterative Agent Guidance
por: Mondal, Anindya, et al.
Publicado: (2025)
por: Mondal, Anindya, et al.
Publicado: (2025)
Learnable Multi-level Discrete Wavelet Transforms for 3D Gaussian Splatting Frequency Modulation
por: Nguyen, Hung, et al.
Publicado: (2026)
por: Nguyen, Hung, et al.
Publicado: (2026)
Every Shot Counts: Using Exemplars for Repetition Counting in Videos
por: Sinha, Saptarshi, et al.
Publicado: (2024)
por: Sinha, Saptarshi, et al.
Publicado: (2024)
Decomposition, Compression, and Synthesis (DCS)-based Video Coding: A Neural Exploration via Resolution-Adaptive Learning
por: Lu, Ming, et al.
Publicado: (2020)
por: Lu, Ming, et al.
Publicado: (2020)
Multichannel Orthogonal Transform-Based Perceptron Layers for Efficient ResNets
por: Pan, Hongyi, et al.
Publicado: (2023)
por: Pan, Hongyi, et al.
Publicado: (2023)
Guided MRI Reconstruction via Schrödinger Bridge
por: Wang, Yue, et al.
Publicado: (2024)
por: Wang, Yue, et al.
Publicado: (2024)
P-Count: Persistence-based Counting of White Matter Hyperintensities in Brain MRI
por: Hu, Xiaoling, et al.
Publicado: (2024)
por: Hu, Xiaoling, et al.
Publicado: (2024)
Saving Foundation Flow-Matching Priors for Inverse Problems
por: Wan, Yuxiang, et al.
Publicado: (2025)
por: Wan, Yuxiang, et al.
Publicado: (2025)
FMPlug: Plug-In Foundation Flow-Matching Priors for Inverse Problems
por: Wan, Yuxiang, et al.
Publicado: (2025)
por: Wan, Yuxiang, et al.
Publicado: (2025)
Interpretable Lightweight Transformer via Unrolling of Learned Graph Smoothness Priors
por: Do, Tam Thuc, et al.
Publicado: (2024)
por: Do, Tam Thuc, et al.
Publicado: (2024)
Self-supervised Deep Learning for Denoising in Ultrasound Microvascular Imaging
por: Huang, Lijie, et al.
Publicado: (2025)
por: Huang, Lijie, et al.
Publicado: (2025)
Generative Video Semantic Communication via Multimodal Semantic Fusion with Large Model
por: Yin, Hang, et al.
Publicado: (2025)
por: Yin, Hang, et al.
Publicado: (2025)
Adaptive Transform Coding for Semantic Compression
por: Enttsel, Andriy, et al.
Publicado: (2026)
por: Enttsel, Andriy, et al.
Publicado: (2026)
Implementation of Licensed Plate Detection and Noise Removal in Image Processing
por: Gao, Yiquan
Publicado: (2026)
por: Gao, Yiquan
Publicado: (2026)
DWTGS: Rethinking Frequency Regularization for Sparse-view 3D Gaussian Splatting
por: Nguyen, Hung, et al.
Publicado: (2025)
por: Nguyen, Hung, et al.
Publicado: (2025)
A UAV-Based VNIR Hyperspectral Benchmark Dataset for Landmine and UXO Detection
por: Lekhak, Sagar, et al.
Publicado: (2025)
por: Lekhak, Sagar, et al.
Publicado: (2025)
Score-Based Turbo Message Passing for Plug-and-Play Compressive Image Recovery
por: Cai, Chang, et al.
Publicado: (2025)
por: Cai, Chang, et al.
Publicado: (2025)
Dimensional Coactivation for Representational Consistency in Frozen Vision Foundation Models
por: Saddik, Izaldein Al-Zyoud Abdulmotaleb El
Publicado: (2026)
por: Saddik, Izaldein Al-Zyoud Abdulmotaleb El
Publicado: (2026)
Compressive Recovery of Signals Defined on Perturbed Graphs
por: Ghosh, Sabyasachi, et al.
Publicado: (2024)
por: Ghosh, Sabyasachi, et al.
Publicado: (2024)
Tunable Wavelet Unit based Convolutional Neural Network in Optical Coherence Tomography Analysis Enhancement for Classifying Type of Epiretinal Membrane Surgery
por: Le, An, et al.
Publicado: (2025)
por: Le, An, et al.
Publicado: (2025)
Benchmarking ResNet for Short-Term Hypoglycemia Classification with DiaData
por: Cinar, Beyza, et al.
Publicado: (2025)
por: Cinar, Beyza, et al.
Publicado: (2025)
A CUBS-Compatible Ultrasound Morphology and Uncertainty-Aware Baseline for Carotid Intima-Media Segmentation and Preliminary Risk Prediction
por: Aueawatthanaphisut, Aueaphum
Publicado: (2026)
por: Aueawatthanaphisut, Aueaphum
Publicado: (2026)
Sim2Real SAR Image Restoration: Metadata-Driven Models for Joint Despeckling and Sidelobes Reduction
por: De Paepe, Antoine, et al.
Publicado: (2026)
por: De Paepe, Antoine, et al.
Publicado: (2026)
Integration of Communication and Computational Imaging
por: Yu, Zhenming, et al.
Publicado: (2024)
por: Yu, Zhenming, et al.
Publicado: (2024)
Biorthogonal Tunable Wavelet Unit with Lifting Scheme in Convolutional Neural Network
por: Le, An, et al.
Publicado: (2025)
por: Le, An, et al.
Publicado: (2025)
Surrogate-based cross-correlation for particle image velocimetry
por: Lee, Yong, et al.
Publicado: (2021)
por: Lee, Yong, et al.
Publicado: (2021)
Deep Learning Based Speckle Filtering for Polarimetric SAR Images. Application to Sentinel-1
por: Mestre-Quereda, Alejandro, et al.
Publicado: (2024)
por: Mestre-Quereda, Alejandro, et al.
Publicado: (2024)
LUM-ViT: Learnable Under-sampling Mask Vision Transformer for Bandwidth Limited Optical Signal Acquisition
por: Liu, Lingfeng, et al.
Publicado: (2024)
por: Liu, Lingfeng, et al.
Publicado: (2024)
Ejemplares similares
-
Actor-agnostic Multi-label Action Recognition with Multi-modal Query
por: Mondal, Anindya, et al.
Publicado: (2023) -
A Single-Parameter Factor-Graph Image Prior
por: Wang, Tianyang, et al.
Publicado: (2026) -
Gabor is Enough: Interpretable Deep Denoising with a Gabor Synthesis Dictionary Prior
por: Janjušević, Nikola, et al.
Publicado: (2022) -
SPAT: A Semantic Port-Aware Adaptive-Rate Transmission Protocol for Semantic Communication
por: Wang, Yunhao, et al.
Publicado: (2026) -
Sparse Signal Reconstruction for Overdispersed Low-photon Count Biomedical Imaging Using $\ell_p$ Total Variation
por: Lu, Yu, et al.
Publicado: (2024)