Boosting Quantitive and Spatial Awareness for Zero-Shot Object Counting
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Da, Li, Bingyu, Wang, Feiyu, Zhao, Zhiyuan, Gao, Junyu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MARIS: Marine Open-Vocabulary Instance Segmentation with Geometric Enhancement and Semantic Alignment
di: Li, Bingyu, et al.
Pubblicazione: (2025)
di: Li, Bingyu, et al.
Pubblicazione: (2025)
UWBench: A Comprehensive Vision-Language Benchmark for Underwater Understanding
di: Zhang, Da, et al.
Pubblicazione: (2025)
di: Zhang, Da, et al.
Pubblicazione: (2025)
IntroSVG: Learning from Rendering Feedback for Text-to-SVG Generation via an Introspective Generator-Critic Framework
di: Wang, Feiyu, et al.
Pubblicazione: (2026)
di: Wang, Feiyu, et al.
Pubblicazione: (2026)
U3M: Unbiased Multiscale Modal Fusion Model for Multimodal Semantic Segmentation
di: Li, Bingyu, et al.
Pubblicazione: (2024)
di: Li, Bingyu, et al.
Pubblicazione: (2024)
FGAseg: Fine-Grained Pixel-Text Alignment for Open-Vocabulary Semantic Segmentation
di: Li, Bingyu, et al.
Pubblicazione: (2025)
di: Li, Bingyu, et al.
Pubblicazione: (2025)
StitchFusion: Weaving Any Visual Modalities to Enhance Multimodal Semantic Segmentation
di: Li, Bingyu, et al.
Pubblicazione: (2024)
di: Li, Bingyu, et al.
Pubblicazione: (2024)
Quantum-inspired Interpretable Deep Learning Architecture for Text Sentiment Analysis
di: Li, Bingyu, et al.
Pubblicazione: (2024)
di: Li, Bingyu, et al.
Pubblicazione: (2024)
An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation
di: Li, Bingyu, et al.
Pubblicazione: (2026)
di: Li, Bingyu, et al.
Pubblicazione: (2026)
Towards Realistic Open-Vocabulary Remote Sensing Segmentation: Benchmark and Baseline
di: Li, Bingyu, et al.
Pubblicazione: (2026)
di: Li, Bingyu, et al.
Pubblicazione: (2026)
Exploring Efficient Open-Vocabulary Segmentation in the Remote Sensing
di: Li, Bingyu, et al.
Pubblicazione: (2025)
di: Li, Bingyu, et al.
Pubblicazione: (2025)
Exploring the Underwater World Segmentation without Extra Training
di: Li, Bingyu, et al.
Pubblicazione: (2025)
di: Li, Bingyu, et al.
Pubblicazione: (2025)
Prototype-Based Low Altitude UAV Semantic Segmentation
di: Zhang, Da, et al.
Pubblicazione: (2026)
di: Zhang, Da, et al.
Pubblicazione: (2026)
NWPU-MOC: A Benchmark for Fine-grained Multi-category Object Counting in Aerial Images
di: Gao, Junyu, et al.
Pubblicazione: (2024)
di: Gao, Junyu, et al.
Pubblicazione: (2024)
Expanding Zero-Shot Object Counting with Rich Prompts
di: Zhu, Huilin, et al.
Pubblicazione: (2025)
di: Zhu, Huilin, et al.
Pubblicazione: (2025)
SVGen: Interpretable Vector Graphics Generation with Large Language Models
di: Wang, Feiyu, et al.
Pubblicazione: (2025)
di: Wang, Feiyu, et al.
Pubblicazione: (2025)
One-Shot Crowd Counting With Density Guidance For Scene Adaptation
di: Chen, Jiwei, et al.
Pubblicazione: (2026)
di: Chen, Jiwei, et al.
Pubblicazione: (2026)
Dynamic Proxy Domain Generalizes the Crowd Localization by Better Binary Segmentation
di: Gao, Junyu, et al.
Pubblicazione: (2024)
di: Gao, Junyu, et al.
Pubblicazione: (2024)
Text-guided Zero-Shot Object Localization
di: Wang, Jingjing, et al.
Pubblicazione: (2024)
di: Wang, Jingjing, et al.
Pubblicazione: (2024)
VLCounter: Text-aware Visual Representation for Zero-Shot Object Counting
di: Kang, Seunggu, et al.
Pubblicazione: (2023)
di: Kang, Seunggu, et al.
Pubblicazione: (2023)
Boosting Object Detection with Zero-Shot Day-Night Domain Adaptation
di: Du, Zhipeng, et al.
Pubblicazione: (2023)
di: Du, Zhipeng, et al.
Pubblicazione: (2023)
CountZES: Counting via Zero-Shot Exemplar Selection
di: Siddiqui, Muhammad Ibraheem, et al.
Pubblicazione: (2025)
di: Siddiqui, Muhammad Ibraheem, et al.
Pubblicazione: (2025)
Text-promptable Object Counting via Quantity Awareness Enhancement
di: Shi, Miaojing, et al.
Pubblicazione: (2025)
di: Shi, Miaojing, et al.
Pubblicazione: (2025)
Locality-Aware Zero-Shot Human-Object Interaction Detection
di: Kim, Sanghyun, et al.
Pubblicazione: (2025)
di: Kim, Sanghyun, et al.
Pubblicazione: (2025)
UNICBench: UNIfied Counting Benchmark for MLLM
di: Rong, Chenggang, et al.
Pubblicazione: (2026)
di: Rong, Chenggang, et al.
Pubblicazione: (2026)
Zero-shot Object Counting with Good Exemplars
di: Zhu, Huilin, et al.
Pubblicazione: (2024)
di: Zhu, Huilin, et al.
Pubblicazione: (2024)
Mutually-Aware Feature Learning for Few-Shot Object Counting
di: Jeon, Yerim, et al.
Pubblicazione: (2024)
di: Jeon, Yerim, et al.
Pubblicazione: (2024)
Liberating Seen Classes: Boosting Few-Shot and Zero-Shot Text Classification via Anchor Generation and Classification Reframing
di: Liu, Han, et al.
Pubblicazione: (2024)
di: Liu, Han, et al.
Pubblicazione: (2024)
Fine-Grained Zero-Shot Object Detection
di: Ma, Hongxu, et al.
Pubblicazione: (2025)
di: Ma, Hongxu, et al.
Pubblicazione: (2025)
Real-Time Crowd Counting for Embedded Systems with Lightweight Architecture
di: Zhao, Zhiyuan, et al.
Pubblicazione: (2025)
di: Zhao, Zhiyuan, et al.
Pubblicazione: (2025)
T2ICount: Enhancing Cross-modal Understanding for Zero-Shot Counting
di: Qian, Yifei, et al.
Pubblicazione: (2025)
di: Qian, Yifei, et al.
Pubblicazione: (2025)
VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing
di: Couairon, Paul, et al.
Pubblicazione: (2023)
di: Couairon, Paul, et al.
Pubblicazione: (2023)
UPRE: Zero-Shot Domain Adaptation for Object Detection via Unified Prompt and Representation Enhancement
di: Zhang, Xiao, et al.
Pubblicazione: (2025)
di: Zhang, Xiao, et al.
Pubblicazione: (2025)
Zero-Shot Co-salient Object Detection Framework
di: Xiao, Haoke, et al.
Pubblicazione: (2023)
di: Xiao, Haoke, et al.
Pubblicazione: (2023)
Learning Spatial Similarity Distribution for Few-shot Object Counting
di: Xu, Yuanwu, et al.
Pubblicazione: (2024)
di: Xu, Yuanwu, et al.
Pubblicazione: (2024)
Task-Specific Zero-shot Quantization-Aware Training for Object Detection
di: Li, Changhao, et al.
Pubblicazione: (2025)
di: Li, Changhao, et al.
Pubblicazione: (2025)
CREST: Cross-modal Resonance through Evidential Deep Learning for Enhanced Zero-Shot Learning
di: Huang, Haojian, et al.
Pubblicazione: (2024)
di: Huang, Haojian, et al.
Pubblicazione: (2024)
Motion-Zero: Zero-Shot Moving Object Control Framework for Diffusion-Based Video Generation
di: Chen, Changgu, et al.
Pubblicazione: (2024)
di: Chen, Changgu, et al.
Pubblicazione: (2024)
TeDA: Boosting Vision-Lanuage Models for Zero-Shot 3D Object Retrieval via Testing-time Distribution Alignment
di: Wang, Zhichuan, et al.
Pubblicazione: (2025)
di: Wang, Zhichuan, et al.
Pubblicazione: (2025)
Boosting 3D Object Detection with Semantic-Aware Multi-Branch Framework
di: Jing, Hao, et al.
Pubblicazione: (2024)
di: Jing, Hao, et al.
Pubblicazione: (2024)
SpatialFormer: Semantic and Target Aware Attentions for Few-Shot Learning
di: Lai, Jinxiang, et al.
Pubblicazione: (2023)
di: Lai, Jinxiang, et al.
Pubblicazione: (2023)
Documenti analoghi
-
MARIS: Marine Open-Vocabulary Instance Segmentation with Geometric Enhancement and Semantic Alignment
di: Li, Bingyu, et al.
Pubblicazione: (2025) -
UWBench: A Comprehensive Vision-Language Benchmark for Underwater Understanding
di: Zhang, Da, et al.
Pubblicazione: (2025) -
IntroSVG: Learning from Rendering Feedback for Text-to-SVG Generation via an Introspective Generator-Critic Framework
di: Wang, Feiyu, et al.
Pubblicazione: (2026) -
U3M: Unbiased Multiscale Modal Fusion Model for Multimodal Semantic Segmentation
di: Li, Bingyu, et al.
Pubblicazione: (2024) -
FGAseg: Fine-Grained Pixel-Text Alignment for Open-Vocabulary Semantic Segmentation
di: Li, Bingyu, et al.
Pubblicazione: (2025)