Bootstrapping MLLM for Weakly-Supervised Class-Agnostic Object Counting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Xiaowen, Yue, Zijie, Luo, Yong, Zhao, Cairong, Chen, Qijun, Shi, Miaojing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Text-promptable Object Counting via Quantity Awareness Enhancement
von: Shi, Miaojing, et al.
Veröffentlicht: (2025)
von: Shi, Miaojing, et al.
Veröffentlicht: (2025)
Weakly-Supervised Referring Video Object Segmentation through Text Supervision
von: Shi, Miaojing, et al.
Veröffentlicht: (2026)
von: Shi, Miaojing, et al.
Veröffentlicht: (2026)
LoSh: Long-Short Text Joint Prediction Network for Referring Video Object Segmentation
von: Yuan, Linfeng, et al.
Veröffentlicht: (2023)
von: Yuan, Linfeng, et al.
Veröffentlicht: (2023)
Bootstrapping Vision-language Models for Self-supervised Remote Physiological Measurement
von: Yue, Zijie, et al.
Veröffentlicht: (2024)
von: Yue, Zijie, et al.
Veröffentlicht: (2024)
Enhancing Space-time Video Super-resolution via Spatial-temporal Feature Interaction
von: Yue, Zijie, et al.
Veröffentlicht: (2022)
von: Yue, Zijie, et al.
Veröffentlicht: (2022)
Weakly and Self-Supervised Class-Agnostic Motion Prediction for Autonomous Driving
von: Li, Ruibo, et al.
Veröffentlicht: (2025)
von: Li, Ruibo, et al.
Veröffentlicht: (2025)
FocalCount: Towards Class-Count Imbalance in Class-Agnostic Counting
von: Zhu, Huilin, et al.
Veröffentlicht: (2025)
von: Zhu, Huilin, et al.
Veröffentlicht: (2025)
OCCAM: Class-Agnostic, Training-Free, Prior-Free and Multi-Class Object Counting
von: Spanakis, Michail, et al.
Veröffentlicht: (2026)
von: Spanakis, Michail, et al.
Veröffentlicht: (2026)
Partial Weakly-Supervised Oriented Object Detection
von: Liu, Mingxin, et al.
Veröffentlicht: (2025)
von: Liu, Mingxin, et al.
Veröffentlicht: (2025)
A Simple-but-effective Baseline for Training-free Class-Agnostic Counting
von: Lin, Yuhao, et al.
Veröffentlicht: (2024)
von: Lin, Yuhao, et al.
Veröffentlicht: (2024)
Boosting Object Detection with Zero-Shot Day-Night Domain Adaptation
von: Du, Zhipeng, et al.
Veröffentlicht: (2023)
von: Du, Zhipeng, et al.
Veröffentlicht: (2023)
Large Model driven Radiology Report Generation with Clinical Quality Reinforcement Learning
von: Zhou, Zijian, et al.
Veröffentlicht: (2024)
von: Zhou, Zijian, et al.
Veröffentlicht: (2024)
DiPEx: Dispersing Prompt Expansion for Class-Agnostic Object Detection
von: Lim, Jia Syuen, et al.
Veröffentlicht: (2024)
von: Lim, Jia Syuen, et al.
Veröffentlicht: (2024)
CountFormer: A Transformer Framework for Learning Visual Repetition and Structure in Class-Agnostic Object Counting
von: Hossain, Md Tanvir, et al.
Veröffentlicht: (2025)
von: Hossain, Md Tanvir, et al.
Veröffentlicht: (2025)
TRAIL: Transferable Robust Adversarial Images via Latent diffusion
von: Xue, Yuhao, et al.
Veröffentlicht: (2025)
von: Xue, Yuhao, et al.
Veröffentlicht: (2025)
UNICBench: UNIfied Counting Benchmark for MLLM
von: Rong, Chenggang, et al.
Veröffentlicht: (2026)
von: Rong, Chenggang, et al.
Veröffentlicht: (2026)
VideoJudge: Bootstrapping Enables Scalable Supervision of MLLM-as-a-Judge for Video Understanding
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
Does it Really Count? Assessing Semantic Grounding in Text-Guided Class-Agnostic Counting
von: Pacini, Giacomo, et al.
Veröffentlicht: (2026)
von: Pacini, Giacomo, et al.
Veröffentlicht: (2026)
A Recipe for CAC: Mosaic-based Generalized Loss for Improved Class-Agnostic Counting
von: Chou, Tsung-Han, et al.
Veröffentlicht: (2024)
von: Chou, Tsung-Han, et al.
Veröffentlicht: (2024)
SQLNet: Scale-Modulated Query and Localization Network for Few-Shot Class-Agnostic Counting
von: Wu, Hefeng, et al.
Veröffentlicht: (2023)
von: Wu, Hefeng, et al.
Veröffentlicht: (2023)
CountingDINO: A Training-free Pipeline for Class-Agnostic Counting using Unsupervised Backbones
von: Pacini, Giacomo, et al.
Veröffentlicht: (2025)
von: Pacini, Giacomo, et al.
Veröffentlicht: (2025)
Entropy Bootstrapping for Weakly Supervised Nuclei Detection
von: Willoughby, James, et al.
Veröffentlicht: (2024)
von: Willoughby, James, et al.
Veröffentlicht: (2024)
Memory-guided Network with Uncertainty-based Feature Augmentation for Few-shot Semantic Segmentation
von: Chen, Xinyue, et al.
Veröffentlicht: (2024)
von: Chen, Xinyue, et al.
Veröffentlicht: (2024)
Advancing Weakly-Supervised Change Detection in Satellite Images via Adversarial Class Prompting
von: Zhao, Zhenghui, et al.
Veröffentlicht: (2025)
von: Zhao, Zhenghui, et al.
Veröffentlicht: (2025)
Weakly Supervised Object Detection for Automatic Tooth-marked Tongue Recognition
von: Zhang, Yongcun, et al.
Veröffentlicht: (2024)
von: Zhang, Yongcun, et al.
Veröffentlicht: (2024)
Mind the Prompt: A Novel Benchmark for Prompt-based Class-Agnostic Counting
von: Ciampi, Luca, et al.
Veröffentlicht: (2024)
von: Ciampi, Luca, et al.
Veröffentlicht: (2024)
SurgPIS: Surgical-instrument-level Instances and Part-level Semantics for Weakly-supervised Part-aware Instance Segmentation
von: Wei, Meng, et al.
Veröffentlicht: (2025)
von: Wei, Meng, et al.
Veröffentlicht: (2025)
Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervision
von: Kuckreja, Kartik, et al.
Veröffentlicht: (2026)
von: Kuckreja, Kartik, et al.
Veröffentlicht: (2026)
General Geometry-aware Weakly Supervised 3D Object Detection
von: Zhang, Guowen, et al.
Veröffentlicht: (2024)
von: Zhang, Guowen, et al.
Veröffentlicht: (2024)
Contrastive Prompt Clustering for Weakly Supervised Semantic Segmentation
von: Wu, Wangyu, et al.
Veröffentlicht: (2025)
von: Wu, Wangyu, et al.
Veröffentlicht: (2025)
Weakly Supervised 3D Object Detection with Multi-Stage Generalization
von: He, Jiawei, et al.
Veröffentlicht: (2023)
von: He, Jiawei, et al.
Veröffentlicht: (2023)
DExTeR: Weakly Semi-Supervised Object Detection with Class and Instance Experts for Medical Imaging
von: Meyer, Adrien, et al.
Veröffentlicht: (2026)
von: Meyer, Adrien, et al.
Veröffentlicht: (2026)
Region-CAM: Towards Accurate Object Regions in Class Activation Maps for Weakly Supervised Learning Tasks
von: Cai, Qingdong, et al.
Veröffentlicht: (2025)
von: Cai, Qingdong, et al.
Veröffentlicht: (2025)
MultiCounter: Multiple Action Agnostic Repetition Counting in Untrimmed Videos
von: Tang, Yin, et al.
Veröffentlicht: (2024)
von: Tang, Yin, et al.
Veröffentlicht: (2024)
When Multi-Task Learning Meets Partial Supervision: A Computer Vision Review
von: Fontana, Maxime, et al.
Veröffentlicht: (2023)
von: Fontana, Maxime, et al.
Veröffentlicht: (2023)
AdaTreeFormer: Few Shot Domain Adaptation for Tree Counting from a Single High-Resolution Image
von: Amirkolaee, Hamed Amini, et al.
Veröffentlicht: (2024)
von: Amirkolaee, Hamed Amini, et al.
Veröffentlicht: (2024)
Dual-Thresholded Heatmap-Guided Proposal Clustering and Negative Certainty Supervision with Enhanced Base Network for Weakly Supervised Object Detection
von: Guo, Yuelin, et al.
Veröffentlicht: (2025)
von: Guo, Yuelin, et al.
Veröffentlicht: (2025)
ResAD++: Towards Class Agnostic Anomaly Detection via Residual Feature Learning
von: Yao, Xincheng, et al.
Veröffentlicht: (2025)
von: Yao, Xincheng, et al.
Veröffentlicht: (2025)
CamoSAM2: Motion-Appearance Induced Auto-Refining Prompts for Video Camouflaged Object Detection
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
FAAR: Efficient Frequency-Aware Multi-Task Fine-Tuning via Automatic Rank Selection
von: Fontana, Maxime, et al.
Veröffentlicht: (2026)
von: Fontana, Maxime, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Text-promptable Object Counting via Quantity Awareness Enhancement
von: Shi, Miaojing, et al.
Veröffentlicht: (2025) -
Weakly-Supervised Referring Video Object Segmentation through Text Supervision
von: Shi, Miaojing, et al.
Veröffentlicht: (2026) -
LoSh: Long-Short Text Joint Prediction Network for Referring Video Object Segmentation
von: Yuan, Linfeng, et al.
Veröffentlicht: (2023) -
Bootstrapping Vision-language Models for Self-supervised Remote Physiological Measurement
von: Yue, Zijie, et al.
Veröffentlicht: (2024) -
Enhancing Space-time Video Super-resolution via Spatial-temporal Feature Interaction
von: Yue, Zijie, et al.
Veröffentlicht: (2022)