Boltzmann Attention Sampling for Image Analysis with Small Objects
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhao, Theodore, Kiblawi, Sid, Usuyama, Naoto, Lee, Ho Hin, Preston, Sam, Poon, Hoifung, Wei, Mu |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Learning Sparse Visual Representations via Spatial-Semantic Factorization
par: Zhao, Theodore Zhengde, et autres
Publié: (2026)
par: Zhao, Theodore Zhengde, et autres
Publié: (2026)
Foundation Models for Biomedical Image Segmentation: A Survey
par: Lee, Ho Hin, et autres
Publié: (2024)
par: Lee, Ho Hin, et autres
Publié: (2024)
BiomedParse: a biomedical foundation model for image parsing of everything everywhere all at once
par: Zhao, Theodore, et autres
Publié: (2024)
par: Zhao, Theodore, et autres
Publié: (2024)
Scaling medical imaging report generation with multimodal reinforcement learning
par: Liu, Qianchu, et autres
Publié: (2026)
par: Liu, Qianchu, et autres
Publié: (2026)
BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs
par: Zhang, Sheng, et autres
Publié: (2023)
par: Zhang, Sheng, et autres
Publié: (2023)
Small Object Detection Model with Spatial Laplacian Pyramid Attention and Multi-Scale Features Enhancement in Aerial Images
par: Ji, Zhangjian, et autres
Publié: (2026)
par: Ji, Zhangjian, et autres
Publié: (2026)
Find your Needle: Small Object Image Retrieval via Multi-Object Attention Optimization
par: Green, Michael, et autres
Publié: (2025)
par: Green, Michael, et autres
Publié: (2025)
From Introspection to Best Practices: Principled Analysis of Demonstrations in Multimodal In-Context Learning
par: Xu, Nan, et autres
Publié: (2024)
par: Xu, Nan, et autres
Publié: (2024)
AURAD: Anatomy-Pathology Unified Radiology Synthesis with Progressive Representations
par: Ding, Shuhan, et autres
Publié: (2025)
par: Ding, Shuhan, et autres
Publié: (2025)
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection
par: Cheung, Tsun-Hin, et autres
Publié: (2024)
par: Cheung, Tsun-Hin, et autres
Publié: (2024)
Pareto Optimal Learning for Estimating Large Language Model Errors
par: Zhao, Theodore, et autres
Publié: (2023)
par: Zhao, Theodore, et autres
Publié: (2023)
LAM-YOLO: Drones-based Small Object Detection on Lighting-Occlusion Attention Mechanism YOLO
par: Zheng, Yuchen, et autres
Publié: (2024)
par: Zheng, Yuchen, et autres
Publié: (2024)
Video Models Can Reason with Verifiable Rewards
par: Zhu, Tinghui, et autres
Publié: (2026)
par: Zhu, Tinghui, et autres
Publié: (2026)
CountCluster: Training-Free Object Quantity Guidance with Cross-Attention Map Clustering for Text-to-Image Generation
par: Lee, Joohyeon, et autres
Publié: (2025)
par: Lee, Joohyeon, et autres
Publié: (2025)
AD-Det: Boosting Object Detection in UAV Images with Focused Small Objects and Balanced Tail Classes
par: Li, Zhenteng, et autres
Publié: (2025)
par: Li, Zhenteng, et autres
Publié: (2025)
Guided Slot Attention for Unsupervised Video Object Segmentation
par: Lee, Minhyeok, et autres
Publié: (2023)
par: Lee, Minhyeok, et autres
Publié: (2023)
SORCE: Small Object Retrieval in Complex Environments
par: Liu, Chunxu, et autres
Publié: (2025)
par: Liu, Chunxu, et autres
Publié: (2025)
Cross-Layer Feature Pyramid Transformer for Small Object Detection in Aerial Images
par: Du, Zewen, et autres
Publié: (2024)
par: Du, Zewen, et autres
Publié: (2024)
Dual Prototype Attention for Unsupervised Video Object Segmentation
par: Cho, Suhwan, et autres
Publié: (2022)
par: Cho, Suhwan, et autres
Publié: (2022)
Better Sampling, towards Better End-to-end Small Object Detection
par: Huang, Zile, et autres
Publié: (2024)
par: Huang, Zile, et autres
Publié: (2024)
Leveraging Image Augmentation for Object Manipulation: Towards Interpretable Controllability in Object-Centric Learning
par: Kim, Jinwoo, et autres
Publié: (2023)
par: Kim, Jinwoo, et autres
Publié: (2023)
MAFE R-CNN: Selecting More Samples to Learn Category-aware Features for Small Object Detection
par: Li, Yichen, et autres
Publié: (2025)
par: Li, Yichen, et autres
Publié: (2025)
Efficient Oriented Object Detection with Enhanced Small Object Recognition in Aerial Images
par: Shi, Zhifei, et autres
Publié: (2024)
par: Shi, Zhifei, et autres
Publié: (2024)
Exploiting Scale-Variant Attention for Segmenting Small Medical Objects
par: Dai, Wei, et autres
Publié: (2024)
par: Dai, Wei, et autres
Publié: (2024)
Task-wise Sampling Convolutions for Arbitrary-Oriented Object Detection in Aerial Images
par: Huang, Zhanchao, et autres
Publié: (2022)
par: Huang, Zhanchao, et autres
Publié: (2022)
Small Object Detection in Complex Backgrounds with Multi-Scale Attention and Global Relation Modeling
par: Tao, Wenguang, et autres
Publié: (2026)
par: Tao, Wenguang, et autres
Publié: (2026)
CountSteer: Steering Attention for Object Counting in Diffusion Models
par: Boo, Hyemin, et autres
Publié: (2025)
par: Boo, Hyemin, et autres
Publié: (2025)
ESOD: Efficient Small Object Detection on High-Resolution Images
par: Liu, Kai, et autres
Publié: (2024)
par: Liu, Kai, et autres
Publié: (2024)
Multi-Modal Mamba Modeling for Survival Prediction (M4Survive): Adapting Joint Foundation Model Representations
par: Lee, Ho Hin, et autres
Publié: (2025)
par: Lee, Ho Hin, et autres
Publié: (2025)
MKSNet: Advanced Small Object Detection in Remote Sensing Imagery with Multi-Kernel and Dual Attention Mechanisms
par: Zhang, Jiahao, et autres
Publié: (2025)
par: Zhang, Jiahao, et autres
Publié: (2025)
Vector Scaffolding: Inter-Scale Orchestration for Differentiable Image Vectorization
par: Lee, Jaerin, et autres
Publié: (2026)
par: Lee, Jaerin, et autres
Publié: (2026)
SOEDiff: Efficient Distillation for Small Object Editing
par: Wu, Yiming, et autres
Publié: (2024)
par: Wu, Yiming, et autres
Publié: (2024)
ViCrop-Det: Spatial Attention Entropy Guided Cropping for Training-Free Small-Object Detection
par: Wang, Hui, et autres
Publié: (2026)
par: Wang, Hui, et autres
Publié: (2026)
Particle Diffusion Matching: Random Walk Correspondence Search for the Alignment of Standard and Ultra-Widefield Fundus Images
par: Lee, Kanggeon, et autres
Publié: (2026)
par: Lee, Kanggeon, et autres
Publié: (2026)
SMILEtrack: SiMIlarity LEarning for Occlusion-Aware Multiple Object Tracking
par: Wang, Yu-Hsiang, et autres
Publié: (2022)
par: Wang, Yu-Hsiang, et autres
Publié: (2022)
Controllable 3D Object Generation with Single Image Prompt
par: Lee, Jaeseok, et autres
Publié: (2025)
par: Lee, Jaeseok, et autres
Publié: (2025)
Putting the Object Back into Video Object Segmentation
par: Cheng, Ho Kei, et autres
Publié: (2023)
par: Cheng, Ho Kei, et autres
Publié: (2023)
SAMF: Small-Area-Aware Multi-focus Image Fusion for Object Detection
par: Li, Xilai, et autres
Publié: (2024)
par: Li, Xilai, et autres
Publié: (2024)
Beyond Image Super-Resolution for Image Recognition with Task-Driven Perceptual Loss
par: Kim, Jaeha, et autres
Publié: (2024)
par: Kim, Jaeha, et autres
Publié: (2024)
FIAS: Feature Imbalance-Aware Medical Image Segmentation with Dynamic Fusion and Mixing Attention
par: Liu, Xiwei, et autres
Publié: (2024)
par: Liu, Xiwei, et autres
Publié: (2024)
Documents similaires
-
Learning Sparse Visual Representations via Spatial-Semantic Factorization
par: Zhao, Theodore Zhengde, et autres
Publié: (2026) -
Foundation Models for Biomedical Image Segmentation: A Survey
par: Lee, Ho Hin, et autres
Publié: (2024) -
BiomedParse: a biomedical foundation model for image parsing of everything everywhere all at once
par: Zhao, Theodore, et autres
Publié: (2024) -
Scaling medical imaging report generation with multimodal reinforcement learning
par: Liu, Qianchu, et autres
Publié: (2026) -
BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs
par: Zhang, Sheng, et autres
Publié: (2023)