BoQ: A Place is Worth a Bag of Learnable Queries
Fuente:
arXiv
Guardado en:
| Autores principales: | Ali-Bey, Amar, Chaib-draa, Brahim, Giguère, Philippe |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Image Aesthetics Assessment via Learnable Queries
por: Xiong, Zhiwei, et al.
Publicado: (2023)
por: Xiong, Zhiwei, et al.
Publicado: (2023)
HBRB-BoW: A Retrained Bag-of-Words Vocabulary for ORB-SLAM via Hierarchical BRB-KMeans
por: Lee, Minjae, et al.
Publicado: (2026)
por: Lee, Minjae, et al.
Publicado: (2026)
Multi-modal Learnable Queries for Image Aesthetics Assessment
por: Xiong, Zhiwei, et al.
Publicado: (2024)
por: Xiong, Zhiwei, et al.
Publicado: (2024)
Bag-of-Word-Groups (BoWG): A Robust and Efficient Loop Closure Detection Method Under Perceptual Aliasing
por: Fei, Xiang, et al.
Publicado: (2025)
por: Fei, Xiang, et al.
Publicado: (2025)
Micro-gesture Online Recognition using Learnable Query Points
por: Liu, Pengyu, et al.
Publicado: (2024)
por: Liu, Pengyu, et al.
Publicado: (2024)
Learnable Query Aggregation with KV Routing for Cross-view Geo-localisation
por: Ye, Hualin, et al.
Publicado: (2025)
por: Ye, Hualin, et al.
Publicado: (2025)
UniEmo: Unifying Emotional Understanding and Generation with Learnable Expert Queries
por: Zhu, Yijie, et al.
Publicado: (2025)
por: Zhu, Yijie, et al.
Publicado: (2025)
Sparse Semi-DETR: Sparse Learnable Queries for Semi-Supervised Object Detection
por: Shehzadi, Tahira, et al.
Publicado: (2024)
por: Shehzadi, Tahira, et al.
Publicado: (2024)
Bag of Bags: Adaptive Visual Vocabularies for Genizah Join Image Retrieval
por: Gogawale, Sharva, et al.
Publicado: (2026)
por: Gogawale, Sharva, et al.
Publicado: (2026)
VOLoc: Visual Place Recognition by Querying Compressed Lidar Map
por: Cai, Xudong, et al.
Publicado: (2024)
por: Cai, Xudong, et al.
Publicado: (2024)
CQVPR: Landmark-aware Contextual Queries for Visual Place Recognition
por: Li, Dongyue, et al.
Publicado: (2025)
por: Li, Dongyue, et al.
Publicado: (2025)
A Video Is Not Worth a Thousand Words
por: Pollard, Sam, et al.
Publicado: (2025)
por: Pollard, Sam, et al.
Publicado: (2025)
MaskBEV: Joint Object Detection and Footprint Completion for Bird's-eye View 3D Point Clouds
por: Guimont-Martin, William, et al.
Publicado: (2023)
por: Guimont-Martin, William, et al.
Publicado: (2023)
LQ-Adapter: ViT-Adapter with Learnable Queries for Gallbladder Cancer Detection from Ultrasound Image
por: Madan, Chetan, et al.
Publicado: (2024)
por: Madan, Chetan, et al.
Publicado: (2024)
A LoRA is Worth a Thousand Pictures
por: Liu, Chenxi, et al.
Publicado: (2024)
por: Liu, Chenxi, et al.
Publicado: (2024)
A Bag of Tricks for Few-Shot Class-Incremental Learning
por: Roy, Shuvendu, et al.
Publicado: (2024)
por: Roy, Shuvendu, et al.
Publicado: (2024)
FLIM Networks with Bag of Feature Points
por: Martinelli, João Deltregia, et al.
Publicado: (2026)
por: Martinelli, João Deltregia, et al.
Publicado: (2026)
Towards Test-time Efficient Visual Place Recognition via Asymmetric Query Processing
por: Kim, Jaeyoon, et al.
Publicado: (2025)
por: Kim, Jaeyoon, et al.
Publicado: (2025)
Vript: A Video Is Worth Thousands of Words
por: Yang, Dongjie, et al.
Publicado: (2024)
por: Yang, Dongjie, et al.
Publicado: (2024)
DC-VLAQ: Query-Residual Aggregation for Robust Visual Place Recognition
por: Zhu, Hanyu, et al.
Publicado: (2026)
por: Zhu, Hanyu, et al.
Publicado: (2026)
Q-Align: Alleviating Attention Leakage in Zero-Shot Appearance Transfer via Query-Query Alignment
por: Kim, Namu, et al.
Publicado: (2025)
por: Kim, Namu, et al.
Publicado: (2025)
A Creative Agent is Worth a 64-Token Template
por: Shi, Ruixiao, et al.
Publicado: (2026)
por: Shi, Ruixiao, et al.
Publicado: (2026)
Images are Worth Variable Length of Representations
por: Mao, Lingjun, et al.
Publicado: (2025)
por: Mao, Lingjun, et al.
Publicado: (2025)
BREEN: Bridge Data-Efficient Encoder-Free Multimodal Learning with Learnable Queries
por: Li, Tianle, et al.
Publicado: (2025)
por: Li, Tianle, et al.
Publicado: (2025)
Pushing the Limits of Sparsity: A Bag of Tricks for Extreme Pruning
por: Li, Andy, et al.
Publicado: (2024)
por: Li, Andy, et al.
Publicado: (2024)
MaxQ: Multi-Axis Query for N:M Sparsity Network
por: Xiang, Jingyang, et al.
Publicado: (2023)
por: Xiang, Jingyang, et al.
Publicado: (2023)
Learnable Quantum Efficiency Filters for Urban Hyperspectral Segmentation
por: Shah, Imad Ali, et al.
Publicado: (2026)
por: Shah, Imad Ali, et al.
Publicado: (2026)
An Image is Worth 32 Tokens for Reconstruction and Generation
por: Yu, Qihang, et al.
Publicado: (2024)
por: Yu, Qihang, et al.
Publicado: (2024)
Reproducible Evaluation of Camera Auto-Exposure Methods in the Field: Platform, Benchmark and Lessons Learned
por: Gamache, Olivier, et al.
Publicado: (2025)
por: Gamache, Olivier, et al.
Publicado: (2025)
Optimizing Multispectral Object Detection: A Bag of Tricks and Comprehensive Benchmarks
por: Zhou, Chen, et al.
Publicado: (2024)
por: Zhou, Chen, et al.
Publicado: (2024)
Trick-GS: A Balanced Bag of Tricks for Efficient Gaussian Splatting
por: Armagan, Anil, et al.
Publicado: (2025)
por: Armagan, Anil, et al.
Publicado: (2025)
DecQ: Detail-Condensing Queries for Enhanced Reconstruction and Generation in Representation Autoencoders
por: Wang, Tianhang, et al.
Publicado: (2026)
por: Wang, Tianhang, et al.
Publicado: (2026)
What is Point Supervision Worth in Video Instance Segmentation?
por: Huang, Shuaiyi, et al.
Publicado: (2024)
por: Huang, Shuaiyi, et al.
Publicado: (2024)
Good Enough: Is it Worth Improving your Label Quality?
por: Jaus, Alexander, et al.
Publicado: (2025)
por: Jaus, Alexander, et al.
Publicado: (2025)
An Item is Worth a Prompt: Versatile Image Editing with Disentangled Control
por: Feng, Aosong, et al.
Publicado: (2024)
por: Feng, Aosong, et al.
Publicado: (2024)
MAG-VLAQ: Multi-modal Aerial-Ground Query Aggregation for Cross-View Place Recognition
por: Xu, Zhengyi, et al.
Publicado: (2026)
por: Xu, Zhengyi, et al.
Publicado: (2026)
PaQ-DETR: Learning Pattern and Quality-Aware Dynamic Queries for Object Detection
por: Kang, Zhengjian, et al.
Publicado: (2026)
por: Kang, Zhengjian, et al.
Publicado: (2026)
Q-Adapter: Visual Query Adapter for Extracting Textually-related Features in Video Captioning
por: Chen, Junan, et al.
Publicado: (2025)
por: Chen, Junan, et al.
Publicado: (2025)
Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs
por: Zhang, Shaojie, et al.
Publicado: (2025)
por: Zhang, Shaojie, et al.
Publicado: (2025)
A Dataset is Worth 1 MB
por: Shoshani, Elad Kimchi, et al.
Publicado: (2026)
por: Shoshani, Elad Kimchi, et al.
Publicado: (2026)
Ejemplares similares
-
Image Aesthetics Assessment via Learnable Queries
por: Xiong, Zhiwei, et al.
Publicado: (2023) -
HBRB-BoW: A Retrained Bag-of-Words Vocabulary for ORB-SLAM via Hierarchical BRB-KMeans
por: Lee, Minjae, et al.
Publicado: (2026) -
Multi-modal Learnable Queries for Image Aesthetics Assessment
por: Xiong, Zhiwei, et al.
Publicado: (2024) -
Bag-of-Word-Groups (BoWG): A Robust and Efficient Loop Closure Detection Method Under Perceptual Aliasing
por: Fei, Xiang, et al.
Publicado: (2025) -
Micro-gesture Online Recognition using Learnable Query Points
por: Liu, Pengyu, et al.
Publicado: (2024)