3rd Place Solution to Large-scale Fine-grained Food Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhong, Yang, Yao, Yifan, Luo, Tong, Zhang, Youcai, Li, Yaqian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
3rd Place Solution to ICCV LargeFineFoodAI Retrieval
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
3rd Place Solution for VisDA 2021 Challenge -- Universally Domain Adaptive Image Recognition
von: Liao, Haojin, et al.
Veröffentlicht: (2021)
von: Liao, Haojin, et al.
Veröffentlicht: (2021)
SAMSON: 3rd Place Solution of LSVOS 2025 VOS Challenge
von: Xie, Yujie, et al.
Veröffentlicht: (2025)
von: Xie, Yujie, et al.
Veröffentlicht: (2025)
FVOS for MOSE Track of 4th PVUW Challenge: 3rd Place Solution
von: Wang, Mengjiao, et al.
Veröffentlicht: (2025)
von: Wang, Mengjiao, et al.
Veröffentlicht: (2025)
3rd Place Solution for PVUW Challenge 2024: Video Panoptic Segmentation
von: Wu, Ruipu, et al.
Veröffentlicht: (2024)
von: Wu, Ruipu, et al.
Veröffentlicht: (2024)
The Instance-centric Transformer for the RVOS Track of LSVOS Challenge: 3rd Place Solution
von: Cao, Bin, et al.
Veröffentlicht: (2024)
von: Cao, Bin, et al.
Veröffentlicht: (2024)
Democratizing Fine-grained Visual Recognition with Large Language Models
von: Liu, Mingxuan, et al.
Veröffentlicht: (2024)
von: Liu, Mingxuan, et al.
Veröffentlicht: (2024)
Towards Fine-grained Large Object Segmentation 1st Place Solution to 3D AI Challenge 2020 -- Instance Segmentation Track
von: Chen, Zehui, et al.
Veröffentlicht: (2020)
von: Chen, Zehui, et al.
Veröffentlicht: (2020)
3rd Place Solution for MOSE Track in CVPR 2024 PVUW workshop: Complex Video Object Segmentation
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
First Place Solution to the ECCV 2024 ROAD++ Challenge @ ROAD++ Spatiotemporal Agent Detection 2024
von: Zhang, Tengfei, et al.
Veröffentlicht: (2024)
von: Zhang, Tengfei, et al.
Veröffentlicht: (2024)
4th PVUW MeViS 3rd Place Report: Sa2VA
von: Yuan, Haobo, et al.
Veröffentlicht: (2025)
von: Yuan, Haobo, et al.
Veröffentlicht: (2025)
Storyboard guided Alignment for Fine-grained Video Action Recognition
von: Liu, Enqi, et al.
Veröffentlicht: (2024)
von: Liu, Enqi, et al.
Veröffentlicht: (2024)
Robust Saliency-Aware Distillation for Few-shot Fine-grained Visual Recognition
von: Liu, Haiqi, et al.
Veröffentlicht: (2023)
von: Liu, Haiqi, et al.
Veröffentlicht: (2023)
3rd Place Solution for MeViS Track in CVPR 2024 PVUW workshop: Motion Expression guided Video Segmentation
von: Pan, Feiyu, et al.
Veröffentlicht: (2024)
von: Pan, Feiyu, et al.
Veröffentlicht: (2024)
LSVOS Challenge 3rd Place Report: SAM2 and Cutie based VOS
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
Large-scale and Fine-grained Vision-language Pre-training for Enhanced CT Image Understanding
von: Shui, Zhongyi, et al.
Veröffentlicht: (2025)
von: Shui, Zhongyi, et al.
Veröffentlicht: (2025)
Enriched Feature Representation and Motion Prediction Module for MOSEv2 Track of 7th LSVOS Challenge: 3rd Place Solution
von: Lim, Chang Soo, et al.
Veröffentlicht: (2025)
von: Lim, Chang Soo, et al.
Veröffentlicht: (2025)
Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time
von: Cheng, Jintao, et al.
Veröffentlicht: (2025)
von: Cheng, Jintao, et al.
Veröffentlicht: (2025)
First Place Solution to the ECCV 2024 ROAD++ Challenge @ ROAD++ Atomic Activity Recognition 2024
von: Li, Ruyang, et al.
Veröffentlicht: (2024)
von: Li, Ruyang, et al.
Veröffentlicht: (2024)
VXP: Voxel-Cross-Pixel Large-scale Image-LiDAR Place Recognition
von: Li, Yun-Jin, et al.
Veröffentlicht: (2024)
von: Li, Yun-Jin, et al.
Veröffentlicht: (2024)
VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation
von: Zhang, Shi-Xue, et al.
Veröffentlicht: (2025)
von: Zhang, Shi-Xue, et al.
Veröffentlicht: (2025)
Fine-grained Image-to-LiDAR Contrastive Distillation with Visual Foundation Models
von: Zhang, Yifan, et al.
Veröffentlicht: (2024)
von: Zhang, Yifan, et al.
Veröffentlicht: (2024)
SelFLoc: Selective Feature Fusion for Large-scale Point Cloud-based Place Recognition
von: Qiu, Qibo, et al.
Veröffentlicht: (2023)
von: Qiu, Qibo, et al.
Veröffentlicht: (2023)
vFusedSeg3D: 3rd Place Solution for 2024 Waymo Open Dataset Challenge in Semantic Segmentation
von: Amjad, Osama, et al.
Veröffentlicht: (2024)
von: Amjad, Osama, et al.
Veröffentlicht: (2024)
OpenSatMap: A Fine-grained High-resolution Satellite Dataset for Large-scale Map Construction
von: Zhao, Hongbo, et al.
Veröffentlicht: (2024)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2024)
3rd Place of MeViS-Audio Track of the 5th PVUW: VIRST-Audio
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
A Unified Hierarchical Framework for Fine-grained Cross-view Geo-localization over Large-scale Scenarios
von: Song, Zhuo, et al.
Veröffentlicht: (2025)
von: Song, Zhuo, et al.
Veröffentlicht: (2025)
MinkUNeXt: Point Cloud-based Large-scale Place Recognition using 3D Sparse Convolutions
von: Cabrera, J. J., et al.
Veröffentlicht: (2024)
von: Cabrera, J. J., et al.
Veröffentlicht: (2024)
Vision Mamba Distillation for Low-resolution Fine-grained Image Classification
von: Chen, Yao, et al.
Veröffentlicht: (2024)
von: Chen, Yao, et al.
Veröffentlicht: (2024)
SARE: Sample-wise Adaptive Reasoning for Training-free Fine-grained Visual Recognition
von: Yang, Jingxiao, et al.
Veröffentlicht: (2026)
von: Yang, Jingxiao, et al.
Veröffentlicht: (2026)
Breaking the Frame: Visual Place Recognition by Overlap Prediction
von: Wei, Tong, et al.
Veröffentlicht: (2024)
von: Wei, Tong, et al.
Veröffentlicht: (2024)
Enhancing Fine-grained Object Detection in Aerial Images via Orthogonal Mapping
von: Zhu, Haoran, et al.
Veröffentlicht: (2024)
von: Zhu, Haoran, et al.
Veröffentlicht: (2024)
Trajectory Attention for Fine-grained Video Motion Control
von: Xiao, Zeqi, et al.
Veröffentlicht: (2024)
von: Xiao, Zeqi, et al.
Veröffentlicht: (2024)
EDTformer: An Efficient Decoder Transformer for Visual Place Recognition
von: Jin, Tong, et al.
Veröffentlicht: (2024)
von: Jin, Tong, et al.
Veröffentlicht: (2024)
Distillation Improves Visual Place Recognition for Low Quality Images
von: Yang, Anbang, et al.
Veröffentlicht: (2023)
von: Yang, Anbang, et al.
Veröffentlicht: (2023)
Language-driven Fine-grained Retrieval
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
Every Subtlety Counts: Fine-grained Person Independence Micro-Action Recognition via Distributionally Robust Optimization
von: Cui, Feng-Qi, et al.
Veröffentlicht: (2025)
von: Cui, Feng-Qi, et al.
Veröffentlicht: (2025)
Tag2Text: Guiding Vision-Language Model via Image Tagging
von: Huang, Xinyu, et al.
Veröffentlicht: (2023)
von: Huang, Xinyu, et al.
Veröffentlicht: (2023)
AffordBot: 3D Fine-grained Embodied Reasoning via Multimodal Large Language Models
von: Wang, Xinyi, et al.
Veröffentlicht: (2025)
von: Wang, Xinyi, et al.
Veröffentlicht: (2025)
SuperPlace: The Renaissance of Classical Feature Aggregation for Visual Place Recognition in the Era of Foundation Models
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
3rd Place Solution to ICCV LargeFineFoodAI Retrieval
von: Zhong, Yang, et al.
Veröffentlicht: (2025) -
3rd Place Solution for VisDA 2021 Challenge -- Universally Domain Adaptive Image Recognition
von: Liao, Haojin, et al.
Veröffentlicht: (2021) -
SAMSON: 3rd Place Solution of LSVOS 2025 VOS Challenge
von: Xie, Yujie, et al.
Veröffentlicht: (2025) -
FVOS for MOSE Track of 4th PVUW Challenge: 3rd Place Solution
von: Wang, Mengjiao, et al.
Veröffentlicht: (2025) -
3rd Place Solution for PVUW Challenge 2024: Video Panoptic Segmentation
von: Wu, Ruipu, et al.
Veröffentlicht: (2024)