DiVE-k: Differential Visual Reasoning for Fine-grained Image Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Kumar, Raja, Sadhu, Arka, Nevatia, Ram |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DiVE: DiT-based Video Generation with Enhanced Control
di: Jiang, Junpeng, et al.
Pubblicazione: (2024)
di: Jiang, Junpeng, et al.
Pubblicazione: (2024)
DiVE: Efficient Multi-View Driving Scenes Generation Based on Video Diffusion Transformer
di: Jiang, Junpeng, et al.
Pubblicazione: (2025)
di: Jiang, Junpeng, et al.
Pubblicazione: (2025)
GaitSTR: Gait Recognition with Sequential Two-stream Refinement
di: Zheng, Wanrong, et al.
Pubblicazione: (2024)
di: Zheng, Wanrong, et al.
Pubblicazione: (2024)
Large Language Models are Good Prompt Learners for Low-Shot Image Classification
di: Zheng, Zhaoheng, et al.
Pubblicazione: (2023)
di: Zheng, Zhaoheng, et al.
Pubblicazione: (2023)
FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation
di: Fang, Xueji, et al.
Pubblicazione: (2026)
di: Fang, Xueji, et al.
Pubblicazione: (2026)
Democratizing Fine-grained Visual Recognition with Large Language Models
di: Liu, Mingxuan, et al.
Pubblicazione: (2024)
di: Liu, Mingxuan, et al.
Pubblicazione: (2024)
SARE: Sample-wise Adaptive Reasoning for Training-free Fine-grained Visual Recognition
di: Yang, Jingxiao, et al.
Pubblicazione: (2026)
di: Yang, Jingxiao, et al.
Pubblicazione: (2026)
CaesarNeRF: Calibrated Semantic Representation for Few-shot Generalizable Neural Rendering
di: Zhu, Haidong, et al.
Pubblicazione: (2023)
di: Zhu, Haidong, et al.
Pubblicazione: (2023)
Robust Saliency-Aware Distillation for Few-shot Fine-grained Visual Recognition
di: Liu, Haiqi, et al.
Pubblicazione: (2023)
di: Liu, Haiqi, et al.
Pubblicazione: (2023)
Open Eyes, Then Reason: Fine-grained Visual Mathematical Understanding in MLLMs
di: Zhang, Shan, et al.
Pubblicazione: (2025)
di: Zhang, Shan, et al.
Pubblicazione: (2025)
DiLightNet: Fine-grained Lighting Control for Diffusion-based Image Generation
di: Zeng, Chong, et al.
Pubblicazione: (2024)
di: Zeng, Chong, et al.
Pubblicazione: (2024)
FiVE: A Fine-grained Video Editing Benchmark for Evaluating Emerging Diffusion and Rectified Flow Models
di: Li, Minghan, et al.
Pubblicazione: (2025)
di: Li, Minghan, et al.
Pubblicazione: (2025)
Towards Generative Class Prompt Learning for Fine-grained Visual Recognition
di: Chattopadhyay, Soumitri, et al.
Pubblicazione: (2024)
di: Chattopadhyay, Soumitri, et al.
Pubblicazione: (2024)
BaFTA: Backprop-Free Test-Time Adaptation For Zero-Shot Vision-Language Models
di: Hu, Xuefeng, et al.
Pubblicazione: (2024)
di: Hu, Xuefeng, et al.
Pubblicazione: (2024)
Can Textual Reasoning Improve the Performance of MLLMs on Fine-grained Visual Classification?
di: Zhu, Jie, et al.
Pubblicazione: (2026)
di: Zhu, Jie, et al.
Pubblicazione: (2026)
AttriStory: Fine-grained Attribute Realization for Visual Storytelling with Diffusion Models
di: Sreenivas, Manogna, et al.
Pubblicazione: (2026)
di: Sreenivas, Manogna, et al.
Pubblicazione: (2026)
Fine-grained Image-to-LiDAR Contrastive Distillation with Visual Foundation Models
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
Vocabulary-free Fine-grained Visual Recognition via Enriched Contextually Grounded Vision-Language Model
di: Demidov, Dmitry, et al.
Pubblicazione: (2025)
di: Demidov, Dmitry, et al.
Pubblicazione: (2025)
FiVA: Fine-grained Visual Attribute Dataset for Text-to-Image Diffusion Models
di: Wu, Tong, et al.
Pubblicazione: (2024)
di: Wu, Tong, et al.
Pubblicazione: (2024)
Fine-grained Text to Image Synthesis
di: Ouyang, Xu, et al.
Pubblicazione: (2024)
di: Ouyang, Xu, et al.
Pubblicazione: (2024)
Storyboard guided Alignment for Fine-grained Video Action Recognition
di: Liu, Enqi, et al.
Pubblicazione: (2024)
di: Liu, Enqi, et al.
Pubblicazione: (2024)
FineRS: Fine-grained Reasoning and Segmentation of Small Objects with Reinforcement Learning
di: Zhang, Lu, et al.
Pubblicazione: (2025)
di: Zhang, Lu, et al.
Pubblicazione: (2025)
FIRE-CIR: Fine-grained Reasoning for Composed Fashion Image Retrieval
di: Gardères, François, et al.
Pubblicazione: (2026)
di: Gardères, François, et al.
Pubblicazione: (2026)
Data-free Knowledge Distillation for Fine-grained Visual Categorization
di: Shao, Renrong, et al.
Pubblicazione: (2024)
di: Shao, Renrong, et al.
Pubblicazione: (2024)
DreamVE: Unified Instruction-based Image and Video Editing
di: Xia, Bin, et al.
Pubblicazione: (2025)
di: Xia, Bin, et al.
Pubblicazione: (2025)
3rd Place Solution to Large-scale Fine-grained Food Recognition
di: Zhong, Yang, et al.
Pubblicazione: (2025)
di: Zhong, Yang, et al.
Pubblicazione: (2025)
FOCUS: Fine-grained Optimization with Semantic Guided Understanding for Pedestrian Attributes Recognition
di: An, Hongyan, et al.
Pubblicazione: (2025)
di: An, Hongyan, et al.
Pubblicazione: (2025)
HiVG: Hierarchical Multimodal Fine-grained Modulation for Visual Grounding
di: Xiao, Linhui, et al.
Pubblicazione: (2024)
di: Xiao, Linhui, et al.
Pubblicazione: (2024)
Multi-modal Instruction Tuned LLMs with Fine-grained Visual Perception
di: He, Junwen, et al.
Pubblicazione: (2024)
di: He, Junwen, et al.
Pubblicazione: (2024)
Towards an Effective Action-Region Tracking Framework for Fine-grained Video Action Recognition
di: Sun, Baoli, et al.
Pubblicazione: (2025)
di: Sun, Baoli, et al.
Pubblicazione: (2025)
Rethinking Rotation-Invariant Recognition of Fine-grained Shapes from the Perspective of Contour Points
di: Xu, Yanjie, et al.
Pubblicazione: (2025)
di: Xu, Yanjie, et al.
Pubblicazione: (2025)
Fine-grained Defocus Blur Control for Generative Image Models
di: Shrivastava, Ayush, et al.
Pubblicazione: (2025)
di: Shrivastava, Ayush, et al.
Pubblicazione: (2025)
Unveiling Chain of Step Reasoning for Vision-Language Models with Fine-grained Rewards
di: Chen, Honghao, et al.
Pubblicazione: (2025)
di: Chen, Honghao, et al.
Pubblicazione: (2025)
BEAR: A Video Dataset For Fine-grained Behaviors Recognition Oriented with Action and Environment Factors
di: Hu, Chengyang, et al.
Pubblicazione: (2025)
di: Hu, Chengyang, et al.
Pubblicazione: (2025)
Body-Hand Modality Expertized Networks with Cross-attention for Fine-grained Skeleton Action Recognition
di: Cho, Seungyeon, et al.
Pubblicazione: (2025)
di: Cho, Seungyeon, et al.
Pubblicazione: (2025)
DAPlankton: Benchmark Dataset for Multi-instrument Plankton Recognition via Fine-grained Domain Adaptation
di: Batrakhanov, Daniel, et al.
Pubblicazione: (2024)
di: Batrakhanov, Daniel, et al.
Pubblicazione: (2024)
Multi-modal Reference Learning for Fine-grained Text-to-Image Retrieval
di: Ma, Zehong, et al.
Pubblicazione: (2025)
di: Ma, Zehong, et al.
Pubblicazione: (2025)
Vision Mamba Distillation for Low-resolution Fine-grained Image Classification
di: Chen, Yao, et al.
Pubblicazione: (2024)
di: Chen, Yao, et al.
Pubblicazione: (2024)
FITA: Fine-grained Image-Text Aligner for Radiology Report Generation
di: Yang, Honglong, et al.
Pubblicazione: (2024)
di: Yang, Honglong, et al.
Pubblicazione: (2024)
Enhancing Fine-grained Image Classification through Attentive Batch Training
di: Le, Duy M., et al.
Pubblicazione: (2024)
di: Le, Duy M., et al.
Pubblicazione: (2024)
Documenti analoghi
-
DiVE: DiT-based Video Generation with Enhanced Control
di: Jiang, Junpeng, et al.
Pubblicazione: (2024) -
DiVE: Efficient Multi-View Driving Scenes Generation Based on Video Diffusion Transformer
di: Jiang, Junpeng, et al.
Pubblicazione: (2025) -
GaitSTR: Gait Recognition with Sequential Two-stream Refinement
di: Zheng, Wanrong, et al.
Pubblicazione: (2024) -
Large Language Models are Good Prompt Learners for Low-Shot Image Classification
di: Zheng, Zhaoheng, et al.
Pubblicazione: (2023) -
FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation
di: Fang, Xueji, et al.
Pubblicazione: (2026)