Saved in:
| Main Authors: | Perez, Gustavo, Sheldon, Daniel, Van Horn, Grant, Maji, Subhransu |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2312.05287 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Merlin L48 Spectrogram Dataset
by: Sun, Aaron, et al.
Published: (2025)
by: Sun, Aaron, et al.
Published: (2025)
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
by: Saha, Oindrila, et al.
Published: (2024)
by: Saha, Oindrila, et al.
Published: (2024)
Consensus-Driven Active Model Selection
by: Kay, Justin, et al.
Published: (2025)
by: Kay, Justin, et al.
Published: (2025)
Active Measurement of Two-Point Correlations
by: Hamilton, Max, et al.
Published: (2026)
by: Hamilton, Max, et al.
Published: (2026)
Generate, Transduct, Adapt: Iterative Transduction with VLMs
by: Saha, Oindrila, et al.
Published: (2025)
by: Saha, Oindrila, et al.
Published: (2025)
Masked Autoencoders with Limited Data: Does It Work? A Fine-Grained Bioacoustics Case Study
by: Liu, Wuao, et al.
Published: (2026)
by: Liu, Wuao, et al.
Published: (2026)
You May Speak Freely: Improving the Fine-Grained Visual Recognition Capabilities of Multimodal Large Language Models with Answer Extraction
by: Lawrence, Logan, et al.
Published: (2025)
by: Lawrence, Logan, et al.
Published: (2025)
Active Measurement: Efficient Estimation at Scale
by: Hamilton, Max, et al.
Published: (2025)
by: Hamilton, Max, et al.
Published: (2025)
WildSAT: Learning Satellite Image Representations from Wildlife Observations
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
Moment Sampling in Video LLMs for Long-Form Video QA
by: Chasmai, Mustafa, et al.
Published: (2025)
by: Chasmai, Mustafa, et al.
Published: (2025)
RealBirdID: Benchmarking Bird Species Identification in the Era of MLLMs
by: Lawrence, Logan, et al.
Published: (2026)
by: Lawrence, Logan, et al.
Published: (2026)
Feedforward Few-shot Species Range Estimation
by: Lange, Christian, et al.
Published: (2025)
by: Lange, Christian, et al.
Published: (2025)
Not All Birds Look The Same: Identity-Preserving Generation For Birds
by: Sun, Aaron, et al.
Published: (2025)
by: Sun, Aaron, et al.
Published: (2025)
Task2Box: Box Embeddings for Modeling Asymmetric Task Relationships
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
3D Space as a Scratchpad for Editable Text-to-Image Generation
by: Saha, Oindrila, et al.
Published: (2026)
by: Saha, Oindrila, et al.
Published: (2026)
SIGMA-GEN: Structure and Identity Guided Multi-subject Assembly for Image Generation
by: Saha, Oindrila, et al.
Published: (2025)
by: Saha, Oindrila, et al.
Published: (2025)
CATRF: Codec-Adaptive TriPlane Radiance Fields for Volumetric Content Delivery
by: Chen, Tung-I, et al.
Published: (2026)
by: Chen, Tung-I, et al.
Published: (2026)
CleverBirds: A Multiple-Choice Benchmark for Fine-grained Human Knowledge Tracing
by: Bossemeyer, Leonie, et al.
Published: (2025)
by: Bossemeyer, Leonie, et al.
Published: (2025)
Perceptually Optimized Color Selection for Visualization
by: Maji, Subhrajyoti, et al.
Published: (2022)
by: Maji, Subhrajyoti, et al.
Published: (2022)
Improving Satellite Imagery Masking using Multi-task and Transfer Learning
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
Human Re-ID Meets LVLMs: What can we expect?
by: Hambarde, Kailash, et al.
Published: (2025)
by: Hambarde, Kailash, et al.
Published: (2025)
Counting Fish with Temporal Representations of Sonar Video
by: Van Brunt, Kai, et al.
Published: (2025)
by: Van Brunt, Kai, et al.
Published: (2025)
LKA-ReID:Vehicle Re-Identification with Large Kernel Attention
by: Xiang, Xuezhi, et al.
Published: (2024)
by: Xiang, Xuezhi, et al.
Published: (2024)
Beyond Visual Cues: Semantic-Driven Token Filtering and Expert Routing for Anytime Person ReID
by: Li, Jiaxuan, et al.
Published: (2026)
by: Li, Jiaxuan, et al.
Published: (2026)
Synthetic-To-Real Video Person Re-ID
by: Zhang, Xiangqun, et al.
Published: (2024)
by: Zhang, Xiangqun, et al.
Published: (2024)
CMCC-ReID: Cross-Modality Clothing-Change Person Re-Identification
by: Xu, Haoxuan, et al.
Published: (2026)
by: Xu, Haoxuan, et al.
Published: (2026)
MSP-ReID: Hairstyle-Robust Cloth-Changing Person Re-Identification
by: He, Xiangyang, et al.
Published: (2026)
by: He, Xiangyang, et al.
Published: (2026)
OC4-ReID: Occluded Cloth-Changing Person Re-Identification
by: Chen, Zhihao, et al.
Published: (2024)
by: Chen, Zhihao, et al.
Published: (2024)
PCD-ReID: Occluded Person Re-Identification for Base Station Inspection
by: Gao, Ge, et al.
Published: (2025)
by: Gao, Ge, et al.
Published: (2025)
QA-ReID: Quality-Aware Query-Adaptive Convolution Leveraging Fused Global and Structural Cues for Clothes-Changing ReID
by: Wang, Yuxiang, et al.
Published: (2026)
by: Wang, Yuxiang, et al.
Published: (2026)
LoopViT: Scaling Visual ARC with Looped Transformers
by: Shu, Wen-Jie, et al.
Published: (2026)
by: Shu, Wen-Jie, et al.
Published: (2026)
Instruct-ReID: A Multi-purpose Person Re-identification Task with Instructions
by: He, Weizhen, et al.
Published: (2023)
by: He, Weizhen, et al.
Published: (2023)
PS-ReID: Advancing Person Re-Identification and Precise Segmentation with Multimodal Retrieval
by: Yan, Jincheng, et al.
Published: (2025)
by: Yan, Jincheng, et al.
Published: (2025)
FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-Identification
by: Sun, Zhen, et al.
Published: (2025)
by: Sun, Zhen, et al.
Published: (2025)
Instruct-ReID++: Towards Universal Purpose Instruction-Guided Person Re-identification
by: He, Weizhen, et al.
Published: (2024)
by: He, Weizhen, et al.
Published: (2024)
Sports Re-ID: Improving Re-Identification Of Players In Broadcast Videos Of Team Sports
by: Comandur, Bharath
Published: (2022)
by: Comandur, Bharath
Published: (2022)
Multi-person Physics-based Pose Estimation for Combat Sports
by: Feiz, Hossein, et al.
Published: (2025)
by: Feiz, Hossein, et al.
Published: (2025)
Colo-ReID: Discriminative Representation Embedding with Meta-learning for Colonoscopic Polyp Re-Identification
by: Xiang, Suncheng, et al.
Published: (2023)
by: Xiang, Suncheng, et al.
Published: (2023)
Visual Autoregressive Modelling for Monocular Depth Estimation
by: El-Ghoussani, Amir, et al.
Published: (2025)
by: El-Ghoussani, Amir, et al.
Published: (2025)
AG-ReID.v2: Bridging Aerial and Ground Views for Person Re-identification
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Similar Items
-
Merlin L48 Spectrogram Dataset
by: Sun, Aaron, et al.
Published: (2025) -
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
by: Saha, Oindrila, et al.
Published: (2024) -
Consensus-Driven Active Model Selection
by: Kay, Justin, et al.
Published: (2025) -
Active Measurement of Two-Point Correlations
by: Hamilton, Max, et al.
Published: (2026) -
Generate, Transduct, Adapt: Iterative Transduction with VLMs
by: Saha, Oindrila, et al.
Published: (2025)