Gamified crowd-sourcing of high-quality data for visual fine-tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Yadav, Shashank, Tomar, Rohan, Jain, Garvit, Ahooja, Chirag, Chaudhary, Shubham, Elkan, Charles |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
QYOLO: Lightweight Object Detection via Quantum Inspired Shared Channel Mixing
by: Mittal, Garvit Kumar, et al.
Published: (2026)
by: Mittal, Garvit Kumar, et al.
Published: (2026)
Visual RAG: Expanding MLLM visual knowledge without fine-tuning
by: Bonomo, Mirco, et al.
Published: (2025)
by: Bonomo, Mirco, et al.
Published: (2025)
Enhancing targeted transferability via feature space fine-tuning
by: Zeng, Hui, et al.
Published: (2024)
by: Zeng, Hui, et al.
Published: (2024)
Gamified AI Approch for Early Detection of Dementia
by: Maji, Paramita Kundu, et al.
Published: (2024)
by: Maji, Paramita Kundu, et al.
Published: (2024)
Uncertainty modeling for fine-tuned implicit functions
by: Susmelj, Anna, et al.
Published: (2024)
by: Susmelj, Anna, et al.
Published: (2024)
Unsupervised Domain Adaptation for Action Recognition via Self-Ensembling and Conditional Embedding Alignment
by: Ghosh, Indrajeet, et al.
Published: (2024)
by: Ghosh, Indrajeet, et al.
Published: (2024)
LICA: Layered Image Composition Annotations for Graphic Design Research
by: Hirsch, Elad, et al.
Published: (2026)
by: Hirsch, Elad, et al.
Published: (2026)
RoboSignature: Robust Signature and Watermarking on Network Attacks
by: Shaan, Aryaman, et al.
Published: (2024)
by: Shaan, Aryaman, et al.
Published: (2024)
Detecting Multiple Diseases in Multiple Crops Using Deep Learning
by: Yadav, Vivek, et al.
Published: (2025)
by: Yadav, Vivek, et al.
Published: (2025)
Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play
by: Wang, Qinsi, et al.
Published: (2025)
by: Wang, Qinsi, et al.
Published: (2025)
Color histogram equalization and fine-tuning to improve expression recognition of (partially occluded) faces on sign language datasets
by: Nunnari, Fabrizio, et al.
Published: (2025)
by: Nunnari, Fabrizio, et al.
Published: (2025)
Lost in Translation and Noise: A Deep Dive into the Failure Modes of VLMs on Real-World Tables
by: Singh, Anshul, et al.
Published: (2025)
by: Singh, Anshul, et al.
Published: (2025)
Optimising CSRNet with parameter-free attention mechanisms for crowd counting in public transport
by: Rostamza, Aida, et al.
Published: (2026)
by: Rostamza, Aida, et al.
Published: (2026)
LookupViT: Compressing visual information to a limited number of tokens
by: Koner, Rajat, et al.
Published: (2024)
by: Koner, Rajat, et al.
Published: (2024)
MTCNET: Multi-task Learning Paradigm for Crowd Count Estimation
by: Kumar, Abhay, et al.
Published: (2019)
by: Kumar, Abhay, et al.
Published: (2019)
Advanced Smart City Monitoring: Real-Time Identification of Indian Citizen Attributes
by: Kale, Shubham, et al.
Published: (2024)
by: Kale, Shubham, et al.
Published: (2024)
RandLoRA: Full-rank parameter-efficient fine-tuning of large models
by: Albert, Paul, et al.
Published: (2025)
by: Albert, Paul, et al.
Published: (2025)
Demographic-aware fine-grained visual recognition of pediatric wrist pathologies
by: Ahmed, Ammar, et al.
Published: (2025)
by: Ahmed, Ammar, et al.
Published: (2025)
What explains the success of cross-modal fine-tuning with ORCA?
by: García-de-Herreros, Paloma, et al.
Published: (2024)
by: García-de-Herreros, Paloma, et al.
Published: (2024)
Towards a Systematic Evaluation of Hallucinations in Large-Vision Language Models
by: Seth, Ashish, et al.
Published: (2024)
by: Seth, Ashish, et al.
Published: (2024)
CAD2DMD-SET: Synthetic Generation Tool of Digital Measurement Device CAD Model Datasets for fine-tuning Large Vision-Language Models
by: Valente, João, et al.
Published: (2025)
by: Valente, João, et al.
Published: (2025)
Oracle-RLAIF: An Improved Fine-Tuning Framework for Multi-modal Video Models through Reinforcement Learning from Ranking Feedback
by: Shi, Derek, et al.
Published: (2025)
by: Shi, Derek, et al.
Published: (2025)
A Simple and Effective Reinforcement Learning Method for Text-to-Image Diffusion Fine-tuning
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
DARN: Dynamic Adaptive Regularization Networks for Efficient and Robust Foundation Model Adaptation
by: Yadav, Dhenenjay, et al.
Published: (2025)
by: Yadav, Dhenenjay, et al.
Published: (2025)
Structured Unrestricted-Rank Matrices for Parameter Efficient Fine-tuning
by: Sehanobish, Arijit, et al.
Published: (2024)
by: Sehanobish, Arijit, et al.
Published: (2024)
Can GPT-4o mini and Gemini 2.0 Flash Predict Fine-Grained Fashion Product Attributes? A Zero-Shot Analysis
by: Shukla, Shubham, et al.
Published: (2025)
by: Shukla, Shubham, et al.
Published: (2025)
How Culturally Aware are Vision-Language Models?
by: Burda-Lassen, Olena, et al.
Published: (2024)
by: Burda-Lassen, Olena, et al.
Published: (2024)
Transfer learning optimization based on evolutionary selective fine tuning
by: Colan, Jacinto, et al.
Published: (2025)
by: Colan, Jacinto, et al.
Published: (2025)
Segmentation of Maya hieroglyphs through fine-tuned foundation models
by: Shivam, FNU, et al.
Published: (2024)
by: Shivam, FNU, et al.
Published: (2024)
Herd Mentality in Augmentation -- Not a Good Idea! A Robust Multi-stage Approach towards Deepfake Detection
by: Monu, et al.
Published: (2024)
by: Monu, et al.
Published: (2024)
Can visual language models resolve textual ambiguity with visual cues? Let visual puns tell you!
by: Chung, Jiwan, et al.
Published: (2024)
by: Chung, Jiwan, et al.
Published: (2024)
Improved visual-information-driven model for crowd simulation and its modular application
by: Liang, Xuanwen, et al.
Published: (2025)
by: Liang, Xuanwen, et al.
Published: (2025)
Feature Fusion Attention Network with CycleGAN for Image Dehazing, De-Snowing and De-Raining
by: Jain, Akshat
Published: (2025)
by: Jain, Akshat
Published: (2025)
Detect an Object At Once without Fine-tuning
by: Hao, Junyu, et al.
Published: (2024)
by: Hao, Junyu, et al.
Published: (2024)
VladVA: Discriminative Fine-tuning of LVLMs
by: Ouali, Yassine, et al.
Published: (2024)
by: Ouali, Yassine, et al.
Published: (2024)
ONRW: Optimizing inversion noise for high-quality and robust watermark
by: Ding, Xuan, et al.
Published: (2026)
by: Ding, Xuan, et al.
Published: (2026)
OpenT2M: No-frill Motion Generation with Open-source,Large-scale, High-quality Data
by: Cao, Bin, et al.
Published: (2026)
by: Cao, Bin, et al.
Published: (2026)
Ridgeformer: Mutli-Stage Contrastive Training For Fine-grained Cross-Domain Fingerprint Recognition
by: Pandey, Shubham, et al.
Published: (2025)
by: Pandey, Shubham, et al.
Published: (2025)
Technical note: ShinyAnimalCV: open-source cloud-based web application for object detection, segmentation, and three-dimensional visualization of animals using computer vision
by: Wang, Jin, et al.
Published: (2023)
by: Wang, Jin, et al.
Published: (2023)
Transfer-LMR: Heavy-Tail Driving Behavior Recognition in Diverse Traffic Scenarios
by: Parikh, Chirag, et al.
Published: (2024)
by: Parikh, Chirag, et al.
Published: (2024)
Similar Items
-
QYOLO: Lightweight Object Detection via Quantum Inspired Shared Channel Mixing
by: Mittal, Garvit Kumar, et al.
Published: (2026) -
Visual RAG: Expanding MLLM visual knowledge without fine-tuning
by: Bonomo, Mirco, et al.
Published: (2025) -
Enhancing targeted transferability via feature space fine-tuning
by: Zeng, Hui, et al.
Published: (2024) -
Gamified AI Approch for Early Detection of Dementia
by: Maji, Paramita Kundu, et al.
Published: (2024) -
Uncertainty modeling for fine-tuned implicit functions
by: Susmelj, Anna, et al.
Published: (2024)