Reducing Annotation Burden in Physical Activity Research Using Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Schonfeldt, Abram, Maylor, Benjamin, Chen, Xiaofang, Clark, Ronald, Doherty, Aiden |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reducing Annotation Burden for Femoral Cartilage Segmentation in Knee MRI via Cross-Sequence Transfer Learning
by: Chiumento, Francesco, et al.
Published: (2026)
by: Chiumento, Francesco, et al.
Published: (2026)
Prompt-Based Caption Generation for Single-Tooth Dental Images Using Vision-Language Models
by: Sukhanova, Anastasiia, et al.
Published: (2026)
by: Sukhanova, Anastasiia, et al.
Published: (2026)
Debiased Prompt Tuning in Vision-Language Model without Annotations
by: Jiang, Chaoquan, et al.
Published: (2025)
by: Jiang, Chaoquan, et al.
Published: (2025)
Effortless Vision-Language Model Specialization in Histopathology without Annotation
by: Qiu, Jingna, et al.
Published: (2025)
by: Qiu, Jingna, et al.
Published: (2025)
Leveraging Vision-Language Models as Weak Annotators in Active Learning
by: Nguyen, Phuong Ngoc, et al.
Published: (2026)
by: Nguyen, Phuong Ngoc, et al.
Published: (2026)
Reducing Annotation Burden: Exploiting Image Knowledge for Few-Shot Medical Video Object Segmentation via Spatiotemporal Consistency Relearning
by: Zheng, Zixuan, et al.
Published: (2025)
by: Zheng, Zixuan, et al.
Published: (2025)
Image as an IMU: Estimating Camera Motion from a Single Motion-Blurred Image
by: Chen, Jerred, et al.
Published: (2025)
by: Chen, Jerred, et al.
Published: (2025)
Annotation Free Spacecraft Detection and Segmentation using Vision Language Models
by: Hicsonmez, Samet, et al.
Published: (2026)
by: Hicsonmez, Samet, et al.
Published: (2026)
DIO: Dataset of 3D Mesh Models of Indoor Objects for Robotics and Computer Vision Applications
by: Nimal, Nillan, et al.
Published: (2024)
by: Nimal, Nillan, et al.
Published: (2024)
Pushing the Limits of Vision-Language Models in Remote Sensing without Human Annotations
by: Cha, Keumgang, et al.
Published: (2024)
by: Cha, Keumgang, et al.
Published: (2024)
Visually Grounded Narratives: Reducing Cognitive Burden in Researcher-Participant Interaction
by: Wu, Runtong, et al.
Published: (2025)
by: Wu, Runtong, et al.
Published: (2025)
Towards Zero-Shot Annotation of the Built Environment with Vision-Language Models (Vision Paper)
by: Han, Bin, et al.
Published: (2024)
by: Han, Bin, et al.
Published: (2024)
Generate, but Verify: Reducing Hallucination in Vision-Language Models with Retrospective Resampling
by: Wu, Tsung-Han, et al.
Published: (2025)
by: Wu, Tsung-Han, et al.
Published: (2025)
HiMix: Reducing Computational Complexity in Large Vision-Language Models
by: Zhang, Xuange, et al.
Published: (2025)
by: Zhang, Xuange, et al.
Published: (2025)
Pre-Trained Vision-Language Models as Partial Annotators
by: Wang, Qian-Wei, et al.
Published: (2024)
by: Wang, Qian-Wei, et al.
Published: (2024)
MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models
by: Zhao, Qiyan, et al.
Published: (2025)
by: Zhao, Qiyan, et al.
Published: (2025)
Bridging Hidden States in Vision-Language Models
by: Fein-Ashley, Benjamin, et al.
Published: (2025)
by: Fein-Ashley, Benjamin, et al.
Published: (2025)
Annotation Free Semantic Segmentation with Vision Foundation Models
by: Seifi, Soroush, et al.
Published: (2024)
by: Seifi, Soroush, et al.
Published: (2024)
RetFiner: A Vision-Language Refinement Scheme for Retinal Foundation Models
by: Fecso, Ronald, et al.
Published: (2025)
by: Fecso, Ronald, et al.
Published: (2025)
CLIP-VAD: Exploiting Vision-Language Models for Voice Activity Detection
by: Appiani, Andrea, et al.
Published: (2024)
by: Appiani, Andrea, et al.
Published: (2024)
Accelerating Multimodal Large Language Models by Searching Optimal Vision Token Reduction
by: Zhao, Shiyu, et al.
Published: (2024)
by: Zhao, Shiyu, et al.
Published: (2024)
Instant Uncertainty Calibration of NeRFs Using a Meta-Calibrator
by: Amini-Naieni, Niki, et al.
Published: (2023)
by: Amini-Naieni, Niki, et al.
Published: (2023)
Generating Vision-Language Navigation Instructions Incorporated Fine-Grained Alignment Annotations
by: Cui, Yibo, et al.
Published: (2025)
by: Cui, Yibo, et al.
Published: (2025)
The Urban Vision Hackathon Dataset and Models: Towards Image Annotations and Accurate Vision Models for Indian Traffic
by: Sharma, Akash, et al.
Published: (2025)
by: Sharma, Akash, et al.
Published: (2025)
Optimizing Vision-Language Interactions Through Decoder-Only Models
by: Tanaka, Kaito, et al.
Published: (2024)
by: Tanaka, Kaito, et al.
Published: (2024)
Bootstrapping Sign Language Annotations with Sign Language Models
by: Lea, Colin, et al.
Published: (2026)
by: Lea, Colin, et al.
Published: (2026)
VLM-CPL: Consensus Pseudo Labels from Vision-Language Models for Annotation-Free Pathological Image Classification
by: Zhong, Lanfeng, et al.
Published: (2024)
by: Zhong, Lanfeng, et al.
Published: (2024)
Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
by: Geigle, Gregor, et al.
Published: (2024)
by: Geigle, Gregor, et al.
Published: (2024)
Longitudinal Assessment of Lung Lesion Burden in CT
by: Mathai, Tejas Sudharshan, et al.
Published: (2025)
by: Mathai, Tejas Sudharshan, et al.
Published: (2025)
SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living
by: Sinha, Arkaprava, et al.
Published: (2025)
by: Sinha, Arkaprava, et al.
Published: (2025)
MERA: Multimodal and Multiscale Self-Explanatory Model with Considerably Reduced Annotation for Lung Nodule Diagnosis
by: Lu, Jiahao, et al.
Published: (2025)
by: Lu, Jiahao, et al.
Published: (2025)
Using Vision Language Models for Safety Hazard Identification in Construction
by: Adil, Muhammad, et al.
Published: (2025)
by: Adil, Muhammad, et al.
Published: (2025)
3D Vision-Language Gaussian Splatting
by: Peng, Qucheng, et al.
Published: (2024)
by: Peng, Qucheng, et al.
Published: (2024)
A Vision-Centric Approach for Static Map Element Annotation
by: Zhang, Jiaxin, et al.
Published: (2023)
by: Zhang, Jiaxin, et al.
Published: (2023)
Revisiting Prompt Pretraining of Vision-Language Models
by: Chen, Zhenyuan, et al.
Published: (2024)
by: Chen, Zhenyuan, et al.
Published: (2024)
MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting
by: Li, Yuheng, et al.
Published: (2025)
by: Li, Yuheng, et al.
Published: (2025)
ViTamin: Designing Scalable Vision Models in the Vision-Language Era
by: Chen, Jieneng, et al.
Published: (2024)
by: Chen, Jieneng, et al.
Published: (2024)
Reg4Pru: Regularisation Through Random Token Routing for Token Pruning
by: Wyatt, Julian, et al.
Published: (2026)
by: Wyatt, Julian, et al.
Published: (2026)
Can Unsupervised Segmentation Reduce Annotation Costs for Video Semantic Segmentation?
by: Some, Samik, et al.
Published: (2026)
by: Some, Samik, et al.
Published: (2026)
The Role of Background Information in Reducing Object Hallucination in Vision-Language Models: Insights from Cutoff API Prompting
by: Tomita, Masayo, et al.
Published: (2025)
by: Tomita, Masayo, et al.
Published: (2025)
Similar Items
-
Reducing Annotation Burden for Femoral Cartilage Segmentation in Knee MRI via Cross-Sequence Transfer Learning
by: Chiumento, Francesco, et al.
Published: (2026) -
Prompt-Based Caption Generation for Single-Tooth Dental Images Using Vision-Language Models
by: Sukhanova, Anastasiia, et al.
Published: (2026) -
Debiased Prompt Tuning in Vision-Language Model without Annotations
by: Jiang, Chaoquan, et al.
Published: (2025) -
Effortless Vision-Language Model Specialization in Histopathology without Annotation
by: Qiu, Jingna, et al.
Published: (2025) -
Leveraging Vision-Language Models as Weak Annotators in Active Learning
by: Nguyen, Phuong Ngoc, et al.
Published: (2026)