(LiFT) Lightweight Fitness Transformer: A language-vision model for Remote Monitoring of Physical Training
Fuente:
arXiv
Saved in:
| Main Authors: | Postlmayr, A., Cosman, P., Dey, S. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LiFT: A Surprisingly Simple Lightweight Feature Transform for Dense ViT Descriptors
by: Suri, Saksham, et al.
Published: (2024)
by: Suri, Saksham, et al.
Published: (2024)
LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment
by: Wang, Yibin, et al.
Published: (2024)
by: Wang, Yibin, et al.
Published: (2024)
LiFT: Lightweight, FPGA-tailored 3D object detection based on LiDAR data
by: Lis, Konrad, et al.
Published: (2025)
by: Lis, Konrad, et al.
Published: (2025)
LiFT: Lifted Inter-slice Feature Trajectories for 3D Image Generation from 2D Generators
by: Zhang, Xinhe, et al.
Published: (2026)
by: Zhang, Xinhe, et al.
Published: (2026)
Body and Head Orientation Estimation from Low-Resolution Point Clouds in Surveillance Settings
by: Tepencelik, Onur N., et al.
Published: (2023)
by: Tepencelik, Onur N., et al.
Published: (2023)
Can We Remove the Ground? Obstacle-aware Point Cloud Compression for Remote Object Detection
by: Zeng, Pengxi, et al.
Published: (2024)
by: Zeng, Pengxi, et al.
Published: (2024)
A Lightweight Sparse Focus Transformer for Remote Sensing Image Change Captioning
by: Sun, Dongwei, et al.
Published: (2024)
by: Sun, Dongwei, et al.
Published: (2024)
A Novel Lightweight Transformer with Edge-Aware Fusion for Remote Sensing Image Captioning
by: Das, Swadhin, et al.
Published: (2025)
by: Das, Swadhin, et al.
Published: (2025)
An analysis of vision-language models for fabric retrieval
by: Giuliari, Francesco, et al.
Published: (2025)
by: Giuliari, Francesco, et al.
Published: (2025)
Lightweight, Pre-trained Transformers for Remote Sensing Timeseries
by: Tseng, Gabriel, et al.
Published: (2023)
by: Tseng, Gabriel, et al.
Published: (2023)
Lightweight Change Detection in Heterogeneous Remote Sensing Images with Online All-Integer Pruning Training
by: Zhang, Chengyang, et al.
Published: (2024)
by: Zhang, Chengyang, et al.
Published: (2024)
KeyNode-Driven Geometry Coding for Real-World Scanned Human Dynamic Mesh Compression
by: Hoang, Huong, et al.
Published: (2025)
by: Hoang, Huong, et al.
Published: (2025)
Efficient Progressive Image Compression with Variance-aware Masking
by: Presta, Alberto, et al.
Published: (2024)
by: Presta, Alberto, et al.
Published: (2024)
Frame Sampling Strategies Matter: A Benchmark for small vision language models
by: Brkic, Marija, et al.
Published: (2025)
by: Brkic, Marija, et al.
Published: (2025)
Visual symbolic mechanisms: Emergent symbol processing in vision language models
by: Assouel, Rim, et al.
Published: (2025)
by: Assouel, Rim, et al.
Published: (2025)
Do large language vision models understand 3D shapes?
by: Eppel, Sagi
Published: (2024)
by: Eppel, Sagi
Published: (2024)
Are vision-language models ready to zero-shot replace supervised classification models in agriculture?
by: Ranario, Earl, et al.
Published: (2025)
by: Ranario, Earl, et al.
Published: (2025)
A multi-modal vision-language model for generalizable annotation-free pathology localization
by: Yang, Hao, et al.
Published: (2024)
by: Yang, Hao, et al.
Published: (2024)
Are vision language models robust to uncertain inputs?
by: Wang, Xi, et al.
Published: (2025)
by: Wang, Xi, et al.
Published: (2025)
LoFT: LoRA-fused Training Dataset Generation with Few-shot Guidance
by: Kim, Jae Myung, et al.
Published: (2025)
by: Kim, Jae Myung, et al.
Published: (2025)
Quantifying the human visual exposome with vision language models
by: Rominger, Christian, et al.
Published: (2026)
by: Rominger, Christian, et al.
Published: (2026)
bi-modal textual prompt learning for vision-language models in remote sensing
by: Kashyap, Pankhi, et al.
Published: (2026)
by: Kashyap, Pankhi, et al.
Published: (2026)
RhythmMamba: Fast, Lightweight, and Accurate Remote Physiological Measurement
by: Zou, Bochao, et al.
Published: (2024)
by: Zou, Bochao, et al.
Published: (2024)
Thinker: A vision-language foundation model for embodied intelligence
by: Pan, Baiyu, et al.
Published: (2026)
by: Pan, Baiyu, et al.
Published: (2026)
LDGNet: A Lightweight Difference Guiding Network for Remote Sensing Change Detection
by: Xu, Chenfeng
Published: (2025)
by: Xu, Chenfeng
Published: (2025)
A vision-language model and platform for temporally mapping surgery from video
by: Kiyasseh, Dani
Published: (2026)
by: Kiyasseh, Dani
Published: (2026)
Lightweight RGB-T Tracking with Mobile Vision Transformers
by: Falaki, Mahdi, et al.
Published: (2025)
by: Falaki, Mahdi, et al.
Published: (2025)
Enhancing medical vision-language contrastive learning via inter-matching relation modelling
by: Li, Mingjian, et al.
Published: (2024)
by: Li, Mingjian, et al.
Published: (2024)
Enhancing the vision-language foundation model with key semantic knowledge-emphasized report refinement
by: Huang, Weijian, et al.
Published: (2024)
by: Huang, Weijian, et al.
Published: (2024)
Initialization matters in few-shot adaptation of vision-language models for histopathological image classification
by: Meseguer, Pablo, et al.
Published: (2026)
by: Meseguer, Pablo, et al.
Published: (2026)
Expert-level vision-language foundation model for real-world radiology and comprehensive evaluation
by: Liu, Xiaohong, et al.
Published: (2024)
by: Liu, Xiaohong, et al.
Published: (2024)
What matters when building vision-language models?
by: Laurençon, Hugo, et al.
Published: (2024)
by: Laurençon, Hugo, et al.
Published: (2024)
RSDehamba: Lightweight Vision Mamba for Remote Sensing Satellite Image Dehazing
by: Zhou, Huiling, et al.
Published: (2024)
by: Zhou, Huiling, et al.
Published: (2024)
Automated Road Extraction and Centreline Fitting in LiDAR Point Clouds
by: Wang, Xinyu, et al.
Published: (2025)
by: Wang, Xinyu, et al.
Published: (2025)
SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion
by: Nassar, Ahmed, et al.
Published: (2025)
by: Nassar, Ahmed, et al.
Published: (2025)
Interpreting the linear structure of vision-language model embedding spaces
by: Papadimitriou, Isabel, et al.
Published: (2025)
by: Papadimitriou, Isabel, et al.
Published: (2025)
Zero-shot segmentation of skin tumors in whole-slide images with vision-language foundation models
by: Moreno, Santiago, et al.
Published: (2025)
by: Moreno, Santiago, et al.
Published: (2025)
Generalizing vision-language models to novel domains: A comprehensive survey
by: Li, Xinyao, et al.
Published: (2025)
by: Li, Xinyao, et al.
Published: (2025)
DreamFit: Garment-Centric Human Generation via a Lightweight Anything-Dressing Encoder
by: Lin, Ente, et al.
Published: (2024)
by: Lin, Ente, et al.
Published: (2024)
An Experimental Study on Exploring Strong Lightweight Vision Transformers via Masked Image Modeling Pre-Training
by: Gao, Jin, et al.
Published: (2024)
by: Gao, Jin, et al.
Published: (2024)
Similar Items
-
LiFT: A Surprisingly Simple Lightweight Feature Transform for Dense ViT Descriptors
by: Suri, Saksham, et al.
Published: (2024) -
LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment
by: Wang, Yibin, et al.
Published: (2024) -
LiFT: Lightweight, FPGA-tailored 3D object detection based on LiDAR data
by: Lis, Konrad, et al.
Published: (2025) -
LiFT: Lifted Inter-slice Feature Trajectories for 3D Image Generation from 2D Generators
by: Zhang, Xinhe, et al.
Published: (2026) -
Body and Head Orientation Estimation from Low-Resolution Point Clouds in Surveillance Settings
by: Tepencelik, Onur N., et al.
Published: (2023)