Vision-based module for accurately reading linear scales in a laboratory
Fuente:
arXiv
Saved in:
| Main Authors: | Saini, Parvesh, Maiti, Soumyadipta, Rai, Beena |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Quality Detection of Stored Potatoes via Transfer Learning: A CNN and Vision Transformer Approach
by: Kapse, Shrikant, et al.
Published: (2026)
by: Kapse, Shrikant, et al.
Published: (2026)
Privacy Preserving Ordinal-Meta Learning with VLMs for Fine-Grained Fruit Quality Prediction
by: Jain, Riddhi, et al.
Published: (2025)
by: Jain, Riddhi, et al.
Published: (2025)
When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models
by: Saini, Harshvardhan, et al.
Published: (2026)
by: Saini, Harshvardhan, et al.
Published: (2026)
Mixed Non-linear Quantization for Vision Transformers
by: Kim, Gihwan, et al.
Published: (2024)
by: Kim, Gihwan, et al.
Published: (2024)
A Computer Vision-Based Quality Assessment Technique for the automatic control of consumables for analytical laboratories
by: Zribi, Meriam, et al.
Published: (2024)
by: Zribi, Meriam, et al.
Published: (2024)
Quantum Inverse Contextual Vision Transformers (Q-ICVT): A New Frontier in 3D Object Detection for AVs
by: Dharavath, Sanjay Bhargav, et al.
Published: (2024)
by: Dharavath, Sanjay Bhargav, et al.
Published: (2024)
Latent fingerprint enhancement for accurate minutiae detection
by: Wahab, Abdul, et al.
Published: (2024)
by: Wahab, Abdul, et al.
Published: (2024)
EPBC-YOLOv8: An efficient and accurate improved YOLOv8 underwater detector based on an attention mechanism
by: Jiang, Xing, et al.
Published: (2025)
by: Jiang, Xing, et al.
Published: (2025)
InstAP: Instance-Aware Vision-Language Pre-Train for Spatial-Temporal Understanding
by: Kumar, Ashutosh, et al.
Published: (2026)
by: Kumar, Ashutosh, et al.
Published: (2026)
Combined concurrent Physical and Chemical model for accelerated weathering damages of polyurethane-based coatings
by: Gupta, Ambesh, et al.
Published: (2024)
by: Gupta, Ambesh, et al.
Published: (2024)
ECViT: Efficient Convolutional Vision Transformer with Local-Attention and Multi-scale Stages
by: Qian, Zhoujie
Published: (2025)
by: Qian, Zhoujie
Published: (2025)
MSVIT: Improving Spiking Vision Transformer Using Multi-scale Attention Fusion
by: Hua, Wei, et al.
Published: (2025)
by: Hua, Wei, et al.
Published: (2025)
IPTQ-ViT: Post-Training Quantization of Non-linear Functions for Integer-only Vision Transformers
by: Kim, Gihwan, et al.
Published: (2025)
by: Kim, Gihwan, et al.
Published: (2025)
TRAX: TRacking Axles for Accurate Axle Count Estimation
by: Rai, Avinash, et al.
Published: (2025)
by: Rai, Avinash, et al.
Published: (2025)
FLNet: Flood-Induced Agriculture Damage Assessment using Super Resolution of Satellite Images
by: Ghosal, Sanidhya, et al.
Published: (2026)
by: Ghosal, Sanidhya, et al.
Published: (2026)
VajraV1 -- The most accurate Real Time Object Detector of the YOLO family
by: Makkar, Naman Balbir Singh
Published: (2025)
by: Makkar, Naman Balbir Singh
Published: (2025)
3DCoMPaT$^{++}$: An improved Large-scale 3D Vision Dataset for Compositional Recognition
by: Slim, Habib, et al.
Published: (2023)
by: Slim, Habib, et al.
Published: (2023)
Multi-Branch Auxiliary Fusion YOLO with Re-parameterization Heterogeneous Convolutional for accurate object detection
by: Yang, Zhiqiang, et al.
Published: (2024)
by: Yang, Zhiqiang, et al.
Published: (2024)
SurgPose: Generalisable Surgical Instrument Pose Estimation using Zero-Shot Learning and Stereo Vision
by: Rai, Utsav, et al.
Published: (2025)
by: Rai, Utsav, et al.
Published: (2025)
Latent Guidance in Diffusion Models for Perceptual Evaluations
by: Saini, Shreshth, et al.
Published: (2025)
by: Saini, Shreshth, et al.
Published: (2025)
VARS: Vision-based Assessment of Risk in Security Systems
by: Gupta, Pranav, et al.
Published: (2024)
by: Gupta, Pranav, et al.
Published: (2024)
Computer Vision based group activity detection and action spotting
by: Sivalingam, Narthana, et al.
Published: (2025)
by: Sivalingam, Narthana, et al.
Published: (2025)
RAU: Reference-based Anatomical Understanding with Vision Language Models
by: Li, Yiwei, et al.
Published: (2025)
by: Li, Yiwei, et al.
Published: (2025)
Towards Two-Stream Foveation-based Active Vision Learning
by: Ibrayev, Timur, et al.
Published: (2024)
by: Ibrayev, Timur, et al.
Published: (2024)
CHUG: Crowdsourced User-Generated HDR Video Quality Dataset
by: Saini, Shreshth, et al.
Published: (2025)
by: Saini, Shreshth, et al.
Published: (2025)
VFM-VLM: Vision Foundation Model and Vision Language Model based Visual Comparison for 3D Pose Estimation
by: Sarowar, Md Selim, et al.
Published: (2025)
by: Sarowar, Md Selim, et al.
Published: (2025)
AppleGrowthVision: A large-scale stereo dataset for phenological analysis, fruit detection, and 3D reconstruction in apple orchards
by: von Hirschhausen, Laura-Sophia, et al.
Published: (2025)
by: von Hirschhausen, Laura-Sophia, et al.
Published: (2025)
Small Object Few-shot Segmentation for Vision-based Industrial Inspection
by: Zhang, Zilong, et al.
Published: (2024)
by: Zhang, Zilong, et al.
Published: (2024)
Case-Enhanced Vision Transformer: Improving Explanations of Image Similarity with a ViT-based Similarity Metric
by: Zhao, Ziwei, et al.
Published: (2024)
by: Zhao, Ziwei, et al.
Published: (2024)
A Survey of Adversarial Defenses in Vision-based Systems: Categorization, Methods and Challenges
by: Chattopadhyay, Nandish, et al.
Published: (2025)
by: Chattopadhyay, Nandish, et al.
Published: (2025)
VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models
by: Balakrishnan, Ravikumar, et al.
Published: (2025)
by: Balakrishnan, Ravikumar, et al.
Published: (2025)
VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models
by: Phute, Mansi, et al.
Published: (2025)
by: Phute, Mansi, et al.
Published: (2025)
DiffCAP: Diffusion-based Cumulative Adversarial Purification for Vision Language Models
by: Fu, Jia, et al.
Published: (2025)
by: Fu, Jia, et al.
Published: (2025)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
by: Wang, Xiao, et al.
Published: (2023)
by: Wang, Xiao, et al.
Published: (2023)
GTP-ViT: Efficient Vision Transformers via Graph-based Token Propagation
by: Xu, Xuwei, et al.
Published: (2023)
by: Xu, Xuwei, et al.
Published: (2023)
Conceptualizing Multi-scale Wavelet Attention and Ray-based Encoding for Human-Object Interaction Detection
by: Pay, Quan Bi, et al.
Published: (2025)
by: Pay, Quan Bi, et al.
Published: (2025)
Graph Classification and Radiomics Signature for Identification of Tuberculous Meningitis
by: Agarwal, Snigdha, et al.
Published: (2025)
by: Agarwal, Snigdha, et al.
Published: (2025)
DLEBench: Evaluating Small-scale Object Editing Ability for Instruction-based Image Editing Model
by: Hong, Shibo, et al.
Published: (2026)
by: Hong, Shibo, et al.
Published: (2026)
Seeing Beyond 8bits: Subjective and Objective Quality Assessment of HDR-UGC Videos
by: Saini, Shreshth, et al.
Published: (2026)
by: Saini, Shreshth, et al.
Published: (2026)
LumaFlux: Lifting 8-Bit Worlds to HDR Reality with Physically-Guided Diffusion Transformers
by: Saini, Shreshth, et al.
Published: (2026)
by: Saini, Shreshth, et al.
Published: (2026)
Similar Items
-
Quality Detection of Stored Potatoes via Transfer Learning: A CNN and Vision Transformer Approach
by: Kapse, Shrikant, et al.
Published: (2026) -
Privacy Preserving Ordinal-Meta Learning with VLMs for Fine-Grained Fruit Quality Prediction
by: Jain, Riddhi, et al.
Published: (2025) -
When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models
by: Saini, Harshvardhan, et al.
Published: (2026) -
Mixed Non-linear Quantization for Vision Transformers
by: Kim, Gihwan, et al.
Published: (2024) -
A Computer Vision-Based Quality Assessment Technique for the automatic control of consumables for analytical laboratories
by: Zribi, Meriam, et al.
Published: (2024)