Generation of Indian Sign Language Letters, Numbers, and Words
Fuente:
arXiv
Salvato in:
| Autori principali: | Yadav, Ajeet Kumar, Kumar, Nishant, N, Rathna G |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Face Detection: Present State and Research Directions
di: Prabhat, Purnendu, et al.
Pubblicazione: (2024)
di: Prabhat, Purnendu, et al.
Pubblicazione: (2024)
ISLR101: an Iranian Word-Level Sign Language Recognition Dataset
di: Ranjbar, Hossein, et al.
Pubblicazione: (2025)
di: Ranjbar, Hossein, et al.
Pubblicazione: (2025)
iSign: A Benchmark for Indian Sign Language Processing
di: Joshi, Abhinav, et al.
Pubblicazione: (2024)
di: Joshi, Abhinav, et al.
Pubblicazione: (2024)
MTCNET: Multi-task Learning Paradigm for Crowd Count Estimation
di: Kumar, Abhay, et al.
Pubblicazione: (2019)
di: Kumar, Abhay, et al.
Pubblicazione: (2019)
From Tokens to Numbers: Continuous Number Modeling for SVG Generation
di: Ogezi, Michael, et al.
Pubblicazione: (2026)
di: Ogezi, Michael, et al.
Pubblicazione: (2026)
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
di: Gueuwou, Shester, et al.
Pubblicazione: (2024)
di: Gueuwou, Shester, et al.
Pubblicazione: (2024)
BdSLW401: Transformer-Based Word-Level Bangla Sign Language Recognition Using Relative Quantization Encoding (RQE)
di: Rubaiyeat, Husne Ara, et al.
Pubblicazione: (2025)
di: Rubaiyeat, Husne Ara, et al.
Pubblicazione: (2025)
Real Time American Sign Language Detection Using Yolo-v9
di: Imran, Amna, et al.
Pubblicazione: (2024)
di: Imran, Amna, et al.
Pubblicazione: (2024)
Advanced Arabic Alphabet Sign Language Recognition Using Transfer Learning and Transformer Models
di: Balat, Mazen, et al.
Pubblicazione: (2024)
di: Balat, Mazen, et al.
Pubblicazione: (2024)
FusionEnsemble-Net: An Attention-Based Ensemble of Spatiotemporal Networks for Multimodal Sign Language Recognition
di: Islam, Md. Milon, et al.
Pubblicazione: (2025)
di: Islam, Md. Milon, et al.
Pubblicazione: (2025)
One Prompt Word is Enough to Boost Adversarial Robustness for Pre-trained Vision-Language Models
di: Li, Lin, et al.
Pubblicazione: (2024)
di: Li, Lin, et al.
Pubblicazione: (2024)
Towards Privacy-Aware Sign Language Translation at Scale
di: Rust, Phillip, et al.
Pubblicazione: (2024)
di: Rust, Phillip, et al.
Pubblicazione: (2024)
LAuReL: Learned Augmented Residual Layer
di: Menghani, Gaurav, et al.
Pubblicazione: (2024)
di: Menghani, Gaurav, et al.
Pubblicazione: (2024)
Vision-Language Models Provide Promptable Representations for Reinforcement Learning
di: Chen, William, et al.
Pubblicazione: (2024)
di: Chen, William, et al.
Pubblicazione: (2024)
Enhancing Financial VQA in Vision Language Models using Intermediate Structured Representations
di: Srivastava, Archita, et al.
Pubblicazione: (2025)
di: Srivastava, Archita, et al.
Pubblicazione: (2025)
AutoCLIP: Auto-tuning Zero-Shot Classifiers for Vision-Language Models
di: Metzen, Jan Hendrik, et al.
Pubblicazione: (2023)
di: Metzen, Jan Hendrik, et al.
Pubblicazione: (2023)
FairGen: Controlling Sensitive Attributes for Fair Generations in Diffusion Models via Adaptive Latent Guidance
di: Kang, Mintong, et al.
Pubblicazione: (2025)
di: Kang, Mintong, et al.
Pubblicazione: (2025)
A Multimodal, Multitask System for Generating E Commerce Text Listings from Images
di: Singh, Nayan Kumar
Pubblicazione: (2025)
di: Singh, Nayan Kumar
Pubblicazione: (2025)
Advancing Autonomous Vehicle Intelligence: Deep Learning and Multimodal LLM for Traffic Sign Recognition and Robust Lane Detection
di: Sah, Chandan Kumar, et al.
Pubblicazione: (2025)
di: Sah, Chandan Kumar, et al.
Pubblicazione: (2025)
A Markovian View of Iterative-Feedback Loops in Image Generative Models: Neural Resonance and Model Collapse
di: Vats, Vibhas Kumar, et al.
Pubblicazione: (2026)
di: Vats, Vibhas Kumar, et al.
Pubblicazione: (2026)
When Better Eyes Lead to Blindness: A Diagnostic Study of the Information Bottleneck in CNN-LSTM Image Captioning Models
di: Gupta, Hitesh Kumar
Pubblicazione: (2025)
di: Gupta, Hitesh Kumar
Pubblicazione: (2025)
Towards Real-Time 2D Mapping: Harnessing Drones, AI, and Computer Vision for Advanced Insights
di: Agnur, Bharath Kumar
Pubblicazione: (2024)
di: Agnur, Bharath Kumar
Pubblicazione: (2024)
A More Word-like Image Tokenization for MLLMs
di: Lee, Hyun, et al.
Pubblicazione: (2026)
di: Lee, Hyun, et al.
Pubblicazione: (2026)
Revolutionizing Communication with Deep Learning and XAI for Enhanced Arabic Sign Language Recognition
di: Balat, Mazen, et al.
Pubblicazione: (2025)
di: Balat, Mazen, et al.
Pubblicazione: (2025)
Continuous Sign Language Recognition System using Deep Learning with MediaPipe Holistic
di: Srivastava, Sharvani, et al.
Pubblicazione: (2024)
di: Srivastava, Sharvani, et al.
Pubblicazione: (2024)
Mechanisms of Non-Monotonic Scaling in Vision Transformers
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
Measuring the (Un)Faithfulness of Concept-Based Explanations
di: Kumar, Shubham, et al.
Pubblicazione: (2025)
di: Kumar, Shubham, et al.
Pubblicazione: (2025)
Parameter Reduction Improves Vision Transformers: A Comparative Study of Sharing and Width Reduction
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
Developing Lightweight DNN Models With Limited Data For Real-Time Sign Language Recognition
di: Nikitin, Nikita, et al.
Pubblicazione: (2025)
di: Nikitin, Nikita, et al.
Pubblicazione: (2025)
POSESTITCH-SLT: Linguistically Inspired Pose-Stitching for End-to-End Sign Language Translation
di: Joshi, Abhinav, et al.
Pubblicazione: (2025)
di: Joshi, Abhinav, et al.
Pubblicazione: (2025)
TempoSyncDiff: Distilled Temporally-Consistent Diffusion for Low-Latency Audio-Driven Talking Head Generation
di: Mazumdar, Soumya, et al.
Pubblicazione: (2026)
di: Mazumdar, Soumya, et al.
Pubblicazione: (2026)
Test-Time Spectrum-Aware Latent Steering for Zero-Shot Generalization in Vision-Language Models
di: Dafnis, Konstantinos M., et al.
Pubblicazione: (2025)
di: Dafnis, Konstantinos M., et al.
Pubblicazione: (2025)
DistortBench: Benchmarking Vision Language Models on Image Distortion Identification
di: Goyal, Divyanshu, et al.
Pubblicazione: (2026)
di: Goyal, Divyanshu, et al.
Pubblicazione: (2026)
Analysis of Hyperparameter Optimization Effects on Lightweight Deep Models for Real-Time Image Classification
di: Rakesh, Vineet Kumar, et al.
Pubblicazione: (2025)
di: Rakesh, Vineet Kumar, et al.
Pubblicazione: (2025)
Gradient-Sign Masking for Task Vector Transport Across Pre-Trained Models
di: Rinaldi, Filippo, et al.
Pubblicazione: (2025)
di: Rinaldi, Filippo, et al.
Pubblicazione: (2025)
MRD-LiNet: A Novel Lightweight Hybrid CNN with Gradient-Guided Unlearning for Improved Drought Stress Identification
di: Patra, Aswini Kumar, et al.
Pubblicazione: (2025)
di: Patra, Aswini Kumar, et al.
Pubblicazione: (2025)
Improved Classification of Nitrogen Stress Severity in Plants Under Combined Stress Conditions Using Spatio-Temporal Deep Learning Framework
di: Patra, Aswini Kumar, et al.
Pubblicazione: (2025)
di: Patra, Aswini Kumar, et al.
Pubblicazione: (2025)
Deep Reinforcement Learning for Urban Air Quality Management: Multi-Objective Optimization of Pollution Mitigation Booth Placement in Metropolitan Environments
di: Rajesh, Kirtan, et al.
Pubblicazione: (2025)
di: Rajesh, Kirtan, et al.
Pubblicazione: (2025)
Advanced Gesture Recognition for Autism Spectrum Disorder Detection: Integrating YOLOv7, Video Augmentation, and VideoMAE for Naturalistic Video Analysis
di: Singh, Amit Kumar, et al.
Pubblicazione: (2024)
di: Singh, Amit Kumar, et al.
Pubblicazione: (2024)
Improved Cotton Leaf Disease Classification Using Parameter-Efficient Deep Learning Framework
di: Patra, Aswini Kumar, et al.
Pubblicazione: (2024)
di: Patra, Aswini Kumar, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Face Detection: Present State and Research Directions
di: Prabhat, Purnendu, et al.
Pubblicazione: (2024) -
ISLR101: an Iranian Word-Level Sign Language Recognition Dataset
di: Ranjbar, Hossein, et al.
Pubblicazione: (2025) -
iSign: A Benchmark for Indian Sign Language Processing
di: Joshi, Abhinav, et al.
Pubblicazione: (2024) -
MTCNET: Multi-task Learning Paradigm for Crowd Count Estimation
di: Kumar, Abhay, et al.
Pubblicazione: (2019) -
From Tokens to Numbers: Continuous Number Modeling for SVG Generation
di: Ogezi, Michael, et al.
Pubblicazione: (2026)