AugmentGest: Can Random Data Cropping Augmentation Boost Gesture Recognition Performance?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Aboudeshish, Nada, Ignatov, Dmitry, Timofte, Radu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Virtually Enriched NYU Depth V2 Dataset for Monocular Depth Estimation: Do We Need Artificial Augmentation?
von: Ignatov, Dmitry, et al.
Veröffentlicht: (2024)
von: Ignatov, Dmitry, et al.
Veröffentlicht: (2024)
A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks
von: Khalid, Waleed, et al.
Veröffentlicht: (2025)
von: Khalid, Waleed, et al.
Veröffentlicht: (2025)
From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs
von: Shrestha, Usha, et al.
Veröffentlicht: (2026)
von: Shrestha, Usha, et al.
Veröffentlicht: (2026)
From Code to Prediction: Fine-Tuning LLMs for Neural Network Performance Classification in NNGPT
von: Hanouneh, Mahmoud, et al.
Veröffentlicht: (2026)
von: Hanouneh, Mahmoud, et al.
Veröffentlicht: (2026)
Closed-Loop LLM Discovery of Non-Standard Channel Priors in Vision Models
von: Uzun, Tolgay Atinc, et al.
Veröffentlicht: (2026)
von: Uzun, Tolgay Atinc, et al.
Veröffentlicht: (2026)
Preparation of Fractal-Inspired Computational Architectures for Advanced Large Language Model Analysis
von: Mittal, Yash, et al.
Veröffentlicht: (2025)
von: Mittal, Yash, et al.
Veröffentlicht: (2025)
From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures
von: Khalid, Waleed, et al.
Veröffentlicht: (2026)
von: Khalid, Waleed, et al.
Veröffentlicht: (2026)
LLM as a Neural Architect: Controlled Generation of Image Captioning Models Under Strict API Contracts
von: Jesani, Krunal, et al.
Veröffentlicht: (2025)
von: Jesani, Krunal, et al.
Veröffentlicht: (2025)
Real Image Denoising with Knowledge Distillation for High-Performance Mobile NPUs
von: Kayani, Faraz, et al.
Veröffentlicht: (2026)
von: Kayani, Faraz, et al.
Veröffentlicht: (2026)
MobileAgeNet: Lightweight Facial Age Estimation for Mobile Deployment
von: Kumar, Arun, et al.
Veröffentlicht: (2026)
von: Kumar, Arun, et al.
Veröffentlicht: (2026)
Delta-Based Neural Architecture Search: LLM Fine-Tuning via Code Diffs
von: Adhikari, Santosh Premi, et al.
Veröffentlicht: (2026)
von: Adhikari, Santosh Premi, et al.
Veröffentlicht: (2026)
Enhancing LLM-Based Neural Network Generation: Few-Shot Prompting and Efficient Validation for Automated Architecture Design
von: Duvvuri, Raghuvir, et al.
Veröffentlicht: (2025)
von: Duvvuri, Raghuvir, et al.
Veröffentlicht: (2025)
GestFormer: Multiscale Wavelet Pooling Transformer Network for Dynamic Hand Gesture Recognition
von: Garg, Mallika, et al.
Veröffentlicht: (2024)
von: Garg, Mallika, et al.
Veröffentlicht: (2024)
VIST-GPT: Ushering in the Era of Visual Storytelling with LLMs?
von: Gado, Mohamed, et al.
Veröffentlicht: (2025)
von: Gado, Mohamed, et al.
Veröffentlicht: (2025)
The Return of Structural Handwritten Mathematical Expression Recognition
von: Seitz, Jakob, et al.
Veröffentlicht: (2025)
von: Seitz, Jakob, et al.
Veröffentlicht: (2025)
Boosting Gesture Recognition with an Automatic Gesture Annotation Framework
von: Shen, Junxiao, et al.
Veröffentlicht: (2024)
von: Shen, Junxiao, et al.
Veröffentlicht: (2024)
HoloGest: Decoupled Diffusion and Motion Priors for Generating Holisticly Expressive Co-speech Gestures
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
Pose2Gest: A Few-Shot Model-Free Approach Applied In South Indian Classical Dance Gesture Recognition
von: Raju, Kavitha, et al.
Veröffentlicht: (2024)
von: Raju, Kavitha, et al.
Veröffentlicht: (2024)
Learned Lightweight Smartphone ISP with Unpaired Data
von: Arhire, Andrei, et al.
Veröffentlicht: (2025)
von: Arhire, Andrei, et al.
Veröffentlicht: (2025)
Practical Manipulation Model for Robust Deepfake Detection
von: Hopf, Benedikt, et al.
Veröffentlicht: (2025)
von: Hopf, Benedikt, et al.
Veröffentlicht: (2025)
PersonaGest: Personalized Co-Speech Gesture Generation with Semantic-Guided Hierarchical Motion Representation
von: Zhao, Junchuan, et al.
Veröffentlicht: (2026)
von: Zhao, Junchuan, et al.
Veröffentlicht: (2026)
Experts-Guided Unbalanced Optimal Transport for ISP Learning from Unpaired and/or Paired Data
von: Perevozchikov, Georgy, et al.
Veröffentlicht: (2025)
von: Perevozchikov, Georgy, et al.
Veröffentlicht: (2025)
Generative Data Augmentation for Skeleton Action Recognition
von: Dong, Xu, et al.
Veröffentlicht: (2026)
von: Dong, Xu, et al.
Veröffentlicht: (2026)
MuDreamer: Learning Predictive World Models without Reconstruction
von: Burchi, Maxime, et al.
Veröffentlicht: (2024)
von: Burchi, Maxime, et al.
Veröffentlicht: (2024)
Data Augmentation Through Random Style Replacement
von: Yang, Qikai, et al.
Veröffentlicht: (2025)
von: Yang, Qikai, et al.
Veröffentlicht: (2025)
Cat: Post-Training Quantization Error Reduction via Cluster-based Affine Transformation
von: Zoljodi, Ali, et al.
Veröffentlicht: (2025)
von: Zoljodi, Ali, et al.
Veröffentlicht: (2025)
The Regularizing Power of Language-Training Deepfake Detectors
von: Hopf, Benedikt, et al.
Veröffentlicht: (2026)
von: Hopf, Benedikt, et al.
Veröffentlicht: (2026)
Sample-aware RandAugment: Search-free Automatic Data Augmentation for Effective Image Recognition
von: Xiao, Anqi, et al.
Veröffentlicht: (2025)
von: Xiao, Anqi, et al.
Veröffentlicht: (2025)
AugGen: Synthetic Augmentation using Diffusion Models Can Improve Recognition
von: Rahimi, Parsa, et al.
Veröffentlicht: (2025)
von: Rahimi, Parsa, et al.
Veröffentlicht: (2025)
Higher fidelity perceptual image and video compression with a latent conditioned residual denoising diffusion model
von: Brenig, Jonas, et al.
Veröffentlicht: (2025)
von: Brenig, Jonas, et al.
Veröffentlicht: (2025)
The Ultimate Combo: Boosting Adversarial Example Transferability by Composing Data Augmentations
von: Yun, Zebin, et al.
Veröffentlicht: (2023)
von: Yun, Zebin, et al.
Veröffentlicht: (2023)
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
von: Geigle, Gregor, et al.
Veröffentlicht: (2023)
von: Geigle, Gregor, et al.
Veröffentlicht: (2023)
Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
SASG-DA: Sparse-Aware Semantic-Guided Diffusion Augmentation For Myoelectric Gesture Recognition
von: Liu, Chen, et al.
Veröffentlicht: (2025)
von: Liu, Chen, et al.
Veröffentlicht: (2025)
Thinking Beyond Labels: Vocabulary-Free Fine-Grained Recognition using Reasoning-Augmented LMMs
von: Demidov, Dmitry, et al.
Veröffentlicht: (2025)
von: Demidov, Dmitry, et al.
Veröffentlicht: (2025)
You Only Need Half: Boosting Data Augmentation by Using Partial Content
von: Hu, Juntao, et al.
Veröffentlicht: (2024)
von: Hu, Juntao, et al.
Veröffentlicht: (2024)
Learning Transformer-based World Models with Contrastive Predictive Coding
von: Burchi, Maxime, et al.
Veröffentlicht: (2025)
von: Burchi, Maxime, et al.
Veröffentlicht: (2025)
Accurate and Efficient World Modeling with Masked Latent Transformers
von: Burchi, Maxime, et al.
Veröffentlicht: (2025)
von: Burchi, Maxime, et al.
Veröffentlicht: (2025)
MixCut:A Data Augmentation Method for Facial Expression Recognition
von: Yu, Jiaxiang, et al.
Veröffentlicht: (2024)
von: Yu, Jiaxiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Virtually Enriched NYU Depth V2 Dataset for Monocular Depth Estimation: Do We Need Artificial Augmentation?
von: Ignatov, Dmitry, et al.
Veröffentlicht: (2024) -
A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks
von: Khalid, Waleed, et al.
Veröffentlicht: (2025) -
From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs
von: Shrestha, Usha, et al.
Veröffentlicht: (2026) -
From Code to Prediction: Fine-Tuning LLMs for Neural Network Performance Classification in NNGPT
von: Hanouneh, Mahmoud, et al.
Veröffentlicht: (2026) -
Closed-Loop LLM Discovery of Non-Standard Channel Priors in Vision Models
von: Uzun, Tolgay Atinc, et al.
Veröffentlicht: (2026)