MM-WLAuslan: Multi-View Multi-Modal Word-Level Australian Sign Language Recognition Dataset
Fuente:
arXiv
Guardado en:
| Autores principales: | Shen, Xin, Du, Heming, Sheng, Hongwei, Wang, Shuyun, Chen, Hui, Chen, Huiqiang, Wu, Zhuojie, Du, Xiaobiao, Ying, Jiaying, Lu, Ruihan, Xu, Qingzheng, Yu, Xin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
3DRealCar: An In-the-wild RGB-D Car Dataset with 360-degree Views
por: Du, Xiaobiao, et al.
Publicado: (2024)
por: Du, Xiaobiao, et al.
Publicado: (2024)
MVGS: Multi-view Regulated Gaussian Splatting for Novel View Synthesis
por: Du, Xiaobiao, et al.
Publicado: (2024)
por: Du, Xiaobiao, et al.
Publicado: (2024)
CMamba: Learned Image Compression with State Space Models
por: Wu, Zhuojie, et al.
Publicado: (2025)
por: Wu, Zhuojie, et al.
Publicado: (2025)
Affective Behaviour Analysis via Integrating Multi-Modal Knowledge
por: Zhang, Wei, et al.
Publicado: (2024)
por: Zhang, Wei, et al.
Publicado: (2024)
Diverse Sign Language Translation
por: Shen, Xin, et al.
Publicado: (2024)
por: Shen, Xin, et al.
Publicado: (2024)
ResiHMR: Residual-Limb Aware Single-Image 3D Human Mesh Recovery for Individuals with Limb Loss
por: Ying, Jiaying, et al.
Publicado: (2026)
por: Ying, Jiaying, et al.
Publicado: (2026)
ISLR101: an Iranian Word-Level Sign Language Recognition Dataset
por: Ranjbar, Hossein, et al.
Publicado: (2025)
por: Ranjbar, Hossein, et al.
Publicado: (2025)
The NGT200 Dataset: Geometric Multi-View Isolated Sign Recognition
por: Ranum, Oline, et al.
Publicado: (2024)
por: Ranum, Oline, et al.
Publicado: (2024)
Dynamic Orchestration of Multi-Agent System for Real-World Multi-Image Agricultural VQA
por: Ke, Yan, et al.
Publicado: (2025)
por: Ke, Yan, et al.
Publicado: (2025)
Mobile-GS: Real-time Gaussian Splatting for Mobile Devices
por: Du, Xiaobiao, et al.
Publicado: (2026)
por: Du, Xiaobiao, et al.
Publicado: (2026)
Gesture Recognition for Prosthetic Hand Control Using Surface Electromyography Time-Frequency Images and Convolutional Neural Networks (Demonstration video and paper data)
por: Chen, Qingzheng
Publicado: (2025)
por: Chen, Qingzheng
Publicado: (2025)
MM-LIMA: Less Is More for Alignment in Multi-Modal Datasets
por: Wei, Lai, et al.
Publicado: (2023)
por: Wei, Lai, et al.
Publicado: (2023)
Generative Sign-description Prompts with Multi-positive Contrastive Learning for Sign Language Recognition
por: Liang, Siyu, et al.
Publicado: (2025)
por: Liang, Siyu, et al.
Publicado: (2025)
SignVTCL: Multi-Modal Continuous Sign Language Recognition Enhanced by Visual-Textual Contrastive Learning
por: Chen, Hao, et al.
Publicado: (2024)
por: Chen, Hao, et al.
Publicado: (2024)
Evidence-based Match-status-Aware Gait Recognition for Out-of-Gallery Gait Identification
por: Du, Heming, et al.
Publicado: (2022)
por: Du, Heming, et al.
Publicado: (2022)
MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing
por: Zheng, Junjie, et al.
Publicado: (2025)
por: Zheng, Junjie, et al.
Publicado: (2025)
DreamCar: Leveraging Car-specific Prior for in-the-wild 3D Car Reconstruction
por: Du, Xiaobiao, et al.
Publicado: (2024)
por: Du, Xiaobiao, et al.
Publicado: (2024)
MVIP -- A Dataset and Methods for Application Oriented Multi-View and Multi-Modal Industrial Part Recognition
por: Koch, Paul, et al.
Publicado: (2025)
por: Koch, Paul, et al.
Publicado: (2025)
GTPBD-MM: A Global Terraced Parcel and Boundary Dataset with Multi-Modality
por: Zhang, Zhiwei, et al.
Publicado: (2026)
por: Zhang, Zhiwei, et al.
Publicado: (2026)
Para-Lane: Multi-Lane Dataset Registering Parallel Scans for Benchmarking Novel View Synthesis
por: Ni, Ziqian, et al.
Publicado: (2025)
por: Ni, Ziqian, et al.
Publicado: (2025)
Multi-View Depth Consistent Image Generation Using Generative AI Models: Application on Architectural Design of University Buildings
por: Du, Xusheng, et al.
Publicado: (2025)
por: Du, Xusheng, et al.
Publicado: (2025)
When One Moment Isn't Enough: Multi-Moment Retrieval with Cross-Moment Interactions
por: Cao, Zhuo, et al.
Publicado: (2025)
por: Cao, Zhuo, et al.
Publicado: (2025)
CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition
por: Wang, Xu, et al.
Publicado: (2026)
por: Wang, Xu, et al.
Publicado: (2026)
MM-Point: Multi-View Information-Enhanced Multi-Modal Self-Supervised 3D Point Cloud Understanding
por: Yu, Hai-Tao, et al.
Publicado: (2024)
por: Yu, Hai-Tao, et al.
Publicado: (2024)
Pip-Stereo: Progressive Iterations Pruner for Iterative Optimization based Stereo Matching
por: Zheng, Jintu, et al.
Publicado: (2026)
por: Zheng, Jintu, et al.
Publicado: (2026)
MM-Mixing: Multi-Modal Mixing Alignment for 3D Understanding
por: Wang, Jiaze, et al.
Publicado: (2024)
por: Wang, Jiaze, et al.
Publicado: (2024)
Breaking the Barriers: Video Vision Transformers for Word-Level Sign Language Recognition
por: Brettmann, Alexander, et al.
Publicado: (2025)
por: Brettmann, Alexander, et al.
Publicado: (2025)
DiffMM: Multi-Modal Diffusion Model for Recommendation
por: Jiang, Yangqin, et al.
Publicado: (2024)
por: Jiang, Yangqin, et al.
Publicado: (2024)
BdSLW60: A Word-Level Bangla Sign Language Dataset
por: Rubaiyeat, Husne Ara, et al.
Publicado: (2024)
por: Rubaiyeat, Husne Ara, et al.
Publicado: (2024)
Hypergraph-based Multi-View Action Recognition using Event Cameras
por: Gao, Yue, et al.
Publicado: (2024)
por: Gao, Yue, et al.
Publicado: (2024)
Multi-Modal Manipulation via Multi-Modal Policy Consensus
por: Chen, Haonan, et al.
Publicado: (2025)
por: Chen, Haonan, et al.
Publicado: (2025)
Multi-Masked Querying Network for Robust Emotion Recognition from Incomplete Multi-Modal Physiological Signals
por: Xu, Geng-Xin, et al.
Publicado: (2025)
por: Xu, Geng-Xin, et al.
Publicado: (2025)
MM-SEAL: A Large-scale Video Dataset of Multi-person Multi-grained Spatio-temporally Action Localization
por: Chen, Shimin, et al.
Publicado: (2022)
por: Chen, Shimin, et al.
Publicado: (2022)
MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning
por: Xu, Tianyu, et al.
Publicado: (2025)
por: Xu, Tianyu, et al.
Publicado: (2025)
FedMM: Federated Multi-Modal Learning with Modality Heterogeneity in Computational Pathology
por: Peng, Yuanzhe, et al.
Publicado: (2024)
por: Peng, Yuanzhe, et al.
Publicado: (2024)
From Numbers to Words: Multi-Modal Bankruptcy Prediction Using the ECL Dataset
por: Arno, Henri, et al.
Publicado: (2024)
por: Arno, Henri, et al.
Publicado: (2024)
MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models
por: Yan, Bei, et al.
Publicado: (2024)
por: Yan, Bei, et al.
Publicado: (2024)
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
por: Gueuwou, Shester, et al.
Publicado: (2024)
por: Gueuwou, Shester, et al.
Publicado: (2024)
GLAM: Geometry-Guided Local Alignment for Multi-View VLP in Mammography
por: Du, Yuexi, et al.
Publicado: (2025)
por: Du, Yuexi, et al.
Publicado: (2025)
MM-HSD: Multi-Modal Hate Speech Detection in Videos
por: Céspedes-Sarrias, Berta, et al.
Publicado: (2025)
por: Céspedes-Sarrias, Berta, et al.
Publicado: (2025)
Ejemplares similares
-
3DRealCar: An In-the-wild RGB-D Car Dataset with 360-degree Views
por: Du, Xiaobiao, et al.
Publicado: (2024) -
MVGS: Multi-view Regulated Gaussian Splatting for Novel View Synthesis
por: Du, Xiaobiao, et al.
Publicado: (2024) -
CMamba: Learned Image Compression with State Space Models
por: Wu, Zhuojie, et al.
Publicado: (2025) -
Affective Behaviour Analysis via Integrating Multi-Modal Knowledge
por: Zhang, Wei, et al.
Publicado: (2024) -
Diverse Sign Language Translation
por: Shen, Xin, et al.
Publicado: (2024)