Exploring VQ-VAE with Prosody Parameters for Speaker Anonymization
Fuente:
arXiv
Guardado en:
| Autores principales: | Leang, Sotheara, Augusma, Anderson, Castelli, Eric, Letué, Frédérique, Sam, Sethserey, Vaufreydaz, Dominique |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Exploring Dynamic Parameters for Vietnamese Gender-Independent ASR
por: Leang, Sotheara, et al.
Publicado: (2025)
por: Leang, Sotheara, et al.
Publicado: (2025)
Variational Encoder--Multi-Decoder (VE-MD) for Privacy-by-functional-design (Group) Emotion Recognition
por: Augusma, Anderson, et al.
Publicado: (2026)
por: Augusma, Anderson, et al.
Publicado: (2026)
A PolSAR Scattering Power Factorization Framework and Novel Roll-Invariant Parameters Based Unsupervised Classification Scheme Using a Geodesic Distance
por: Ratha, Debanshu, et al.
Publicado: (2019)
por: Ratha, Debanshu, et al.
Publicado: (2019)
Towards Task-Compatible Compressible Representations
por: de Andrade, Anderson, et al.
Publicado: (2024)
por: de Andrade, Anderson, et al.
Publicado: (2024)
Neural Electromagnetic Fields for High-Resolution Material Parameter Reconstruction
por: Chen, Zhe, et al.
Publicado: (2026)
por: Chen, Zhe, et al.
Publicado: (2026)
Exploring Textual Semantics Diversity for Image Transmission in Semantic Communication Systems using Visual Language Model
por: Huang, Peishan, et al.
Publicado: (2025)
por: Huang, Peishan, et al.
Publicado: (2025)
A Single-Parameter Factor-Graph Image Prior
por: Wang, Tianyang, et al.
Publicado: (2026)
por: Wang, Tianyang, et al.
Publicado: (2026)
A Renderer-Enabled Framework for Computing Parameter Estimation Lower Bounds in Plenoptic Imaging Systems
por: Sambasivan, Abhinav V., et al.
Publicado: (2026)
por: Sambasivan, Abhinav V., et al.
Publicado: (2026)
PCE-GAN: A Generative Adversarial Network for Point Cloud Attribute Quality Enhancement based on Optimal Transport
por: Guo, Tian, et al.
Publicado: (2025)
por: Guo, Tian, et al.
Publicado: (2025)
D-CNN and VQ-VAE Autoencoders for Compression and Denoising of Industrial X-ray Computed Tomography Images
por: Hejazi, Bardia, et al.
Publicado: (2025)
por: Hejazi, Bardia, et al.
Publicado: (2025)
Exploring Sparsity for Parameter Efficient Fine Tuning Using Wavelets
por: Bilican, Ahmet, et al.
Publicado: (2025)
por: Bilican, Ahmet, et al.
Publicado: (2025)
A Wearable Gait Monitoring System for 17 Gait Parameters Based on Computer Vision
por: Chen, Jiangang, et al.
Publicado: (2024)
por: Chen, Jiangang, et al.
Publicado: (2024)
A Comprehensive Multi-scale Approach for Speech and Dynamics Synchrony in Talking Head Generation
por: Airale, Louis, et al.
Publicado: (2023)
por: Airale, Louis, et al.
Publicado: (2023)
SRViT: Vision Transformers for Estimating Radar Reflectivity from Satellite Observations at Scale
por: Stock, Jason, et al.
Publicado: (2024)
por: Stock, Jason, et al.
Publicado: (2024)
Exploring Challenges in Deep Learning of Single-Station Ground Motion Records
por: Çağlar, Ümit Mert, et al.
Publicado: (2024)
por: Çağlar, Ümit Mert, et al.
Publicado: (2024)
OG-PCL: Efficient Sparse Point Cloud Processing for Human Activity Recognition
por: Yan, Jiuqi, et al.
Publicado: (2025)
por: Yan, Jiuqi, et al.
Publicado: (2025)
3D Field of Junctions: A Noise-Robust, Training-Free Structural Prior for Volumetric Inverse Problems
por: Kim, Namhoon, et al.
Publicado: (2026)
por: Kim, Namhoon, et al.
Publicado: (2026)
Scene Understanding Enabled Semantic Communication with Open Channel Coding
por: Xiang, Zhe, et al.
Publicado: (2025)
por: Xiang, Zhe, et al.
Publicado: (2025)
Hierarchical Attention Networks for Lossless Point Cloud Attribute Compression
por: Chen, Yueru, et al.
Publicado: (2025)
por: Chen, Yueru, et al.
Publicado: (2025)
Collaborative Perception for Connected and Autonomous Driving: Challenges, Possible Solutions and Opportunities
por: Hu, Senkang, et al.
Publicado: (2024)
por: Hu, Senkang, et al.
Publicado: (2024)
Radar-Based NLoS Pedestrian Localization for Darting-Out Scenarios Near Parked Vehicles with Camera-Assisted Point Cloud Interpretation
por: Kim, Hee-Yeun, et al.
Publicado: (2025)
por: Kim, Hee-Yeun, et al.
Publicado: (2025)
Perspective-aware fusion of incomplete depth maps and surface normals for accurate 3D reconstruction
por: Hlinka, Ondrej, et al.
Publicado: (2026)
por: Hlinka, Ondrej, et al.
Publicado: (2026)
AI- Enhanced Stethoscope in Remote Diagnostics for Cardiopulmonary Diseases
por: Ghouse, Hania, et al.
Publicado: (2025)
por: Ghouse, Hania, et al.
Publicado: (2025)
BenchHAR: Benchmarking Self-Supervised Learning for Generalizable Sensor-based Activity Recognition
por: Cai, Yize, et al.
Publicado: (2026)
por: Cai, Yize, et al.
Publicado: (2026)
Role and Integration of Image Processing Systems in Maritime Target Tracking
por: Zardoua, Yassir, et al.
Publicado: (2022)
por: Zardoua, Yassir, et al.
Publicado: (2022)
A Hybrid Quantum-Classical Approach based on the Hadamard Transform for the Convolutional Layer
por: Pan, Hongyi, et al.
Publicado: (2023)
por: Pan, Hongyi, et al.
Publicado: (2023)
mmAnomaly: Leveraging Visual Context for Robust Anomaly Detection in the Non-Visual World with mmWave Radar
por: Toha, Tarik Reza, et al.
Publicado: (2026)
por: Toha, Tarik Reza, et al.
Publicado: (2026)
Surface Recognition for e-Scooter Using Smartphone IMU Sensor
por: Eweida, Areej, et al.
Publicado: (2023)
por: Eweida, Areej, et al.
Publicado: (2023)
Deep Inertial Pose: A deep learning approach for human pose estimation
por: Cerqueira, Sara M., et al.
Publicado: (2025)
por: Cerqueira, Sara M., et al.
Publicado: (2025)
Neural-HAR: A Dimension-Gated CNN Accelerator for Real-Time Radar Human Activity Recognition
por: Wu, Yizhuo, et al.
Publicado: (2025)
por: Wu, Yizhuo, et al.
Publicado: (2025)
Quality-controlled registration of urban MLS point clouds reducing drift effects by adaptive fragmentation
por: Rincon, Marco Antonio Ortiz, et al.
Publicado: (2025)
por: Rincon, Marco Antonio Ortiz, et al.
Publicado: (2025)
The frame-level leakage trap: rethinking evaluation protocols for intrinsic image decomposition, with source-separable uncertainty as a case study
por: Woo, Jihwan
Publicado: (2026)
por: Woo, Jihwan
Publicado: (2026)
Visible Light Positioning With Lamé Curve LEDs: A Generic Approach for Camera Pose Estimation
por: Pan, Wenxuan, et al.
Publicado: (2026)
por: Pan, Wenxuan, et al.
Publicado: (2026)
Vision-Language Based Expert Reporting for Painting Authentication and Defect Detection
por: Ouda, Eman, et al.
Publicado: (2026)
por: Ouda, Eman, et al.
Publicado: (2026)
Knowledge Distillation of Convolutional Neural Networks through Feature Map Transformation using Decision Trees
por: Srinivas, Maddimsetti, et al.
Publicado: (2024)
por: Srinivas, Maddimsetti, et al.
Publicado: (2024)
A Robust Pipeline for Classification and Detection of Bleeding Frames in Wireless Capsule Endoscopy using Swin Transformer and RT-DETR
por: Alavala, Sasidhar, et al.
Publicado: (2024)
por: Alavala, Sasidhar, et al.
Publicado: (2024)
Wi-CBR: Salient-aware Adaptive WiFi Sensing for Cross-domain Behavior Recognition
por: Zhang, Ruobei, et al.
Publicado: (2025)
por: Zhang, Ruobei, et al.
Publicado: (2025)
Quality-Aware Framework for Video-Derived Respiratory Signals
por: Nguyen, Nhi, et al.
Publicado: (2025)
por: Nguyen, Nhi, et al.
Publicado: (2025)
Efficiency vs. Efficacy: Assessing the Compression Ratio-Dice Score Relationship through a Simple Benchmarking Framework for Cerebrovascular 3D Segmentation
por: Elbana, Shimaa, et al.
Publicado: (2025)
por: Elbana, Shimaa, et al.
Publicado: (2025)
DeepEyeNet: Adaptive Genetic Bayesian Algorithm Based Hybrid ConvNeXtTiny Framework For Multi-Feature Glaucoma Eye Diagnosis
por: Roy, Angshuman, et al.
Publicado: (2025)
por: Roy, Angshuman, et al.
Publicado: (2025)
Ejemplares similares
-
Exploring Dynamic Parameters for Vietnamese Gender-Independent ASR
por: Leang, Sotheara, et al.
Publicado: (2025) -
Variational Encoder--Multi-Decoder (VE-MD) for Privacy-by-functional-design (Group) Emotion Recognition
por: Augusma, Anderson, et al.
Publicado: (2026) -
A PolSAR Scattering Power Factorization Framework and Novel Roll-Invariant Parameters Based Unsupervised Classification Scheme Using a Geodesic Distance
por: Ratha, Debanshu, et al.
Publicado: (2019) -
Towards Task-Compatible Compressible Representations
por: de Andrade, Anderson, et al.
Publicado: (2024) -
Neural Electromagnetic Fields for High-Resolution Material Parameter Reconstruction
por: Chen, Zhe, et al.
Publicado: (2026)