A Simple HMM with Self-Supervised Representations for Phone Segmentation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yang, Gene-Ping, Tang, Hao |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
PhoneWorld: Scaling Phone-Use Agent Environments
par: Tang, Zhengyang, et autres
Publié: (2026)
par: Tang, Zhengyang, et autres
Publié: (2026)
Maximizing Data Efficiency for Cross-Lingual TTS Adaptation by Self-Supervised Representation Mixing and Embedding Initialization
par: Huang, Wei-Ping, et autres
Publié: (2024)
par: Huang, Wei-Ping, et autres
Publié: (2024)
Probing for Phonology in Self-Supervised Speech Representations: A Case Study on Accent Perception
par: Venkateswaran, Nitin, et autres
Publié: (2025)
par: Venkateswaran, Nitin, et autres
Publié: (2025)
TIPAA-SSL: Text Independent Phone-to-Audio Alignment based on Self-Supervised Learning and Knowledge Transfer
par: Tits, Noé, et autres
Publié: (2024)
par: Tits, Noé, et autres
Publié: (2024)
DenoSent: A Denoising Objective for Self-Supervised Sentence Representation Learning
par: Wang, Xinghao, et autres
Publié: (2024)
par: Wang, Xinghao, et autres
Publié: (2024)
Existing LLMs Are Not Self-Consistent For Simple Tasks
par: Lin, Zhenru, et autres
Publié: (2025)
par: Lin, Zhenru, et autres
Publié: (2025)
The Role of Prosodic and Lexical Cues in Turn-Taking with Self-Supervised Speech Representations
par: Russell, Sam OConnor, et autres
Publié: (2026)
par: Russell, Sam OConnor, et autres
Publié: (2026)
PRiSM: Benchmarking Phone Realization in Speech Models
par: Bharadwaj, Shikhar, et autres
Publié: (2026)
par: Bharadwaj, Shikhar, et autres
Publié: (2026)
InT: Self-Proposed Interventions Enable Credit Assignment in LLM Reasoning
par: Yang, Matthew Y. R., et autres
Publié: (2026)
par: Yang, Matthew Y. R., et autres
Publié: (2026)
Syllable based DNN-HMM Cantonese Speech to Text System
par: Wong, Timothy, et autres
Publié: (2024)
par: Wong, Timothy, et autres
Publié: (2024)
Property Neurons in Self-Supervised Speech Transformers
par: Lin, Tzu-Quan, et autres
Publié: (2024)
par: Lin, Tzu-Quan, et autres
Publié: (2024)
Generating Equivalent Representations of Code By A Self-Reflection Approach
par: Li, Jia, et autres
Publié: (2024)
par: Li, Jia, et autres
Publié: (2024)
ESURF: Simple and Effective EDU Segmentation
par: Sediqin, Mohammadreza, et autres
Publié: (2025)
par: Sediqin, Mohammadreza, et autres
Publié: (2025)
Efficient Sentiment Analysis: A Resource-Aware Evaluation of Feature Extraction Techniques, Ensembling, and Deep Learning Models
par: Kamruzzaman, Mahammed, et autres
Publié: (2023)
par: Kamruzzaman, Mahammed, et autres
Publié: (2023)
Safe, or Simply Incapable? Rethinking Safety Evaluation for Phone-Use Agents
par: Tang, Zhengyang, et autres
Publié: (2026)
par: Tang, Zhengyang, et autres
Publié: (2026)
SPIN: Self-Supervised Prompt INjection
par: Zhou, Leon, et autres
Publié: (2024)
par: Zhou, Leon, et autres
Publié: (2024)
Identifying Speaker Information in Feed-Forward Layers of Self-Supervised Speech Transformers
par: Lin, Tzu-Quan, et autres
Publié: (2025)
par: Lin, Tzu-Quan, et autres
Publié: (2025)
SDA: Simple Discrete Augmentation for Contrastive Sentence Representation Learning
par: Zhu, Dongsheng, et autres
Publié: (2022)
par: Zhu, Dongsheng, et autres
Publié: (2022)
The Flip Side of RLHF: On-Policy Feedback for Reward Model Self-Supervised Improvement
par: Wang, Xiaobo, et autres
Publié: (2026)
par: Wang, Xiaobo, et autres
Publié: (2026)
Self-Supervised Speech Representations are More Phonetic than Semantic
par: Choi, Kwanghee, et autres
Publié: (2024)
par: Choi, Kwanghee, et autres
Publié: (2024)
Representation Learning for Weakly Supervised Relation Extraction
par: Li, Zhuang
Publié: (2021)
par: Li, Zhuang
Publié: (2021)
Do Discrete Self-Supervised Representations of Speech Capture Tone Distinctions?
par: Osakuade, Opeyemi, et autres
Publié: (2024)
par: Osakuade, Opeyemi, et autres
Publié: (2024)
Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision
par: He, Yinghui, et autres
Publié: (2026)
par: He, Yinghui, et autres
Publié: (2026)
Exploring CLIP's Dense Knowledge for Weakly Supervised Semantic Segmentation
par: Yang, Zhiwei, et autres
Publié: (2025)
par: Yang, Zhiwei, et autres
Publié: (2025)
Graph Integrated Language Transformers for Next Action Prediction in Complex Phone Calls
par: Marani, Amin Hosseiny, et autres
Publié: (2024)
par: Marani, Amin Hosseiny, et autres
Publié: (2024)
Embarrassingly Simple Self-Distillation Improves Code Generation
par: Zhang, Ruixiang, et autres
Publié: (2026)
par: Zhang, Ruixiang, et autres
Publié: (2026)
No Train, all Gain: Self-Supervised Gradients Improve Deep Frozen Representations
par: Simoncini, Walter, et autres
Publié: (2024)
par: Simoncini, Walter, et autres
Publié: (2024)
Prompting Techniques for Reducing Social Bias in LLMs through System 1 and System 2 Cognitive Processes
par: Kamruzzaman, Mahammed, et autres
Publié: (2024)
par: Kamruzzaman, Mahammed, et autres
Publié: (2024)
The Impact of Disability Disclosure on Fairness and Bias in LLM-Driven Candidate Selection
par: Kamruzzaman, Mahammed, et autres
Publié: (2025)
par: Kamruzzaman, Mahammed, et autres
Publié: (2025)
SSTAG: Structure-Aware Self-Supervised Learning Method for Text-Attributed Graphs
par: Liu, Ruyue, et autres
Publié: (2025)
par: Liu, Ruyue, et autres
Publié: (2025)
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation
par: Barsellotti, Luca, et autres
Publié: (2024)
par: Barsellotti, Luca, et autres
Publié: (2024)
Transformer-Lite: High-efficiency Deployment of Large Language Models on Mobile Phone GPUs
par: Li, Luchang, et autres
Publié: (2024)
par: Li, Luchang, et autres
Publié: (2024)
Self-Supervised Prompt Optimization
par: Xiang, Jinyu, et autres
Publié: (2025)
par: Xiang, Jinyu, et autres
Publié: (2025)
SHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction
par: Gueuwou, Shester, et autres
Publié: (2024)
par: Gueuwou, Shester, et autres
Publié: (2024)
Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization
par: Lin, Tzu-Quan, et autres
Publié: (2025)
par: Lin, Tzu-Quan, et autres
Publié: (2025)
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
par: Abdin, Marah, et autres
Publié: (2024)
par: Abdin, Marah, et autres
Publié: (2024)
Re-Examine Distantly Supervised NER: A New Benchmark and a Simple Approach
par: Li, Yuepei, et autres
Publié: (2024)
par: Li, Yuepei, et autres
Publié: (2024)
SSFO: Self-Supervised Faithfulness Optimization for Retrieval-Augmented Generation
par: Tang, Xiaqiang, et autres
Publié: (2025)
par: Tang, Xiaqiang, et autres
Publié: (2025)
Is Smaller Always Faster? Tradeoffs in Compressing Self-Supervised Speech Transformers
par: Lin, Tzu-Quan, et autres
Publié: (2022)
par: Lin, Tzu-Quan, et autres
Publié: (2022)
RulePrompt: Weakly Supervised Text Classification with Prompting PLMs and Self-Iterative Logical Rules
par: Li, Miaomiao, et autres
Publié: (2024)
par: Li, Miaomiao, et autres
Publié: (2024)
Documents similaires
-
PhoneWorld: Scaling Phone-Use Agent Environments
par: Tang, Zhengyang, et autres
Publié: (2026) -
Maximizing Data Efficiency for Cross-Lingual TTS Adaptation by Self-Supervised Representation Mixing and Embedding Initialization
par: Huang, Wei-Ping, et autres
Publié: (2024) -
Probing for Phonology in Self-Supervised Speech Representations: A Case Study on Accent Perception
par: Venkateswaran, Nitin, et autres
Publié: (2025) -
TIPAA-SSL: Text Independent Phone-to-Audio Alignment based on Self-Supervised Learning and Knowledge Transfer
par: Tits, Noé, et autres
Publié: (2024) -
DenoSent: A Denoising Objective for Self-Supervised Sentence Representation Learning
par: Wang, Xinghao, et autres
Publié: (2024)