Saved in:
| Main Authors: | Yang, Gene-Ping, Tang, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2409.09646 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PhoneWorld: Scaling Phone-Use Agent Environments
by: Tang, Zhengyang, et al.
Published: (2026)
by: Tang, Zhengyang, et al.
Published: (2026)
Probing for Phonology in Self-Supervised Speech Representations: A Case Study on Accent Perception
by: Venkateswaran, Nitin, et al.
Published: (2025)
by: Venkateswaran, Nitin, et al.
Published: (2025)
TIPAA-SSL: Text Independent Phone-to-Audio Alignment based on Self-Supervised Learning and Knowledge Transfer
by: Tits, Noé, et al.
Published: (2024)
by: Tits, Noé, et al.
Published: (2024)
Maximizing Data Efficiency for Cross-Lingual TTS Adaptation by Self-Supervised Representation Mixing and Embedding Initialization
by: Huang, Wei-Ping, et al.
Published: (2024)
by: Huang, Wei-Ping, et al.
Published: (2024)
DenoSent: A Denoising Objective for Self-Supervised Sentence Representation Learning
by: Wang, Xinghao, et al.
Published: (2024)
by: Wang, Xinghao, et al.
Published: (2024)
InT: Self-Proposed Interventions Enable Credit Assignment in LLM Reasoning
by: Yang, Matthew Y. R., et al.
Published: (2026)
by: Yang, Matthew Y. R., et al.
Published: (2026)
Syllable based DNN-HMM Cantonese Speech to Text System
by: Wong, Timothy, et al.
Published: (2024)
by: Wong, Timothy, et al.
Published: (2024)
Existing LLMs Are Not Self-Consistent For Simple Tasks
by: Lin, Zhenru, et al.
Published: (2025)
by: Lin, Zhenru, et al.
Published: (2025)
The Role of Prosodic and Lexical Cues in Turn-Taking with Self-Supervised Speech Representations
by: Russell, Sam OConnor, et al.
Published: (2026)
by: Russell, Sam OConnor, et al.
Published: (2026)
Property Neurons in Self-Supervised Speech Transformers
by: Lin, Tzu-Quan, et al.
Published: (2024)
by: Lin, Tzu-Quan, et al.
Published: (2024)
PRiSM: Benchmarking Phone Realization in Speech Models
by: Bharadwaj, Shikhar, et al.
Published: (2026)
by: Bharadwaj, Shikhar, et al.
Published: (2026)
Safe, or Simply Incapable? Rethinking Safety Evaluation for Phone-Use Agents
by: Tang, Zhengyang, et al.
Published: (2026)
by: Tang, Zhengyang, et al.
Published: (2026)
Efficient Sentiment Analysis: A Resource-Aware Evaluation of Feature Extraction Techniques, Ensembling, and Deep Learning Models
by: Kamruzzaman, Mahammed, et al.
Published: (2023)
by: Kamruzzaman, Mahammed, et al.
Published: (2023)
Generating Equivalent Representations of Code By A Self-Reflection Approach
by: Li, Jia, et al.
Published: (2024)
by: Li, Jia, et al.
Published: (2024)
Identifying Speaker Information in Feed-Forward Layers of Self-Supervised Speech Transformers
by: Lin, Tzu-Quan, et al.
Published: (2025)
by: Lin, Tzu-Quan, et al.
Published: (2025)
ESURF: Simple and Effective EDU Segmentation
by: Sediqin, Mohammadreza, et al.
Published: (2025)
by: Sediqin, Mohammadreza, et al.
Published: (2025)
Exploring CLIP's Dense Knowledge for Weakly Supervised Semantic Segmentation
by: Yang, Zhiwei, et al.
Published: (2025)
by: Yang, Zhiwei, et al.
Published: (2025)
SPIN: Self-Supervised Prompt INjection
by: Zhou, Leon, et al.
Published: (2024)
by: Zhou, Leon, et al.
Published: (2024)
Self-Supervised Speech Representations are More Phonetic than Semantic
by: Choi, Kwanghee, et al.
Published: (2024)
by: Choi, Kwanghee, et al.
Published: (2024)
Prompting Techniques for Reducing Social Bias in LLMs through System 1 and System 2 Cognitive Processes
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
The Impact of Disability Disclosure on Fairness and Bias in LLM-Driven Candidate Selection
by: Kamruzzaman, Mahammed, et al.
Published: (2025)
by: Kamruzzaman, Mahammed, et al.
Published: (2025)
Do Discrete Self-Supervised Representations of Speech Capture Tone Distinctions?
by: Osakuade, Opeyemi, et al.
Published: (2024)
by: Osakuade, Opeyemi, et al.
Published: (2024)
Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization
by: Lin, Tzu-Quan, et al.
Published: (2025)
by: Lin, Tzu-Quan, et al.
Published: (2025)
No Train, all Gain: Self-Supervised Gradients Improve Deep Frozen Representations
by: Simoncini, Walter, et al.
Published: (2024)
by: Simoncini, Walter, et al.
Published: (2024)
The Flip Side of RLHF: On-Policy Feedback for Reward Model Self-Supervised Improvement
by: Wang, Xiaobo, et al.
Published: (2026)
by: Wang, Xiaobo, et al.
Published: (2026)
SDA: Simple Discrete Augmentation for Contrastive Sentence Representation Learning
by: Zhu, Dongsheng, et al.
Published: (2022)
by: Zhu, Dongsheng, et al.
Published: (2022)
Is Smaller Always Faster? Tradeoffs in Compressing Self-Supervised Speech Transformers
by: Lin, Tzu-Quan, et al.
Published: (2022)
by: Lin, Tzu-Quan, et al.
Published: (2022)
Graph Integrated Language Transformers for Next Action Prediction in Complex Phone Calls
by: Marani, Amin Hosseiny, et al.
Published: (2024)
by: Marani, Amin Hosseiny, et al.
Published: (2024)
Self-Supervised Prompt Optimization
by: Xiang, Jinyu, et al.
Published: (2025)
by: Xiang, Jinyu, et al.
Published: (2025)
An Empirical Recipe for Universal Phone Recognition
by: Bharadwaj, Shikhar, et al.
Published: (2026)
by: Bharadwaj, Shikhar, et al.
Published: (2026)
Do Phone-Use Agents Respect Your Privacy?
by: Tang, Zhengyang, et al.
Published: (2026)
by: Tang, Zhengyang, et al.
Published: (2026)
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation
by: Barsellotti, Luca, et al.
Published: (2024)
by: Barsellotti, Luca, et al.
Published: (2024)
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
by: Abdin, Marah, et al.
Published: (2024)
by: Abdin, Marah, et al.
Published: (2024)
Representation Learning for Weakly Supervised Relation Extraction
by: Li, Zhuang
Published: (2021)
by: Li, Zhuang
Published: (2021)
Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision
by: He, Yinghui, et al.
Published: (2026)
by: He, Yinghui, et al.
Published: (2026)
Embarrassingly Simple Self-Distillation Improves Code Generation
by: Zhang, Ruixiang, et al.
Published: (2026)
by: Zhang, Ruixiang, et al.
Published: (2026)
Transformer-Lite: High-efficiency Deployment of Large Language Models on Mobile Phone GPUs
by: Li, Luchang, et al.
Published: (2024)
by: Li, Luchang, et al.
Published: (2024)
SHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction
by: Gueuwou, Shester, et al.
Published: (2024)
by: Gueuwou, Shester, et al.
Published: (2024)
SSTAG: Structure-Aware Self-Supervised Learning Method for Text-Attributed Graphs
by: Liu, Ruyue, et al.
Published: (2025)
by: Liu, Ruyue, et al.
Published: (2025)
Hierarchical Self-Supervised Representation Learning for Depression Detection from Speech
by: Li, Yuxin, et al.
Published: (2025)
by: Li, Yuxin, et al.
Published: (2025)
Similar Items
-
PhoneWorld: Scaling Phone-Use Agent Environments
by: Tang, Zhengyang, et al.
Published: (2026) -
Probing for Phonology in Self-Supervised Speech Representations: A Case Study on Accent Perception
by: Venkateswaran, Nitin, et al.
Published: (2025) -
TIPAA-SSL: Text Independent Phone-to-Audio Alignment based on Self-Supervised Learning and Knowledge Transfer
by: Tits, Noé, et al.
Published: (2024) -
Maximizing Data Efficiency for Cross-Lingual TTS Adaptation by Self-Supervised Representation Mixing and Embedding Initialization
by: Huang, Wei-Ping, et al.
Published: (2024) -
DenoSent: A Denoising Objective for Self-Supervised Sentence Representation Learning
by: Wang, Xinghao, et al.
Published: (2024)