专题:Speech Recognition and Synthesis

This cluster of papers focuses on the advances in speech recognition technology, covering topics such as acoustic modeling using deep neural networks, speaker verification, convolutional neural networks for speech recognition, end-to-end speech recognition systems, hidden Markov models, sequence-to-sequence models, automatic speech recognition, speaker diarization, and statistical language modeling.
最新文献
近5年高被引文献
Kaldi Speech Recognition Toolkit

article Full Text OpenAlex 4894 FWCI9.1983

LLaMA: Open and Efficient Foundation Language Models

preprint Full Text OpenAlex 3949 FWCI0

WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing

article Full Text OpenAlex 1814 FWCI170.2159

Enhancements in Immediate Speech Emotion Detection: Harnessing Prosodic and Spectral Characteristics

article Full Text OpenAlex 1663 FWCI717.999

Robust Speech Recognition via Large-Scale Weak Supervision

preprint Full Text OpenAlex 1176 FWCI0

ICASSP 2022 - 2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

paratext Full Text OpenAlex 926 FWCI0

Perceptual Signal Analysis and Eigenvalue-Weighted Metrics for Musical Audio Quality Assessment

conference-paper Full Text OpenAlex 920 FWCI0

Spoken Language Processing

book-chapter Full Text OpenAlex 803 FWCI1.7723

Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?

conference-paper Full Text OpenAlex 675 FWCI100.0294

Autoencoders and their applications in machine learning: a survey

article Full Text OpenAlex 562 FWCI128.7764