Repository logo
Communities & Collections
Research Outputs
Fundings & Projects
People
Statistics
New user? Click here to register.Have you forgotten your password?
  1. Home
  2. KMITL
  3. Publication
  4. A hybrid of fractal code descriptor and harmonic pattern generator for improving speech recognition of different sampling rates
Loading...
Thumbnail Image

A hybrid of fractal code descriptor and harmonic pattern generator for improving speech recognition of different sampling rates

Author(s)
Hokking, Rattaphon
Woraratpanya, Kuntpong
Date Issued
January 1, 2018
Type
Conference Paper
DOI
10.1007/978-3-319-60663-7_4
Abstract
Currently, the different sampling rate for speech recognition is a grand challenge due to supporting applications of divergent platform devices, such as mobile device interaction, interactive voice response system, voice search, voice dictation and voice identification. Furthermore, such applications require efficient speech features to represent input signals. However, the different sampling rates of speech signals lead to the different features. This phenomenon comes from speech harmonic signal lost. It becomes a key factor that decreases the speech recognition rate. Therefore, this paper proposes a hybrid of fractal code descriptor and harmonic pattern generator to convert all different sampling rate signals to standardized signals. In this method, an independent resolution property of fractal code descriptor is applied to training and testing speech signals. Then, the pitches of such signals are used to recover harmonic pattern of lost signals. This method can effectively reconstruct speech signals at any sampling rates. When its performance is evaluated with AN4 corpus of CMU Sphinx speech recognition engine, the experimental results show that the proposed method can significantly improve the speech recognition rate, even if the sampling rate of testing speeches differs from that of training speeches.
Citation
Advances in Intelligent Systems and Computing, 566, 32-42, 2018
Subjects

Different sampling ra...

Fractal code descript...

Harmonic reconstructi...

Mel frequency cepstra...

Resolution independen...

Speech recognition

Metrics
Get Involved!
  • Source Code
  • Documentation
  • Slack Channel
Make it your own

DSpace-CRIS can be extensively configured to meet your needs. Decide which information need to be collected and available with fine-grained security. Start updating the theme to match your Institution's web identity.

Need professional help?

The original creators of DSpace-CRIS at 4Science can take your project to the next level, get in touch!

Built with DSpace-CRIS software - Extension maintained and optimized by 4Science

  • Accessibility settings
  • Privacy policy
  • End User Agreement
  • Send Feedback