Acoustic Model - Desktop-based Speech Recognition

Desktop-based Speech Recognition

For speech recognition on a standard desktop PC, the limiting factor is the sound card. Most sound cards today can record at sampling rates of between 16 kHz-48 kHz of audio, with bit rates of 8 to 16-bits per sample, and playback at up to 96 kHz.

As a general rule, a speech recognition engine works better with acoustic models trained with speech audio data recorded at higher sampling rates/bits per sample. But using audio with too high a sampling rate/bits per sample can slow the recognition engine down. A compromise is needed. Thus for desktop speech recognition, the current standard is acoustic models trained with speech audio data recorded at sampling rates of 16 kHz/16bits per sample.

Read more about this topic:  Acoustic Model

Famous quotes containing the words speech and/or recognition:

    His speech is a burning fire;
    With his lips he travaileth;
    In his heart is a blind desire,
    In his eyes foreknowledge of death:
    He weaves, and is clothed with derision;
    Sows, and he shall not reap;
    His life is a watch or a vision
    Between a sleep and a sleep.
    —A.C. (Algernon Charles)

    By now, legions of tireless essayists and op-ed columnists have dressed feminists down for making such a fuss about entering the professions and earning equal pay that everyone’s attention has been distracted from the important contributions of mothers working at home. This judgment presumes, of course, that prior to the resurgence of feminism in the ‘70s, housewives and mothers enjoyed wide recognition and honor. This was not exactly the case.
    Mary Kay Blakely (20th century)