US2003225578A1PendingUtilityA1

System and method for improving the accuracy of a speech recognition program

Priority: Jul 28, 1999Filed: Jun 13, 2003Published: Dec 4, 2003
Est. expiryJul 28, 2019(expired)· nominal 20-yr term from priority
G10L 15/26G10L 2015/0631
43
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

A system and method for improving the accuracy of a speech recognition program. The system is based on a speech recognition program that automatically converts a pre-recorded audio file into a written text. The system parses the written text into segments, each of which can be corrected by the system and saved in a retrievable manner in association with the computer. The standard speech files are saved towards improving accuracy in speech-to-text conversion by the speech recognition program. The system further includes facilities to repetitively establish an independent instance of the written text from the pre-recorded audio file using the speech recognition program. This independent instance can then be broken into segments and each erroneous segment in said independent instance replaced with the corrected segment associated with that segment. In this manner, repetitive instruction of a speech recognition program can be facilitated.

Claims

exact text as granted — not AI-modified
What is claimed is:  
     
         1 . A system for improving the accuracy of a speech recognition program operating on a computer, said system comprising: 
 means for automatically converting a pre-recorded audio file into a written text;    means for parsing said written text into segments;    means for correcting each and every segment of said written text;    means for saving each corrected segment in a retrievable manner in association with said computer;    means for saving speech files associated with a substantially corrected written text and used by said speech recognition program towards improving accuracy in speech-to-text conversion by said speech recognition program; and    means for repetitively establishing an independent instance of said written text from said pre-recorded audio file using said speech recognition program and for replacing each erroneous segment in said independent instance of said written text with said corrected segment associated therewith.    
     
     
         2 . The invention according to  claim 1  wherein said parsing means includes means for directly accessing functions of said speech recognition program.  
     
     
         3 . The invention according to  claim 2  wherein said parsing means further include means to determine a character count to the beginning of each of said segments and means to determine a character count to the end of each of said segments.  
     
     
         4 . The invention according to  claim 3  wherein said means to determine the character count to the beginning of each of said segments includes UtteranceBegin function from the Dragon Naturally Speaking™, and said means to determine the character count to the end of each of said segments includes UtteranceEnd function from the Dragon Naturally Speaking™.  
     
     
         5 . The invention according to  claim 1  wherein said means for automatically converting includes means for directly accessing functions of said speech recognition program.  
     
     
         6 . The invention according to  claim 5  wherein said means for automatically converting further includes TranscribeFile function of Dragon Naturally Speaking™.  
     
     
         7 . The invention according to  claim 1  wherein said correcting means further includes means for highlighting likely errors in said written text.  
     
     
         8 . The invention according to  claim 7  wherein said written text is at least temporarily synchronized to said pre-recorded audio file, said highlighting means comprises: 
 means for sequentially comparing a copy of said written text with a second written text resulting in a sequential list of unmatched words culled from said copy of said written text, said sequential list having a beginning, an end and a current unmatched word, said current unmatched word being successively advanced from said beginning to said end;  
 means for incrementally searching for said current unmatched word contemporaneously within a first buffer associated with the speech recognition program containing said written text and a second buffer associated with said sequential list; and  
 means for correcting said current unmatched word in said second buffer, said correcting means including means for displaying said current unmatched word in a manner substantially visually isolated from other text in said copy of said written text and means for playing a portion of said synchronized voice dictation recording from said first buffer associated with said current unmatched word.  
 
     
     
         9 . The invention according to  claim 8  wherein said second written text is established by a second speech recognition program having at least one conversion variable different from said speech recognition program.  
     
     
         10 . The invention according to  claim 8  wherein said second written text is established by one or more human beings.  
     
     
         11 . The invention according to  claim 8  wherein said correcting means further includes means for alternatively viewing said current unmatched word in context within said copy of said written text.  
     
     
         12 . A method for improving the accuracy of a speech recognition program operating on a computer comprising: 
 (a) automatically converting a pre-recorded audio file into a written text;    (b) parsing the written text into segments;    (c) correcting each and every segment of the written text;    (d) saving each corrected segment in a retrievable manner;    (e) saving speech files associated with a substantially corrected written text and used by the speech recognition program towards improving accuracy in speech-to-text conversion by the speech recognition program;    (f) establishing an independent instance of the written text from the pre-recorded audio file using the speech recognition program;    (g) replacing each erroneous segment in the independent instance of the written text with the corrected segment associated therewith;    (h) saving speech files associated with the independent instance of the written text used by the speech recognition program towards improving accuracy in speech-to-text conversion by the speech recognition program; and    (i) repeating steps (f) through (i) a predetermined number of times.    
     
     
         13 . The method according to  claim 12  further comprising highlighting likely errors is said written text.  
     
     
         14 . The method according to  claim 13  wherein highlighting includes: 
 comparing sequentially a copy of said written text with a second written text resulting in a sequential list of unmatched words culled from said copy of said written text, said sequential list having a beginning, an end and a current unmatched word, said current unmatched word being successively advanced from said beginning to said end;  
 searching incrementally for said current unmatched word contemporaneously within a first buffer associated with the speech recognition program containing said written text and a second buffer associated with said sequential list; and  
 correcting said current unmatched word in said second buffer, said correcting means including means for displaying said current unmatched word in a manner substantially visually isolated from other text in said copy of said written text and means for playing a portion of said synchronized voice dictation recording from said first buffer associated with said current unmatched word.

Join the waitlist — get patent alerts

Track US2003225578A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.