US2009112587A1PendingUtilityA1
System and method for generating a phrase pronunciation
Est. expiryFeb 27, 2024(expired)· nominal 20-yr term from priority
G10L 13/08G10L 15/187
49
PatentIndex Score
0
Cited by
0
References
0
Claims
Abstract
A system and method for a speech recognition technology that allows language models to be customized through the addition of special pronunciations for components of phrases, which are added to the factory language models during customization. It allows components of a phrase to have different pronunciations inside customer-added phrases than are specified for those isolated components in the factory language models.
Claims
exact text as granted — not AI-modified1 .- 17 . (canceled)
18 . A method in a computer system comprising a language model, a background dictionary, and at least one preexisting pron component list, for adding phrase pronunciations to a language model, said method comprising the steps of:
inputting at least one phrase to be added to the language model; determining if said at least one phrase is contained in said language model; if said at least one phrase is not contained in said language model, determining if said at least one phrase is contained in said background dictionary, and, if so, adding background dictionary pronunciation to said language model; if said at least one phrase is not contained in said language model or said background dictionary, parsing said at least one phrase into an ordered set of tokens in accordance with certain rules sequentially associating prons with each said token of said ordered set of tokens, generating a pron derived phrase pronunciation from said prons and adding said pron derived phrase pronunciation to said language model.
19 . A method, in accordance with claim 18 , wherein the step of generating said pron derived phrase pronunciation, further comprises, for each said token of said ordered set of tokens:
a) sequentially determining if each said associated pron is in said preexisting pron component list, and, if so, obtaining pronunciation from said preexisting pron component list; b) if said associated pron is not in said preexisting pron component list, determining if said associated pron is in a preexisting language model, and, if so, adding said language model pron to said preexisting pron component list; c) if said associated pron is not in preexisting pron component list or said preexisting language model, determining if said associated pron is in preexisting background dictionary, and, if so, adding said background dictionary pron to said preexisting pron component list; d) if said associated pron is not in said preexisting pron component list, said preexisting language model or said preexisting background dictionary, generating a guess pron, and adding said guess pron to said preexisting pron component list; e) if there is an additional token in said ordered set of tokens, repeating steps a) to d); and, f) if there are no additional tokens in said ordered set of tokens, generating said pron derived phrase pronunciation by combining said associated pron pronunciations as obtained from said preexisting pron component list, in sequence.
20 . A method for adding phrase pronunciations to a language model, in accordance with claim 18 , wherein said pron component list includes punctuations or formatting that is present in the said at least one phrase but is silent in the pronunciation of said at least one phrase.
21 . A method for adding phrase pronunciations to a language model, in accordance with claim 19 , wherein said pron component list selected from a plurality of lists in accordance with the position of the said token within the said at least one phrase.
22 . A method for adding phrase pronunciations to a language model, in accordance with claim 18 , wherein said certain rules comprise breaking up the said phrase into tokens at certain boundaries.
23 . A method for adding phrase pronunciations to a language model, in accordance with claim 22 , wherein said certain boundaries comprise white spaces and/or punctuation.
24 . A method for adding phrase pronunciations to a language model, in accordance with claim 18 , wherein said certain rules comprise looking for the longest match in said preexisting language model or said preexisting background dictionary.
25 . A method for adding phrase pronunciations to a language model, in accordance with claim 19 , wherein said preexisting pron component lists comprise an initial pron component list and a non-initial pron component list.
26 . A tangible computer usable medium having computer readable instructions stored thereon for execution by a processor and comprising a language model, a background dictionary, and at least one preexisting pron component list to perform a method comprising:
inputting at least one phrase to be added to the language model; determining if said at least one phrase is contained in said language model; if said at least one phrase is not contained in said language model, determining if said at least one phrase is contained in said background dictionary, and, if so, adding background dictionary pronunciation to said language model; if said at least one phrase is not contained in said language model or said background dictionary, parsing said at least one phrase into an ordered set of tokens in accordance with certain rules, sequentially associating prons with each said token of said ordered set of tokens, generating a pron derived phrase pronunciation from said prons, and adding said pron derived phrase pronunciation to said language model.
27 . A tangible computer usable medium, in accordance with claim 26 , to perform a method wherein the step of generating said pron derived phrase pronunciation further comprises, for each said token of said ordered set of tokens:
a) sequentially determining if each said associated pron is in said preexisting pron component list, and, if so, obtaining pronunciation from said preexisting pron component list; b) if said associated pron is not in said preexisting pron component list, determining if said associated pron is in a preexisting language model, and, if so, adding said language model pron to said preexisting pron component list; c) if said associated pron is not in preexisting pron component list or said preexisting language model, determining if said associated pron is in preexisting background dictionary, and, if so, adding said background dictionary pron to said preexisting pron component list; d) if said associated pron is not in said preexisting pron component list, said preexisting language model or said preexisting background dictionary, generating a guess pron, and adding said guess pron to said preexisting pron component list; e) if there is an additional token in said ordered set of tokens, repeating steps a) to d); and, f) if there are no additional tokens in said ordered set of tokens, generating said pron derived phrase pronunciation by combining said associated pron pronunciations as obtained from said preexisting pron component list, in sequence.
28 . A computer usable medium, in accordance with claim 27 , wherein said pron component list includes punctuations or formatting that is present in the text but is silent in the pronunciation of said at least one phrase.
29 . A computer usable medium, in accordance with claim 27 , wherein said pron component list selected from a plurality of lists in accordance with the position of the said token within the said at least one phrase.
30 . A computer usable medium, in accordance with claim 26 , wherein said certain rules comprise breaking up the said phrase into tokens at certain boundaries.
31 . A computer usable medium, in accordance with claim 30 , wherein said certain boundaries comprise white spaces and/or punctuation.
32 . A computer usable medium, in accordance with claim 26 , wherein said certain rules comprise looking for the longest match in said preexisting language model or said preexisting background dictionary.
33 . A computer usable medium, in accordance with claim 27 , wherein said preexisting pron component lists comprise an initial pron component list and a non-initial pron component list.Join the waitlist — get patent alerts
Track US2009112587A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.