US2021343286A1PendingUtilityA1

Auto-completion for Multi-modal User Input in Assistant Systems

Assignee: FACEBOOK TECH LLCPriority: Apr 20, 2018Filed: Jul 6, 2021Published: Nov 4, 2021
Est. expiryApr 20, 2038(~11.7 yrs left)· nominal 20-yr term from priority
G06F 16/3329G06F 3/017G06F 3/011G06V 10/764G06N 7/01G06F 18/2411G06Q 10/40G06N 3/045G06N 3/0442G06N 3/0464G06N 3/09G06F 40/205G06F 16/90335H04L 67/75H04L 51/18H04L 51/216H04L 67/53H04L 67/5651H04L 51/52H04L 67/535H04L 51/222G06F 9/453G06F 40/295H04L 63/102G06V 10/82G06V 40/28G06V 20/10G06F 16/24578G06F 9/44505G06F 16/3323G10L 15/063G06F 16/248H04L 12/2816G06N 20/00G06F 40/30G06F 40/40G10L 13/00H04L 41/22G10L 2015/223G06F 16/2365G06F 16/2255H04L 67/306G06F 16/90332G06N 3/006G10L 15/16G06F 16/24575G06N 3/08G06F 16/9532G06F 16/338G06F 8/31G06F 3/013G10L 17/06G10L 15/187H04L 51/02G06Q 10/00G06F 7/14G10L 15/07G06F 3/167G10L 15/26G06F 16/951G10L 2015/225H04L 43/0894G06F 2216/13G06F 16/3344G06N 5/027H04W 12/08G10L 17/22G06F 16/24552G06N 5/022G06F 16/176H04L 43/0882G06F 16/243G10L 15/1815G10L 15/22H04L 67/10G06F 16/9535G06F 40/274G06F 16/3322G06F 16/904G10L 15/02G10L 13/04G06F 16/9038H04L 41/20G10L 17/00G10L 15/183G06F 9/4451H04L 51/046G10L 15/1822G06Q 10/42G06Q 10/48H04L 5/02G06F 21/6245
78
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

In one embodiment, a method includes receiving an initial input in a first modality from a first user at a client system, determining intents and slots corresponding to the initial input, wherein the slots are conditioned on the intents, generating one or more candidate continuation-inputs based on the intents and slots, where the one or more candidate continuation-inputs are in one or more candidate modalities, respectively, wherein the candidate modalities are different from the first modality, and wherein each of the candidate continuation-inputs references entities represented by the slots, and presenting one or more suggested inputs corresponding to one or more of the candidate continuation-inputs at the client system.

Claims

exact text as granted — not AI-modified
What is claimed is: 
     
         1 . A method comprising, by a client system:
 receiving, at the client system, an initial input from a first user, wherein the initial input is in a first modality;   determining one or more intents and one or more slots corresponding to the initial input, wherein the one or more slots are conditioned on the one or more intents;   generating, based on the one or more intents and the one or more slots, one or more candidate continuation-inputs, where the one or more candidate continuation-inputs are in one or more candidate modalities, respectively, wherein the candidate modalities are different from the first modality, and wherein each of the candidate continuation-inputs references one or more entities represented by the one or more slots; and   presenting, at the client system, one or more suggested inputs corresponding to one or more of the candidate continuation-inputs.   
     
     
         2 . The method of  claim 1 , wherein the first modality comprises one of audio, text, image, video, motion, or orientation. 
     
     
         3 . The method of  claim 2 , wherein the first modality comprises motion, and wherein the initial input comprises a gesture. 
     
     
         4 . The method of  claim 2 , wherein the first modality comprises orientation, and wherein the initial input comprises a gaze on an object. 
     
     
         5 . The method of  claim 1 , further comprising:
 determining that the first user needs one or more suggested inputs.   
     
     
         6 . The method of  claim 5 , wherein determining that the first user needs one or more suggested inputs is based on a wake-up input from the first user. 
     
     
         7 . The method of  claim 6 , wherein the wake-up input comprises one or more of a voice utterance, a character string, an image, a video clip, a gesture, or a gaze. 
     
     
         8 . The method of  claim 5 , wherein the initial input comprises a gaze on an object, and wherein determining that the first user needs one or more suggested inputs is further based on the gaze on the object. 
     
     
         9 . The method of  claim 8 , wherein generating the one or more candidate continuation-inputs is further based on the object. 
     
     
         10 . The method of  claim 5 , wherein determining that the first user needs one or more suggested inputs is further based on contextual information associated with the initial input. 
     
     
         11 . The method of  claim 5 , wherein determining that the first user needs one or more suggested inputs is further based on the one or more intents. 
     
     
         12 . The method of  claim 1 , further comprising:
 identifying one or more entities associated with the one or more intents.   
     
     
         13 . The method of  claim 12 , wherein generating the one or more candidate continuation-inputs is further based the one or more entities. 
     
     
         14 . The method of  claim 1 , further comprising:
 receiving, at the client system, a user-selected input from the first user, wherein the user-selected input comprises one of the suggested inputs; and   executing one or more tasks based on the user-selected input.   
     
     
         15 . The method of  claim 1 , further comprising:
 receiving, at the client system, a first user-selected input from the first user, wherein the first user-selected input comprises one of the suggested inputs, and wherein the first user-selected input is associated with a first intent;   generating, based on the first user-selected input, one or more additional candidate continuation-inputs, wherein each of the one or more additional candidate continuation-inputs is associated with the first intent;   presenting, at the client system, one or more additional suggested inputs corresponding to one or more of the additional candidate continuation-inputs;   receiving, at the client system, a second user-selected input from the first user, wherein the second user-selected input comprises one of the additional suggested inputs; and   executing one or more tasks based on the second user-selected input.   
     
     
         16 . The method of  claim 1 , further comprising:
 determining the initial input comprises an incomplete input for triggering an execution of one or more tasks corresponding to the one or more intents.   
     
     
         17 . The method of  claim 16 , wherein each of the candidate continuation-inputs comprises a complete instruction to trigger the execution of a respective task of the one or more tasks. 
     
     
         18 . The method of  claim 17 , wherein each of the presented suggested inputs comprises a guidance to the first user for executing the complete instruction corresponding to the respective candidate continuation-input. 
     
     
         19 . One or more computer-readable non-transitory storage media embodying software that is operable when executed to:
 receive, at a client system, an initial input from a first user, wherein the initial input is in a first modality;   determine one or more intents and one or more slots corresponding to the initial input, wherein the one or more slots are conditioned on the one or more intents;   generate, based on the one or more intents and the one or more slots, one or more candidate continuation-inputs, where the one or more candidate continuation-inputs are in one or more candidate modalities, respectively, wherein the candidate modalities are different from the first modality, and wherein each of the candidate continuation-inputs references one or more entities represented by the one or more slots; and   present, at the client system, one or more suggested inputs corresponding to one or more of the candidate continuation-inputs.   
     
     
         20 . A system comprising: one or more processors; and a non-transitory memory coupled to the processors comprising instructions executable by the processors, the processors operable when executing the instructions to:
 receive, at a client system, an initial input from a first user, wherein the initial input is in a first modality;   determine one or more intents and one or more slots corresponding to the initial input, wherein the one or more slots are conditioned on the one or more intents;   generate, based on the one or more intents and the one or more slots, one or more candidate continuation-inputs, where the one or more candidate continuation-inputs are in one or more candidate modalities, respectively, wherein the candidate modalities are different from the first modality, and wherein each of the candidate continuation-inputs references one or more entities represented by the one or more slots; and   present, at the client system, one or more suggested inputs corresponding to one or more of the candidate continuation-inputs.

Join the waitlist — get patent alerts

Track US2021343286A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.