US2022075829A1PendingUtilityA1
Voice searching metadata through media content
Est. expiryOct 3, 2034(~8.2 yrs left)· nominal 20-yr term from priority
G06F 16/433G06F 16/7867H04N 21/4828G10L 25/54G06F 16/48G10L 15/26G06F 16/7844H04N 21/4826G06F 16/44G06F 16/683G06F 16/90332G06F 16/4387
62
PatentIndex Score
0
Cited by
0
References
0
Claims
Abstract
Techniques for searching metadata through media content. User input identifying a search criteria is received from a user device. Metadata associated with media content files is searched to identify a subset of the media content files. Search results identifying the subset of the media content files are provided to the user device. The metadata is generated by an originator of each media content file and describes each scene.
Claims
exact text as granted — not AI-modifiedWhat is claimed is:
1 . A computer-implemented method, comprising:
receiving, from a user device, user input identifying a search criteria; searching, by operation of one or more computer processors, metadata associated with a plurality of media content files to identify a subset of the plurality of media content files, the subset of the plurality of media content files comprising one or more media content files of the plurality of media content files that includes one or more scenes that match the search criteria; and providing, to the user device, search results identifying the subset of the plurality of media content files, wherein the metadata is generated by a respective originator of each media content file of the plurality of media content files and describes each scene of the plurality of media content files.
2 . The computer-implemented method of claim 1 , wherein the metadata comprises visual metadata describing an actor, an actress, a character, an object, a location, an emotion, an action, a theme, or a plot point associated with each scene of the plurality of media content files.
3 . The computer-implemented method of claim 1 , wherein the metadata comprises audio metadata describing a dialog or a song associated with each scene of the plurality of media content files.
4 . The computer-implemented method of claim 1 , wherein the metadata comprises subtitle metadata describing a subtitle associated with each scene of the plurality of media content files.
5 . The computer-implemented method of claim 1 , wherein the metadata comprises visual metadata describing an actor, an actress, a character, an object, a location, an emotion, an action, a theme, or a plot point associated with each scene of the plurality of media content files, and
wherein the metadata further comprises at least one of:
audio metadata describing a dialog or a song associated with each scene of the plurality of media content files, or
subtitle metadata describing a subtitle associated with each scene of the plurality of media content files.
6 . The computer-implemented method of claim 1 , wherein the user input comprises vocal user input, and wherein the computer-implemented method further comprises:
initiating a speech-to-text recognition process to ascertain the search criterion from the vocal user input.
7 . The computer-implemented method of claim 1 , further comprising:
generating derivative media content by stitching together the search results into a single media content file based on a stitching option.
8 . A non-transitory computer-readable medium containing a program executable to perform an operation comprising:
receiving, from a user device, user input identifying a search criteria; searching, by one or more computer processors when executing the program, metadata associated with a plurality of media content files to identify a subset of the plurality of media content files, the subset of the plurality of media content files comprising one or more media content files of the plurality of media content files that includes one or more scenes that match the search criteria; and providing, to the user device, search results identifying the subset of the plurality of media content files, wherein the metadata is generated by a respective originator of each media content file of the plurality of media content files and describes each scene of the plurality of media content files.
9 . The non-transitory computer-readable medium of claim 8 , wherein the metadata comprises visual metadata describing an actor, an actress, a character, an object, a location, an emotion, an action, a theme, or a plot point associated with each scene of the plurality of media content files.
10 . The non-transitory computer-readable medium of claim 8 , wherein the metadata comprises audio metadata describing a dialog or a song associated with each scene of the plurality of media content files.
11 . The non-transitory computer-readable medium of claim 8 , wherein the metadata comprises subtitle metadata describing a subtitle associated with each scene of the plurality of media content files.
12 . The non-transitory computer-readable medium of claim 8 , wherein the metadata comprises visual metadata describing an actor, an actress, a character, an object, a location, an emotion, an action, a theme, or a plot point associated with each scene of the plurality of media content files, and
wherein the metadata further comprises at least one of:
audio metadata describing a dialog or a song associated with each scene of the plurality of media content files, or
subtitle metadata describing a subtitle associated with each scene of the plurality of media content files.
13 . The non-transitory computer-readable medium of claim 8 , wherein the user input comprises vocal user input, and wherein the operation further comprises:
initiating a speech-to-text recognition process to ascertain the search criterion from the vocal user input.
14 . The non-transitory computer-readable medium of claim 8 , wherein the operation further comprises:
generating derivative media content by stitching together the search results into a single media content file based on a stitching option.
15 . A system comprising:
one or more computer processors; a memory containing a program executable by the one or more computer processors to perform an operation comprising:
receiving, from a user device, user input identifying a search criteria;
searching metadata associated with a plurality of media content files to identify a subset of the plurality of media content files, the subset of the plurality of media content files comprising one or more media content files of the plurality of media content files that includes one or more scenes that match the search criteria; and
providing, to the user device, search results identifying the subset of the plurality of media content files,
wherein the metadata is generated by a respective originator of each media content file of the plurality of media content files and describes each scene of the plurality of media content files.
16 . The system of claim 15 , wherein the metadata comprises visual metadata describing an actor, an actress, a character, an object, a location, an emotion, an action, a theme, or a plot point associated with each scene of the plurality of media content files.
17 . The system of claim 15 , wherein the metadata comprises audio metadata describing a dialog or a song associated with each scene of the plurality of media content files.
18 . The system of claim 15 , wherein the metadata comprises subtitle metadata describing a subtitle associated with each scene of the plurality of media content files.
19 . The system of claim 15 , wherein the metadata comprises visual metadata describing an actor, an actress, a character, an object, a location, an emotion, an action, a theme, or a plot point associated with each scene of the plurality of media content files, and
wherein the metadata further comprises at least one of:
audio metadata describing a dialog or a song associated with each scene of the plurality of media content files, or
subtitle metadata describing a subtitle associated with each scene of the plurality of media content files.
20 . The system of claim 15 , wherein the user input comprises vocal user input, and wherein the operation further comprises:
initiating a speech-to-text recognition process to ascertain the search criterion from the vocal user input.
21 . The system of claim 15 , wherein the operation further comprises:
generating derivative media content by stitching together the search results into a single media content file based on a stitching option.Join the waitlist — get patent alerts
Track US2022075829A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.