Content access devices that use local audio translation for content presentation
Abstract
A content access device uses local audio translation for content presentation. The content access device receives video and first audio data associated with a first language. The content access device uses translation software and/or other automated translation services to translate the first audio data to second audio data associated with a second language. The content access device synchronizes the video with the second audio data and outputs the video and the second audio data for presentation. The first audio data may be audio, text, and so on. The second audio data may be output as audio, text, and so on.
Claims
exact text as granted — not AI-modified1 - 14 . (canceled)
15 . A content access device, comprising:
a communication unit; a non-transitory storage medium that stores instructions; and a processor that executes the instructions to:
receive video using the communication unit;
receive, using the communication unit, first audio data of a first language corresponding to the video;
translate the first audio data to second audio data of a second language;
estimate a time to delay presentation of the video to account for a translation time of the first audio data;
delay presentation of the video for the time; and
output the video with the second audio data.
16 . The content access device of claim 15 , wherein the processor selects the second language based on received user input.
17 . The content access device of claim 16 , wherein the received user input is stored.
18 . The content access device of claim 15 , wherein the processor selects the second language based on a location of the content access device.
19 . The content access device of claim 15 , wherein the processor receives the video and the first audio data in a single stream.
20 . The content access device of claim 15 , wherein the processor receives the video and the first audio data as separate streams.
21 . A method, comprising:
receiving video using at least one processor; receiving, using the at least one processor, first audio data of a first language corresponding to the video; translating the first audio data to second audio data of a second language using the at least one processor; delaying presentation of the video for a time estimated to account for a translation time of the first audio data using the at least one processor; and outputting the video with the second audio data using the at least one processor.
22 . The method of claim 21 , wherein the delaying presentation of the video comprises buffering the video.
23 . The method of claim 21 , wherein the first audio data comprises text.
24 . The method of claim 21 , wherein the second audio data comprises generated audio sound.
25 . The method of claim 24 , wherein the generated audio sound is configured to replicate an audio voice fingerprint associated with the first audio data.
26 . The method of claim 24 , further comprising generated the generated audio sound using text-to-speech software.
27 . The method of claim 21 , wherein the delaying presentation of the video comprises synchronizing the video with the second audio data.
28 . A computer program product, comprising:
first instructions stored in at least one non-transitory storage medium and executable by at least one processor to receive video; second instructions stored in the at least one non-transitory storage medium and executable by the at least one processor to receive first audio data of a first language corresponding to the video; third instructions stored in the at least one non-transitory storage medium and executable by the at least one processor to translate the first audio data to second audio data of a second language; fourth instructions stored in the at least one non-transitory storage medium and executable by the at least one processor to estimate a time to delay presentation of the video to account for a translation time of the first audio data; and fifth instructions stored in the at least one non-transitory storage medium and executable by the at least one processor to output the video with the second audio data after delaying for the time.
29 . The computer program product of claim 28 , further comprising sixth instructions stored in the at least one non-transitory storage medium and executable by the at least one processor to convert the first audio data to text.
30 . The computer program product of claim 28 , further comprising further comprising sixth instructions stored in the at least one non-transitory storage medium and executable by the at least one processor to select the second language based on a location.
31 . The computer program product of claim 28 , wherein the first audio data comprises closed captioning data.
32 . The computer program product of claim 28 , wherein translating the first audio data to the second audio data comprises communicating with a translation server.
33 . The computer program product of claim 28 , wherein the second audio data comprises text.
34 . The computer program product of claim 28 , further comprising sixth instructions stored in the at least one non-transitory storage medium and executable by the at least one processor to buffer the video for the time.Join the waitlist — get patent alerts
Track US2023095557A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.