US2020259945A1PendingUtilityA1

Apparatus, method, and system of cognitive communication assistant for enhancing ability and efficiency of users communicating comprehension

Assignee: FUVI COGNITIVE NETWORK CORPPriority: May 9, 2018Filed: Apr 30, 2020Published: Aug 13, 2020
Est. expiryMay 9, 2038(~11.8 yrs left)· nominal 20-yr term from priority
Inventors:Phu Nguyen
H04M 1/72427H04M 1/7243A61B 5/021A61B 5/165H04M 2250/74H04M 1/656A61B 5/6803H04M 2250/16A61B 5/01A61B 5/02055H04M 1/0216A61B 5/0476H04M 1/72544
59
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

A communication apparatus, a method, a computer readable medium, and a system providing communication with cognitive and visual assistance. The cognitive assistance and visual assistance is provided during a communication between a first communication apparatus with at least one second communication apparatus via a network. The first communication apparatus captures communication data comprising visual and audio information obtained from the communication and captures synchronized cognitive and emotional data generated from the user during the communication with the second communication apparatus. The communication data and the synchronized cognitive and emotional data is stored and converted into a visual form comprising at least one of synchronized text, symbols, sketches, images, and animation. The visual form is displayed on a display of the first communication apparatus.

Claims

exact text as granted — not AI-modified
What is claimed is: 
     
         1 . A method of providing cognitive assistance during a communication comprising:
 establishing, by a first communication apparatus operated by a user, the communication with a second communication apparatus of a participant;   receiving, by the first communication apparatus, first multimedia data generated by the second communication apparatus during the established communication;   capturing, by the first communication apparatus, second multimedia data generated by the user during the established communication;   extracting, during the established communication, audio content from the first multimedia data and the second multimedia data;   dividing the audio content to generate a plurality of audio data blocks;   converting each of the plurality of audio data blocks into a text format to generate a respective text block; and   during the established communication:
 displaying, in a first display area of the first communication apparatus, the respective text block for each of the first multimedia data and the second multimedia data in a form of a plurality of scripts, and 
 displaying, in a second display area of the first communication apparatus, the first multimedia data including the participant and the second multimedia data including the user. 
   
     
     
         2 . The method of  claim 1 , further comprising:
 displaying, during the established communication, in the first display area, third multimedia data generated during the established communication,   wherein the third multimedia data includes an environment of the user or the participant, or data downloaded via a network.   
     
     
         3 . The method of  claim 1 , wherein displaying the respective text block for each of the first multimedia data and the second multimedia data includes:
 displaying, during the established communication, a plurality of text blocks of preceding and current audio data, wherein the plurality of text blocks are color coded based on a cognitive state of the user.   
     
     
         4 . The method of  claim 1 , further comprising:
 receiving, via a user interface of the first communication apparatus, user input related to whether to start converting each of the plurality of audio data blocks into the respective text block.   
     
     
         5 . The method of  claim 1 , further comprising:
 displaying, after the established communication, in the first display area, the respective text block for each of the first multimedia data and the second multimedia data and a corresponding section for adding notes by the user.   
     
     
         6 . The method of  claim 5 , further comprising:
 displaying the plurality of scripts and the notes in a distinguishable manner from one another in the first display area.   
     
     
         7 . The method of  claim 1 , further comprising:
 displaying the plurality of scripts such that each of a first text block corresponding to the first multimedia data is visually distinguishable from a second text block corresponding to the second multimedia data.   
     
     
         8 . The method of  claim 1 , further comprising:
 based on user input, switching the first communication apparatus from a communication mode in which the established communication occurs, to a review mode in which the established communication has ended.   
     
     
         9 . An apparatus for providing cognitive assistance during a communication, the apparatus comprising:
 a network interface configured to establish the communication with a communication apparatus of a participant and to receive first multimedia data generated by the communication apparatus during the communication;   an image capturer configured to capture second multimedia data generated by a user of the apparatus during the communication;   a memory configured to store computer executable instructions;   a processor configured to execute the computer executable instructions, which when executed by the processor causes the processor to:
 extract, during the communication, audio content from the first multimedia data and the second multimedia data, 
 divide the audio content to generate a plurality of audio data blocks, and 
 convert each of the plurality of audio data blocks into a text format to generate a respective text block; and 
   at least one display configured to display, during the communication:
 the respective text block for each of the first multimedia data and the second multimedia data in a form of a plurality of scripts, in a first display area, and 
 the first multimedia data including the participant and the second multimedia data including the user, in a second display area. 
   
     
     
         10 . The apparatus of  claim 9 , wherein the at least one display is further configured to display, during the communication, in the first display area, third multimedia data generated during the communication and wherein the third multimedia data includes an environment of the user or the participant, or data downloaded via a network. 
     
     
         11 . The apparatus of  claim 9 , wherein the at least one display is configured to display the respective text block for each of the first multimedia data and the second multimedia data by:
 displaying, during the communication, a plurality of text blocks of preceding and current audio data, the plurality of text blocks being color coded based on a cognitive state of the user.   
     
     
         12 . The apparatus of  claim 9 , further comprising:
 a user interface configured to receive user input related to whether to start converting each of the plurality of audio data blocks into the respective text block.   
     
     
         13 . The apparatus of  claim 9 , further comprising:
 a user interface configured to receive input from a user in a form of notes or comments,   wherein the at least one display is further configured to display, after the communication, in the first display area, the respective text block for each of the first multimedia data and the second multimedia data and a section with the notes that are input by the user, and   wherein each one of the notes is displayed corresponding to the respective text block.   
     
     
         14 . The apparatus of  claim 13 , wherein the at least one display is further configured to display the plurality of scripts and the notes in a distinguishable manner from one another in the first display area. 
     
     
         15 . The apparatus of  claim 9 , wherein the at least one display is further configured to display the plurality of scripts such that each of a first text block corresponding to the first multimedia data is visually distinguishable from a second text block corresponding to the second multimedia data. 
     
     
         16 . The apparatus of  claim 9 , further comprising:
 a user interface configured to receive user input for switching the apparatus from a communication mode in which the communication occurs, to a review mode in which the communication has ended.   
     
     
         17 . A non-transitory computer readable medium configured to store instructions for providing cognitive assistance during a communication, the instructions are executed by a processor and cause the processor to execute the following operations:
 establish the communication with a communication apparatus of a participant;   receive first multimedia data generated by the communication apparatus during the established communication;   control an image capturer to capture second multimedia data generated by a user during the established communication;   extract, during the established communication, audio content from the first multimedia data and the second multimedia data;   divide the audio content to generate a plurality of audio data blocks;   convert each of the plurality of audio data blocks into a text format to generate a respective text block; and   during the established communication, control at least one display to display:
 the respective text block for each of the first multimedia data and the second multimedia data in a form of a plurality of scripts, in a first display area, and 
 the first multimedia data including the participant and the second multimedia data including the user, in a second display area. 
   
     
     
         18 . The non-transitory computer readable medium of  claim 17 , wherein the instructions further cause the processor to control the at least one display to display, during the established communication, in the first display area, third multimedia data generated during the established communication, wherein the third multimedia data includes an environment of the user or the participant, or data downloaded via a network. 
     
     
         19 . The non-transitory computer readable medium of  claim 17 , wherein the instructions further cause the processor to control the at least one display to display the respective text block for each of the first multimedia data and the second multimedia data by displaying, during the established communication, a plurality of text blocks of preceding and current audio data, and
 wherein the plurality of text blocks are color coded based on a cognitive state of the user.   
     
     
         20 . The non-transitory computer readable medium of  claim 17 , wherein the instructions further cause the processor to receive, via a user interface, user input related to whether to start converting each of the plurality of audio data blocks into the respective text block.

Join the waitlist — get patent alerts

Track US2020259945A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.