IP Library › Granted Patent US 12,393,713
Granted Patent B2
US 12,393,713 · App. 18/354,730 · Granted Aug 19, 2025

Automatic processing of meetings for confidential content

Inventors: Yu-Siang Chen (Minxiong Township, TW); Ching-Chun Liu (Taipei, TW); Joey H. Y. Tseng (Taipei, TW); Amanda P L Yang (Taipei, TW)
Assignee: International Business Machines Corporation
G06F21/6218G06N20/00G10L15/063G10L15/18G10L2015/088
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,393,713
App. No.
18/354,730
Granted
Aug 19, 2025
Kind
B2
Abstract

A method, computer system, and a computer program product for processing digital content is provided. The present invention may include building a sensitive sentence classification model. The present invention may include receiving a digital content, wherein the digital content is intended for one or more groups of content consumers. The present invention may include processing the digital content using the sensitive sentence classification model. The present invention may include generating a consumer specific digital content for each of the one or more groups of content consumers.

Claims (62)

1. A method performed by a processing device for processing digital content, the method comprising:

building a sensitive sentence classification model;

receiving a digital content, wherein the digital content is intended for one or more groups of content consumers identified by a content producer within a user interface;

assigning a plurality of consumers within an organization to the one or more groups of content consumers based on a hierarchal structure of the organization constructed based on an analysis of organizational information;

processing the digital content using the sensitive sentence classification model;

generating a consumer specific digital content for each of the one or more groups of content consumers; and

displaying the consumer specific digital content to a plurality of consumers, wherein each of the plurality of consumers may access the consumer specific digital content corresponding to their content consumer group.

2. The method of claim 1 , wherein building the sensitive sentence classification model further comprises:

training the sensitive sentence classification model using a customized training dataset and an existing dataset, wherein the customized training dataset and the existing dataset are preprocessed using a Word2Vec algorithm, and wherein the customized training dataset is comprised of keywords and entities derived from organizational information, and wherein the existing dataset is comprised of one or more existing sensitive word datasets sourced from one or more publicly available resources.

3. The method of claim 2 , wherein the customized training dataset is further comprised of a plurality of manually identified spoken language examples.

4. The method of claim 1 , wherein processing the digital content using the sensitive sentence classification model further comprises:

storing the digital content received in its original form in an original content database;

identifying one or more sensitive data segments using the sensitive sentence classification model; and

storing the one or more sensitive data segments in a processed content database according to data type, timestamp, and corresponding consumer roles, wherein the corresponding consumer roles include at least one of the one or more groups of content consumers.

5. The method of claim 4 , wherein identifying the one or more sensitive data segments further comprises:

receiving one or more additional sensitive data segments manually identified by the content producer;

storing the one or more additional sensitive data segments in the processed content database; and

retraining the sensitive sentence classification model based on the one or more additional sensitive data segments manually identified by the content producer.

6. The method of claim 1 , wherein an Image Recognition Neural Network deep learning model is utilized in identifying visual content within the digital content which may be filtered using the sensitive sentence classification model.

7. The method of claim 6 , further comprising:

blocking at least a portion of the visual content of the digital content for at least one of the one or more groups of content consumers.

8. The method of claim 1 , further comprising:

defining consumer role settings for each of the one or more groups of content consumers based on the hierarchal structure, wherein the consumer role settings designates information which is considered sensitive for each of the one or more groups based on roles within the organization.

9. The method of claim 1 , wherein the digital content is comprised of at least audio content and visual content.

10. The method of claim 9 , further comprising:

designating, by the content producer within the user interface, a masking period for at least one of the or more groups of content consumers.

11. The method of claim 10 , further comprising:

masking, by an output content processor, at least a portion of the audio content or video content for the masking period of the at least one of the one or more groups of content consumers.

12. The method of claim 1 , wherein the processing of the digital content further comprises:

delaying the processing of the digital content received in real time for an intervening time period, wherein the digital content is processed using the sensitive sentence classification model during the intervening time period.

13. A computer system for processing digital content, comprising:

one or more processors, one or more computer-readable memories, one or more computer-readable tangible storage medium, and program instructions stored on at least one of the one or more tangible storage medium for execution by at least one of the one or more processors via at least one of the one or more memories, wherein the computer system is capable of performing a method comprising:

building a sensitive sentence classification model;

receiving a digital content, wherein the digital content is intended for one or more groups of content consumers identified by a content producer within a user interface;

assigning a plurality of consumers within an organization to the one or more groups of content consumers based on a hierarchal structure of the organization constructed based on an analysis of organizational information;

processing the digital content using the sensitive sentence classification model;

generating a consumer specific digital content for each of the one or more groups of content consumers; and

displaying the consumer specific digital content to a plurality of consumers, wherein each of the plurality of consumers may access the consumer specific digital content corresponding to their content consumer group.

14. The computer system of claim 13 , wherein building the sensitive sentence classification model further comprises:

program instructions, stored on at least one of the one or more computer-readable storage media for execution by at least one of the one or more processors via at least one of the one or more memories, to train the sensitive sentence classification model using a customized training dataset and an existing dataset, wherein the customized training dataset and the existing dataset are preprocessed using a Word2Vec algorithm, and wherein the customized training dataset is comprised of keywords and entities derived from organizational information, and wherein the existing dataset is comprised of one or more existing sensitive word datasets sourced from one or more publicly available resources.

15. The computer system of claim 13 , wherein processing the digital content using the sensitive sentence classification model further comprises:

program instructions, stored on at least one of the one or more computer-readable storage media for execution by at least one of the one or more processors via at least one of the one or more memories, to store the digital content received in its original form in an original content database;

program instructions, stored on at least one of the one or more computer-readable storage media for execution by at least one of the one or more processors via at least one of the one or more memories, to identify one or more sensitive data segments using the sensitive sentence classification model; and

program instructions, stored on at least one of the one or more computer-readable storage media for execution by at least one of the one or more processors via at least one of the one or more memories, to store the one or more sensitive data segments in a processed content database according to data type, timestamp, and corresponding consumer roles, wherein the corresponding consumer roles include at least one of the one or more groups of content consumers.

16. The computer system of claim 15 , wherein identifying the one or more sensitive data segments further comprises:

program instructions, stored on at least one of the one or more computer-readable storage media for execution by at least one of the one or more processors via at least one of the one or more memories, to receive one or more additional sensitive data segments manually identified by the content producer;

program instructions, stored on at least one of the one or more computer-readable storage media for execution by at least one of the one or more processors via at least one of the one or more memories, to store the one or more additional sensitive data segments in the processed content database; and

program instructions, stored on at least one of the one or more computer-readable storage media for execution by at least one of the one or more processors via at least one of the one or more memories, to retrain the sensitive sentence classification model based on the one or more additional sensitive data segments manually identified by the content producer.

17. A computer program product for processing digital content, comprising:

one or more computer readable storage media, and program instructions collectively stored on the one or more computer readable storage media, the program instructions comprising:

building a sensitive sentence classification model;

receiving a digital content, wherein the digital content is intended for one or more groups of content consumers identified by a content producer within a user interface;

assigning a plurality of consumers within an organization to the one or more groups of content consumers based on a hierarchal structure of the organization constructed based on an analysis of organizational information;

processing the digital content using the sensitive sentence classification model;

generating a consumer specific digital content for each of the one or more groups of content consumers; and

displaying the consumer specific digital content to a plurality of consumers, wherein each of the plurality of consumers may access the consumer specific digital content corresponding to their content consumer group.

18. The computer program product of claim 17 , wherein building the sensitive sentence classification model further comprises:

program instructions, stored on at least one of the one or more computer-readable storage media, to train the sensitive sentence classification model using a customized training dataset and an existing dataset, wherein the customized training dataset and the existing dataset are preprocessed using a Word2Vec algorithm, and wherein the customized training dataset is comprised of keywords and entities derived from organizational information, and wherein the existing dataset is comprised of one or more existing sensitive word datasets sourced from one or more publicly available resources.

19. The computer program product of claim 17 , wherein processing the digital content using the sensitive sentence classification model further comprises:

program instructions, stored on at least one of the one or more computer-readable storage media, to store the digital content received in its original form in an original content database;

program instructions, stored on at least one of the one or more computer-readable storage media, to identify one or more sensitive data segments using the sensitive sentence classification model; and

program instructions, stored on at least one of the one or more computer-readable storage media, to store the one or more sensitive data segments in a processed content database according to data type, timestamp, and corresponding consumer roles, wherein the corresponding consumer roles include at least one of the one or more groups of content consumers.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 19, 2023
From: CHEN, YU-SIANG; LIU, CHING-CHUN; TSENG, JOEY H.Y.; YANG, AMANDA PL
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 064308/0692 →
Continuity (1)
Related Publication 20250028847A1 · Jan 23, 2025
References Cited (42)
US 7840487B2 · Hatano · 2010 [cited by applicant]
US 9736214B2 · Mendez · 2017 [cited by applicant]
US 11977515B1 · Pham · 2024 [cited by examiner]
US 12020265B1 · Mao · 2024 [cited by examiner]
US 12164503B1 · Khan · 2024 [cited by examiner]
US 20140108542A1 · Cheng · 2014 [cited by applicant]
US 20160080485A1 · Hamedi · 2016 [cited by examiner]
US 20160188725A1 · Wang · 2016 [cited by examiner]
US 20160203208A1 · Anderson · 2016 [cited by examiner]
US 20160335674A1 · Plourde · 2016 [cited by examiner]
US 20160350310A1 · Jurka · 2016 [cited by examiner]
US 20180137203A1 · Hennekey · 2018 [cited by examiner]
US 20180225710A1 · Kar · 2018 [cited by examiner]
US 20180293678A1 · Shanahan · 2018 [cited by examiner]
US 20180337871A1 · Matta · 2018 [cited by examiner]
US 20190034976A1 · Hamedi · 2019 [cited by examiner]
US 20190052819A1 · Pogorelik · 2019 [cited by applicant]
US 20190282155A1 · St Amant · 2019 [cited by examiner]
US 20190354777A1 · Beck · 2019 [cited by examiner]
US 20200042531A1 · Bhojarajan · 2020 [cited by examiner]
US 20200065857A1 · Lagi · 2020 [cited by examiner]
US 20200379787A1 · Martin · 2020 [cited by examiner]
US 20210173959A1 · Gueta · 2021 [cited by applicant]
US 20210281569A1 · Soon-Shiong · 2021 [cited by examiner]
US 20220012268A1 · Ghoshal · 2022 [cited by examiner]
US 20220237368A1 · Tran · 2022 [cited by examiner]
US 20230418857A1 · Pickens · 2023 [cited by examiner]
US 20240193913A1 · Saraee · 2024 [cited by examiner]
CN 119337169A · 2025 [cited by applicant]
TW 570711B · 2016 [cited by applicant]
WO 2011120573A1 · 2011 [cited by applicant]
Ali, “A simple Word2vec tutorial”, Medium.com, https://medium.com/@zafaralibagh6/a-simple-word2vec-tutorial-61e64e38a6a1, Jan. 6, 2019, pp. 1-20. [cited by applicant]
Camacho, “CNNs for Text Classification”, Github, May 27, 2019, pp. 1-13, https://cezannec.github.io/CNN_Text_Classification/. [cited by applicant]
Davidson, “Hate-Speech-and-Offensive_Language,” Github, Mar. 29, 2019, 1 pg., https://github.com/t-davidson/hate-speech-and-offensive-language/tree/master/data. [cited by applicant]
IBM, “Interested in Watson Natural Language Understanding?”, IBM.com, [Accessed Jan. 5, 2023], 1 pg., Retrieved from the Internet: <https://www.ibm.com/demos/live/natural-language-understanding/self-service/home>. [cited by applicant]
IBM, “Watson Natural Language Understanding—Details—Singapore”, IBM.com, [Accessed Dec. 28, 2022], pp. 1-10, Retrieved from the Internet: <https://www.ibm.com/sg-en/cloud/watson-natural-language-understanding/details>. [cited by applicant]
IBM, “Watson Natural Language Understanding,” IBM.com, [Accessed Jan. 5, 2023], 9 pgs., Retrieved from the Internet: <https://www.ibm.com/cloud/watson-natural-language-understanding>. [cited by applicant]
Kaggle, “Toxic Comment Classification Challenge,” Kaggle.com, [Accessed Jan. 5, 2023], 3 pgs., Retrieved from the Internet: <https://www.kaggle.com/c/jigsaw-toxic-comment-classification-challenge>. [cited by applicant]
Kim, “Understanding how Convolutional Neural Network (CNN) perform text classification with word embeddings”, Towards Data Science, Dec. 2, 2017, pp. 1-6, https://towardsdatascience.com/understanding-how-convolutional-n… [cited by applicant]
Lee, et. al., “Making sense of text: artificial intelligence-enabled content analysis”, ResearchGate, European Journal of Marketing, vol. 54, No. 3, pp. 615-644, Feb. 2020, https://www.researchgate.net/publication/33947… [cited by applicant]
Reddit, “Complete Public Reddit Comments Corpus,” Reddit.com, Jul. 2015, 2 pgs., https://archive.org/details/2015_reddit_comments_corpus. [cited by applicant]
Tripathi, et. al., “Detecting Sensitive Content in Spoken Language”, 6th IEEE International Conference on Data Science and Advanced Analytics (DSAA), IEEE, Oct. 5-8, 2019, pp. 374-381, https://assets.amazon.science/f4/a… [cited by applicant]