International Business Machines Corporation
USING CLOSED CAPTIONS AS PARALLEL TRAINING DATA FOR CUSTOMIZATION OF CLOSED CAPTIONING SYSTEMS

Last updated:

Abstract:

Method, apparatus, and computer program product are provided for customizing an automatic closed captioning system. In some embodiments, at a data use (DU) location, an automatic closed captioning system that includes a base model is provided, search criteria are defined to request from one or more data collection (DC) locations, a search request based on the search criteria is sent to the one or more DC locations, relevant closed caption data from the one or more DC locations are received responsive to the search request, the received relevant closed caption data are processed by computing a confidence score for each of a plurality of data sub-sets of the received relevant closed caption data and selecting one or more of the data sub-sets based on the confidence scores, and the automatic closed captioning system is customized by using the selected one or more data sub-sets to train the base model.

Status:
Application
Type:

Utility

Filling date:

14 Dec 2019

Issue date:

17 Jun 2021