The word images must be able to be described by parameters extracted from a Con-volutional Neural Network (CNN). This to determine what words look similar. The similarity of word images is what is used to create an alignment. The algorithm finds words that are common in the text and identifies these by the shape of the word image and its position on the page. It uses the features it creates for each image to return an alignment, linking words in the transcript to word images given the similarity of reoc-curring words in the text. These words are added as labelled images to a data set. A labelled data set of word images subsequently grows with each page shown to the algo-rithm and is used to facilitate better alignments on future pages. This yields not only a probable alignment of words to word images on a page but also a data set of labelled words.
Noch keine Bewertungen vorhanden
Verfassen Sie die erste Bewertung zu diesem Artikel
Helfen Sie anderen Kundinnen und Kunden durch Ihre Meinung.
Kurze Frage zu unserer Seite
Vielen Dank für dein Feedback
Wir nutzen dein Feedback, um unsere Produktseiten zu verbessern. Bitte habe Verständnis, dass wir dir keine Rückmeldung geben können. Falls du Kontakt mit uns aufnehmen möchtest, kannst du dich aber gerne an unseren Kund*innenservice wenden.