VISUAL ANALYSIS FOR DOCUMENT IMPORT
案件概要
発明者
Nathaniel McConathy
IPC分類
CPC分類
Embodiments extract a layout from a digital image of a document, including performing an analysis of image data to identify areas of content and storing the identified areas as design elements of an electronic document template. Analyzing the image data to identify areas of interest includes testing a plurality of lines of pixels from the digital image against a background color definition to identify boundaries of a content area of interest. Content from the content area of interest is processed using a machine learning model to assign a content type for the content area of interest, where the machine learning model represents multiple types of content and is trained to assign content types to input content. The content area of interest is stored as a design element of a digital page template.
原文(中国語)
Embodiments extract a layout from a digital image of a document, including performing an analysis of image data to identify areas of content and storing the identified areas as design elements of an electronic document template. Analyzing the image data to identify areas of interest includes testing a plurality of lines of pixels from the digital image against a background color definition to identify boundaries of a content area of interest. Content from the content area of interest is processed using a machine learning model to assign a content type for the content area of interest, where the machine learning model represents multiple types of content and is trained to assign content types to input content. The content area of interest is stored as a design element of a digital page template.
外部リソース