VISUAL ANALYSIS FOR DOCUMENT IMPORT
卷宗概要
发明人
Nathaniel McConathy
IPC 分类
CPC 分类
Embodiments extract a layout from a digital image of a document, including performing an analysis of image data to identify areas of content and storing the identified areas as design elements of an electronic document template. Analyzing the image data to identify areas of interest includes testing a plurality of lines of pixels from the digital image against a background color definition to identify boundaries of a content area of interest. Content from the content area of interest is processed using a machine learning model to assign a content type for the content area of interest, where the machine learning model represents multiple types of content and is trained to assign content types to input content. The content area of interest is stored as a design element of a digital page template.
原文(中文)
Embodiments extract a layout from a digital image of a document, including performing an analysis of image data to identify areas of content and storing the identified areas as design elements of an electronic document template. Analyzing the image data to identify areas of interest includes testing a plurality of lines of pixels from the digital image against a background color definition to identify boundaries of a content area of interest. Content from the content area of interest is processed using a machine learning model to assign a content type for the content area of interest, where the machine learning model represents multiple types of content and is trained to assign content types to input content. The content area of interest is stored as a design element of a digital page template.