Processing

Please wait...

Settings

Settings

1. EP1894144 - GRAMMATICAL PARSING OF DOCUMENT VISUAL STRUCTURES

Office European Patent Office
Application Number 06786329
Application Date 30.06.2006
Publication Number 1894144
Publication Date 05.03.2008
Publication Kind A4
IPC
G PHYSICS
06
COMPUTING; CALCULATING; COUNTING
K
RECOGNITION OF DATA; PRESENTATION OF DATA; RECORD CARRIERS; HANDLING RECORD CARRIERS
9
Methods or arrangements for reading or recognising printed or written characters or for recognising patterns, e.g. fingerprints
62
Methods or arrangements for recognition using electronic means
72
using context analysis based on the provisionally recognised identity of a number of successive patterns, e.g. a word
[IPC code unknown for G06F 40]
G06K 9/72
G06F 40/00
CPC
G06F 40/211
G06K 9/726
G06K 2209/01
Applicants MICROSOFT CORP
Inventors VIOLA PAUL A
SHILMAN MICHAEL
Designated States
Priority Data 17328005 01.07.2005 US
2006026140 30.06.2006 US
Title
(DE) GRAMMATISCHES ANALYSIEREN VON VISUELLEN STRUKTUREN EINES DOKUMENTS
(EN) GRAMMATICAL PARSING OF DOCUMENT VISUAL STRUCTURES
(FR) ANALYSE GRAMMATICALE DE STRUCTURES VISUELLES DE DOCUMENT
Abstract
(EN)
A two-dimensional representation of a document is leveraged to extract a hierarchical structure that facilitates recognition of the document. The visual structure is grammatically parsed utilizing two-dimensional adaptations of statistical parsing algorithms. This allows recognition of layout structures (e.g., columns, authors, titles, footnotes, etc.) and the like such that structural components of the document can be accurately interpreted. Additional techniques can also be employed to facilitate document layout recognition. For example, grammatical parsing techniques that utilize machine learning, parse scoring based on image representations, boosting techniques, and/or 'fast features' and the like can be employed to facilitate in document recognition.

(FR)
Représentation 2D de document étayée pour l'extraction d'une structure hiérarchique qui facilite la reconnaissance du document. On effectue une analyse grammaticale de la structure visuelle par des adaptations 2D d'algorithmes d'analyse statistique. Cela permet la reconnaissance de structures de présentation (par exemple, colonnes, auteurs, titres, notes de bas de page, etc.) et autres éléments, permettant l'interprétation précise de composantes structurelles du document. On peut aussi utiliser des techniques additionnelles pour faciliter la reconnaissance de présentation. Par exemple, les techniques d'analyse grammaticale qui font intervenir l'apprentissage machine, le score d'analyse sur la base de représentations d'images, les techniques de renforcement, et/ou les «fonctions rapides» et autres éléments, peuvent être utilisées pour faciliter la reconnaissance de document.