PATENTSCOPE will be unavailable a few hours for maintenance reason on Tuesday 19.11.2019 at 4:00 PM CET
Search International and National Patent Collections
Some content of this application is unavailable at the moment.
If this situation persists, please contact us atFeedback&Contact
1. (IN40/DELNP/2008) GRAMMATICAL PARSING OF DOCUMENT VISUAL STRUCTURES

Office : India
Application Number: 40/DELNP/2008 Application Date: 01.01.2008
Publication Number: 40/DELNP/2008 Publication Date: 04.04.2008
Publication Kind : A
Prior PCT appl.: Application Number:PCTUS2006026140 ; Publication Number:WO2007005937 Click to see the data
IPC:
G06K 9/72
G PHYSICS
06
COMPUTING; CALCULATING; COUNTING
K
RECOGNITION OF DATA; PRESENTATION OF DATA; RECORD CARRIERS; HANDLING RECORD CARRIERS
9
Methods or arrangements for reading or recognising printed or written characters or for recognising patterns, e.g. fingerprints
62
Methods or arrangements for recognition using electronic means
72
using context analysis based on the provisionally recognised identity of a number of successive patterns, e.g. a word
Applicants: MICROSOFT CORPORATION
Inventors: VIOLA, PAUL A
SHILMAN, MICHAEL
Priority Data: 11/173,280 01.07.2005 US
Title: (EN) GRAMMATICAL PARSING OF DOCUMENT VISUAL STRUCTURES
Abstract:
(EN) A two-dimensional representation of a document is leveraged to extract a hierarchical structure that facilitates recognition of the document. The visual structure is grammatically parsed utilizing two-dimensional adaptations of statistical parsing algorithms. This allows recognition of layout structures (e.g., columns, authors, titles, footnotes, etc.) and the like such that structural components of the document can be accurately interpreted. Additional techniques can also be employed to facilitate document layout recognition. For example, grammatical parsing techniques that utilize machine learning, parse scoring based on image representations, boosting techniques, and/or "fast features" and the like can be employed to facilitate in docume...
Also published as:
NO20080090NZ565147MXMX/a/2008/000180KR1020080026128EP1894144ZA2008/00041
JP2009500755RU0002421810CN101253514CA2614177WO/2007/005937