US2002184188A1PendingUtilityA1

Method for extracting content from structured or unstructured text documents

Priority: Jan 22, 2001Filed: Jan 22, 2002Published: Dec 5, 2002
Est. expiryJan 22, 2021(expired)· nominal 20-yr term from priority
G06F 8/20
43
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

A method for selecting textual content within a document. Text is selected using mechanisms of pattern recognition on the document's structure or content itself. A pattern recognition rule selects the desired text by identifying the start and/or end positions of the content in the document. The delineated contents is then said to be enclosed in an envelope. A series of envelopes may be used to identify the desired content. Successive envelopes are defined relative to a previous envelope. The contents of any envelope within a series, including the final envelope, may be extracted for use by other documents.

Claims

exact text as granted — not AI-modified
What is claimed is:  
     
         1 ) A method for extracting content from a document, comprising the step of: 
 creating at least one selection envelope based upon a plurality of selection commands for locating specific content within said document; and    selecting content from said document based upon said at least one selection envelope.

Join the waitlist — get patent alerts

Track US2002184188A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.