Article and method of automatically determining text genre using surface features of untagged texts
US6973423B1 · kind B1 · utility
Assignee
Inventors
Key dates
| Filing date | Jun 18, 1998 |
| Grant date | Dec 6, 2005 |
| Priority date | — |
| Expiry date | Jun 18, 2018 |
Classification
- Technology area (CPC G)Physics
- CPC primaryG06F40/253
- WIPO fieldComputer technology
- WIPO sectorElectrical engineering
Abstract
A processor implemented method of identifying the text genre of a machine-readable, untagged text. The processor implemented method begins by generating a cue vector from the text, which represents occurrences in the text of a first set of nonstructural, surface cues, which are easily computable. Afterward, the processor determines whether the text is an instance of a first text genre using the cue vector and a weighting vector associated with the first text genre.
Source: USPTO / EPO open patent data. Objective bibliographic and citation counts.