Computer detection of typographical errors
Abstract
Describes a computer program written for the UNIX time-sharing system which reduces by several orders of magnitude the task of finding words in a document which contain typographical errors. The program is adaptive in the sense that it uses statistics from the document itself for its analysis. In a first pass through the document, a table of diagram and trigram frequencies is prepared. The second pass through the document breaks out individual words and compares the diagrams and trigrams in each word with the frequencies from the table. An index is given to each word which reflects the hypothesis that the trigrams in the given word were produced from the same source that produced the trigram table. The words are sorted in decreasing order of their indices and printed. Printing is suppressed for words appearing in a table of 2726 common technical English words.
- Journal
- IEEE Transactions on Professional Communication
- Published
- 1975-03-01
- DOI
- 10.1109/tpc.1975.6593963
- CompPile
- Open Access
- Closed
- Topics
-
- Export
- BibTeX RIS
Citation context
Cited by in this index (0)
No articles in this index cite this work.
References (0)
No references on file for this article.
Related articles
-
Business and Professional Communication Quarterly Nov 2025Matthew J. Baker; Grant Eckstein; Ana Barraza; Benjamin Duffield
-
Business and Professional Communication Quarterly Jun 2025Mohammad Sadegh Khorshidi; José M. Merigó; Ghassan Beydoun
-
Business and Professional Communication Quarterly Jun 2025Nelson Lamar Reinsch; Jeanine Warisse Turner
-
Writing and Pedagogy Apr 2025Sarah Williams; Amy Seely Flint; Rebecca Rohloff
-
Res Rhetorica Mar 2025Jacek Wasilewski; Agata Kostrzewa