OSCAR & OpenMRS Core
-
Distribution of all the terms in software lexicon.
-
10 most frequent terms in corpus (report tf).
-
10 terms with highest document frequency (report df).
-
3 documents with highest number of unique terms.
-
3 documents with lowest number of unique terms.
-
Libraries required -Lucene version 4.8 - https://archive.apache.org/dist/lucene/java/4.8.0/ Add Lucene Core and Lucene Analyzer commons