Goal of this project is to have a NLP tool that would give statistical analysis results based on Google Ngram data.

Furthermore, it is now just a NetBeans project without a final JAR.
Furthermore, there will be a github version for anyone who wishes to contribute.
In the future versions, user will be able to convert a single word to numerical data, to be able to compare two words and get the comparison data, and to be able to do the same for the sentences, paragraphs and documents.

I will JAR-it once I decide that it can be called a final release.

This project was made by creating a corpus from the Google Ngrams data for English Language, version 20120701.
EOWL list of English words was used to filter-out the words from Ngrams data.
For each year, per word, the data was added and calculated to describe the average appearance of a word per document for a given year.
Before using this program, you MUST download the corpus.

Project Samples

Project Activity

See All Activity >

Follow Natural Language Analysis with Ngrams

Natural Language Analysis with Ngrams Web Site

Other Useful Business Software
$300 Free Credits to Build on Google Cloud Icon
$300 Free Credits to Build on Google Cloud

New customers can spin up VMs, build with AI, and query data at no cost.

Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
Start Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of Natural Language Analysis with Ngrams!

Additional Project Details

Registered

2015-01-25