Skip to content

A data pipeline for information retrieval using the TF-IDF algorithm, a corpus made up of BBC News, Spark and Scala.

Notifications You must be signed in to change notification settings

raymondklutse/TF-IDF

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

12 Commits
 
 
 
 
 
 
 
 

Repository files navigation

TF-IDF

The aim of this project is to create a data pipeline for information retrieval using TF-IDF. This project was developed using Spark, Scala and a corpus containing BBC news

About

A data pipeline for information retrieval using the TF-IDF algorithm, a corpus made up of BBC News, Spark and Scala.

Topics

Resources

Stars

Watchers

Forks

Releases

No releases published

Packages

No packages published