You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
#my plans {have a search engine that can fine stuff in a query }
#example documten{
#{"id": "document"
#"text": """this is a blop of text that has blah blah blah words and some words, its a typical document"
#"created: "2012-02_18T20:18:00-0000
#tokenization [taking our big document text blob and breaking into usable word sizes{filtering meaningless words using text blob spliting white space.}]
#stemming it tokenizes stemming finds the root words (avoid manually searching through the whole blob... )
#ngrams solves most of the shortcomming of stemming
#inver index is key value its terms we just got from edge n grams