Vectorization with stopwords removal

While writing the myTokenizer function how will the document passed to this function as we are fitting the CountVectorizer object on a different corpus.

Hey @srinjoyghosh0702, CountVectorizer, first passes all documents throught the my Tokenizer function and recieves all the unique words. Than it will iterative go one by one on each document to convert it into vectorized output.

Hope this resolved your doubt.
Plz mark the doubt as resolved in my doubts section. :blush:

I hope I’ve cleared your doubt. I ask you to please rate your experience here
Your feedback is very important. It helps us improve our platform and hence provide you
the learning experience you deserve.

On the off chance, you still have some questions or not find the answers satisfactory, you may reopen
the doubt.