Tokenization, text cleaning, preprocessing (stemming/lemmatization), vectorization (BoW/TF-IDF), problem with RNN/LSTM (sequential bottleneck, vanishing gradient, no long-range context), self-attention as the solution
Mini text analysis (tokenize + vectorize) as foundation for understanding how NLP processes text.
Outcomes:
LAB title
Lab title: