Reference 2 stops to get here

Stop Words

Common words (the, is, at) often removed in NLP preprocessing as they carry little semantic meaning.

Your route here

2 stops · basics first
  1. Token ✓ understood

    The basic unit of text that a language model processes, typically representing a word, subword, or character. Tokens are the fundamental building blocks for LLM input and output.

  2. Tokenization ✓ understood

    Splitting text into tokens, usually subword pieces, and mapping each to an integer ID so a language model can process it.

  3. Stop Words · you are here ✓ understood

Where it sits

Before this

Tokenization
Stop Words

Leads to

Nothing yet: a destination in its own right.

Explore nearby