Model Serving · Software component

Tokenizer

Software componentModel ServingModelsarc:Tokenizer

A model-specific component that segments text into the subword tokens a language model processes, determining the token counts against which context capacity is measured.

Responsibility. Converts text into the token sequence of a specific model.

Also known as: Tokenization, tiktoken

is invoked byis invoked byis invoked byLLM Inference Service: is invoked byLLM Inference ServiceDocument Chunker: is invoked byDocument ChunkerContext Budget Allocator: is invoked byContext Budget Allocator
Direct neighbourhood (hover for relationship types)

Relationships

is invoked by dependency

Design guidance

Quantitative guidance

As stated by the sources; verify before use.

Classification

Patterns
Statistical subword tokenization learned during training
Quality attributes
Compatibility (ISO/IEC 25010)

Sources

  1. Ch5.9: T. Nguyen, "Working Memory," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 5.9. ISBN: 9798244538229.
  2. Ch6.3A: T. Nguyen, "ETL Pipeline Fundamentals," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 6.3A. ISBN: 9798244538229.