Loading
Models, tokens, context, sampling, embeddings, limitations, model selection, and dependable application boundaries.
Recommended start
Defines a large language model through next-token prediction and connects that mechanism to useful language behavior and its limits.