How a language model works: tokens, context window and temperature
What a token is, what the context window limits, what temperature does and why language models are confidently wrong. Explained for developers, no maths.
Category
Generative AI has changed how code is written, information is searched and tasks are automated. In this category we explain the concepts behind it (language models, tokens, context, agents) and how to use these tools with judgement.
We avoid empty promises. When a claim depends on a vendor announcement or documentation, we link to the original source, and we point out limitations as clearly as advantages.
What a token is, what the context window limits, what temperature does and why language models are confidently wrong. Explained for developers, no maths.
Why models charge per token, how to count tokens exactly for GPT and approximately for Claude, and how to cut the cost of each call without losing quality.
An AI agent combines a language model with tools and a decision loop to complete tasks. How it differs from a chatbot, how it works, its risks and when to use one.