#Tokens

Posts tagged Tokens.

What a large language model actually predicts
What a large language model actually predicts
12/09/2026 — admin@byqreal.test

A model does not look anything up and does not decide what is true. It estimates which token comes next. Almost everythi...

Tokens, not words: how a model reads your text
Tokens, not words: how a model reads your text
10/09/2026 — admin@byqreal.test

Models do not see characters or words. They see tokens — and once you know how text becomes tokens, several odd behaviou...

The context window is a budget, not a memory
The context window is a budget, not a memory
02/09/2026 — admin@byqreal.test

Bigger context windows did not give models memory. They gave you a larger envelope to fill on every single request — and...

Prompt caching: the cheapest speedup you are not using
Prompt caching: the cheapest speedup you are not using
30/07/2026 — admin@byqreal.test

If every request begins with the same two thousand tokens of instructions, you are paying full price to resend them. Ord...

The cost model of an AI feature
The cost model of an AI feature
22/07/2026 — admin@byqreal.test

Per-token pricing looks trivial until you multiply by retries, conversation history and the context you resend on every...