From our blog

Check out our latest news and updates.

Prompt caching: the cheapest speedup you are not using
Prompt caching: the cheapest speedup you are not using
30/07/2026 — [email protected]

If every request begins with the same two thousand tokens of instructions, you are paying full price to resend them. Ord...

Rate limits, retries and the backoff you actually need
Rate limits, retries and the backoff you actually need
26/07/2026 — [email protected]

Every AI feature meets a 429 eventually. Whether that is a blip or an outage depends on retry logic written before you n...

How to evaluate an AI feature before you ship it
How to evaluate an AI feature before you ship it
24/07/2026 — [email protected]

You cannot unit test "is this a good answer", but you can build a set of real cases with known-good outputs. An afternoo...

The cost model of an AI feature
The cost model of an AI feature
22/07/2026 — [email protected]

Per-token pricing looks trivial until you multiply by retries, conversation history and the context you resend on every...

Writing prompts your team can maintain
Writing prompts your team can maintain
20/07/2026 — [email protected]

Prompts are code with none of the tooling. Treat them like code anyway — version them, review them, test them, and expla...

AI adoption without a rewrite
AI adoption without a rewrite
16/07/2026 — [email protected]

The best first AI feature is small, measurable and easy to switch off. Start where a wrong answer is cheap and the value...