[ Journal ]
All essays

Product
·
6 min read
Introducing Lyra-3
Our most expressive model yet: one million tokens of context, 38 ms to first token, and a voice you can tune in an afternoon.

Engineering
·
12 min read
Speculative decoding on custom silicon
A field guide to how we cut latency by 71% without trading away a single point on our quality evals.

Product
·
7 min read
The style guide is the prompt
How brand teams are replacing thousand-word prompts with a single living document the model actually reads.

Research
·
10 min read
Evaluating taste
Benchmarks measure correctness. We built a panel of poets, editors and support agents to measure something harder.

Culture
·
5 min read
Notes from our first offsite in Lisbon
Forty people, three models, one very long table. What we argued about, what we agreed on, and what we shipped the week after.
