Insights & writing
/ LATEST FROM THE TEAM
What we're thinking and shipping.
How we cut model latency by 47% in one quarter
A walkthrough of the infra changes our team shipped between October and December — with the actual numbers, the dead ends, and the wins that surprised us.
Maya Chen · 1 min read · Apr 24
Why we shipped multi-region by default
Deploying AI workloads across regions shouldn't be the customer's problem. Here's what we built.
Alex Reece · 1 min read · Apr 18
Evals are the new unit tests
A practical guide to running evals against every prompt change before it hits production.
Priya Desai · 1 min read · Apr 12
Aurora ships with Pagewise. Here’s what changed.
Pagewise's support team rolled out Aurora last quarter. The post-mortem is good news for anyone in the same boat.
James Grey · 1 min read · Apr 06
A better way to do streaming inference
We rewrote our inference layer over a long weekend. Here's the architecture and why it matters.
Maya Chen · 1 min read · Mar 30
Building a remote-first team that ships weekly
How we structure our weeks, our reviews, and our async writing to keep the bar high without burning out.
Alex Reece · 1 min read · Mar 24
How we approach data privacy at the model layer
A deep dive into the controls we built so customer data never leaves their tenant.
Priya Desai · 1 min read · Mar 18/ NEWSLETTER
Get our writing in your inbox.
Once a fortnight — engineering, research, and customer stories. No spam, unsubscribe in one click.