RESOURCES
Everything we've written down.
Blog posts, answers to common questions, and a glossary of the terms we use, all in one place.
LATEST
From the blog
Pydantic AI for Building AI Agents: What It's For and When It's Worth Using
Pydantic AI brings the type-safety and validation discipline of Pydantic to agent building. Here's the problem it actually solves, how it compares to the alternatives, and whether it holds up in production.
AI Frameworks vs AI Toolkits: What Actually Sets Them Apart
LangChain, CrewAI, LlamaIndex, and the Vercel AI SDK all get lumped together as "the AI stack." They solve different problems. Here's what an AI framework is actually for, and where a toolkit like the AI SDK fits instead.
Where Langfuse and Harbor Fit Into How We Evaluate Our Own AI Agents
Our own self-monitoring pass catches an agent quietly going off the rails. It was never built to grade whether an answer was actually good, or to prove a change is safe before it ships. That's the gap Langfuse and Harbor fill.