AI development AI development AI development AI development
GPT vs Claude vs Llama in production: choose by task type
Four kinds of LLM work, the model tier that wins each one, and the volume at which running open weights yourself starts to make sense.
AI voice agent architecture: the four parts that decide if it ships
Where each layer of a production voice agent breaks on live phone calls, and the order we'd build them in, from voice agents we run in production.
pgvector vs Pinecone: when Postgres stops being enough
The four signals that tell you pgvector has hit its limit, and the fixes to try inside Postgres before you pay for a second database.
English-to-SQL at a million users: what breaks, in order
The five things that broke as FormulaBot's English-to-SQL pipeline grew to 1.5M+ users, and the order we'd build the fixes in.
Prashant Abbi