AI & EngineeringWhat a production RAG pipeline actually costs to run
Retrieval sounds cheap until you price embeddings, storage and re-indexing at real volume. Here is the arithmetic we use before quoting a client.
What we're learning building AI systems, MVPs and digital products — written by the team doing the work.
AI & EngineeringRetrieval sounds cheap until you price embeddings, storage and re-indexing at real volume. Here is the arithmetic we use before quoting a client.
Product & MVPThe hardest part of an MVP is not building it. It is the conversation where you agree what is not in it.
AI & EngineeringSelf-hosted, Postgres-backed, and the content model lives in your repo. Here is the trade we made and when we would not make it.
GrowthWe measured a cold Next.js build at 3.8 GB and the running app at 157 MB. That 24x gap decides your hosting bill.
GrowthLCP is usually one image and one font. Fix those two and most sites pass without a rewrite.
AI & EngineeringGuardrails, evals and the escalation path matter more than the model. What we learned putting one into production.