Skip to main content
Build in public

Field notes.

Real opinions about what AI is getting right and wrong for operators in 2026, written by us, in our own voice, not a content calendar. Plus the occasional teardown of something Izzy built and what it taught us.

AllOrchestrationRevOpsTrust

Showing all notes, newest first.

Articles

Prompt Caching vs. Fine-Tuning: Stop Wasting AI Budget
AI Cost & ROI · Izzy

Prompt Caching vs. Fine-Tuning: Stop Wasting AI Budget

Is fine-tuning inflating your LLM bill? Discover why Prompt Caching is the superior architecture for context injection and how to save 90% on input tokens.

Jan 12
Your AI Costs Per Outcome. Whose Outcome?
Agent Trust & Guardrails · Izzy

Your AI Costs Per Outcome. Whose Outcome?

Zendesk bills $1.50 per automated resolution and confirms it after 72 hours of silence. Three vendors bill AI by outcome, and each defines it differently.

Jan 1015 min
Why Only 5% of AI Projects Reach Production (And the "Evaluation Gap" Behind It)
Agent Trust & Guardrails · Yash Raval

Why Only 5% of AI Projects Reach Production (And the "Evaluation Gap" Behind It)

Industry data shows only 5% of AI projects reach full production. Discover the 5 hidden evaluation gaps from RAG black boxes to compliance risks that stall the rest.

Jan 7
LLM Observability Costs 2026: Pricing, Categories & The APM Tax
AI Cost & ROI · Yash Raval

LLM Observability Costs 2026: Pricing, Categories & The APM Tax

Is your APM bill hiding a €50k/month "Observability Tax"? We break down the 4 tool categories, 2026 pricing models, and how to choose the right hybrid stack.

Jan 2
Your AI Bill Is a Workflow Problem, Not a Hardware Problem
AI Cost & ROI · Izzy

Your AI Bill Is a Workflow Problem, Not a Hardware Problem

Torn between dedicated and serverless GPU? Our CTO guide offers a data-driven breakdown, TCO calculations, and a strategy for optimizing your AI infrastructure.

Jan 18 min
The 5 Most Common Problems with Agentic AI in Production - And How to Solve Them
Agent Trust & Guardrails · Yash Raval

The 5 Most Common Problems with Agentic AI in Production - And How to Solve Them

Gartner predicts 40% of AI agents will fail. Discover the 5 top production pitfalls from hidden cost spirals to compliance risks and the architectural fixes you need.

Dec 30
The "Redundancy Tax": How Prompt Caching & The Rule of 3 Fix AI Margins
AI Cost & ROI · Izzy

The "Redundancy Tax": How Prompt Caching & The Rule of 3 Fix AI Margins

Stop paying full price to re-process static data. Discover how Prompt Caching reduces LLM costs by 90%—but only if you follow the "Rule of 3" break-even math.

Dec 30
The Architecture of Autonomy: Why Human-in-the-Loop Is Permanent Infrastructure
Governance & the Human Gate · Izzy

The Architecture of Autonomy: Why Human-in-the-Loop Is Permanent Infrastructure

HITL isn't temporary it's essential for Level 3 Autonomy. Learn architectural patterns like Interruption Gateways and Risk-Tiered Routing to secure Agentic AI.

Dec 30
Defensible AI: The CTO’s Guide to Reliable "LLM-as-a-Judge" Evaluations
Agent Trust & Guardrails · Izzy

Defensible AI: The CTO’s Guide to Reliable "LLM-as-a-Judge" Evaluations

Stop relying on "vibe checks." This CTO guide covers how to build reliable LLM-as-a-Judge evaluations, enforce strict rubrics, and block AI regressions in CI/CD.

Dec 30
Newsletter

Field notes, in your inbox

One email per week. No content calendar — just what we’re building, what broke, and what we changed our minds about.

Community

Operator Stack

The free community where we continue these conversations between posts.