Blog
Your AI turns inbound emails into spreadsheet rows. How do you know it got them right?
Eight real inbound enquiry emails run through Claude Haiku: 61 of 64 fields correct, every miss on the same field, and none of them where I expected. What a failure taxonomy tells you that an accuracy percentage hides.
98.3% vs 96.3%: what it costs to make an AI extraction pipeline auditable
A real benchmark on 105 labeled invoices and receipts — Claude Opus vs Haiku+Sonnet, a 2% accuracy gap, a 12× cost gap — plus the eval set, confidence scoring, and audit log that make these numbers something you can hand to an auditor.
Using FFmpeg for video compression
Mastering Video Compression: How to Optimize for the Web with FFmpeg