FROM THE TRENCHES Issue №244

Opinions, Not Press Releases

Browse a curated front page, then move through focused topic and cluster hubs without losing older work to chronology.

Explore by topic
06
Focused clusters
12
Blog posts
244

Latest articles

Nine articles per page, ordered by publication date.

LLM cost calculator 2026: cost per task, prompt caching, batching, routing and self-hosting AI & Agents

LLM Cost Calculator 2026: Cost per Task, Not Cost per Token

The spreadsheet logic behind a real AI bill: count completed tasks, not token prices. Formulas for cached input, batch discounts, routing escalation, self-host utilization, retries, human rework and eval quality.

The Factory Returns: AI revives the software-factory dream, and governance decides whether agility survives it Delivery & QA

The Factory Returns

How agentic AI revives the software-factory dream that failed twice, and whether agility can survive it. Alexandre Kotcherguine and Kevin Riedl weigh the technical-debt and productivity evidence on both sides, trace the craft moving up the stack into specifications and quality gates, and read Stripe's agent fleet as the thesis observed in production.

Building real applications with zero-knowledge proofs and FHE in 2026: a pragmatic guide Web3 & Privacy

Building Real Applications With ZK and FHE in 2026: A Pragmatic Guide

Apple documents production HE lookups, Google Wallet announced ZK age assurance, and Ethereum proving now has multi-GPU real-time benchmarks. The decision tree, current evidence, and five common privacy-tech failure modes.

Zero-knowledge proofs in 2026: zkVMs, client-side proving and what is production-ready Web3 & Privacy

Zero-Knowledge Proofs in 2026: What Is Actually Production-Ready

Proving an Ethereum block fell from 1.69 dollars to under 4 cents in one year, and ZK identity landed in Google Wallet. Which zkVMs to build on, what proving costs, what works on a phone, and where the security bodies are buried.

Fully homomorphic encryption in 2026: what ships in production and what is still hype Web3 & Privacy

Fully Homomorphic Encryption in 2026: What Ships and What Is Still Hype

FHE powers private lookups on Apple devices and encrypted transactions on Ethereum, but performance depends on the scheme, operation mix, parameters, and hardware. The production pattern that works and the encrypted-LLM reality check.

ZK vs FHE vs MPC vs TEE: the 2026 decision framework for architects Web3 & Privacy

ZK vs FHE vs MPC vs TEE: How to Choose in 2026

Four privacy technologies, four trust models, four price tags. The four questions that pick the right one, a side-by-side comparison with honest numbers, and the EU regulations that increasingly force the choice.

Open USD explained: an announced shared stablecoin backed by more than 140 companies Web3 & Privacy

Open USD Explained: Shared Economics and Open Questions

Open USD promises partner governance and shared reserve economics, but remained pre-launch on 2 September 2026. Here is what is confirmed, what is not, and what teams should verify before integrating it.

Rendering your prompt as an image to cut LLM costs: the pxpipe trick, explained honestly AI & Agents

Pxpipe Review: Can Images Cut Claude Code Token Costs 60%?

Pxpipe turns bulky Claude Code context into dense PNGs. The measured token savings are real, but exact IDs and numbers can fail silently. See where it works and where text must stay text.

Cost per token versus cost per task: a lower unit price can still produce a higher total bill AI & Agents

Cheaper Per Token Can Still Cost More Per Task

A benchmark once put Sonnet 5 above Opus 4.8 in cost per task, but Anthropic cancelled the price increase behind that result. Learn how to compare current rates with observed usage.

Inbox, without the noise

Follow the work that matters to you

Get a short email when we publish something new. Follow the whole blog or only the problems you care about.

What would you like to receive?
Choose your topics

Free, double opt-in, no tracking pixels.