Opinions, Not Press Releases
Browse a curated front page, then move through focused topic and cluster hubs without losing older work to chronology.
Search every article
Use site search to search the full archive, not only this page.
No posts match your search. Try another term or pick a topic.
Latest articles
Nine articles per page, ordered by publication date.
Can an AI Agent Use Your Product, or Only Read About It?
Readable means an assistant can cite you. Usable means it can complete a task and be refused when it should be. The second is an authorization problem, not a model problem.
Graft Review 2026: Do Agent Repo Maps Belong in Git?
Graft writes your codebase into linked markdown so agents stop re-exploring it. We checked whether that map really travels through git, what its benchmarks prove and what belongs in version control instead.
LLM Pseudonymization Gateways: Does the Prompt Leave GDPR Scope?
Platforms in this category promise that pseudonymizing a prompt makes a hosted model safe to use. The 2025 CJEU ruling and the EDPB guidance say something narrower. Here is the evidence a buyer actually has to collect.
How Coding Agents Keep Token Bills in Check with Output Compression
Coding agents can waste many tokens on oversized tool output. Learn how output compression and decision-first observability make AI execution costs measurable again.
Smarter Token Usage with Your AI Coding Agent
Build a predictable token budget for coding agents by combining cache design, model routing, and context compression in the order that protects quality.
DeepSeek Harness Review: Is the Plugin Stack Production-Ready?
A buyer-focused review of DeepSeek Harness, its replaceable Cordis architecture, four agent presets, security boundaries and enterprise pilot fit.
Netflix's vLLM and Triton Stack: 7 Production Lessons
Netflix chose vLLM for operational fit, then found the real bottlenecks in constrained decoding, deployment, model loading and metrics. Here is what smaller teams should copy, and what they should buy instead.
OpenSandbox Review: Is Self-Hosting Worth It?
A buyer-focused review of OpenSandbox isolation, Credential Vault, Kubernetes operations, hidden costs, production gaps and a 30-day pilot plan.
Transformers.js Browser AI: When Local Inference Belongs in Your Product
A production guide to private in-browser AI: verified adoption, WebGPU and WASM, real cost, privacy boundaries, use cases, fallbacks and a 30-day pilot.
Follow the work that matters to you
Get a short email when we publish something new. Follow the whole blog or only the problems you care about.