Opinions, Not Press Releases
Browse a curated front page, then move through focused topic and cluster hubs without losing older work to chronology.
Search every article
Use site search to search the full archive, not only this page.
No posts match your search. Try another term or pick a topic.
Get the next field note
One concise email when we publish. No tracking pixels, and no inbox filler.
Explore by topic
Six durable entry points keep every article close to the main blog hub.
AI and agents
Engineering, operating and evaluating AI systems, coding agents and model infrastructure.
Product and MVP
Validate demand, shape scope and turn uncertain product bets into testable releases.
Delivery and QA
Architecture, production readiness and quality practices for software that must keep working.
Leadership and teams
Fractional leadership, AI-enabled teams and the operating decisions behind delivery.
Business and regulation
Buying software, funding delivery and navigating governance, contracts and regulation.
Web3 and privacy
Blockchain infrastructure, cryptography and privacy technologies evaluated for production use.
Evergreen field notes
- Claude Code Design System: 4 Parts for On-Brand UIStop re-explaining your brand in every prompt. Build a reusable Claude Code design system with three Markdown files, an examples directory and a review loop that teams can govern.
- AI Agent Harness, Explained: The Reliability Layer Around an LLMA visual architecture, failure map and buyer checklist for the context, policy, tools, verification and observability that make an AI agent dependable.
- LLM Pseudonymization Gateways: Does the Prompt Leave GDPR Scope?Platforms in this category promise that pseudonymizing a prompt makes a hosted model safe to use. The 2025 CJEU ruling and the EDPB guidance say something narrower. Here is the evidence a buyer actually has to collect.
- Micro Agency vs Mid-Size Agency: How Big Should Your Software Partner Be in 2026?The one-person agency versus micro agency argument is a debate between sellers. Here is the buyer's version, with the team-size research, the 2026 agency margin data, and the four clauses that make a small partner safe to hire.
- Qwen3.8-27B: Self-Hosted Computer-Use Agents Without Exporting ScreenshotsAn open-weight model now leads desktop-operation benchmarks. What that changes for back-office automation you cannot send to a hosted API.
- Transformers.js Browser AI: When Local Inference Belongs in Your ProductA production guide to private in-browser AI: verified adoption, WebGPU and WASM, real cost, privacy boundaries, use cases, fallbacks and a 30-day pilot.
- Stateful LLM Platform Production ArchitectureAn anonymized field guide to persistence, tenant access, long-running jobs, external interfaces, deployment controls and continuous QA without a full rewrite.
- AI 3D Model Generators for Game Development: 2026 GuideWhich AI 3D tools can help a game team, and which outputs still need an artist? Compare hosted and self-hosted options, quality gates, licensing and engine handoff.
- AI-Ready Company Wiki: Architecture and Build GuideA buyer-focused blueprint for company knowledge that humans can edit and AI agents can retrieve with the same permissions, sources and review loop.
- AI Agent Contract Signing: eIDAS QES Integration GuideA developer and buyer guide to binding AI-generated contracts to human approval, remote QES, ID Austria or EUDI Wallet flows, validation and audit evidence.
- How to Self-Host LiteLLM in Production: 2026 GuideA security-first LiteLLM deployment guide for teams choosing their own AI gateway: architecture, Postgres, Redis, keys, upgrades, cost and go-live checks.
- Internal AI Agent Marketplace: A 2026 Enterprise Build GuideBuild a governed internal agent catalog employees will use. Architecture, listing contract, identity, context, evaluations, approvals, metrics and a 90-day rollout.
- Is Linux the Best OS for AI Agents? A 2026 Infrastructure GuideLinux gives AI agents a mature shell, open interfaces and several isolation layers. This guide shows where it wins, where it does not, and how to deploy agents without handing them the host.
- How to Make AI Writing Sound Human with Agent SkillsNo AI Slop, i-have-adhd and book-to-skill solve different writing problems. Here is how to combine, test and govern them without erasing your voice.
- PII Redaction Before LLM Prompts: A Practical PipelineCompare Presidio, OpenAI Privacy Filter and managed DLP services. Then build a reversible prompt gateway that protects personal data without destroying task context.
- Why AI Agent Projects Get CancelledWhat Gartner's forecast does and does not show, plus an evidence-based review of value, evaluation, tools, cost, handoff, data, scope and governance risks.
- The MVP Is Dead. Build a Minimum Credible Product.AI can reduce prototype effort, not the need for product judgment. Avoidable roughness may damage the signal an MVP should collect. A practical case for less scope, one trustworthy path, and decision-grade evidence.
- What Test-Driven Development Can and Cannot ProveTDD can tighten feedback and make intended behavior executable, but it is not a guarantee of defect-free or secure software. Use it as one layer in a risk-based testing strategy.
- Fractional CTPO vs CTO and CPOWhen one combined product-and-technology lead can reduce handoffs, when to split the remit, and how to assess capacity, independence, decision rights, and continuity without a fixed team-size rule.
- Choosing a Software Project Pricing ModelCompare time-based, fixed-price, and staged hybrid models by uncertainty, change control, acceptance criteria, budget limits, and delivery evidence.
- Smart Contract Pre-Audit Checklist (30 Questions)Thirty risk-based questions for Solidity teams covering specifications, toolchains, privileges, external calls, upgrades, economics, testing, and operations.
Latest articles
Nine articles per page, ordered by publication date.
SwarmLLM Review 2026: Browser P2P LLM Inference Across Phones and Laptops
A buyer-focused review of SwarmLLM: browser-native 27B inference across phones and laptops, WebGPU/WebRTC architecture, benchmark limits, peer trust, and when to pilot it.
Ramp Inspect Architecture 2026: Background Coding Agents at Scale
How Ramp Inspect uses full-stack Modal sandboxes, filesystem snapshots, coordination and verification, plus the controls a regulated engineering team needs before scaling background coding agents.
Model Hardware Standard: Enterprise Guide to Physical AI
Evaluate Anthropic's MHS for lab and factory automation: architecture, MHS vs MCP, safety boundaries, pilot economics and a practical adoption checklist.
Fonio AI Review 2026: Pricing, API, GDPR & Build vs Buy
A buyer-focused Fonio AI review with current phone pricing, SIP and API integration, EU recording duties, operational limits, and a practical build-versus-buy scorecard.
Mosaic (YC S26) Review: Shared Memory for Team AI Agents
Mosaic centralizes coding-agent sessions across tools and teammates. See how shared session memory differs from Claude Code Agent Teams, agent memory, Git, and a real team knowledge layer.
Ripwire Review 2026: AI Repo Context Without Embeddings?
A buyer-focused review of Red Hat ET's Ripwire: deterministic repository context, ranked call graphs, token claims, CLI vs MCP, benchmark limits, and when it beats grep, Graft or Graphify.
Atomic Multi-File Edits for AI Coding Agents: The Semaprax Lesson
Why sequential file writes can expose a broken half-update, and how immutable generations plus one active pointer create a safer publication boundary.
Phonely Alma Review: Is the Voice LLM Ready for Production?
Alma claims 182 ms TTFT, a 206 ms P99 and $0.55 per blended million tokens. See what the benchmark proves, what it omits and how to run a buyer-side voice-agent pilot.
Utopia Review: Temporal Knowledge Graph for Enterprise
Evaluate Utopia for enterprise knowledge: two-clock history, source provenance, self-hosting limits and a practical pilot plan before you invest.
Follow the work that matters to you
Get a short email when we publish something new. Follow the whole blog or only the problems you care about.