Opinions, Not Press Releases
Browse a curated front page, then move through focused topic and cluster hubs without losing older work to chronology.
Search every article
Use site search to search the full archive, not only this page.
No posts match your search. Try another term or pick a topic.
Latest articles
Nine articles per page, ordered by publication date.
EU AI Vendor Security Questionnaire: 45 Questions Before You Sign
An evidence-led AI procurement checklist covering training, retention, subprocessors, regions, logs, permissions, model changes, incidents, evals, deletion, exit and EU AI Act roles. Includes a multilingual spreadsheet.
DACH AI Adoption Benchmark 2026: What SMEs Put Into Production
Source-backed adoption and use-case data for Austria, Germany and Switzerland, plus the production, budget, ownership and shelfware metrics public surveys still do not measure.
Graphify Review 2026: Is a Codebase Knowledge Graph Worth It?
Buyer review of Graphify vs search and RAG: architecture, privacy boundaries, benchmark limits, real adoption cost and a two-week pilot scorecard.
Bonsai 27B Review: Can a 27B LLM Really Run on a Phone?
Verified buyer's review of 1-bit vs ternary, real memory, uneven benchmark losses, phone and WebGPU speed, use cases and pilot gates.
Soofi S: Is Germany's Sovereign LLM Ready for Business?
Germany's 31.6B sparse model is strong on German and code, but the current checkpoint is a selected-partner closed beta with no final license. Benchmarks, openness, infrastructure, alternatives and the pilot checklist EU buyers need.
AI Pilot Kill-or-Scale Scorecard: 12 Metrics to Check After 30 Days
A decision-grade scorecard with formulas, hard gates and one worked example across baseline cost, successful actions, straight-through completion, correction time, failures, latency, adoption, auditability, data readiness and payback.
T3MP3ST Review 2026: Can It Replace a Penetration Test?
A buyer-focused review separating benchmarked single-agent capabilities from the unproven swarm, mapped against OWASP APTS with a safe pilot decision for CTOs.
Are Vibe Coders the New Junior Developers?
The traditional ticket-taking junior is shrinking, but the junior developer is not extinct. Current labour data, the difference between prompting and engineering, and a hiring model for AI-native entry-level talent.
External QA Benchmark: What We Find in the First 30 Days
A research-backed defect benchmark for permissions, regressions, browser/device combinations, data integrity, AI failures and production escapes, and why no honest universal bug-count median exists.
Follow the work that matters to you
Get a short email when we publish something new. Follow the whole blog or only the problems you care about.