Research library
Insights
Analysis of AI, automation, and software: what is changing, what holds up, and what it means for practical decisions.
155 articles · Page 6 of 13
What It Costs to Run AI in Production: A 2026 Pricing Breakdown
What LLM inference actually costs in 2026: per-token rates, reasoning-token overhead, prompt caching, and the self-host crossover point.
GLM 5.2: An Open-Weight Coder That Beats GPT-5.5 on Price, Not the Frontier
Z.ai's MIT-licensed GLM 5.2 edges GPT-5.5 on agentic coding at one-sixth the API cost, but trails Opus 4.8 and loses on terminal coding.
Workflow Automation ROI: A Calculator and a Worked Example
How to compute automation ROI: loaded labor cost, error-reduction value, build cost, and payback period — with a concrete invoice-processing example.
Kimi K2.7 Code: Moonshot Cuts Thinking Tokens 30%, But the Benchmarks Are All Its Own
Moonshot's Kimi K2.7 Code cuts reasoning tokens ~30% and undercuts the frontier on price, but every headline benchmark is first-party and unverified.
AI Governance for Small Businesses: A Practical Checklist
A concrete, non-bureaucratic AI governance checklist for SMBs: data handling, vendor due diligence, shadow AI, human-in-the-loop, and audit logging.
Apple WWDC 2026: Siri Runs on Gemini Now, and Apple Is Fine With That
Apple's rebuilt Siri runs on a custom 1.2T-parameter Google Gemini model at ~$1B/year — in Tim Cook's last keynote, Apple conceded the model race.
Why Every AI Lab Suddenly Ships a Coding Agent
In two weeks Anthropic, OpenAI, Microsoft, and xAI all shipped or expanded terminal coding agents. The convergence tells you where AI's value now sits.
When to Build an AI Agent vs. Buy a SaaS Tool
Custom AI agents vs. off-the-shelf SaaS: a decision framework covering control, data, differentiation, maintenance, cost curve, and time-to-value.
Microsoft Build 2026: Seven MAI Models and the Quiet Exit From OpenAI
At Build 2026 Microsoft shipped seven in-house MAI models — reasoning, coding, voice, image — built to cut OpenAI dependence and undercut rivals.
Anthropic Files to Go Public at a $965B Valuation, Ahead of OpenAI
Anthropic confidentially filed for an IPO on June 1 at a $965B valuation, past OpenAI, days after a $65B Series H. Revenue is now at a $47B run rate.
Claude Opus 4.8: Anthropic Bets on Honesty and Subagent Orchestration
Opus 4.8 is 4x less likely to let its own code flaws slide, adds dynamic workflows orchestrating up to 1,000 subagents, and cuts fast mode pricing.
Project Glasswing's First Month: 10,000 Vulnerabilities Found, and Why That Was the Easy Part
Anthropic's Project Glasswing found 10,000+ high-severity flaws in a month via Claude Mythos Preview. Finding bugs got cheap; fixing them didn't.
