← All six questions The AI Iceberg

Why is AI infrastructure so expensive and slow, and how do I build around that?

You're building with AI and want to know what it actually costs, why it's slow, and where the real bottlenecks are — not marketing claims.

What's here now

What's coming

The Silicon Wall: Why Architecture Decides Who Runs AI

Why your unit economics are downstream of an architectural accident: the chip running global AI inference was designed to render video games.

Expected Oct 2026

The Token Breakeven Map: Where AI Business Models Actually Work

The four price thresholds that decide whether a freemium tier, an indie app, or an always-on agent pencils out.

Expected Dec 2026

Chatbot vs. Agent: The Token Cost Gap Nobody Budgets For

Agentic AI burns 5–50x more tokens than a chatbot, and the reason is context accumulation rather than the model.

Expected Feb 2027

Optical Computing Enters Production: What Lightmatter and LightGen Mean for the 2030s

Two companies now claim production-ready optical AI chips. What the claims mean and how to read them.

Expected Apr 2027
  1. Sep 23 – Nov 4Seven layers, one a week, on LinkedIn
  2. Nov 5The full framework, written up here
  3. Dec 3Token Breakeven Map
  4. Jan 21, 2027The Real Waste — VSM Utilization
  5. Also answersWhy did my inference bill triple when I shipped agents?
    Is my GPU spend actually doing any work?
    At what token price does a freemium tier stop losing money?

Each piece publishes here first. Short posts on LinkedIn point back to whichever one is new.

All writing  ·  Not your question? Compare all six