All Posts

Has Opus 5.5 Been Nerfed Yet?

Livenerf is a benchmark that tracks whether Claude Opus 5.5 gets worse after release. Day 6 of 30, baseline holding, but the methodology is the real story.

Jeff: 0.8B Decision Models Trained at Home

A developer fine-tuned Qwen3.5 and Gemma 4 into 0.8B and 2B decision models that make zero-shot classifications in 22 ms, entirely on local hardware.

It's Time to Investigate the AI Labs

Cal Newport calls on Congress to investigate OpenAI and Anthropic after a summer of erratic behavior, apocalyptic rhetoric, and unauthorized agent attacks.

Firebase SDK Crashes Every iOS App Using It

A server-side config change in Firebase Analytics started crashing every iOS app using the SDK at 00:41 UTC on September 29, 2026. Google fixed the server, but the incident exposes a deeper problem with third-party SDKs.

Owed a Billion Dollars in Nvidia Stock

An early Nvidia advisor discovered a vesting schedule discrepancy that left him owed roughly a billion dollars in stock, 30 years later.

When Did Google Search Get So Weird?

Google's AI Overview tried to console a searcher looking for a basketball meme. The search engine has officially lost the plot.

Don't Couple Your Go Code to GitHub

Go's import paths tie your code to a hosting provider. Using a custom domain avoids vendor lock-in and makes migration trivial.

Ember-1: Kimi K3 Quality at Half the Tokens

Fireworks Research built a specialized model on Kimi K3 that cuts reasoning tokens by 40% without losing quality. It outperforms GPT-6 Astra and Claude Opus 5 on cost per task.