Blog 388 posts

Notes from the trenches on AI engineering, LLM apps, and the full-stack work that holds it all together.

Self-Harness: how agents rewrite their own operating rules without retraining
5 min read

Self-Harness: how agents rewrite their own operating…

Shanghai AI Lab's Self-Harness lets a fixed model improve its own agent scaffolding through weakness mining, targeted edits, and regression gates. Here is what the Terminal-Bench numbers mean and how to run a lightweight version today.

AI agent credentials belong in a vault, not a .env file on someone's laptop
6 min read

AI agent credentials belong in a vault,…

Twelve million servers leak .env files to the open web. When you give an agent Gmail, CRM, and Slack access, local plaintext tokens turn a config mistake into a company-wide breach.

Anthropic put $10M CAD in Claude credits into eight Canadian research labs
4 min read

Anthropic put $10M CAD in Claude credits…

The July 2026 commitment funds Amii, Mila, Vector, CHEO, CAMH, Université Laval, U of T, and U of Saskatchewan with no-strings Claude API credits plus startup program access for affiliated founders.