Blog 388 posts

Notes from the trenches on AI engineering, LLM apps, and the full-stack work that holds it all together.

Nvidia's 1B Nemotron embed model is built for multilingual RAG at 8K context
4 min read

Nvidia's 1B Nemotron embed model is built…

llama-nemotron-embed-1b-v2 ships Matryoshka 2048-dim vectors, 26-language eval coverage, and commercial-friendly NeMo Retriever licensing for long-document QA retrieval.

Nvidia's $6B Poolside deal is a bet on open-weight Nemotron, not another chat app
3 min read

Nvidia's $6B Poolside deal is a bet…

Nvidia licensed Poolside's Model Factory for $6 billion, invested $1 billion at a $12B valuation, and hired 109 engineers to chase frontier open-weight models that compete with DeepSeek and Kimi K3.

Outer Bio keeps human skin alive for four weeks to train its AI
3 min read

Outer Bio keeps human skin alive for…

Lady Gaga co-founder Michael Polansky's startup Outer Bio emerged from stealth with Yuna, a platform that feeds living skin experiments into an AI loop that now proposes a new skincare compound every six weeks.

Ox Alpha is free on OpenRouter and nobody will say who built it
4 min read

Ox Alpha is free on OpenRouter and…

A stealth coding model with a 1M-token window landed on OpenRouter August 20 with zero lab name attached. Community forensics point at Zhipu GLM infrastructure, and that raises real routing questions for production code.