All posts

Research 23 posts

Every post filed under Research, newest first.

Anthropic put $10M CAD in Claude credits into eight Canadian research labs
4 min read

Anthropic put $10M CAD in Claude credits…

The July 2026 commitment funds Amii, Mila, Vector, CHEO, CAMH, Université Laval, U of T, and U of Saskatchewan with no-strings Claude API credits plus startup program access for affiliated founders.

Inherent's Faraday beats frontier models at paper replication with a 27B orchestrator
4 min read

Inherent's Faraday beats frontier models at paper…

Faraday is a 27B agent trained with long-horizon RL to replicate research figures. Inherent reports it beats Claude Opus 4.8 and GPT-5.5 on its Replica benchmark by directing Codex as a tool.

Every population may carry DNA from a ghost hominin lineage
4 min read

Every population may carry DNA from a…

A Science study analyzed 503 modern genomes with a new TRACE model and found archaic ancestry that matches no known Neanderthal or Denisovan sequence in every population tested. Roughly 0.5 to 1 percent of non-African genomes may come from a branch that split off more than 500,000 years ago.

Claude Science is Anthropic's bet that researchers need a workbench, not a chatbot
3 min read

Claude Science is Anthropic's bet that researchers…

Anthropic launched Claude Science in beta on June 30: a macOS and Linux desktop app with 60+ database connectors, live code execution, HPC orchestration, and full provenance on every artifact. Here's what it actually does.

GLOSSOPETRAE proves LLMs code better in alien languages than in English
4 min read

GLOSSOPETRAE proves LLMs code better in alien…

GLOSSOPETRAE generates procedural coding languages from a seed. At full opacity, human legibility drops to ~15% while Opus and GPT hit 97-100% task accuracy. Human readability hurts model performance.

All five major LLMs show pro-female hiring bias on Japanese resumes
4 min read

All five major LLMs show pro-female hiring…

A 43,200-call study on rirekisho-format resumes finds significant pro-female bias across Claude, GPT-4o, DeepSeek, Gemini, and Llama. Prompt fixes failed. Name removal helped but broke GPT-4o safety filters 42% of the time.

VIMPO beats GRPO on hard math benchmarks without training a critic
4 min read

VIMPO beats GRPO on hard math benchmarks…

VIMPO derives a policy-implied value function from KL-regularized RL optimality conditions. It improves over GRPO on AIME and OlympiadBench while staying critic-free. Code on GitHub.

MIT's SMT trains RNNs in parallel without backpropagation through time
4 min read

MIT's SMT trains RNNs in parallel without…

Supervised Memory Training uses a Transformer teacher to label optimal memory states, then trains nonlinear RNNs with one-step supervision. You get O(1) gradient paths and time-parallel pretraining without unrolling the full sequence.