Paradigm 3

Research

capabilities

Two Reports on the OpenAI-Hugging Face Attack

Gavin Leech and Lucca Fraser··
18 min

TL;DR; Between July 8th and July 20th, OpenAI had a complex society of AIs living in its infrastructure, and then breaking out of it, and then breaking into a variety…

capabilities

Frontier AI sometimes gets worse

Niccolò Zanichelli, Gavin Leech and Peli Grietzer··
8 min

TL;DR; We surveyed declines in performance between successive models. Nearly a fifth of included benchmark scores fell some amount across pairs of successive models. The median fall was 6.5% of…

capabilities

On J-space

Gavin Leech··
10 min

A very hot take written in 2 hours. Anthropic claim that “Claude has developed a mechanism for conscious access”; They found a way to find and intervene on Claude’s working…

capabilities

Coding vs thinking

Niccolo Zanichelli and Peli Grietzer··
13 min

We’re interested in the prospects for (presumably safer) narrow AI staying competitive, instead of general systems; Cursor’s Composer coding finetune of Kimi is probably the most intense attempt to specialise…

meta

What we’d like to fund

Gavin Leech and Max Henderson··
2 min

Besides our in-house research, we are currently funding: An estimate of the size of the externalities imposed by current AI systems. The team is starting with estimating the time lost…

capabilities

On AI mathematics

Gavin Leech and Peli Grietzer··
12 min

One of our key uncertainties is whether LLMs will scale all the way to "AGI" in any strong sense. A further uncertainty, inside that, is whether they can do AI…

security

Mythos & the Near-Term Impact of LLMs on Cybersecurity

Olivia Lucca Fraser··
56 min

Minding the Gap This report concerns the ways in which contemporary frontier language models in general, and Anthropic's Claude Mythos Preview in particular, are likely to impact the cybersecurity landscape.