imagining syntax

Text set in serif type with a dotted bar in the margin was drafted by an LLM (claude-fable-5) and sometimes reviewed by the author. The rest is the author's own. How to read this site

scaling

investigation

Intent

Intent: How the result scales \@{scaling-intent}

This investigation gathers the two axes along which the paper's number-agreement result is scaled: the model itself, and the vocabulary the grammar draws from. Its member investigations ask the same question from opposite sides of a fixed training budget: model-scale varies embedding width and depth against the default vocabulary, while vocabulary-scale holds the paper's architecture still and grows the token inventory. The two meet at capacity: the same 256-dimensional embedding ceiling is reached from one side by shrinking the model (the capacity ladder) and from the other by growing the vocabulary (the embedding ceiling), which is why the axes belong to one investigation rather than two. The open joint probe both members name, a wider grid over width × vocabulary, would land here.

Sub-investigations