t27.aiРусский

An honest scoreboard

You will learn

How T27 was compared with NVFP4, MX+ and MXFP4 on a real language model, and what the result does not show.

A format is judged by what it does to a model. We stored the weights of one small language model in each format and measured perplexity on text it had not seen: lower is better. The rule was written before the run: T27 beats a rival when its perplexity is below 990 thousandths of the rival's. It beat NVFP4, MX+ and MXFP4, and it lost to E2M2, a wider format. The panel under the bars lists what this test cannot claim. The one claim that holds: T27 beats NVFP4 on a ternary substrate.

Try it

Find the NVFP4 ratio and the threshold line; then name the format T27 loses to and what it pays for that.

Open the interactive lesson →

T27 vs NVFP4: six weight formats on one small language model
T27 vs NVFP4: six weight formats on one small language model ↗

All lessons