flâneur — a map of the web's best reading

Coding vs thinking — Paradigm 3

paradigm3.org · 2,310 words · saved by 1 readers

We’re interested in the prospects for (presumably safer) narrow AI staying competitive, instead of general systems; Cursor’s Composer coding finetune of Kimi is probably the most intense attempt to specialise a general model: probably more than 10^25 FLOPs of post-training; We find that, compared to its base model, Composer shows…

TL;DR # We’re interested in the prospects for (presumably safer) narrow AI staying competitive, instead of general systems. Cursor’s Composer coding finetune of Kimi is probably the most intense attempt to specialise a general model: probably more than 10^25 FLOPs of post-training. We find that, compared to its base model, Composer shows significant gains (+20% to 60%) on visual reasoning benchmarks, RPG-style games, and (as you’d hope) agentic coding. But, surprisingly, we also saw severe losses (-30% to -40%) on mathematical and scientific reasoning benchmarks. This cuts against the old idea

Explore this link on the map →

saved by

related reading