The bet, stated falsifiably
2026-09-14
Historical note. This article records the September model thesis and may describe mechanisms or schedules that are no longer the active POC plan. Read the current status first.
H5
A 1.5B-parameter EOS-augmented model should beat a 70B conventional model on emotional intelligence benchmarks. If that holds, architecture, not scale, is the emotional-intelligence axis, and a foundation-model raise is justified.
If it is false, we say so and stop. Testing it costs under $1,000. Every pass/fail threshold is pre-registered before the experiments run, and the result posts publicly either way.
Why pre-register
Because a benchmark you control after the fact is not a benchmark. The preregistration document fixes the baselines, the thresholds, and the kill criteria before a single experiment runs. This is the part most AI claims skip, and the part we consider non-negotiable.