Pangram verdict · v3.3
We believe that this entire text is AI.
AI likelihood · overall
AIArticle text · 200 words · 1 segments analyzed
fig. i — the other sophont, observed soph·ont · n. one capable of reason Releasing soon We reject the paradigm of scale. 1× Sophontic 60× Competitors, on reasoning evaluations Sophontic cultivates geometric density in the substrate to achieve reasoning performance exceeding that of models up to 60× its size. We have pioneered methods of directly training the internal geometry of the model. fig. ii — thesis Observed signal Reasoning as geometry, not mass. The prototype is being evaluated through paired perturbations, calibrated anchors, and public transfer surfaces. The measurement standard remains stable as the rail set evolves. 60× Larger models exceeded on reasoning evaluations. plate — the long argument Methodology The perturbation paradigm. We measure reasoning by flip rate — a paired-item test that precludes the surface-feature heuristics most models use to game conventional evaluations. The pair is the unit of measurement, not the item. Reasoning, not recall. Inspect the eval kit Flip-rate instrument paired item canonical If a proof depends on a fact, the answer follows. stable perturbed Change the load-bearing fact. The answer must change. flip fig. iii — method The model and the evaluation fig. iv — artefacts Colophon Company Sophontic, Inc. Structure Delaware C-corporation Founded 2026