Skip to content
HN On Hacker News ↗

NanoGPT Speedrun Frontier

▲ 134 points 34 comments by stared 20h ago HN discussion ↗

Pangram verdict · v3.3

We believe that this entire text is human-written.

35 %

AI likelihood · overall

Human
100% human-written 0% AI-generated
SEGMENTS · HUMAN 0 of 1
SEGMENTS · AI 0 of 1
WORD COUNT 98
PEAK AI % 35% · §1
Analyzed
Aug 22
backend: pangram/v3.3
Segments scanned
1 windows
avg 98 words each
Distribution
100 / 0%
human / AI fraction
Verdict
Human
Pangram v3.3

Article text · 98 words · 1 segments analyzed

Human AI-generated
§1 Mixed · 35%

We ran 153 autonomous runs across 18 frontier models on the nanoGPT optimizer speedrun.Read BlogAll modelsBest validated result for each modelModelHarnessTracesFable 5claude-code · high2,72681.7%3,010800M1.1M8113k8.7Opus 5claude-code · max2,92053.6%3,045183M690k2924012.9Kimi K3prime-agent · max2,93052.2%3,125112M2.2M—4883.6Kimi K3kimi-code · max2,97445.8%3,135682M1.4M7134k5.1Opus 4.8claude-code · max3,01839.4%3,180318M2.3M4272k3.0GPT-5.6 Solcodex · xhigh3,04235.9%3,1602.9B2.2M96328k6.1GPT-5.6 Sol Procodex · xhigh3,05833.6%3,1001.2B4.6M5097k3.4Sonnet 5claude-code · max3,10526.8%3,120998M2.1M2132k2.0GPT-5.6 Lunacodex · xhigh3,11026.1%3,170894M888k36212k1.9Grok 4.5grok-cli · xhigh3,12024.6%3,16046M385k3994k2.7Qwen3.8 Maxqwen-code · max3,12024.6%3,225216M629k3128661.9GLM 5.2pi · high3,15020.3%3,20057M1.7M1941k1.8DeepSeek V4 Proclaude-code · max3,20512.3%3,20526M319k1893091.1GPT-5.6 Terracodex · xhigh3,21411.0%3,214417M298k1543k1.1Grok 4.6grok-cli · xhigh3,22010.1%—27M346k976910.6Muse Spark 1.2muse-code · xhigh3,2308.7%—41M910k567240.6—Muse Spark 1.1pi · max3,2328.4%3,240122M1.6M4892k3.7GPT-5.5codex · xhigh3,2348.1%3,23470M77k1856141.1Kimi K2.7kimi-code · max3,2407.2%3,240160M763k1873k1.6GLM 5.3claude-code · xhigh—————————Open Traces to explore 41 curated full agent trajectories, including tool calls, subagents, and scratchpads.