Skip to content
HN On Hacker News ↗

Igor Kotenkov (@stalkermustang) on X

▲ 10 points 4 comments by paulpauper 4w ago HN discussion ↗

Pangram verdict · v3.3

We believe that this entire text is human-written.

0 %

AI likelihood · overall

Human
100% human-written 0% AI-generated
SEGMENTS · HUMAN 1 of 1
SEGMENTS · AI 0 of 1
WORD COUNT 273
PEAK AI % 0% · §1
Analyzed
Aug 1
backend: pangram/v3.3
Segments scanned
1 windows
avg 273 words each
Distribution
100 / 0%
human / AI fraction
Verdict
Human
Pangram v3.3

Article text · 273 words · 1 segments analyzed

Human AI-generated
§1 Human · 0%

It's hard for an ordinary person to understand the complexity of these tasks. I'm no mathematician, and I don't see a difference between e.g., results 3 and 10. So I had GPT-5.6 Sol Pro and Fable 5 Max classify these using @EpochAIResearch OpenMath's rubric: — "Solid Result": A strong researcher in the area would be happy if their median output addressed problems of this caliber. Still, the problem would probably not get much engagement outside of its subfield. — "Major Advance": The median person working in a broad area of mathematics (on the scale of number theory or graph theory) would take note, and would likely make the time to understand at least the outline of the solution. — "Breakthrough": The median mathematician would want to know about this result, even if it was outside their area. It would be a candidate for one of the best results of the year in all of mathematics === Both Fable and Sol agree #3 is a Breakthrough (which explains why @SebastienBubeck opens his tweet with it). They also agree that at least 7 are Major Advancements. Fable thinks #7 is just a Solid Result, while Sol assigns the "Major Advance" label. What's also interesting is that the official Epoch.AI rubrics say this: > When multiple tiers seemed plausible for a problem, we erred in the conservative direction. It would be disappointing to downgrade a problem’s notability after it was solved, whereas we can always highlight any unexpectedly interesting elements of a solution. And Fable 5 thinks that at least 3 of the results are "Borderline Breakthrough" (#1, #4, and #9). @AcerFur any thoughts on this