Skip to content
RCreddit.com·
Not on the current live radar

I ran GPT-6 Luna Max on MathArena's harness

AI summary

A developer tested GPT-6 Luna Max on MathArena's harness, noting its strong performance on Riemann Bench and considering it an underrated model. The model correctly answered 16 out of 19 finite or discrete questions, 14 out of 20 analysis and probability questions, and 9 out of 18 geometry, algebra, and topology questions. The developer abandoned testing another model, xhigh, after it performed worse than Luna Max.

Why this one

This report uniquely offers specific performance metrics for GPT-6 Luna Max across different mathematical domains, unlike general claims of strong performance.

Time & source

Times shown in UTC

Display time zone: UTC

Local time zone unavailable; showing UTC.

IngestedOffset at this time: UTC+0Sep 27, 2026, 05:00 UTC

Ingested
Sep 27, 2026, 05:00
Source type
Dev community

Full text isn't available here.

Read at source →
Source·reddit.com