What was claimed
Neuraxon 2.0 scores 0.21% on the challenging ARC-AGI 3 benchmark (nearly double our previous entry), outperforming most frontier AIs/LLMs while using no LLMs, no GPUs, just CPUs.
Our verdict
Needs CautionNeuraxon's 0.18 score is lower than ChatGPT (0.20) and Claude (0.22), so it does not outperform these major frontier models. This claim is factually incorrect. Public posts reference prior Neuraxon entries around 0.13%–0.16%, which would make 0.21% less than double; however, the exact prior value used by the author is not documented, so the "nearly double" comparison cannot be confirmed.
Check your own claim
Paste any statement, headline, or AI answer — 3 independent AIs verify it in seconds, with sources.
Key findings
Outperforming most frontier AIs/LLMs
This result is nearly double our previous entry.
Using no LLMs, no GPUs, just CPUs
Neuraxon 2.0 outperforms most frontier AIs/LLMs on ARC-AGI-3 while using no LLMs and only CPUs (no GPUs).