HaiAI123

Curated Global AI Tools Directory

NewHot story
97
heat index
New

ISO-AdamW tested: 1000-question benchmark shows only 4 answers over AdamW

10/03/2026 — 10/03, 06:11·1 sources·1 reports

Story overview

On October 3, 2026, a DEV Community write-up described a hands-on test of ISO-AdamW, a new geometric optimizer the author had implemented, benchmarked against a baseline AdamW on the same set of problems. The comparison used a 1,000-question held-out test. ISO-AdamW answered 758 questions correctly; the baseline AdamW answered 754. The gap was four answers.

The author called out that gap directly and imagined the more exciting headline it could have supported — something along the lines of "New geometric optimizer beats AdamW! Time to rewrite the RL pipeline!" — but did not actually make that claim. No conclusion that the new optimizer beats AdamW was offered, and no comparison beyond the single held-out set was reported.

The write-up also documents a separate piece of engineering work: fixing a painfully slow matrix operation encountered during the implementation. That fix is part of the same record, alongside the benchmark numbers.

As of this report, the work stops at the published measurement. The only figures given are the 758 versus 754 result on 1,000 held-out questions and the resulting four-answer difference. The author did not state whether testing would be expanded, whether the RL pipeline would be changed, or whether ISO-AdamW would be evaluated on any other benchmark or task.

AI-generated from 1 reports · updated 1 hour ago

Latest turnA developer implemented a new geometric optimizer, ISO-AdamW, and ran it against baseline AdamW on the same 1,000-question held-out test: 758 correct answers versus 754, a four-question gap. No clear win—and the write-up also covers fixing a painfully slow matrix operation along the way.

Reports on this story headlines open the original

Today
  1. A developer implemented a new geometric optimizer, ISO-AdamW, and ran it against baseline AdamW on the same 1,000-question held-out test: 758 correct answers versus 754, a four-question gap. No clear win—and the write-up also covers fixing a painfully slow matrix operation along the way.

    DEV Community · AIAI score 62

Other stories people are talking about

How is heat calculated?About the method

Heat counts how many independent sources covered a story in the last 48 hours: one source counts once no matter how many posts it published, decaying with a 24-hour half-life. What ranks first is what many people are talking about.

This page aggregates public feeds. Headlines and summaries are machine-organized and remain the property of the original authors; verify important facts at the source.

Surge
Discussion rising fast
New
First report within 6 hours
Rising
Still gathering discussion

Back to the hot board →