ISO-AdamW tested: 1000-question benchmark shows only 4 answers over AdamW
10/03/2026 — 10/03, 06:11·1 sources·1 reports
Story overview
On October 3, 2026, a DEV Community write-up described a hands-on test of ISO-AdamW, a new geometric optimizer the author had implemented, benchmarked against a baseline AdamW on the same set of problems. The comparison used a 1,000-question held-out test. ISO-AdamW answered 758 questions correctly; the baseline AdamW answered 754. The gap was four answers.
The author called out that gap directly and imagined the more exciting headline it could have supported — something along the lines of "New geometric optimizer beats AdamW! Time to rewrite the RL pipeline!" — but did not actually make that claim. No conclusion that the new optimizer beats AdamW was offered, and no comparison beyond the single held-out set was reported.
The write-up also documents a separate piece of engineering work: fixing a painfully slow matrix operation encountered during the implementation. That fix is part of the same record, alongside the benchmark numbers.
As of this report, the work stops at the published measurement. The only figures given are the 758 versus 754 result on 1,000 held-out questions and the resulting four-answer difference. The author did not state whether testing would be expanded, whether the RL pipeline would be changed, or whether ISO-AdamW would be evaluated on any other benchmark or task.
AI-generated from 1 reports · updated 1 hour ago
Latest turnA developer implemented a new geometric optimizer, ISO-AdamW, and ran it against baseline AdamW on the same 1,000-question held-out test: 758 correct answers versus 754, a four-question gap. No clear win—and the write-up also covers fixing a painfully slow matrix operation along the way.

Reports on this story headlines open the original
A developer implemented a new geometric optimizer, ISO-AdamW, and ran it against baseline AdamW on the same 1,000-question held-out test: 758 correct answers versus 754, a four-question gap. No clear win—and the write-up also covers fixing a painfully slow matrix operation along the way.
DEV Community · AIAI score 62
Other stories people are talking about
- 659RisingNVIDIA launches 64GB DGX Spark desktop AI computer at $4,9998 sources
- 539NewApple tightens macOS Full Disk Access over AI agent risks6 sources
- 476NewMeta open-sources Muse Gadgets firmware and SDK5 sources
- 303Google releases Gemini 4 Argon as next-gen frontier AI model8 sources
- 270NewAnthropic launches Claude Frontier Academy with $100M3 sources
- 248SurgeOpenAI DevDay 2026 launches Dots, Decisions API, and more3 sources
How is heat calculated?About the methodHide
Heat counts how many independent sources covered a story in the last 48 hours: one source counts once no matter how many posts it published, decaying with a 24-hour half-life. What ranks first is what many people are talking about.
This page aggregates public feeds. Headlines and summaries are machine-organized and remain the property of the original authors; verify important facts at the source.
- Surge
- Discussion rising fast
- New
- First report within 6 hours
- Rising
- Still gathering discussion
