HaiAI123

Curated Global AI Tools Directory

SurgeHot story
156
heat index
↑ 164%

Cloudflare releases Clef and Clef-flash decision models

10/02/2026 — 10/03, 02:28·2 sources·2 reports

Story overview

On October 2, 2026, Cloudflare released Clef and Clef-flash, a pair of open-weight decision models at 27B and 9B parameters. Instead of generating text, the models return typed probabilities, a design meant to let AI agents make structured decisions without an LLM writing prose first. Both are Jev-API compatible and accept image input, and they run on Workers AI with median latencies of 209.3 ms and 38.8 ms respectively.

The following day, The Decoder added detail: both models are built on Qwen and released under Apache 2.0, and they are aimed at letting AI agents make structured decisions without producing text. According to that report, Clef-flash returns classifications in about 39 ms, making it more than ten times faster than TypeSafe AI's Jev, and the release is framed as a challenge to TypeSafe AI's Jev decision model. The report's headline also states Cloudflare's position that the new Clef model means humans no longer need to be in the loop for agent decisions.

So far the story stops at launch: there is no word in either report on availability beyond Workers AI, pricing, or adoption. The two accounts agree on the model sizes, the open weights, the Apache 2.0 license, the Qwen base, Jev-API compatibility and image input. On speed, the numbers appear as 38.8 ms as a median latency on Workers AI and roughly 39 ms for returning classifications; the reports do not say whether those two figures describe exactly the same measurement.

AI-generated from 2 reports · updated 52 minutes ago

Latest turnCloudflare has introduced Clef and Clef-flash, two Qwen-based models released under Apache 2.0 that let AI agents make structured decisions without generating text. Clef-flash returns classifications in about 39 milliseconds, more than ten times faster than TypeSafe AI's Jev, and Cloudflare argues humans no longer need to sit in the loop.

Related tools
24-hour heatpeak 158 · now
24 hours agonow

Reports on this story headlines open the original

Today
  1. Cloudflare has introduced Clef and Clef-flash, two Qwen-based models released under Apache 2.0 that let AI agents make structured decisions without generating text. Clef-flash returns classifications in about 39 milliseconds, more than ten times faster than TypeSafe AI's Jev, and Cloudflare argues humans no longer need to sit in the loop.

    The DecoderAI score 85
Yesterday
  1. Cloudflare released Clef (27B) and Clef-flash (9B), open-weight decision models that return typed probabilities instead of text. They are Jev-API compatible, accept images, and run on Workers AI at median latencies of 209.3 ms and 38.8 ms.

    MarkTechPostAI score 78

Other stories people are talking about

How is heat calculated?About the method

Heat counts how many independent sources covered a story in the last 48 hours: one source counts once no matter how many posts it published, decaying with a 24-hour half-life. What ranks first is what many people are talking about.

This page aggregates public feeds. Headlines and summaries are machine-organized and remain the property of the original authors; verify important facts at the source.

Surge
Discussion rising fast
New
First report within 6 hours
Rising
Still gathering discussion

Back to the hot board →