Developer benchmarks Cloudflare's new Clef model against hosted Jev API for agent tasks
Cloudflare released Clef, an open-weights decision model designed as an alternative to TypeSafe's hosted Jev API for agent developers. An independent developer benchmarked Clef-Flash 9B running locally on an RTX 3090 against the Jev API using 42 real decisions from an autonomous coding agent. Both models performed identically on computer-use action choices (10/10 correct) and agent supervision tasks (8/12 correct), including safety-critical decisions. Jev showed an advantage only in message triage tasks, scoring 12/20 versus Clef's 10/20, leading to overall accuracy rates of 71.4% for Jev and 66.7% for Clef.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in