Jev AI

Model comparison · Updated October 2, 2026

Jev vs Clef AI

Jev, Clef and Clef Flash turn state and typed questions into decision probabilities. Compare their capabilities and Cloudflare's published evaluation results before choosing a model for your workflow.

TypeSafe

Jev

Keep an established Jev workflow when its measured accuracy and thresholds meet your requirements. Cloudflare's comparison still gives Jev the lead on some tasks.

Cloudflare · 27B

Clef

Evaluate the larger Cloudflare model for classification quality, a larger context window or an open-weight deployment.

Cloudflare · 9B

Clef Flash

Evaluate the smaller model for latency-sensitive decisions. Its reported speed comes with task-dependent quality differences.

Capabilities and integration

Model documentation and the Cloudflare launch comparison, reviewed October 2, 2026
CapabilityJevClefClef Flash
DeveloperTypeSafeCloudflareCloudflare
Typed outputsYes/no, choice, scoreYes/no, choice, scoreYes/no, choice, score
Documented model context32K (Cloudflare launch comparison)65,536 tokens65,536 tokens
Model-level visionText-only native JevEmbedded imagesEmbedded images
Inputs in this Jev AI integrationText and JSONText and JSON; images not enabledText and JSON; images not enabled
HostingHosted Jev APIWorkers AI or self-hosted weightsWorkers AI or self-hosted weights
Jev AI API selectorjev-latest / jev-preview / jev-1.13.0clefclef-flash

The documented context window is a model limit, not a promise that this site's input limit is the same. Jev AI's playground and API limits apply. Cloudflare may truncate long text to fit its model context.

Cloudflare's published decision benchmarks

The following is the complete ten-task table from Cloudflare's October 1 launch article, restricted to the three models compared here. Higher is better within each row. Metrics differ between rows; these values are not averaged into a new overall score.

Vendor-reported results, not measurements taken through Jev AI. This is a different evaluation from the JevBench v1.5.0 snapshot on our other comparison pages. Scores and rankings should not be mixed.

Cloudflare-reported results · higher is better
Task / metricJevClefClef Flash
BFCLCase exact95.7598.4798.76
ToolRetnDCG@1065.2869.1966.43
API-BankAccuracy88.1991.9393.11
Home appliancesCase exact52.2782.9597.73
When2CallAccuracy80.9772.3765.58
BANKING77Macro-F179.7494.2090.93
CLINC150+OOSMacro-F189.2797.4366.77
BRIGHTnDCG@1047.5245.9139.26
Amazon ESCIMacro-F155.2157.4857.39
PhishNChipsAccuracy62.5579.6075.05

Clef leads Jev on eight of these ten rows. Jev leads on When2Call and BRIGHT. Clef Flash also trails Jev on CLINC150+OOS, which is a reason to test out-of-scope detection before replacing a classifier.

Workflow evaluations

Cloudflare also reports these four results against TypeSafe's workflow evaluation suite. The best choice changes with the workflow.

Cloudflare-reported workflow scores · higher is better
WorkflowJevClefClef Flash
Invoice processing61.864.757.1
Customer service76.076.377.0
Security incidents61.762.961.7
Agent trace observability71.668.569.8

Reported latency and upstream cost

Cloudflare's reported latency across 43 evaluation benchmarks · milliseconds, lower is better
MeasurementJevClefClef Flash
Median latency524.1209.338.8
p95 latency536.0238.6122.4

These figures are not an end-to-end Jev AI latency guarantee. Network distance, input length, question count, queues and the serving path affect your result.

Cloudflare lists Clef at $0.24 per million input tokens and Clef Flash at $0.09 per million input tokens. Those are upstream Workers AI prices, not Jev AI plan prices. Requests through this site use the existing Jev AI credit and token billing. View Jev AI pricing.

Evaluate before switching

  1. Keep the same state, questions and allowed labels for each model.
  2. Include difficult examples, missing information and cases outside your normal categories.
  3. Compare decisions against human-reviewed labels; measure both missed cases and false alarms.
  4. Recheck probability thresholds, real request latency and billed input usage before changing a production route.

Clef model guide and API example · Create a Jev AI API key

Common questions

Is Clef the same model as Jev?

No. Clef is a separate Cloudflare model family. It uses a compatible state-and-questions interface, so the request structure is familiar, but model behavior, probabilities and limits can differ.

Is Clef always more accurate?

No. Cloudflare reports stronger Clef results on several classification and tool benchmarks. Its own table also shows Jev ahead on When2Call and BRIGHT, and on agent trace observability in the workflow evaluation. Evaluate the tasks and error types that matter to your application.

Can I use the same Jev AI API key?

Yes, for models configured in this environment. Send the selected model to POST /api/v1/systemone. Check authenticated GET /api/v1/models first. Clef requests are sent to Cloudflare and are never silently replaced with Jev.

Can I send images to Clef here?

This integration currently accepts text and JSON only. Cloudflare documents embedded-image input on Workers AI, but that input path is not enabled through Jev AI. Native Jev and the separate Jev-Omni model should not be confused.

Can I reuse my production thresholds?

Reuse your question schema as a starting point, then evaluate the new model on held-out examples. A threshold chosen for one model can produce a different false-positive or false-negative rate on another.

Sources

Reviewed October 2, 2026. Jev AI is an independent integration, not the official site of TypeSafe or Cloudflare.