
Cloudflare Clef Explained (2026): Decision Models That Do Not Write Text, Plus the RL Fine-Tuning Stack
Cloudflare announced Clef and Clef-flash on October 1, 2026. This guide lines up the official blog post with the Workers AI model pages: how a decision model returns typed probabilities instead of text, the vision encoder and 64k context window, every published benchmark row including where Clef loses, median latency of 209 ms versus 524 ms for Jev, pricing at $0.24 and $0.09 per million input tokens, the Jev-compatible API, training on a frozen Qwen backbone with RLCD, the new reinforcement learning fine-tuning stack, and the caveats of self-hosting the Apache 2.0 weights.









































