Experiential
ModelsLogs
Star us on GitHubDocsSettings
Sign in
Continue with GoogleContinue with GitHub
or
Models

Clef Flash

clef-flashby CloudflareZDR (all rungs)Free

Cloudflare's smaller, faster Clef: a 9B decision model that turns a state and a schema of typed questions (noul, choice, score) into calibrated probabilities in one forward pass. Served natively by Experiential Cloud through /v1/systemone with text state; not a chat or streaming model. Context: 65,536 tokens shared by state and questions.

Context

66K

Max output

no data

Price per million tokens

$0.090$0in$0.090$0cached$0out

Waterfall

  1. Experiential Cloudfree tierFree$0.090/$0100% up--
  2. 1
    Experiential Cloud$0.090/$0100% up--

How to use this model for coding agents

Clef Flash is a decision helper for your coding agent, not a chat model.

Agent prompt
Add clef-flash (Clef Flash) as a decision helper/tool for my coding agent.
Keep my coding agent on its current chat model. Clef Flash is not a primary coding/chat model and does not generate chat completions.
Write a small native HTTP helper; there is no SDK or vendor key for this integration:
- POST only to https://api.experientiallabs.ai/v1/systemone.
- Read the Bearer key only from EXPERIENTIAL_API_KEY. If missing, stop and ask me to set it locally; never ask me to paste or log it.
- Do not request any other provider API key or call any other endpoint.
- Send model "clef-flash" exactly, preserving any selected :free spelling.
Implement a native JSON request with model, state (text, object, or array), and questions (a map keyed by my question IDs). Each question needs type and instructions:
- choice: criteria maps options such as accept/review to their descriptions.
- noul: judge a proposition; optional criteria describe true and false.
- score: criteria is an ordered list of level descriptions.
Use a concise diff and test summary as state, not the entire repository. Ask independent questions; an answer is not fed into another question in the same request.
Clef Flash reads at most 65,536 tokens of context, shared by state and all question definitions.
Gateway limits: at most 32 questions per request, 64 options per choice, and 2-10 levels per score. These are platform admission bounds, not provider limits.
Request shape reference (the same SystemOne wire): https://docs.typesafe.ai/introduction/quickstart
Model card: https://huggingface.co/Cloudflare/clef-flash
Read answers by question ID (for example answers.review), and show the decision and reported usage to me. Probabilities and confidence are not guarantees of correctness. Never execute a returned choice, merge a change, or bypass review automatically.
Use a bounded timeout, surface HTTP and transport errors, and do not add automatic retries or Idempotency-Key: an unknown outcome may already be charged. This endpoint does not stream and has no Chat Completions, Responses, or Messages facade.
Prepare the helper and a mocked test first. Ask before making a potentially charged live call; do not claim it worked without a real response. Tell me which files you changed.

Native API quickstartAbout Clef Flash

CapabilitiesAccepts text

Sampling

Stop sequences

Limits

Max-tokens fieldmax_tokens

Sending an unsupported field? See error reference

Providers serving this model1
Status
Experiential Cloudfree tierno datano data100%66Kno dataFreeFreeFree-ZDRactivemeasured
Experiential Cloudexperientialno datano data100%66Kno data$0.090$0$0.090-ZDRactivemeasured
Hugging Face

Released

-

Promotion

Free