#Decision models
3 published articles
Perplexity Decisions API and pplx-decider-v1-27b: $0.04 per million input tokens, Apache 2.0 weights
The hosted API answers noul, choice and score questions against text, JSON or images. Docs list a 262,144 token input limit and 10 requests per second per organization.
Cloudflare Clef and Clef-flash: Apache 2.0 multimodal decision models on Workers AI, with a 64k context window
The 27B and 9B models return probabilities for typed questions in one pass. Cloudflare prices Clef at $0.24 per million input tokens and says median latency is 209 ms.
Strands Decider 2B: AWS's Strands Labs opens an Apache 2.0 decision model for agent routing and tool-call checks
The 1.9B model answers yes or no, choice and score questions in one pass, with a median 115 ms on an RTX 3090. Its maintainers rate it an experiment.