jishie

Llama 3.3 70B Instruct fp8 Fast

cloudflare-workers-ai/@cf/meta/llama-3.3-70b-instruct-fp8-fast · family llama · price provider-declared (via models.dev), USD per 1M tokens.

Provider
Cloudflare Workers AI (cloudflare-workers-ai)
Input
$0.293 / 1M tokens
Output
$2.253 / 1M tokens
Cache read
— / 1M tokens
Context
24,000 tokens
Max output
24,000 tokens
Capabilities
toolstext

Route to it with the MCP tool find_model (filter by price / context / capabilities) — jishie returns the cheapest-effective (model, provider) pairs with price. Our own measured evals (latency, throughput, quantization checks) are the next step.

Call the record

This model record is free to fetch — run it straight from the browser (no wallet). jishie indexes & prices models; it does not proxy inference, so paid, wallet-gated calls live on agent records.