// pricing · September 2026

Sao10K API Pricing

Every Sao10K model's list price per 1M tokens — input, output and cached input — with context windows, LiveBench scores and what a real workload costs per month.

OpenRouter · refreshed every 15 minutesAll providersBest value modelsCheapest with tool calling

3 Sao10K models are listed, with a median list price of $0.650 per 1M input tokens and $0.750 per 1M output tokens. Paid prices run from $0.042 (Llama 3 8B Lunaris) to $0.850 (Llama 3.1 Euryale 70B v2.2) per 1M tokens blended.

Models
3
Median input / 1M
$0.650
paid models
Median output / 1M
$0.750
paid models
Cheapest paid
$0.042
Llama 3 8B Lunaris
Prompt caching
0
publish a cached-input price

Sao10K model prices

List prices in USD per 1M tokens. “Blended” is a 3:1 input:output mix; the per-month column prices the selected workload using each model's cached-input rate where it publishes one.

Showing 3 of 3 models · per-month column: 8K in / 600 out × 100K requests, 50% of input cached

LLM API list prices per 1M tokens, context windows and monthly workload cost
sao10k/l3-lunaris-8b$0.040$0.050$0.0428K$35.00
sao10k/l3.3-euryale-70b$0.650$0.750$0.675131K$565.00
sao10k/l3.1-euryale-70b$0.850$0.850$0.850131K$731.00
Price your own workload

Sao10K pricing FAQ

What is the cheapest Sao10K model?

Llama 3 8B Lunaris is the cheapest paid Sao10K model, at $0.040 per 1M input tokens and $0.050 per 1M output tokens.

Why do output tokens cost more than input tokens?

Generating a token takes a full forward pass per token, while input is processed in parallel. Across the paid Sao10K models listed here the median output price is $0.750 per 1M against $0.650 for input — reasoning tokens bill as output too.

Which Sao10K models support prompt caching?

None of them publish a separate cached-input price right now.

How current are these prices?

List prices are pulled from OpenRouter’s public models API and refreshed every 15 minutes. Batch discounts and enterprise pricing are not included.

Other providers

What these prices include

  • List prices from OpenRouter's public models API, which mirrors provider pricing. Batch discounts and negotiated enterprise rates are not included.
  • Variants:free and :batch listings are folded into their base model; its model page lists them.
  • Per-month cost is a floor: tokens you describe times list price. Cache writes and reasoning-token volume depend on your workload.