← model directory

sao10k

Sao10K: Llama 3 8B Lunaris

sao10k/l3-lunaris-8b

Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3. It's a strategic merge of multiple models, designed to balance creativity with improved logic and general knowledge....

specs & pricing
Type
text
Provider
sao10k
Model ID
sao10k/l3-lunaris-8b
Context window
8K tokens
Input price
$0.044 / 1M tokens
Output price
$0.055 / 1M tokens

Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.

What it costs to run

cloud · 1M in + 1M out

$0.099

$0.044 for a million input tokens plus $0.055 for a million output tokens, billed from credits only when a request fails over to the cloud.

your hardware

cloud only

Sao10K: Llama 3 8B Lunaris is a hosted model, so it is always served from the cloud. Route easy traffic to an open model on your own hardware and keep this one for the requests that need it.

/// initialize

Call Sao10K: Llama 3 8B Lunaris through one OpenAI-compatible endpoint.

Mix it with open models on your own hardware — the gateway routes each request to the cheapest place that can serve it.

no credit card · 2 nodes free · openai-compatible

Frequently asked questions

What is the context window of Sao10K: Llama 3 8B Lunaris?
Sao10K: Llama 3 8B Lunaris accepts up to 8,192 tokens (8K) of context per request.
How much does Sao10K: Llama 3 8B Lunaris cost per million tokens?
Through Wide Area Intelligence, Sao10K: Llama 3 8B Lunaris costs $0.044 per 1M input tokens and $0.055 per 1M output tokens when served from the cloud, billed from prepaid credits. A workload of 1M input plus 1M output tokens costs $0.099.

More from sao10k

← browse all models