Sao10K's latest Stheno fine-tune optimized for instruction following.
Context Window
16.4K
Max Output
8.2K
Input Price (Auto)
$0.20/1M
Output Price (Auto)
$0.20/1M
Cache Read (Auto)
$0.10/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Sao10K Stheno 8b with similar models from the same provider or model family.
Llama 3.1 70B Euryale
Sao10K/L3.1-70B-Euryale-v2.2A 70B parameter model from SAO10K based on Llama 3.1 70B, offering high-quality text generation.
Llama 3.1 70B Hanami
Sao10K/L3.1-70B-Hanami-x1Euryale v2.2-based finetune.
Llama 3.3 70B Euryale
Sao10K/L3.3-70B-Euryale-v2.3A 70B parameter model from SAO10K based on Llama 3.3 70B, offering high-quality text generation.
Muse Glimmer 30B TEE
TEE/muse-glimmer-30bMeta's Muse Glimmer 30B is a dense, open-weight multimodal model for long-horizon agentic and coding workflows. Running inside a TEE (Trusted Execution Environment), with provider attestation support.
Muse Glimmer 30B
meta/muse-glimmer-30bMeta's Muse Glimmer 30B is a dense, open-weight multimodal model distilled from Muse Spark for long-horizon agents and coding workflows. It supports multi-step reasoning, reliable tool use, failure recovery, image understanding, and more than 100 languages.
Muse Spark 1.2
meta/muse-spark-1.2Meta's Muse Spark 1.2 is a multimodal reasoning model for complex agentic and coding tasks, with tool calling, structured output, and a one-million-token context window.