Pre-launchAPI access is by invitation. Prices shown are estimates.Request access →
Contact sales
ModelsPricingDedicatedDocsEnterpriseContact sales
Models

Docs / Models

Precision & versioning

What weights and numerical precision sit behind each Kurrens model id, and how that changes over time.

View .md

When you call a model on Kurrens, you should know exactly what you’re running. We make three commitments.

1. The id is the Hugging Face repository

Model ids are Hugging Face ids — deepseek-ai/DeepSeek-V4.1-Flash is the model published at huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash. We serve the developer’s released weights with the developer’s official chat template and tokenizer. We don’t fine-tune, distill, or merge them.

2. Precision is declared, never changed silently

Each model lists the numerical precision it runs at — for example FP8. The same value is published as quantization in kurrens.ai/models.json and shown on OpenRouter.

Precision What it means
BF16 / FP16 The weights as the developer trained or released them. Highest fidelity, highest cost.
FP8 8-bit floating point. Standard for serving large models; quality is very close to BF16 on most tasks.
FP4 / NVFP4 / MXFP4 4-bit formats. Cheaper and faster; more quality loss on long reasoning and code.

We never lower a model’s precision behind an existing id. If we offer another precision, it is listed as a separate model or service tier.

3. New releases get a new id

When a developer publishes an updated model, it gets its own id — for example deepseek-ai/DeepSeek-V4-Pro-0813 alongside deepseek-ai/DeepSeek-V4-Pro. We don’t swap the weights behind an id you already use. The older id keeps working until it is retired under the model lifecycle policy, with at least 30 days’ notice.

Current catalog

Model idPrecisionWeights
deepseek-ai/DeepSeek-V4.1-FlashFP8Hugging Face
deepseek-ai/DeepSeek-V4-Pro-0813FP8Hugging Face
deepseek-ai/DeepSeek-V4-Flash-0731FP8Hugging Face
zai-org/GLM-5.3FP8Hugging Face
zai-org/GLM-5.3-FlashFP8Hugging Face
zai-org/GLM-5.2FP8Hugging Face
moonshotai/Kimi-K3FP8Hugging Face
moonshotai/Kimi-K2.7-CodeFP8Hugging Face
moonshotai/Kimi-K2.6FP8Hugging Face
MiniMaxAI/MiniMax-M3FP8Hugging Face
MiniMaxAI/MiniMax-M2.7FP8Hugging Face
Qwen/Qwen3.8-27BFP8Hugging Face
Qwen/Qwen3.8-Flash-NextFP8Hugging Face
Qwen/Qwen3.8-2.4T-A95BFP8Hugging Face
Qwen/Qwen3.6-35B-A3BFP8Hugging Face
Qwen/Qwen3-Coder-NextFP8Hugging Face
XiaomiMiMo/MiMo-V2.6-Pro-RLFP8Hugging Face
XiaomiMiMo/MiMo-V2.6-Flash-RLFP8Hugging Face
openai/gpt-oss-120bFP8Hugging Face
openai/gpt-oss-20bFP8Hugging Face
google/gemma-4-31B-itFP8Hugging Face

    Type to search titles, headings, and page text.

    ↑↓ to move · ↵ to open