Flagship · Apr 2026

Kimi K2.6

Strong reasoning model with excellent long-context understanding.

By Moonshot · Text · Modified MIT

Context

256K tokens

Modality

Text

Architecture

MoE

License

Modified MIT

Per 1M tokens $0.75 in · $3.50 out
Model details

Specs & substance

Developed by
Moonshot
Model family
Kimi
Use case
Flagship
Modality
Text
Context window
256K tokens
Architecture
MoE
Version
K2.6
License
Modified MIT
Pricing
$0.75 in · $3.50 out · $0.16 cache read
Released
Apr 2026
Endpoint
parasail-kimi-k26

Strong reasoning model with excellent long-context understanding.

Kimi K2.6 is part of the Kimi family by Moonshot, and sits in the flagship category. It supports a 256K tokens context window and is built on a MoE architecture.

Key strengths: reasoning, long-context. Parasail serves Kimi K2.6 on a global fleet of current-gen GPUs behind a single OpenAI-compatible endpoint — with per-token pricing, no minimums, and dedicated capacity options when you need guaranteed throughput.

Integrate

Drop-in via the OpenAI SDK

Point any OpenAI-compatible client at Parasail and change the model name. That's it.

python parasail · Kimi K2.6
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.parasail.io/v1"
)

response = client.chat.completions.create(
    model="parasail-kimi-k26",
    messages=[
        {"role": "user", "content": "Hello, what can you do?"}
    ],
    stream=True,
    max_tokens=1000
)

for chunk in response:
    if chunk.choices and chunk.choices[0].delta.content is not None:
        print(chunk.choices[0].delta.content, end="", flush=True)
More models

Explore the library

All models

Start building today

Instantly run any open model — popular or specialized.