> ## Documentation Index
> Fetch the complete documentation index at: https://docs.deepshi.ai/llms.txt
> Use this file to discover all available pages before exploring further.

> ## Agent Instructions
> The Deepshi API is an OpenAI-compatible gateway. Base URL: https://api.deepshi.ai/v1. API keys start with sk-bf- and go in the Authorization: Bearer header. Prefer the official OpenAI SDKs pointed at the Deepshi base URL. Model ids are clean with no provider prefix (e.g. deepshi-3.0, claude-opus-4.8, gpt-5.5). Chat and image requests are synchronous; video and music requests are asynchronous job APIs (create, then poll). Every synchronous response carries usage.cost.total_cost in USD. Do not reference the internal /api/* admin plane or virtual keys.

# Chat models

> Every chat model available through the Deepshi API, with context windows and capabilities.

Use these models with the [chat completions](/capabilities/text-and-chat) endpoint. Reference a model by its **id** in the `model` field.

<Note>
  The exact set your key can call is returned by `GET /v1/models`. For per-model
  rates, see [deepshi.ai](https://deepshi.ai/).
</Note>

## Deepshi models

All of Deepshi's own models are uncensored.

<CardGroup cols={2}>
  <Card title="Deepshi 2.0" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/deepshi.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=2d2758984e3f279a6559cbe7be1c7809" width="36" height="36" data-path="logos/deepshi.svg">
    `deepshi-2.0` · **Context** 128K

    Highly uncensored, with open, minimally-filtered responses. The most direct option.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Deepshi 3.0" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/deepshi.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=2d2758984e3f279a6559cbe7be1c7809" width="36" height="36" data-path="logos/deepshi.svg">
    `deepshi-3.0` · **Context** 256K

    Multimodal flagship with strong visual understanding and agentic coding.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>
</CardGroup>

## Third-party models

The latest models from leading providers, through the same endpoint, key, and balance.
Recently added models appear first.

<CardGroup cols={2}>
  <Card title="Inkling" icon="https://mintcdn.com/evermindlabs/W-vm-DdmM6WsnMY7/logos/thinkingmachines.svg?fit=max&auto=format&n=W-vm-DdmM6WsnMY7&q=85&s=b737f8905b1a9010c208ecb79096e4ab" width="16" height="16" data-path="logos/thinkingmachines.svg">
    `inkling` · **Context** 1M

    Thinking Machines' open-weight MoE model for reasoning, coding, and agentic work, with image and audio input.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Kimi K3" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/moonshot.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=41e89155a0ad5c9ec6ad4487c7bfd8b4" width="16" height="16" data-path="logos/moonshot.svg">
    `kimi-k3` · **Context** 1M

    Moonshot's open-weight reasoning model for large codebases, tool use, and long autonomous runs.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Claude Sonnet 5" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/anthropic.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=15968d8fded62b947be1619c8f33ddcf" width="16" height="16" data-path="logos/anthropic.svg">
    `claude-sonnet-5` · **Context** 1M

    Anthropic's most capable Sonnet, balancing speed and depth for coding and agent workflows.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Claude Fable 5" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/anthropic.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=15968d8fded62b947be1619c8f33ddcf" width="16" height="16" data-path="logos/anthropic.svg">
    `claude-fable-5` · **Context** 1M

    Anthropic's agentic Claude 5 model for long-running, minimally-supervised coding and research.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GPT-5.6 Sol" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-5.6-sol` · **Context** 1M

    The flagship of OpenAI's GPT-5.6 series, built for complex reasoning and multi-step coding.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GPT-5.6 Terra" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-5.6-terra` · **Context** 1M

    The balanced GPT-5.6 tier for everyday coding, reasoning, and agentic tasks.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GPT-5.6 Luna" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-5.6-luna` · **Context** 1M

    The fast, cost-efficient GPT-5.6 tier for high-volume, latency-sensitive work.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Grok 4.5" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/xai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=64caadbafacada32ba0493df97a6575a" width="16" height="16" data-path="logos/xai.svg">
    `grok-4.5` · **Context** 500K

    xAI's most capable model, with frontier performance across coding, knowledge work, and STEM.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Kimi K2.7 Code" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/moonshot.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=41e89155a0ad5c9ec6ad4487c7bfd8b4" width="16" height="16" data-path="logos/moonshot.svg">
    `kimi-k2.7-code` · **Context** 256K

    Moonshot's coding-focused Kimi model for end-to-end programming over long contexts.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Nemotron 3 Ultra" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/nvidia.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b845cf278d049ade9827e6851dcf2b17" width="16" height="16" data-path="logos/nvidia.svg">
    `nemotron-3-ultra-550b-a55b` · **Context** 1M

    NVIDIA's open frontier reasoning and orchestration model with a 1M context.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="MiniMax M3" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/minimax.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b9c2dc9a9285c81193fd647e9ecca1a5" width="16" height="16" data-path="logos/minimax.svg">
    `minimax-m3` · **Context** 1M

    MiniMax multimodal model for long-horizon agentic work, with a 1M context.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Step 3.7 Flash" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/stepfun.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=c85c3f00d79448e7936e6831c90f897e" width="16" height="16" data-path="logos/stepfun.svg">
    `step-3.7-flash` · **Context** 256K

    StepFun's efficient multimodal model with native image and video understanding.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Grok Build 0.1" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/xai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=64caadbafacada32ba0493df97a6575a" width="16" height="16" data-path="logos/xai.svg">
    `grok-build-0.1` · **Context** 256K

    xAI's fast coding model built for agentic software engineering.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Granite 4.1 8B" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/ibm.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=cd4dd37cb46cb9044c26a9969d447c42" width="16" height="16" data-path="logos/ibm.svg">
    `granite-4.1-8b` · **Context** 128K

    IBM's dense 8B model for enterprise retrieval, tools, and structured text.

    <Icon icon="wrench" /> Tools
  </Card>

  <Card title="Qwen3.6 35B A3B" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/qwen.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=bad4d2f5e45baa557078c4ded17e0dde" width="16" height="16" data-path="logos/qwen.svg">
    `qwen3.6-35b-a3b` · **Context** 256K

    Alibaba's open multimodal MoE balancing quality and efficient inference.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Qwen3.6 27B" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/qwen.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=bad4d2f5e45baa557078c4ded17e0dde" width="16" height="16" data-path="logos/qwen.svg">
    `qwen3.6-27b` · **Context** 256K

    Dense 27B Qwen model with multimodal input for general-purpose work.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="DeepSeek V4 Pro" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/deepseek.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=695545f6090253507c4491c9fb9a870c" width="16" height="16" data-path="logos/deepseek.svg">
    `deepseek-v4-pro` · **Context** 1M

    DeepSeek's large MoE for advanced reasoning and coding over a 1M context.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="DeepSeek V4 Flash" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/deepseek.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=695545f6090253507c4491c9fb9a870c" width="16" height="16" data-path="logos/deepseek.svg">
    `deepseek-v4-flash` · **Context** 1M

    Efficiency-tuned DeepSeek MoE for fast, low-cost reasoning at a 1M context.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="MiMo V2.5 Pro" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/xiaomi.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=744bb620a742b4fc38f34eb54aa889d7" width="16" height="16" data-path="logos/xiaomi.svg">
    `mimo-v2.5-pro` · **Context** 1M

    Xiaomi's flagship for agentic capability and complex software engineering.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="GPT-5.4 Mini" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-5.4-mini` · **Context** 400K

    Faster, efficient GPT-5.4 variant for high-throughput workloads.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GPT-5.4 Nano" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-5.4-nano` · **Context** 400K

    The lightest GPT-5.4 model, tuned for low-latency, high-volume use.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GLM 5 Turbo" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/zai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=bba42d0291592cd3fffcd02781315ab4" width="16" height="16" data-path="logos/zai.svg">
    `glm-5-turbo` · **Context** 256K

    Z.ai's fast-inference model tuned for agent-driven workflows.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="Nemotron 3 Super" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/nvidia.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b845cf278d049ade9827e6851dcf2b17" width="16" height="16" data-path="logos/nvidia.svg">
    `nemotron-3-super-120b-a12b` · **Context** 1M

    NVIDIA's 120B hybrid MoE for efficient multi-agent applications.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="MiniMax M2.5" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/minimax.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b9c2dc9a9285c81193fd647e9ecca1a5" width="16" height="16" data-path="logos/minimax.svg">
    `minimax-m2.5` · **Context** 200K

    MiniMax model trained for real-world productivity and coding.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="Qwen3 Coder Next" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/qwen.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=bad4d2f5e45baa557078c4ded17e0dde" width="16" height="16" data-path="logos/qwen.svg">
    `qwen3-coder-next` · **Context** 256K

    Qwen's open coding model for coding agents and local development.

    <Icon icon="wrench" /> Tools
  </Card>

  <Card title="GLM 4.7 Flash" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/zai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=bba42d0291592cd3fffcd02781315ab4" width="16" height="16" data-path="logos/zai.svg">
    `glm-4.7-flash` · **Context** 200K

    Z.ai's 30B-class model balancing performance with agentic coding.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="Gemma 4 26B A4B" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/google.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=10fdc070e84e6698933120255ba20e4d" width="16" height="16" data-path="logos/google.svg">
    `gemma-4-26b-a4b-it` · **Context** 256K

    Google's instruction-tuned MoE delivering near-31B quality at lower cost.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="MiniMax M2.7" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/minimax.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b9c2dc9a9285c81193fd647e9ecca1a5" width="16" height="16" data-path="logos/minimax.svg">
    `minimax-m2.7` · **Context** 200K

    Next-generation MiniMax model for autonomous, agentic productivity.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="GPT-OSS 120B" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-oss-120b` · **Context** 128K

    OpenAI's open-weight 117B MoE for high-reasoning, agentic use.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="GPT-OSS 20B" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-oss-20b` · **Context** 128K

    OpenAI's open-weight 21B MoE under Apache 2.0, tuned for efficient reasoning.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="Claude Opus 4.8" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/anthropic.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=15968d8fded62b947be1619c8f33ddcf" width="16" height="16" data-path="logos/anthropic.svg">
    `claude-opus-4.8` · **Context** 1M

    Anthropic's most capable Opus model, with deep reasoning over long, complex tasks.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Claude Opus 4.7" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/anthropic.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=15968d8fded62b947be1619c8f33ddcf" width="16" height="16" data-path="logos/anthropic.svg">
    `claude-opus-4.7` · **Context** 1M

    Opus-family model built for long-running, asynchronous agents.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Claude Sonnet 4.6" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/anthropic.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=15968d8fded62b947be1619c8f33ddcf" width="16" height="16" data-path="logos/anthropic.svg">
    `claude-sonnet-4.6` · **Context** 1M

    Strong all-rounder for coding, agents, and professional work.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GPT-5.5" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-5.5` · **Context** 1M

    OpenAI's frontier model for complex professional workloads.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GPT-5.4" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-5.4` · **Context** 1M

    Frontier model unifying the GPT and Codex lines, strong at agentic coding.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GPT-4.1" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-4.1` · **Context** 1M

    Tuned for precise instruction-following and software engineering.

    <Icon icon="wrench" /> Tools  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GPT-4o" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/openai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=b86014f84a1f156fd42dc525ae3dbfca" width="16" height="16" data-path="logos/openai.svg">
    `gpt-4o` · **Context** 128K

    GPT-4-class intelligence that runs faster and cheaper. Solid general-purpose pick.

    <Icon icon="wrench" /> Tools  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Gemini 3.5 Flash" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/google.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=10fdc070e84e6698933120255ba20e4d" width="16" height="16" data-path="logos/google.svg">
    `gemini-3.5-flash` · **Context** 1M

    Google's high-efficiency model with near-Pro reasoning at Flash speed and cost.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Grok 4.3" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/xai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=64caadbafacada32ba0493df97a6575a" width="16" height="16" data-path="logos/xai.svg">
    `grok-4.3` · **Context** 1M

    xAI reasoning model suited to agentic workflows and high factual accuracy.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Grok 4.20" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/xai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=64caadbafacada32ba0493df97a6575a" width="16" height="16" data-path="logos/xai.svg">
    `grok-4.20` · **Context** 2M

    Fast xAI reasoning model with strong tool-calling and low hallucination.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="GLM-5.2" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/zai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=bba42d0291592cd3fffcd02781315ab4" width="16" height="16" data-path="logos/zai.svg">
    `glm-5.2` · **Context** 1M

    Large-scale reasoning model for long-horizon agent and engineering work.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="GLM-5.1" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/zai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=bba42d0291592cd3fffcd02781315ab4" width="16" height="16" data-path="logos/zai.svg">
    `glm-5.1` · **Context** 200K

    A major step up in coding ability on long-horizon tasks.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="GLM-5" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/zai.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=bba42d0291592cd3fffcd02781315ab4" width="16" height="16" data-path="logos/zai.svg">
    `glm-5` · **Context** 200K

    Z.ai's flagship open model for complex systems and agent workflows.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning
  </Card>

  <Card title="Kimi K2.6" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/moonshot.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=41e89155a0ad5c9ec6ad4487c7bfd8b4" width="16" height="16" data-path="logos/moonshot.svg">
    `kimi-k2.6` · **Context** 256K

    Built for long-horizon coding, UI/UX generation, and multi-agent orchestration.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>

  <Card title="Gemma 4 31B" icon="https://mintcdn.com/evermindlabs/DGZ60wKm5MW0do78/logos/google.svg?fit=max&auto=format&n=DGZ60wKm5MW0do78&q=85&s=10fdc070e84e6698933120255ba20e4d" width="16" height="16" data-path="logos/google.svg">
    `gemma-4-31b-it` · **Context** 256K

    A \~31B dense open model with an optional reasoning mode and native tools.

    <Icon icon="wrench" /> Tools  <Icon icon="brain" /> Reasoning  <Icon icon="eye" /> Vision
  </Card>
</CardGroup>

<Tip>
  Not sure where to start? Use **`deepshi-2.0`** for highly uncensored
  responses, or **`deepshi-3.0`** for an all-round flagship with tools and
  reasoning.
</Tip>
