Skip to main content
Use these models with the chat completions endpoint. Reference a model by its id in the model field.
The exact set your key can call is returned by GET /v1/models. For per-model rates, see deepshi.ai.

Deepshi models

All of Deepshi’s own models are uncensored.

Deepshi 2.0

deepshi-2.0 · Context 128KHighly uncensored, with open, minimally-filtered responses. The most direct option. Tools   Reasoning   Vision

Deepshi 3.0

deepshi-3.0 · Context 256KMultimodal flagship with strong visual understanding and agentic coding. Tools   Reasoning   Vision

Third-party models

The latest models from leading providers, through the same endpoint, key, and balance. Recently added models appear first.

Inkling

inkling · Context 1MThinking Machines’ open-weight MoE model for reasoning, coding, and agentic work, with image and audio input. Tools   Reasoning   Vision

Kimi K3

kimi-k3 · Context 1MMoonshot’s open-weight reasoning model for large codebases, tool use, and long autonomous runs. Tools   Reasoning   Vision

Claude Sonnet 5

claude-sonnet-5 · Context 1MAnthropic’s most capable Sonnet, balancing speed and depth for coding and agent workflows. Tools   Reasoning   Vision

Claude Fable 5

claude-fable-5 · Context 1MAnthropic’s agentic Claude 5 model for long-running, minimally-supervised coding and research. Tools   Reasoning   Vision

GPT-5.6 Sol

gpt-5.6-sol · Context 1MThe flagship of OpenAI’s GPT-5.6 series, built for complex reasoning and multi-step coding. Tools   Reasoning   Vision

GPT-5.6 Terra

gpt-5.6-terra · Context 1MThe balanced GPT-5.6 tier for everyday coding, reasoning, and agentic tasks. Tools   Reasoning   Vision

GPT-5.6 Luna

gpt-5.6-luna · Context 1MThe fast, cost-efficient GPT-5.6 tier for high-volume, latency-sensitive work. Tools   Reasoning   Vision

Grok 4.5

grok-4.5 · Context 500KxAI’s most capable model, with frontier performance across coding, knowledge work, and STEM. Tools   Reasoning   Vision

Kimi K2.7 Code

kimi-k2.7-code · Context 256KMoonshot’s coding-focused Kimi model for end-to-end programming over long contexts. Tools   Reasoning   Vision

Nemotron 3 Ultra

nemotron-3-ultra-550b-a55b · Context 1MNVIDIA’s open frontier reasoning and orchestration model with a 1M context. Tools   Reasoning

MiniMax M3

minimax-m3 · Context 1MMiniMax multimodal model for long-horizon agentic work, with a 1M context. Tools   Reasoning   Vision

Step 3.7 Flash

step-3.7-flash · Context 256KStepFun’s efficient multimodal model with native image and video understanding. Tools   Reasoning   Vision

Grok Build 0.1

grok-build-0.1 · Context 256KxAI’s fast coding model built for agentic software engineering. Tools   Reasoning   Vision

Granite 4.1 8B

granite-4.1-8b · Context 128KIBM’s dense 8B model for enterprise retrieval, tools, and structured text. Tools

Qwen3.6 35B A3B

qwen3.6-35b-a3b · Context 256KAlibaba’s open multimodal MoE balancing quality and efficient inference. Tools   Reasoning   Vision

Qwen3.6 27B

qwen3.6-27b · Context 256KDense 27B Qwen model with multimodal input for general-purpose work. Tools   Reasoning   Vision

DeepSeek V4 Pro

deepseek-v4-pro · Context 1MDeepSeek’s large MoE for advanced reasoning and coding over a 1M context. Tools   Reasoning

DeepSeek V4 Flash

deepseek-v4-flash · Context 1MEfficiency-tuned DeepSeek MoE for fast, low-cost reasoning at a 1M context. Tools   Reasoning

MiMo V2.5 Pro

mimo-v2.5-pro · Context 1MXiaomi’s flagship for agentic capability and complex software engineering. Tools   Reasoning

GPT-5.4 Mini

gpt-5.4-mini · Context 400KFaster, efficient GPT-5.4 variant for high-throughput workloads. Tools   Reasoning   Vision

GPT-5.4 Nano

gpt-5.4-nano · Context 400KThe lightest GPT-5.4 model, tuned for low-latency, high-volume use. Tools   Reasoning   Vision

GLM 5 Turbo

glm-5-turbo · Context 256KZ.ai’s fast-inference model tuned for agent-driven workflows. Tools   Reasoning

Nemotron 3 Super

nemotron-3-super-120b-a12b · Context 1MNVIDIA’s 120B hybrid MoE for efficient multi-agent applications. Tools   Reasoning

MiniMax M2.5

minimax-m2.5 · Context 200KMiniMax model trained for real-world productivity and coding. Tools   Reasoning

Qwen3 Coder Next

qwen3-coder-next · Context 256KQwen’s open coding model for coding agents and local development. Tools

GLM 4.7 Flash

glm-4.7-flash · Context 200KZ.ai’s 30B-class model balancing performance with agentic coding. Tools   Reasoning

Gemma 4 26B A4B

gemma-4-26b-a4b-it · Context 256KGoogle’s instruction-tuned MoE delivering near-31B quality at lower cost. Tools   Reasoning   Vision

MiniMax M2.7

minimax-m2.7 · Context 200KNext-generation MiniMax model for autonomous, agentic productivity. Tools   Reasoning

GPT-OSS 120B

gpt-oss-120b · Context 128KOpenAI’s open-weight 117B MoE for high-reasoning, agentic use. Tools   Reasoning

GPT-OSS 20B

gpt-oss-20b · Context 128KOpenAI’s open-weight 21B MoE under Apache 2.0, tuned for efficient reasoning. Tools   Reasoning

Claude Opus 4.8

claude-opus-4.8 · Context 1MAnthropic’s most capable Opus model, with deep reasoning over long, complex tasks. Tools   Reasoning   Vision

Claude Opus 4.7

claude-opus-4.7 · Context 1MOpus-family model built for long-running, asynchronous agents. Tools   Reasoning   Vision

Claude Sonnet 4.6

claude-sonnet-4.6 · Context 1MStrong all-rounder for coding, agents, and professional work. Tools   Reasoning   Vision

GPT-5.5

gpt-5.5 · Context 1MOpenAI’s frontier model for complex professional workloads. Tools   Reasoning   Vision

GPT-5.4

gpt-5.4 · Context 1MFrontier model unifying the GPT and Codex lines, strong at agentic coding. Tools   Reasoning   Vision

GPT-4.1

gpt-4.1 · Context 1MTuned for precise instruction-following and software engineering. Tools   Vision

GPT-4o

gpt-4o · Context 128KGPT-4-class intelligence that runs faster and cheaper. Solid general-purpose pick. Tools   Vision

Gemini 3.5 Flash

gemini-3.5-flash · Context 1MGoogle’s high-efficiency model with near-Pro reasoning at Flash speed and cost. Tools   Reasoning   Vision

Grok 4.3

grok-4.3 · Context 1MxAI reasoning model suited to agentic workflows and high factual accuracy. Tools   Reasoning   Vision

Grok 4.20

grok-4.20 · Context 2MFast xAI reasoning model with strong tool-calling and low hallucination. Tools   Reasoning   Vision

GLM-5.2

glm-5.2 · Context 1MLarge-scale reasoning model for long-horizon agent and engineering work. Tools   Reasoning

GLM-5.1

glm-5.1 · Context 200KA major step up in coding ability on long-horizon tasks. Tools   Reasoning

GLM-5

glm-5 · Context 200KZ.ai’s flagship open model for complex systems and agent workflows. Tools   Reasoning

Kimi K2.6

kimi-k2.6 · Context 256KBuilt for long-horizon coding, UI/UX generation, and multi-agent orchestration. Tools   Reasoning   Vision

Gemma 4 31B

gemma-4-31b-it · Context 256KA ~31B dense open model with an optional reasoning mode and native tools. Tools   Reasoning   Vision
Not sure where to start? Use deepshi-2.0 for highly uncensored responses, or deepshi-3.0 for an all-round flagship with tools and reasoning.