coding
23 coding models and apps on inference shell.
run via API, SDK, or belt CLI.

gemini-3-7-flash
Gemini 3.7 Flash via Vertex AI — improved software engineering, web development and agentic workflows over 3.6 Flash. 1M token context window.
chat
gemini-3-8-flash
Gemini 3.8 Flash via Vertex AI — Google's most capable Flash model, built for long-horizon software engineering, autonomous agents and complex enterprise workflows. 1M token context window.
chatchat
23
gemini-3-7-flash
google/gemini-3-7-flash
gemini 3.7 flash via vertex ai — improved software engineering, web development and agentic workflows over 3.6 flash. 1m token context window.

gemini-3-8-flash
google/gemini-3-8-flash
gemini 3.8 flash via vertex ai — google's most capable flash model, built for long-horizon software engineering, autonomous agents and complex enterprise workflows. 1m token context window.

DeepSeek V4.1 Flash
melious/deepseek-v4-1-flash
deepseek v4.1 flash — 552b multimodal moe with 1m context, image input and controllable reasoning. frontier agentic and coding performance. via the melious api.

MiniMax M3
melious/minimax-m3
minimax m3 — 428b multimodal moe with 1m context and image input for long-horizon agentic, coding and cowork tasks. via the melious api.

GLM 5.3
melious/glm-5-3
glm 5.3 — zai's 744b moe post-trained for complex coding and long-horizon agentic tasks. 1m context, thinking on by default. via the melious api.

Kimi K3
melious/kimi-k3
kimi k3 — moonshot's 2.8t open-weight multimodal agentic model with native vision and 1m context for long-horizon coding and reasoning. via the melious api.

Qwen 3.6 27B
melious/qwen3-6-27b
qwen 3.6 27b — dense 27b vision-language model with 262k context and thinking mode. strong coding and multimodal reasoning, apache 2.0. via the melious api.

DeepSeek V4 Pro 0813
melious/deepseek-v4-pro-0813
deepseek v4 pro 0813 — 1.6t moe with 1m context and low/high/max reasoning effort for agentic and coding work. text-only. via the melious api.

GLM 5.2
melious/glm-5-2
glm 5.2 — zai's 744b moe with 1m context and hybrid reasoning for coding and agentic tasks. mit license. via the melious api.

DeepSeek V4 Flash 0731
melious/deepseek-v4-flash-0731
deepseek v4 flash 0731 — 304b moe with 1m context and low/high/max reasoning effort. strong agentic coding and tool use. text-only. via the melious api.

grok-4-7
xai/grok-4-7
grok 4.7 — xai's most capable model for chat, code and agents. 500k context, reasoning effort low to xhigh, image input and tool use via the xai responses api. direct api.

grok-build-0-1
xai/grok-build-0-1
grok build 0.1 — xai's agentic coding model. 256k context, reasoning, image input and tool use via the xai responses api. direct api.

Gemini 2.5 Flash
google/gemini-2-5-flash
gemini 2.5 flash via vertex ai — fast, cost-efficient thinking model with strong reasoning, coding, and multimodal performance. 1m token context window.

Gemini 2.5 Pro
google/gemini-2-5-pro
gemini 2.5 pro via vertex ai — google's most capable thinking model with enhanced reasoning, coding, math, and science performance. 1m token context window.

MiniMax M3
minimax/m3
minimax-m3 — frontier multimodal model with 1m context window. text, image, and video inputs. advanced coding, reasoning, and long-horizon agentic tasks. direct minimax api.

Claude Opus 5
anthropic/claude-opus-5
claude opus 5 — anthropic's frontier opus for complex agentic coding and long-horizon work. 1m context, 128k output, adaptive thinking, vision, tool use. direct api.

Gemini 3.6 Flash
google/gemini-3-6-flash
gemini 3.6 flash via vertex ai — efficient workhorse model with improved coding, knowledge work, multimodal performance, and 17% fewer output tokens than 3.5 flash.

Gemini 3 Flash Preview
openrouter/gemini-3-flash-preview
gemini 3 flash preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance.

Claude Opus 4.6
openrouter/claude-opus-46
opus 4.6 is anthropic’s strongest model for coding and long-running professional tasks. it is built for agents that operate across entire workflows rather than single prompts, making it especially effective for large codebases, complex refactors, and multi-step debugging that unfolds over time.

Claude Sonnet 4.5
openrouter/claude-sonnet-45
claude sonnet 4.5 is anthropic’s most advanced sonnet model to date, optimized for real-world agents and coding workflows. it delivers state-of-the-art performance on coding benchmarks such as swe-bench verified, with improvements across system design, code security, and specification adherence. the model is designed for extended autonomous operation, maintaining task continuity across sessions and providing fact-based progress tracking.

Claude Haiku 4.5
openrouter/claude-haiku-45
a very fast and economical ai designed for real-time uses and everyday business and coding tasks, offering performance similar to much larger, pricier options.

Magistral Small 2506
infsh/magistral-small-2506
a high-performance system for complex reasoning, coding, and math problems, featuring strong multilingual support.

GLM-4.6
openrouter/glm-46
a powerful, open-source language system excelling in advanced coding, complex reasoning, and integrating tools for sophisticated tasks.
explore more on inference shell
coding is one of many things you can run on the grid. discover hundreds of apps across image, video, audio, and more.
we use cookies
we use cookies to ensure you get the best experience on our website. for more information on how we use cookies, please see our cookie policy.
by clicking "accept", you agree to our use of cookies.
learn more.