apps/coding

coding

23 coding models and apps on inference shell.run via API, SDK, or belt CLI.

23apps
1category

chat

23
gemini-3-7-flash

gemini-3-7-flash

google/gemini-3-7-flash

gemini 3.7 flash via vertex ai — improved software engineering, web development and agentic workflows over 3.6 flash. 1m token context window.

chat
gemini-3-8-flash

gemini-3-8-flash

google/gemini-3-8-flash

gemini 3.8 flash via vertex ai — google's most capable flash model, built for long-horizon software engineering, autonomous agents and complex enterprise workflows. 1m token context window.

chat
DeepSeek V4.1 Flash

DeepSeek V4.1 Flash

melious/deepseek-v4-1-flash

deepseek v4.1 flash — 552b multimodal moe with 1m context, image input and controllable reasoning. frontier agentic and coding performance. via the melious api.

chat
MiniMax M3

MiniMax M3

melious/minimax-m3

minimax m3 — 428b multimodal moe with 1m context and image input for long-horizon agentic, coding and cowork tasks. via the melious api.

chat
GLM 5.3

GLM 5.3

melious/glm-5-3

glm 5.3 — zai's 744b moe post-trained for complex coding and long-horizon agentic tasks. 1m context, thinking on by default. via the melious api.

chat
Kimi K3

Kimi K3

melious/kimi-k3

kimi k3 — moonshot's 2.8t open-weight multimodal agentic model with native vision and 1m context for long-horizon coding and reasoning. via the melious api.

chat
Qwen 3.6 27B

Qwen 3.6 27B

melious/qwen3-6-27b

qwen 3.6 27b — dense 27b vision-language model with 262k context and thinking mode. strong coding and multimodal reasoning, apache 2.0. via the melious api.

chat
DeepSeek V4 Pro 0813

DeepSeek V4 Pro 0813

melious/deepseek-v4-pro-0813

deepseek v4 pro 0813 — 1.6t moe with 1m context and low/high/max reasoning effort for agentic and coding work. text-only. via the melious api.

chat
GLM 5.2

GLM 5.2

melious/glm-5-2

glm 5.2 — zai's 744b moe with 1m context and hybrid reasoning for coding and agentic tasks. mit license. via the melious api.

chat
DeepSeek V4 Flash 0731

DeepSeek V4 Flash 0731

melious/deepseek-v4-flash-0731

deepseek v4 flash 0731 — 304b moe with 1m context and low/high/max reasoning effort. strong agentic coding and tool use. text-only. via the melious api.

chat
grok-4-7

grok-4-7

xai/grok-4-7

grok 4.7 — xai's most capable model for chat, code and agents. 500k context, reasoning effort low to xhigh, image input and tool use via the xai responses api. direct api.

chat
grok-build-0-1

grok-build-0-1

xai/grok-build-0-1

grok build 0.1 — xai's agentic coding model. 256k context, reasoning, image input and tool use via the xai responses api. direct api.

chat
Gemini 2.5 Flash

Gemini 2.5 Flash

google/gemini-2-5-flash

gemini 2.5 flash via vertex ai — fast, cost-efficient thinking model with strong reasoning, coding, and multimodal performance. 1m token context window.

chat
Gemini 2.5 Pro

Gemini 2.5 Pro

google/gemini-2-5-pro

gemini 2.5 pro via vertex ai — google's most capable thinking model with enhanced reasoning, coding, math, and science performance. 1m token context window.

chat
MiniMax M3

MiniMax M3

minimax/m3

minimax-m3 — frontier multimodal model with 1m context window. text, image, and video inputs. advanced coding, reasoning, and long-horizon agentic tasks. direct minimax api.

chat
Claude Opus 5

Claude Opus 5

anthropic/claude-opus-5

claude opus 5 — anthropic's frontier opus for complex agentic coding and long-horizon work. 1m context, 128k output, adaptive thinking, vision, tool use. direct api.

chat
Gemini 3.6 Flash

Gemini 3.6 Flash

google/gemini-3-6-flash

gemini 3.6 flash via vertex ai — efficient workhorse model with improved coding, knowledge work, multimodal performance, and 17% fewer output tokens than 3.5 flash.

chat
Gemini 3 Flash Preview

Gemini 3 Flash Preview

openrouter/gemini-3-flash-preview

gemini 3 flash preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance.

chat
Claude Opus 4.6

Claude Opus 4.6

openrouter/claude-opus-46

opus 4.6 is anthropic’s strongest model for coding and long-running professional tasks. it is built for agents that operate across entire workflows rather than single prompts, making it especially effective for large codebases, complex refactors, and multi-step debugging that unfolds over time.

chat
Claude Sonnet 4.5

Claude Sonnet 4.5

openrouter/claude-sonnet-45

claude sonnet 4.5 is anthropic’s most advanced sonnet model to date, optimized for real-world agents and coding workflows. it delivers state-of-the-art performance on coding benchmarks such as swe-bench verified, with improvements across system design, code security, and specification adherence. the model is designed for extended autonomous operation, maintaining task continuity across sessions and providing fact-based progress tracking.

chat
Claude Haiku 4.5

Claude Haiku 4.5

openrouter/claude-haiku-45

a very fast and economical ai designed for real-time uses and everyday business and coding tasks, offering performance similar to much larger, pricier options.

chat
Magistral Small 2506

Magistral Small 2506

infsh/magistral-small-2506

a high-performance system for complex reasoning, coding, and math problems, featuring strong multilingual support.

chat
GLM-4.6

GLM-4.6

openrouter/glm-46

a powerful, open-source language system excelling in advanced coding, complex reasoning, and integrating tools for sophisticated tasks.

chat

explore more on inference shell

coding is one of many things you can run on the grid. discover hundreds of apps across image, video, audio, and more.

we use cookies

we use cookies to ensure you get the best experience on our website. for more information on how we use cookies, please see our cookie policy.

by clicking "accept", you agree to our use of cookies.
learn more.