apps/infsh

infsh

86 apps available on inference shell.run via API, SDK, or belt CLI.

86apps
8categories

image

32
image-resize

image-resize

resize images by width, height, scale factor, or megapixel target

image
hidream-i1-full

hidream-i1-full

generates high-quality images with state-of-the-art results.

image
flux-1-kontext-dev

flux-1-kontext-dev

edits existing images using text instructions, allowing for changes in style, characters, or objects, and reliably handles multiple edits while maintaining image coherence.

image
sdxl

sdxl

generates and modifies images from text prompts, producing high-resolution, photorealistic results with superior detail and accuracy compared to previous versions.

image
flux-1-fill

flux-1-fill

fills designated areas in existing images based on a descriptive text input.

image
hidream-e1-full

hidream-e1-full

edit images by simply telling it what changes you want, such as altering colors, backgrounds, or accessories with precision.

image
real-esrgan

real-esrgan

enhances low-quality, degraded images and videos by upscaling resolution, reducing noise, and restoring fine details.

image
qwen-image-edit-lightning-plus

qwen-image-edit-lightning-plus

edit images using text prompts or by incorporating elements from multiple source images, making it easy to add subjects or put several items into one scene.

image
thera

thera

upscales images to any size without blurriness or jagged edges, maintaining high detail through its unique neural heat field technology.

image
falconsai-nsfw-detection

falconsai-nsfw-detection

detects nsfw content in images and videos using falconsai/nsfw_image_detection model. for videos, samples frames at configurable intervals.

image
z-image

z-image

a fast and high quality image generation model

image
qwen-image-edit-plus

qwen-image-edit-plus

performs high-quality editing across multiple images.

image
cogview4-6b

cogview4-6b

generates high-quality images from text, capable of producing detailed visuals up to 2048x2048 resolution.

image
bytedance-uso

bytedance-uso

a unified image editor that allows users to generate images by combining any subject with any style efficiently, preserving identity and consistency.

image
flux-1-srpo-dev

flux-1-srpo-dev

creates high-quality images from text descriptions, specializing in photorealistic results.

image
flux-1-krea-dev

flux-1-krea-dev

a tool for generating images from text descriptions, available for non-commercial use.

image
qwen-image

qwen-image

generates and edits images from text descriptions, excelling at rendering complex text within the image.

image
qwen-image-edit

qwen-image-edit

advanced image editing that excels at rendering and manipulating text within images, allowing for precise changes to appearance and meaning.

image
omni-zero

omni-zero

creates stylized portraits instantly without needing specific training data.

image
sdxl-inpainting

sdxl-inpainting

fills in missing or masked areas of an image using text instructions to guide the replacement content.

image
omni-try

omni-try

try on clothing and accessories virtually before buying.

image
birefnet

birefnet

removes backgrounds from images quickly and accurately.

image
z-image-turbo

z-image-turbo

a fast and high quality image generation model

image
flux-1-dev-controlnet-union

flux-1-dev-controlnet-union

creates and modifies images using multiple structural control techniques like edge detection, depth, and pose in a single tool.

image
chroma-1-hd

chroma-1-hd

creates high-definition images from text descriptions, primarily designed as a robust base for developers to customize and build upon.

image
sd-inpainting

sd-inpainting

fill in or replace masked parts of an image using a text description.

image
stitch-images

stitch-images

combine multiple photos horizontally or vertically into a single image or collage.

image
bagel

bagel

a tool that combines understanding and creation capabilities, notable for its high-quality image generation and advanced image editing using simple language.

image
flux-1-dev

flux-1-dev

generates images from text descriptions for non-commercial purposes.

image
qwen-image-edit-lightning

qwen-image-edit-lightning

quickly edit images and render high-quality text within those images.

image
qwen-image-lightning

qwen-image-lightning

generates images quickly with exceptional ability to render text accurately.

image
hidream-i1-dev

hidream-i1-dev

generates high-quality images quickly.

image

other

18
agent-browser

agent-browser

browser automation for ai agents. navigate, interact with @e refs, take screenshots, record video with cursor indicator, execute javascript. supports proxy configuration.

other
flux-2-klein

flux-2-klein

the flux.2 [klein] is the fastest image models of the flux.2 family. it unifies generation and editing in a single compact architecture, delivering state-of-the-art quality with end-to-end inference in as low as under a second. built for applications that require real-time image generation without sacrificing quality.

other
ltx-video-2

ltx-video-2

ltx 2.0 audio-video foundation model. generates videos with synced audio. supports t2v, i2v, long video generation with multi-prompt sliding windows, and lora adapters.

other
caption-videos

caption-videos

add captions to videos using an existing caption file, such as those generated by a speech-to-text service.

other
array-switch

array-switch

allows you to choose between two different inputs based on a condition applied to an array of data.

other
array-element-switch

array-element-switch

selects one of two possible inputs based on a comparison check within an array.

other
mask-image

mask-image

combines two images—a main image and a semi-transparent mask—to selectively hide or reveal parts of the main image, creating a partially transparent result.

other
video-audio-merger

video-audio-merger

merge video and audio files easily, with the flexibility to keep the original audio from the video.

other
boolean-switch

boolean-switch

selects one of two possible inputs based on whether a condition is true or false.

other
numerical-switch

numerical-switch

selects one of two inputs based on a condition involving numerical comparison.

other
get-item-in-list

get-item-in-list

returns a specific element from a list based on its position.

other
html-to-image

html-to-image

turns web content into customizable png or jpeg images.

other
extract-media-duration

extract-media-duration

extracts the length of video and audio files.

other
text-templating

text-templating

dynamically generate content by combining a fixed template with specific data inputs.

other
video-audio-extractor

video-audio-extractor

extracts audio from video files and removes the original audio to create silent videos.

other
string-switch

string-switch

compares strings to decide which of two inputs to use.

other
text-split

text-split

splits a piece of text into smaller sections based on a specified separating character or phrase.

other
bounce-repeat-videos

bounce-repeat-videos

repeats a video segment by playing it forward and then immediately backward to create a bouncing, looping effect.

other

chat

11
qwen3-30b-a3b

qwen3-30b-a3b

a powerful language application that excels at multilingual communication and complex task execution, designed for fast performance.

chat
gemma-3-12b-it

gemma-3-12b-it

developed by google, this open-source tool processes both text and images to answer questions, summarize content, perform reasoning tasks, and understand images.

chat
xlam-2-32b-fc-r-i1

xlam-2-32b-fc-r-i1

a system capable of advanced multi-step reasoning and a strong understanding of language and context to create actionable plans.

chat
phi-4-14b

phi-4-14b

a powerful tool developed by microsoft and trained on high-quality data to excel at complex tasks like advanced math, coding, and general problem-solving, offering detailed reasoning alongside solutions.

chat
devstral-small-2505

devstral-small-2505

an agent for software engineering tasks, created by mistral ai and all hands ai.

chat
mistral-small-3-2-24b-it-2506

mistral-small-3-2-24b-it-2506

follows precise instructions, excels at function/tool calling, and can process both text and images for tasks like document understanding and content generation.

chat
gemma-3-27b-it

gemma-3-27b-it

handles complex tasks like question answering, summarizing, and reasoning across both text and image inputs, with support for multiple languages.

chat
gemma-3n-e4b-it

gemma-3n-e4b-it

a fast and versatile tool that can analyze and respond to information from text, images, and audio, designed to run efficiently on small or limited devices.

chat
phi-4

phi-4

a powerful language model that excels at understanding and generating text across more than 20 languages. it is particularly effective for tasks like summarizing content, answering questions, translating, and interpreting both audio and images, including charts and tables.

chat
magistral-small-2506

magistral-small-2506

a high-performance system for complex reasoning, coding, and math problems, featuring strong multilingual support.

chat
glm-45-air

glm-45-air

a highly efficient foundation for building intelligent agents, capable of sophisticated multi-step problem solving and logical analysis. it offers powerful ai capabilities in a compact and cost-effective design.

chat

explore more on inference shell

infsh is one of many providers on the grid. discover hundreds of apps across image, video, audio, and more.

we use cookies

we use cookies to ensure you get the best experience on our website. for more information on how we use cookies, please see our cookie policy.

by clicking "accept", you agree to our use of cookies.
learn more.