app library

the grid

hundreds of tools that run serverless on CPU or GPU.call directly via API or let agents orchestrate them.

100sof apps
1API for everything

all apps

gemini-2-5-flash-lite

google/gemini-2-5-flash-lite

gemini 2.5 flash-lite via vertex ai — most cost-efficient gemini 2.5 model optimized for high-throughput, low-latency workloads. 1m token context window.

gemini-2-5-flash

google/gemini-2-5-flash

gemini 2.5 flash via vertex ai — fast, cost-efficient thinking model with strong reasoning, coding, and multimodal performance. 1m token context window.

gemini-2-5-pro

google/gemini-2-5-pro

gemini 2.5 pro via vertex ai — google's most capable thinking model with enhanced reasoning, coding, math, and science performance. 1m token context window.

avatar-x

mirage/avatar-x

generate an expressive talking-head video from a portrait image or video reference and an audio track using mirage avatar x, with natural lip sync, eye contact and micro-expressions.

post-analytics

x/post-analytics

get engagement analytics for specific posts. returns impressions, engagements, and other metrics over a time period.

trends

x/trends

get trending topics by location. use woeid 1 for worldwide, or a specific location id.

following-list

x/following-list

list accounts a user follows. returns profiles with bios, follower counts, and join dates.

followers-list

x/followers-list

list followers of a user. returns profiles with bios, follower counts, and join dates.

post-likers

x/post-likers

list users who liked a specific post. returns user profiles with follower counts.

post-quotes

x/post-quotes

list posts that quoted a specific post. returns quote tweets with text, author, and engagement metrics.

bookmark-remove

x/bookmark-remove

remove a bookmarked post on x.com.

bookmark-add

x/bookmark-add

bookmark a post on x.com.

bookmarks-list

x/bookmarks-list

list your bookmarked posts. returns saved posts with text, author, and engagement metrics.

dm-list

x/dm-list

list recent dm events. returns direct message text, sender, and timestamps.

user-mentions

x/user-mentions

get your recent mentions. returns posts that mention the authenticated user.

user-timeline

x/user-timeline

get your home timeline. returns recent posts from accounts you follow with engagement metrics.

user-posts

x/user-posts

list posts by a user. returns text, timestamps, engagement metrics, and article metadata.

probe-timeline

x/probe-timeline

probe user posts and timeline for article metadata

article-publish

x/article-publish

publish long-form articles on x.com from markdown. supports headings, lists, blockquotes, code blocks, gfm tables, latex, images, embedded tweets, and inline formatting (bold, italic, strikethrough, links). optional cover image and draft-only mode.

fid-score

eval/fid-score

frechet inception distance — measures distributional similarity between two sets of images

resize

eval/resize

crop and resize images — center crop to square, face-aware crop, or custom resize

clipscore

eval/clipscore

clipscore — cosine similarity between clip text and image embeddings for prompt adherence

arcface

eval/arcface

arcface identity similarity — cosine similarity between face embeddings of two images

post-render

x/post-render

render x/twitter post cards as png images — dark/light theme, profile picture, engagement metrics

pickscore

eval/pickscore

pickscore v1 — scores how well an image matches a text prompt (clip-h finetuned on pick-a-pic)

blur

eval/blur

degrade images with blur, noise, and compression for benchmark testing

charts

eval/charts

research-quality chart generation (bar, grouped bar, scatter) using matplotlib

download

archive/download

download a file from an internet archive item

metadata

archive/metadata

get item metadata from the internet archive

search

archive/search

search the internet archive

wayback

archive/wayback

check url availability in the wayback machine

markdown-to-word

infsh/markdown-to-word

convert markdown text or files to word (.docx) documents

seedance-2-5

bytedance/seedance-2-5

professional multimodal video generation from text, images, video, and audio references using bytedance's seedance 2.5 model via byteplus ark api. supports up to 4k (10-bit color), durations up to 30s, mov output, and multimodal reference-to-video with synchronized audio.

flux-3-video

bfl/flux-3-video

flux 3 video by black forest labs — generate, animate, and extend video up to 20s at hd or full hd with synchronized audio. supports text-to-video, image-to-video with keyframes, video continuation, and draft mode.

minimax-m3

openrouter/minimax-m3

minimax-m3 is a frontier multimodal model with 1m context window. supports text, image, and video inputs for coding, reasoning, and long-horizon agentic tasks.

music-cover

minimax/music-cover

minimax music cover — ai-powered song covers and style transfer. upload a reference track and generate a cover with new style and optional new lyrics.

speech-2-8-hd

minimax/speech-2-8-hd

minimax speech 2.8 hd — high-quality text-to-speech. 40 languages, 9 emotions, custom voices. direct minimax api.

music-3-0

minimax/music-3-0

minimax music 3.0 — ai music generation up to 5 minutes. supports vocals with lyrics, instrumentals, and style prompts. direct minimax api.

speech-2-8-turbo

minimax/speech-2-8-turbo

minimax speech 2.8 turbo — fast text-to-speech. 40 languages, 9 emotions, custom voices. lower latency than hd. direct minimax api.

m-2-7

minimax/m-2-7

minimax-m2.7 — large language model with enhanced reasoning, image understanding, and file processing. 200k context. direct minimax api.

m3

minimax/m3

minimax-m3 — frontier multimodal model with 1m context window. text, image, and video inputs. advanced coding, reasoning, and long-horizon agentic tasks. direct minimax api.

m-2-7-highspeed

minimax/m-2-7-highspeed

minimax-m2.7-highspeed — m2.7 performance with significantly accelerated inference. 200k context. direct minimax api.

h3

minimax/h3

minimax h3 — multimodal video generation with native audio. text-to-video, image-to-video with end_image, and reference-based generation. 2k resolution, 5-15s duration, 24fps. prompt with timelines, audio cues, and negative lists.

krea-2-medium-turbo-train

krea/krea-2-medium-turbo-train

train custom lora styles for krea 2 medium turbo generation

krea-2-medium

krea/krea-2-medium

krea 2 medium — expressive illustrations, ~10s per generation, 1.5k native resolution

krea-2-medium-train

krea/krea-2-medium-train

train custom lora styles for krea 2 medium generation

krea-2-large-train

krea/krea-2-large-train

train custom lora styles for krea 2 large generation

krea-2-large

krea/krea-2-large

krea 2 large — photorealistic generation, ~25s per generation, 2k native resolution

krea-2-medium-turbo

krea/krea-2-medium-turbo

krea 2 medium turbo — fastest k2 model, ~3s per generation, 1.5k native resolution

act-two

runway/act-two

runway act-two — character performance transfer. animate a character image or video using a reference performance video with facial expressions, gestures, and body control. 5 credits/second.

gen-4-image-turbo

runway/gen-4-image-turbo

runway gen-4 image turbo — fast text-to-image generation with reference image support. 2 credits per image at any resolution.

gen-4-image

runway/gen-4-image

runway gen-4 image — high-quality text-to-image generation with optional reference images and @tag syntax. 5 credits per 720p, 8 credits per 1080p.

aleph-2

runway/aleph-2

runway aleph 2.0 — video-to-video transformation with text and keyframe guidance. supports aspect ratio targeting and up to 5 reference keyframes. 28 credits/second.

gen-4-turbo

runway/gen-4-turbo

runway gen-4 turbo — fast image-to-video generation. animate images into video with text guidance. 5 credits/second.

gen-4-5

runway/gen-4-5

runway gen-4.5 — high-quality video generation from text or image. 2-10 second duration, multiple aspect ratios. 12 credits/second.

text-overlays

mirage/text-overlays

render up to 4 static text variants onto one video with per-variant font, size and colour — built for testing ad hooks and headlines against the same footage.

video-1

mirage/video-1

generate an expressive talking-head video from a portrait image and an audio track using mirage video 1, with natural lip sync, eye contact and micro-expressions.

video-captions

mirage/video-captions

burn styled animated captions onto a vertical video using mirage caption templates, from an upload or an existing mirage video id.

magi-1

infsh/magi-1

generates high-quality videos from text descriptions, ensuring smooth, consistent motion across the video, and supports real-time streaming.

fast-wan

infsh/fast-wan

generates high-quality videos quickly by utilizing a technology that distills a complex process into fewer steps, significantly improving generation speed.

claude-opus-5

anthropic/claude-opus-5

claude opus 5 — anthropic's frontier opus for complex agentic coding and long-horizon work. 1m context, 128k output, adaptive thinking, vision, tool use. direct api.

p-image-ideogram

pruna/p-image-ideogram

high-quality text-to-image generation with strong typography and prompt understanding, built with ideogram

modify

pixverse/modify

pixverse modify — edit video content using text prompts. swap subjects, add elements, or restyle regions with reference images.

avatar

pixverse/avatar

pixverse avatar — generate talking avatar video from a portrait image. supports audio file or text-to-speech input at up to 1080p.

lip-sync

pixverse/lip-sync

pixverse lip sync — align speech to mouth movement in video. supports audio file input or text-to-speech with selectable voices.

fusion

pixverse/fusion

pixverse fusion — generate video from reference images. compose 1-3 subject/background images into a video scene with text-guided motion.

extend

pixverse/extend

pixverse extend — extend an existing video with ai-generated continuation. supports 5s or 8s extensions at up to 1080p.

c1

pixverse/c1

pixverse c1 — advanced video generation model. text-to-video and image-to-video with up to 1080p resolution and 15s duration. per-second billing.

v6

pixverse/v6

pixverse v6 — versatile video generation model. text-to-video and image-to-video with up to 1080p resolution and 15s duration. per-second billing.

gemini-3-5-flash-lite

google/gemini-3-5-flash-lite

gemini 3.5 flash-lite via vertex ai — fastest, most cost-effective 3.5-class model at 350 output tokens/s. built for high-throughput agentic workflows.

gemini-3-6-flash

google/gemini-3-6-flash

gemini 3.6 flash via vertex ai — efficient workhorse model with improved coding, knowledge work, multimodal performance, and 17% fewer output tokens than 3.5 flash.

domain-search

fastly/domain-search

domain research api — search for available domains and check registration status

harrier-oss-v1

infsh/harrier-oss-v1

multilingual text embedding using microsoft's harrier oss v1 models. supports retrieval, clustering, semantic similarity, classification, and reranking with state-of-the-art mteb v2 scores.

search

ceramic/search

web search powered by ceramic ai

seedance-2-0-mini

bytedance/seedance-2-0-mini

cost-effective multimodal video generation from text, images, video, and audio references using bytedance's seedance 2.0 mini model via byteplus ark api. ~50% cheaper than seedance 2.0, supports text-to-video, image-to-video, and multimodal reference-to-video with synchronized audio.

seedream-5-pro

bytedance/seedream-5-pro

bytedance's flagship seedream 5.0 pro image model via byteplus ark api. precision creation and editing with pixel-level regional edits, intelligent layer understanding, complex infographic generation, multi-reference blending (up to 10 images), and native text rendering in 14 languages.

video-upscale

topaz/video-upscale

video upscaling and enhancement — proteus family models for precision upscaling, deinterlacing, face recovery, and cgi enhancement

search

arxiv/search

search arxiv papers by query with field prefixes, boolean operators, category filtering, and sorting

search

biorxiv/search

search biorxiv and medrxiv preprints by date range with optional category filtering

search

chemrxiv/search

search chemrxiv preprints via crossref api

paper

arxiv/paper

get a specific arxiv paper by its id with full metadata including title, authors, abstract, and links

paper

chemrxiv/paper

get a chemrxiv paper by doi via crossref api

paper

biorxiv/paper

get a specific biorxiv or medrxiv paper by its doi

astra

topaz/astra

creative video upscaling — ai-guided upscaling with prompt and creativity controls

starlight

topaz/starlight

generative video upscaling — precision, hq, mini, sharp, and fast models

denoise

topaz/denoise

video denoising — nyx family models for noise, compression, and artifact removal

video-utilities

topaz/video-utilities

video utilities — motion deblur, colorization, and sdr to hdr conversion

frame-interpolation

topaz/frame-interpolation

video frame interpolation — slowmo and fps boost with apollo, chronos, and aion models

proteus

topaz/proteus

proteus video upscaling and enhancement — precision upscaling, deinterlacing, face recovery, and cgi enhancement

contents

you/contents

contents api — fetch clean markdown or html from any url, batch up to 10 pages per request

finance-research

you/finance-research

finance research api — agentic financial research with filings, macro data, and institutional-grade sources

research

you/research

deep research api — multi-step web research with source-backed citations and configurable effort levels

search

you/search

web search api — ground your apps in reliable, web-scale knowledge with contextual snippets

claude-sonnet-5

anthropic/claude-sonnet-5

claude sonnet 5 — frontier sonnet with near-opus performance. 1m context, vision, extended thinking, tool use. direct api.

claude-opus-4-8

anthropic/claude-opus-4-8

claude opus 4.8 — anthropic's most capable opus model. 1m context, 128k output, vision, extended thinking, tool use. direct api.

gemini-3-1-flash-lite-image

google/gemini-3-1-flash-lite-image

gemini 3.1 flash lite image (nanobanana 2 lite) via vertex ai — ultra-low latency image generation

gemini-omni-flash

google/gemini-omni-flash

gemini omni flash — text-to-video with synchronized audio, grounded in real-world knowledge

mai-image-2-5

microsoft/mai-image-2-5

mai image 2.5 — microsoft's photorealistic image generation and editing model with fine-grained pixel-level control.

glm-5-2

openrouter/glm-5-2

glm 5.2 - zhipu's latest flagship language model with 1m context via openrouter

remix

reve/remix

reve remix — create images from text and 1-6 reference images combined.

not enough? create new apps fast. templates + coding agents make it insanely extensible.

create your own apps

start from templates. add code, packages, docs. deploy in minutes.

$ infsh app init
my-app/
inference.py
requirements.txt
$ infsh app deploy

schemas become tool parameters automatically. your app shows up in the grid and can be used by agents and workflows.

create workflows

build a graph of apps. deploy as a single callable app.

workflow builder

drag and drop to build the graph. map io to connect steps. deploy as an app.

view all appsexplore what's available
create your own appread the docs & start building

we use cookies

we use cookies to ensure you get the best experience on our website. for more information on how we use cookies, please see our cookie policy.

by clicking "accept", you agree to our use of cookies.
learn more.