Image, video, audio, and avatar generation — each studio runs on your own provider key at provider price, inside the same media library you post from.
36 studios — capabilities from the live app
Postmill — Settings · AI Media
36 studios
Amazon BedrockBeta
AWS's fully managed service for generative-AI apps and agents, offering foundation models from leading providers — including image generation via Amazon Nova, Titan, and Stability AI.
Image
Azure OpenAIBeta
Microsoft Azure's managed access to OpenAI models, including gpt-image generation — backed by Azure's enterprise security, regional deployment, and compliance controls.
Image
Black Forest Labs
Creators of the FLUX model family — a frontier image lab known for state-of-the-art photorealism, precise prompt control, and production-grade character and style consistency.
Image
D-IDBeta
D-ID's Creative Reality Studio creates photorealistic talking-head videos from a single image and deploys real-time conversational avatars — turning static photos into humanlike speaking video.
VideoAvatar
Deepgram
Speech AI platform — fast, accurate speech-to-text transcription and captions (plus text-to-speech).
Audio
DeepInfraBeta
A developer-friendly inference hub serving 100+ models across text, image, video, and speech via simple APIs — pay-as-you-go on its own cost-optimized US infrastructure.
ImageVideoAudio
ElevenLabs
The leading AI audio platform for ultra-realistic text-to-speech, voice cloning, and multilingual dubbing across 70+ languages — known for the most natural-sounding AI voices available.
Audio
Fireworks AIBeta
A high-performance inference platform serving frontier open models — including FLUX image generation — at open-source economics, processing 30T+ tokens per day.
Image
GenviralBeta
Genviral Studio AI generates short-form videos from a prompt, routing to top video models (Sora, Seedance, and more). Configure your Genviral Partner API key to generate straight into your media library.
Video
Google AI StudioBeta
Google's Gemini Developer API for media generation — spanning Nano Banana and Imagen image models plus Veo video, all driven by a single Gemini API key.
ImageVideo
Google VertexBeta
Google Cloud's unified AI platform for building and scaling generative apps — featuring Imagen for photorealistic images and Veo for high-quality text-to-video, with governance built in.
ImageVideo
GroqBeta
Groq runs AI inference on its purpose-built LPU chip for ultra-low latency. For speech it serves Whisper ASR and Orpheus/PlayAI TTS — real-time transcription and voice generation.
Audio
HedraBeta
A multimodal creative platform built on its proprietary Character-3 model, generating expressive character video, image, and audio — known for lifelike performance from a keyframe.
VideoAvatar
HeyGen Featured
AI video platform for lifelike avatar and talking-photo videos generated from a script.
VideoAvatar
HiggsfieldBeta
An AI-native creative suite that generates images, videos, and voice from text or references — Soul for image, DoP for cinematic image-to-video, and Speak for talking-video.
ImageVideo
IdeogramBeta
An AI image generator best known for industry-leading in-image text rendering — crisp, correctly-spelled typography ideal for posters, ads, and social designs.
Image
KlingBeta
A leading video generator by Kuaishou, known for long, cinematic clips with strong realism and physics. Recent models add native audio, voiceovers, and sound in a single pass.
ImageVideoAudio
Leonardo.aiBeta
A creative platform powered by its foundational Phoenix model — known for high-resolution output, coherent text, and a Real-Time Canvas that turns sketches into polished art instantly.
Image
LTX StudioBeta
LTX Studio by Lightricks is an end-to-end AI video production platform. Powered by the open-source LTX-2 model, it takes you from script to storyboard to finished video.
Video
Luma
Luma AI's Dream Machine turns text and images into realistic, fluid video. Its Ray models are known for natural motion, strong physics, and fast, prolific creative iteration.
Video
MiniMax
Hailuo, MiniMax's AI video generator, is known for striking cinematic motion and strong prompt following from text or a single image — with templates for dance, effects, and character animation.
ImageVideoAudio
OpenAI
OpenAI's image generation via gpt-image-1 — the natively multimodal model behind ChatGPT — delivering versatile styles, strong world knowledge, accurate text, and prompt-driven edits.
ImageVideoAudio
OpenRouterBeta
A unified gateway to 400+ models from 70+ providers through a single OpenAI-compatible API — including a dedicated image API spanning 30+ models — with automatic failover.
Image
QwenBeta
Alibaba's Qwen family, served via DashScope / Model Studio — Qwen-Image generates native 2K images from long prompts, and Wan delivers text-to-video and image-to-video.
ImageVideo
RecraftBeta
A design-focused AI image platform best known for generating editable vector/SVG graphics alongside photoreal images, with reusable custom brand styles that need no training.
Image
Reel.FarmBeta
Reel.Farm turns a single prompt into a finished TikTok-style slideshow video — slide text, layout, and pacing are all generated for you. Built for bulk, automated short-form content.
Video
Replicate Featured
Run and fine-tune open-source image, video, and audio models through one API — from FLUX and Stable Diffusion to video upscalers and music generators.
ImageVideoAudioAvatar
Runway
A pioneering generative-AI platform whose Gen-4 family produces cinematic, high-fidelity video from text and images — prized for consistent characters, scenes, and director-grade motion control.
ImageVideo
SiliconFlowBeta
A lightning-fast inference platform serving 200+ open and commercial models — LLMs plus image, video, and audio — through a single OpenAI-compatible API with predictable pricing.
ImageVideoAudio
Stability AI
The company behind Stable Diffusion and the Stable Image models, offering open, enterprise-grade generative media for image creation and editing with full creative control.
ImageVideoAudio
SunoBeta
Suno is a leading generative-AI music model that turns a text prompt — or your own lyrics, style and title — into complete, studio-quality songs with vocals or instrumentals. This studio uses the sunoapi.org gateway.
Audio
TavusBeta
Tavus builds foundational models for face-to-face AI — real-time video replicas and conversational agents that see, hear, and respond with emotion, via its Phoenix rendering and CVI APIs.
VideoAvatar
Together AIBeta
A full-stack inference cloud serving 200+ open-source models through one API — chat, image, audio, and video — with optimized kernels for faster, cheaper generation at scale.
ImageVideoAudio
Vercel AIBeta
Vercel AI Gateway routes requests to hundreds of models across many providers through a single API — text, image, and video — with automatic failover and no platform markup.
ImageVideo
WanBeta
Alibaba's Wan creative platform (Model Studio) lowers the barrier to content creation with the Wan2.x model family — spanning text-to-video, image-to-video, and text-to-image.
ImageVideo
xAI Grok
xAI builds Grok, the AI assistant from Elon Musk's xAI with real-time knowledge of the world. Its image model (Aurora) renders photorealistic, prompt-faithful images directly from the same API key as the Grok chat models.