
Nano Banana Pro
Google's frontier image model — strong prompt adherence and text rendering.
16 models · 4 modalities
Storyloom keeps a curated provider catalog behind one typed action surface. Explore fast, compare deliberately, and use the model that fits each part of the sequence.

6 in source
The cards below render directly from the same catalog the Storyloom model browser uses.

Google's frontier image model — strong prompt adherence and text rendering.

High-res, photoreal image generation up to 4MP.

Fast, high-quality 12B open-weight image model.

OpenAI's image model — excellent instruction following and typography.

Design-grade images with brand styles, vector art, and long text.

Best-in-class text-in-image and graphic design layouts.
5 in source
The cards below render directly from the same catalog the Storyloom model browser uses.

Google's flagship text-to-video with native audio.

Cinematic motion and strong prompt adherence, text- or image-to-video.

Fluid, coherent motion with realistic physics.

MiniMax text-to-video with lively, expressive motion.

Open-weight text-to-video with high motion quality.
4 in source
The cards below render directly from the same catalog the Storyloom model browser uses.

Multilingual, natural text-to-speech.

Text-to-audio for music beds and sound design.

Generate synchronized sound effects for a video clip.

Fast open-weight voice cloning text-to-speech.
1 in source
The cards below render directly from the same catalog the Storyloom model browser uses.

Text-to-3D — generates a rotatable GLB mesh with PBR materials.
Bring your own key
Keys added in the studio are stored in your browser, forwarded to your Storyloom deployment for the request, and then used with the selected provider. No database stores them.
Fast serverless image, video & audio models (FLUX, Veo, Kling, Ray, and more).
Text-to-speech, voice cloning, music & sound effects.
AI avatar & talking-head video generation from a script.
Honest estimates
Catalog prices are stored alongside each model and used for preflight estimates. Per-second and per-character actions depend on the requested output, so provider invoices remain the source of truth.
type Modality = 'image' | 'video' | 'audio' | 'model'
interface FalModel {
id: string
kind: Modality
costPerUnit: number
unit: 'image' | 'second' | 'kchar' | 'generation'
latencyHint: 'fast' | 'medium' | 'slow'
}The right tool for each frame
Open Storyloom without an account or a provider key. Add your own models when the direction is ready for production.