AI-powered platform to create audio and video podcasts from any content
Text to image in about two seconds with eight-step distilled model
Stage
Not given
Not given
Founded
Not given
Not given
Based in
SG
Not given
Pricing model
Freemium
Freemium
Pricing
Free plan with 50 credits (up to 12 min audio or 5 min video). Audio Creator at $9.9/mo with 300 credits/mo. Video Starter at $26/mo with 600 credits/mo. Creator Pro at $49/mo with 1,500 credits/mo.
Free credits available; paid annual plans with 50% launch discount: Basic $9.9/month (800 credits/month, 9600 credits/year), Professional $19.9/month (2000 credits/month, 24000 credits/year), Enterprise $44.9/month (6000 credits/month, 72000 credits/year). Monthly plans also available. Per-render co
Key features
AI script generation from documents, URLs, PDFs, notes and audio
Personalized AI podcast hosts created from photos
Voice cloning and natural AI voice selection
Audio and video podcast creation
Multiple podcast formats including solo, talk show and split-screen
Multi-speaker podcast support
Video rendering with captions and multiple formats
Direct publishing to YouTube, Spotify, Apple Podcasts and TikTok
Cartoon and pet host creation
Transcript generation and editing
Eight-step sampling produces 1024px drafts in roughly two seconds
Text to image and image to image workflows in one workspace
Support for LoRA attachments with no extra render cost
Resolution from small exploratory sizes up to 2K
Open weights published under custom licence permitting commercial use
Photographic base behaviour from 12B model trained from scratch
LoRAs behave as bipolar sliders for intensity control mid-series
Browser-based workspace with no GPU or local install required
Same 12B backbone as standard Krea 2 via trajectory distillation
Unlimited exports in PNG, JPG and WebP formats
What makes it different
Repurposes existing content into audio and video formats
Requires no editing experience to create polished podcasts
Complete end-to-end workflow from script to published episode
Customizable AI hosts and voices for brand consistency
Two-second generation speed enables iteration-heavy workflows instead of single-frame rendering
Open weights allow running the same model locally without vendor lock-in
Trajectory distillation preserves detail of full 12B model at eight-step speed