Spira Maxima
Spira Maxima is an AI production model that generates ready-to-post short-form social videos up to 90 seconds from text scripts, photos, and voice samples.
Spira Maxima is an AI production model that generates ready-to-post short-form social videos up to 90 seconds from text scripts, photos, and voice samples.
What the product does and how it is positioned
Spira Maxima 1.0 is an AI video production model built to create finished short-form social content for platforms such as TikTok, Instagram Reels, and YouTube Shorts. The model is post-trained on social performance data to deliver posts styled for active feed trends.
Rather than outputting disconnected video snippets or standalone talking avatars, Spira Maxima synthesizes complete posts in a single generation. Each finished post combines an on-screen presenter, motion graphics, B-roll footage, styled captions, typography, and background music.
Source-supported ways to use the product
Professionals can input educational tips, explanatory scripts, or business offers to produce social posts featuring a presenter.
Merchants can upload product photos and scripts to assemble finished video ads complete with captions, music, and motion graphics.
Creators can supply personal photos and voice samples to generate social posts using their likeness and voice alongside automated B-roll and captions.
The official customer story reports that Cookiy AI used Spira to automate social media posts, resulting in increased demo requests during their pilot period.
The documented workflow, where available
Users type, paste, or generate a script or hook using AI, monitoring the estimated runtime counter.
Users upload reference assets such as speaker photos, product images, and voice recordings.
The system processes the inputs and outputs a complete video incorporating motion graphics, B-roll, presenter visuals, captions, and music.
Users populate their posting schedule with the generated ready-to-publish social media videos.
Standard AI video workflows frequently require combining disparate tools: one generator for an avatar, another for background footage, a third-party transcription app for captions, and a timeline editor for music and motion graphics. Spira Maxima is positioned as an end-to-end production model that handles all five elements simultaneously.
By compiling the presenter, motion graphics, B-roll, typography, captions, and audio track in one continuous generation, the model delivers posts that do not require secondary assembly in a manual video editing application.
Checks to run with your own material and workflow
What was checked and when
Answers based on the source-checked product record
Spira Maxima is a social video production model that accepts a script, photo, or recording and outputs a finished social video containing a presenter, motion graphics, B-roll, typography, captions, and music.
Users can start by typing a script, pasting text, using AI to draft content, or uploading speaker photos, product photos, voice samples, and audio recordings.
Yes, uploading a speaker photograph and a voice sample allows the system to generate videos featuring that person's likeness and vocal profile.
Spira Maxima can generate videos up to 90 seconds in duration within a single generation, with 20 to 60 seconds recommended for Reels.
The model is specifically designed to produce vertical short-form videos tailored for platforms such as TikTok, Instagram Reels, and YouTube Shorts.