Image and video decision
ShortsAuto.ai
A technically capable developer can build a useful, limited AI-shorts generator and auto-posting flow in about a week and modest monthly costs, but reproducing the vendor's proprietary models, polish, scale, and integrated features (voice-cloning quality, template library, analytics, and reliability) is costly—so building a narrow replacement is realistic but matching the full product is not.
Visit website↗$19/mo
$228/yr
Read off the official pricing page.
$100one-off36 h to build
$100/mo6 h/mo upkeep
On cash alone, building overtakes the subscription at 6 seats.
Open-source builds that already do this
Every project below is open source and already does this job today. Fork one, self-host it, or take the parts you need - the build prompt further down assumes an empty file, and this is the shortcut past that. Licences differ; check the one on each card before you ship. All ShortsAuto.ai alternatives, with the arithmetic →
What a replacement has to do
- Take a text prompt or niche → generate a short script/story → synthesize/cloned voice → assemble visual scenes and motion templates → render short-form video in vertical formats → (optionally) auto-post to social platforms.
What it still won’t have
- Vendor-trained proprietary generative video models and large voice-cloning models
- Polished template library and pre-tuned viral styles
- High-availability hosting, scaling, and built-in credit/usage management
- Integrated analytics, affiliate system, and turnkey marketing materials
What remains hard
- Product polish and ongoing maintenance
First-year cost
Keep paying
Paying is—cheaper in year one.
On cash alone, building overtakes the subscription at 6 seats.
Money you would actually spend
Time you would spend
—
What you would spend
What we assumed
The verdict above measures whether you could build it. This one is only about money.
Runnable build prompt
Build a minimal AI shorts generator as a single-developer project using: Node.js (Express) backend, PostgreSQL, React frontend, FFmpeg for video assembly, and OpenAI (or similar) for text generation + a TTS provider (or open-source voice clone like Coqui). Core features in scope: (1) prompt-based story/script generation, (2) selectable TTS voices and an upload-for-voice-clone flow, (3) templated scene assembly and vertical video rendering pipeline (FFmpeg jobs), (4) web UI to create/edit/preview videos, (5) export in vertical formats and simple scheduled auto-post endpoint for YouTube (OAuth) and TikTok. Out of scope: building custom generative video models, training proprietary voice models at scale, analytics dashboard, affiliate program, and polished marketplace of templates. Require input validation, error handling for all external API calls, background job retries for render/posting, and automated tests for API endpoints and render pipeline.
How we checked
How the score was reached
- Partly verdict base52
- An open-source build was found+5
- 3 cited sources+3
- Price verified on pricing page+3
- 3/3 assessment runs agreed+4
- Evidence score67
The base comes from the verdict. Everything under it is a check that either happened or did not, and each one is a fact frozen in this record rather than a judgement made at render time - so the same evidence always produces the same number.
How scoring works →Cited sources · 3
Every page the run actually retrieved.
- official productShortsAuto product & pricing
- open sourceATH-MaaS/Pixelle-Video
- open sourceHBAI-Ltd/Toonflow-app
Integrity checks
What held up, and what did not.





