Developer Releases Open-Source Pipeline That Generates Full Documentaries from a Topic
A developer has published AI Video Factory, an MIT-licensed Python tool that converts a user-supplied topic into a 20-30 minute documentary with scripted narration, stock visuals, music, captions, and YouTube metadata. The pipeline uses an LLM to write cited scripts, pulls footage from Pexels, Pixabay, and NASA, handles text-to-speech locally, and assembles everything via FFmpeg — requiring no cloud services. During development, the creator encountered and resolved issues including hallucinated citations, runtime drift, and inconsistent AI-generated visuals, opting instead for real stock footage. Seven documentary presets are available, ranging from business autopsies to science and horror formats, with content-addressed caching to avoid full re-renders during edits. The project is publicly available on GitHub, with planned additions including Whisper-based caption alignment, multilingual narration, and optional AI video provider integrations.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in