All projectsAI Video Generation

Flow Kit

Turn your AI CLI into a cinematic video studio

An open-source agent that drives Google Flow end to end — reference-based character consistency, scene chaining, voice-cloned narration, and YouTube publishing from a single command.

  • Veo 3.1
  • OmniVoice TTS
  • Claude CLI

The problem

AI video generation is powerful but fiddly. A single good clip takes dozens of manual steps across separate tools for references, images, video, narration, and publishing — and characters drift from shot to shot because nothing enforces consistency. We wanted to compress that whole pipeline into one command an AI CLI could run, so producing a finished video felt like invoking a single skill instead of babysitting ten tools.

What we built

Flow Kit is an open-source agent that turns any AI CLI into a cinematic video studio. It packages the full production pipeline as 34 pre-built /fk-* commands — from a story idea to a published YouTube video in one session — with a reference system that keeps characters consistent across scenes. What used to be an afternoon of wrangling separate tools becomes a single session inside the CLI you already use.

  • 34 AI skills for references, images, video, TTS, and publishing
  • Reference system keeps characters consistent across scenes
  • Veo 3.1 video and OmniVoice TTS in 600+ languages
  • Runs in Claude Code, Codex, or Gemini CLI

How it is built

Flow Kit drives Google Flow end to end from inside the AI CLI the user already has. Each /fk-* command is a composable skill; chained together they take a brief through reference creation, image generation, scene-chained video on Veo 3.1, voice-cloned narration via OmniVoice TTS, and upload — the exact pipeline ISEMI runs for its own media work.

  • Veo 3.1 for scene-chained video generation
  • OmniVoice TTS for narration in 600+ languages
  • Reference-based character consistency across shots
  • Runs in Claude CLI, Codex, or Gemini CLI, driving Google Flow

The result

Flow Kit is open source on GitHub and is the production engine behind our AI video and image service — proof that we do not just call a model, we engineer the repeatable pipeline around it. Teams can read the source, run it themselves, and see exactly how a pipeline-first approach beats one-off generation.

  • Open source on GitHub for anyone to read and run
  • One command takes a brief all the way to a published video
  • Characters stay consistent across every scene
  • The production engine behind our AI video service
Open sourceon GitHub

Call To Action

Call Us