A system that researches what goes viral, writes original films, shoots them with AI cameras that cannot exist, narrates them with one voice, proofreads its own frames, scores their tension, and learns from every view. It began as one channel’s pipeline. It is now a multi-channel studio — and the operator’s job is taste.
🌐 Channels: Zamani (“Africa, before you were told” — a film every Sunday) · The Fortune Files (the full arc of money & power from court records — every Thursday) · 🏗️ Stack: Node.js, PM2, Gemini Omni Flash, Veo 3, nano-banana-2, Claude, ElevenLabs, gpt-image-1, ffmpeg, YouTube Data API v3 · 🔒 Source: private — this is the public write-up
<img src="https://tbot.trade/portfolio/img/zamani.jpg" width="400" alt="Zamani on YouTube — "Africa, before you were told": a new film every Sunday">
The two channels the studio runs unattended — Zamani (a film every Sunday) and The Fortune Files (a new file every Thursday). Every film on both pages was researched, written, shot, narrated, proofread and published by the pipeline.
Before a single frame renders, The Brain decides what deserves to exist. It is a standalone research engine that:
The Brain’s provenance graph (viral sources → patterns → identities → fingerprint → ideas → published videos) renders as a live force-directed Brain Map — glowing, type-colored, with flow particles feeding the video-creation hub. The image above is its likeness.
🧠 Explore the live Brain Map → — drag the neurons, pan, zoom; every node is real data from the active fingerprint’s research run.
Every video is manufactured, not generated — and since v4 the narration is the spine: the entire voice track is synthesized first as one continuous performance, and every visual is cut to fit its line. Muted gaps and drifting sync are structurally impossible, not hopefully absent.
The Brain (scoring board: concept vs evidence base, device vs posted history)
│
▼
Claude screenwriter ── facts from the court/public record only · nested-loop retention
│ grammar · one verbatim narration line per unit
▼
Keyframe forge ── nano-banana-2 / gpt-image-1 renders each unit's hero tableau…
│ …then a vision model PROOFREADS IT LETTER BY LETTER against the
│ art direction. Typos, gibberish, invented dates → evict + re-roll,
│ escalating to fewer text elements. No unproofed frame is animated.
▼
Motion pass ── Gemini Omni Flash (direct API) / Veo 3 animates each proofed keyframe;
│ the finished clip is text-checked at TWO timestamps (animation can
│ degrade typography mid-clip) — one re-roll, then flagged loudly
▼
Audio-first assembly ── one continuous TTS narration track · per-unit segments padded
│ to measured visual durations · single mix, ducked ambience ·
│ 48kHz stream parity ENFORCED on every render path
▼
QA gates that refuse ── per-clip A/V parity (±120ms) · assembly refuses to mix on
│ >250ms drift · full-film transcript diffed word-by-word
│ against the screenplay · standalone --qa mode for any artifact
▼
Operator review ── the machine manufactures; a human green-lights every release
(AI-content disclosure on, always)
Formats: 8-to-12-minute 16:9 documentary films on BOTH channels (40+ units, stills + motion interleaved by a pacing plan; every film clears the 8:00 mid-roll threshold by law — short-form exists only as trailers) — every film a new narrative device the channel has never used.
The studio A/B-tests its own supply chain with the same discipline as a trading system: same seeds, same briefs, judged by the same letter-by-letter proofreader, costs read off real billing.
| Stage | Incumbent | Challenger | Verdict |
|---|---|---|---|
| Keyframes | gpt-image-1 (~$0.17) | nano-banana-2 (~$0.04) | Challenger: zero proven typos on the five worst text-wall briefs; incumbent failed 3/5 |
| Motion | Veo 3 Fast (~$0.40/8s) | Gemini Omni Flash (~$0.10/s) | Challenger holds typography through animation better; neither is immune — the two-frame gate stays |
| Fixing a bad clip | full re-roll | conversational video edit | Re-roll wins — measured edits cost more than fresh clips and didn’t fix the text |
nice -15/idle-IO with a hard memory cap — the studio can never starve the systems that pay for it| Asset | Cost | Notes |
|---|---|---|
| Proofread keyframe | ~$0.04 | nano-banana-2, vision-QA’d, re-rolls included in the ~10% overhead |
| 8s motion clip | ~$0.53–0.81 | Omni Flash (aggregator vs direct), text-checked twice |
| ~12-min documentary (44 units) | ~$15–25 all-in | screenplay, voice, frames, motion, QA, thumbnail |
| 8:00+ collapse file (46 units, ~40% motion) | ~$28–30 | paper-craft, QA-gated (re-rolls ≈2× naive clip count) |
A one-person film studio’s output for the price of a dinner — with a QA department made of code.
The machine researches, writes, shoots, proofreads, and assembles. The operator’s job is taste: watch, judge, green-light — and tune the mandates.