I Made This Talking Head Video Using Only Claude & Remotion
A YouTube creator maps the three-tool, three-gate content engine, Claude Code, Remotion, and WhisperX, that auto-edits his talking-head videos from raw footage to a finished MP4.
July 9thKev wires the tool-dispatch model Jev into his open-source editor HyperEdit, so a chat prompt triggers real cuts, captions, and media pulled straight from an Obsidian vault.
Routing video-editing prompts through Jev, a non-reasoning tool-selection model, turns editing into instant, free tool calls instead of slow LLM reasoning, so one person can source, cut, caption, and publish without touching a timeline by hand.
Kev connects Jev, a model from TypeSafe AI that resolves every prompt into one of three answer types instead of open-ended reasoning, to his open-source video editor HyperEdit. He demos plain-English prompts triggering real edits: extracting audio, generating yellow captions, and stripping dead air, then shows an Obsidian vault letting Jev pull matching logos and clips onto the timeline by keyword search. He frames this as solving the tradeoff between slow open-source models and token-expensive frontier models, since Jev's tool calls cost nothing, and pitches the endpoint as a pipeline that edits, produces, and posts to social with no human touching the timeline.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Create a free account →
Kev reintroduces HyperEdit, his open-source AI video editor, and teases the Jev AI integration as the biggest update yet.

Plain-English prompts typed into HyperEdit's chat panel execute as real edits: audio extraction, yellow captions, and a dead-air cut that shrinks the timeline from 2:30 to 2:28.

A connected Obsidian vault lets Jev search by keyword ('Claude', 'subway surfer') and drop matching logos and a video clip directly onto the timeline.

Kev frames the real payoff: with Jev wired in, an agent can go from source to cut to caption without a human at the timeline at all.

Jev resolves prompts via three answer types (choice, score, yes/no) fired at once instead of reasoning, and Kev diagrams the 'conductor' architecture routing calls to FFmpeg, Whisper, Remotion, or Claude, closing with the setup ask: grab the codebase and a Jev API key.
Splitting 'decide what to do' into three cheap answer types instead of open-ended reasoning turns routine tool calls into instant, free actions.
“Boy, oh boy, do I have an absolute treat for you guys today.”
“Jev is the first AI model that doesn't think. It just knows answers.”
“we were kind of in between a rock and a hard place... and then Jev comes in and absolutely solves this entire problem”
“the agent... doesn't just do the editing, it'll do post-production, it'll actually post the content to social media”
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
Kev already open-sourced an AI video editor, HyperEdit, that hundreds of people have forked. This video is about the piece he just bolted on: Jev, a model that doesn't reason so much as instantly pick an answer, now sitting in the driver's seat of the whole edit.
The three response shapes Jev is described as using to resolve every prompt simultaneously instead of reasoning step by step.
Diagram showing Jev at the center, routing every prompt to FFmpeg (x2), Whisper, Remotion, or Claude, with Claude invoked only as a fallback.
“all you need to do to set up HyperEdit is to grab the code base and grab the Jev API key and everything else is ready to go”
Soft CTA folded into the sign-off rather than a hard pitch; points to the free GitHub repo and Jev API signup. His own Creator OS / Skool links only appear in the description, not spoken on camera.
00:00
00:06
00:10
00:14
00:18
00:22
00:24
00:28
00:34
00:38
00:40
00:45
00:48
00:52
00:58
01:02
01:06
01:10
01:14
01:18
01:22
01:26
01:30
01:34
01:38
01:42
01:46
01:50
01:54
01:58
02:02
02:06
02:10
02:13
02:17
02:21
02:25
02:28
02:32
02:36
02:40
02:43
02:47
02:51
02:55
02:59
03:03
03:07
03:11
03:15
03:19
03:23
03:27
03:31
03:35
03:39
03:43
03:47
03:51
03:55
03:59
04:03
04:07
04:11
04:15
04:19
04:23
04:27
04:31
04:35
04:39
04:43
04:47
04:51
04:55
04:59
05:03
05:07
05:11
05:15Add Modern Creator as a preferred source and Google shows you more of our breakdowns in Search, Top Stories, and AI Overviews. It only changes what you see, and you can undo it in your Google settings anytime.
A YouTube creator maps the three-tool, three-gate content engine, Claude Code, Remotion, and WhisperX, that auto-edits his talking-head videos from raw footage to a finished MP4.
July 9thOne prompt, one Claude Code project, and two hours forty-two minutes: the intro you just watched was cut entirely by an autonomous Premiere/DaVinci pipeline with zero manual touch-ups.
September 8thA three-step prompt workflow turns Claude Code into a chat-driven motion-design studio that reverse-engineers any video's cuts, motion, and B-roll.
September 2ndA YouTuber hands Claude Code a boring talking-head recording and gets back motion graphics, camera zooms, and a pitch for a $2,500-per-client AI editing service.
July 27thA three-part walkthrough of Reel Studio, a tool built inside Claude Code that turns raw footage and a plain-English prompt into a finished, captioned reel for about 50 cents.
August 4thA five-tool stack — transcription, cuts, AI b-roll, code-built motion graphics, and a self-review loop — that lets an agent finish a rough cut to a client-ready render without a human in the loop.
August 6th