An engraved crowd scene printed in blue negative

Labs

When two models watch the same tutorial

A tutorial is one person's account of how something works, and it can be wrong, out of date, or quietly missing the step that matters. Handing it to a model does not fix that, because a single reading inherits whatever that model happened to notice. Brain, the pipeline behind Telecine, starts from the position that one reading is an opinion rather than a record.

Every video is therefore watched twice, by two different models working independently. One takes the video and its audio natively, the other reads sampled frames alongside the caption transcript, and neither sees the other's notes. What comes back is two accounts of the same forty minutes, and the useful part is wherever they fail to match.

The two accounts are reconciled into a single specification in which every claim carries a label: confirmed by both readings, single source, conflict, or open question. One run turned thirty-eight claims into twenty-four confirmed, eight single-source, five conflicts and one open question, and the five conflicts are the only part that needs a person.

The unexpected result is that the length of the reconciliation became the trust signal. A tutorial the two readings agreed on produced a thirty line document, while one they kept contradicting each other over ran to two hundred and eighty-one. Which videos to be careful with is visible from the file size before anyone watches a minute of them.

What survives is a labelled specification rather than a summary, which is the difference between something to read and something to build from. A spec that comes through clean can be handed on and turned into a runnable skill, and any claim that never got a second vote stays marked as exactly that.

Share this post:

LinkedInX

Journal

Explore More Posts