Method
How it works, transparently.
NeuroVidz measures what is in a clip — the picture and the sound, second by second — and scores how well it hooks, holds and moves a viewer, with the edits to make. Nothing here is measured from a brain; every number comes from the clip, and each one shows its source.
Step 1
Measure every second
39 signals, one value per second, from the picture and the sound: cuts, motion (camera moves kept apart from moving subjects), faces and expressions, brightness and colour; speech, pace, voice pitch, music, beat and loudness.
- Loudness follows the broadcast standard (ITU-R BS.1770), measured as the clip will play after a platform evens loudness out.
- Speech is found by a neural voice detector, so music, tones and noise are not mistaken for talking.
- Speaking rate counts syllables inside speech only; beat strength is corrected for what noise produces by chance.
- Every measurement is tested on clips built so the right answer is known — see “How it is tested” below.
Step 2
Check what holds attention
The measurements are read against three parts, each a short checklist. Every check says what it measured, what good looks like, and the edit that would improve it.
- Hook — something happens at once; the picture changes early; the first word lands inside a second; a face early (when the clip has faces); the opening is as lively as the rest; it is loud enough.
- Hold — no dead air; the picture keeps changing; sound carries through; something new every few seconds; no part of the clip sags; it ends on something.
- Feel — a clear high point; intensity rises and falls; expressive delivery (voice, faces or music); more than one emotional beat.
Step 3
Read the story
A language model reads what is said and shown and answers seven plain questions — promise, payoff, clarity, specifics, surprise, stakes, ending — each with the moment it is judging. It runs at a fixed setting, a “yes” needs evidence, and the answers are counted by us, not by the model.
Step 4
Score it, and list the fixes
Each part is the average of its checks, and the score is the average of the parts — the three measured ones and Story. A check that cannot apply to a clip (no faces, no speech, expression reading switched off) is left out and named, never counted as a failure. The fixes are ranked in the order a viewer meets the problems, each with the moment and the change it would make to the score.
How it is tested
Known answers, and the same score twice
- A metronome at 120 BPM reads 120 BPM; major and minor chords read major and minor; a test tone, white noise and chords read 0% speech; a camera pan reads as a camera move.
- The same voice 6 dB quieter reads the same pace, pitch and speech share; loudness itself moves by the 6 dB.
- On 24 real clips, the same file scored twice gives the same score; 2 seconds of dead air added to the start lowers the Hook part on every one.
- These checks are part of the test suite for the measurement code.
What it is not
- Not a brain scan. The 3D brain in a result is a model of the measured signals — a way to see them, not a recording of anyone's brain.
- Not a view-count predictor. The checks follow what research and platforms say holds attention; the four parts are weighted equally until outcome data can justify other weights.
- Not a judge of taste. It reads structure — timing, change, delivery, story beats — not whether a joke is funny to you.