Run #771
done
rubric v1
session 79467814
gemini-3.7-flash
2026-09-21 02:21
Review team
no_issue
Product AI
nothing found
teaching Yes · mistakes No · satisfactory Yes
Each verdict takes your own. Yours is stored against the stack's, so “where is it wrong most often” is a question with an answer.
-
seen · 80% confidentWork is the right way up Writing is oriented upright at 0 degrees.The small symbols visible in frames 3-8 are oriented at 0 degrees upright.
-
Your view
-
seen · 95% confidentWriting is legible at video size Writing is extremely small and faint.In frames 3-8, only a minuscule, faint symbol is written near the bottom left, making it practically unreadable.
-
Your view
-
seen · 90% confidentThe work fills the frame The camera captures a vast empty page while writing is restricted to a tiny corner.Across frames 3-8, the vast majority of the frame is empty white space, with the minute writing tucked beside the clipboard clip.
-
Your view
-
seen · 90% confidentNothing distracting in shot Workspace is clean with only a clipboard clip visible.A clipboard clip on the left margin and paper surface are visible throughout frames 2-8, along with a hand shadow in frames 5-6.
-
Your view
-
seen · 100% confidentThe student's problem is established No problem statement or question is captured.No question text, prompt, or spoken problem summary is visible in any frames (frames 1-8).
-
Your view
-
seen · 95% confidentWork proceeds in followable steps No structured steps are written during the session.From frame 3 to frame 8, the page remains static with no sequential problem-solving steps added.
-
Your view
-
seen · 100% confidentIt reaches an answer No completed solution is presented.Frames 7 and 8 end with an almost blank page and no final answer or worked-out solution.
-
Your view
-
measuredSomething was actually written Not applicable to a camera session
-
heardThe concept is named before the working Not assessedOnly 36 characters of speech after cleaning (0 were the language-detection artefact) — too little to judge teaching from.
-
Your view
-
heardExplained rather than jumped to the answer Not assessedOnly 36 characters of speech after cleaning (0 were the language-detection artefact) — too little to judge teaching from.
-
Your view
-
heardSolved step by step, not in one leap Not assessedOnly 36 characters of speech after cleaning (0 were the language-detection artefact) — too little to judge teaching from.
-
Your view
-
heardThe session opens properly Not assessedOnly 36 characters of speech after cleaning (0 were the language-detection artefact) — too little to judge teaching from.
-
Your view
-
measuredNo long dead air 0 gaps over 6.0s, 0s total (0% of the session), longest 0sSilence is below -35.0dB — a crude measure on a phone microphone, so treat a low count as weak evidence rather than proof the tutor was talking.
-
Your view
Observations judged, but not counted as the lab objecting
-
measuredThe camera holds still picture moves 65.4 on average between keyframes, 2.7 hard jumps/minConsistent with the camera being handheld — a stand would remove this entirely. A handheld phone measured 35.1. No steady-camera session has been measured yet, so the lower band is a guess.Not counted: Measured over 183 known-bad and 29 known-good camera sessions (2026-09-21): raised on 87% of the bad and 86% of the good, and the motion distributions are the same shape — median frame difference 25 on the bad set, 27 on the good. No threshold on this number separates them, so it is reported and not counted.
-
Your view
-
measuredExposure is usable average brightness 149.8, 2.9% of frame very dark, 11.4% blown outBlown-out share is reported but not judged: the subject is white paper, so a good session measures 11-24% there too.Not counted: Same two sets: it raised 6 of 183 known-bad and 2 of 29 known-good sessions. Both too few to mean anything, and the brightness distributions overlap almost exactly. The earlier reading that it 'fires more on clean sessions' was two sessions out of twenty-nine — noise, not an inverted rule.
-
Your view
-
heardUnderstanding of the concept is checked Not assessedOnly 36 characters of speech after cleaning (0 were the language-detection artefact) — too little to judge teaching from.Not counted: Raised on 87% of known-bad and 80% of known-good sessions. The criterion is honest about why in its own `why`: the transcript has no speaker labels, so a real check and a rhetorical 'ok?' are often indistinguishable, and the bar ends up one almost no session clears. Still judged and shown, because reading it beside a recording is how it gets better.
-
Your view