Interview a structural-biology data owner. Direct an agent. Decide whether the model actually answers the question.
You run the exact forward-deployed-scientist loop from weeks 1–3. What changes: the data is now a protein sequence and a 3D structure with a confidence score. The bar goes up in one specific way — a model can look convincing and still be confident about the wrong thing for the owner's question.
The analysis loop is cheap — a couple of minutes of agent work for a few cents. The clock goes to your thinking, reviewing and checking. A checked partial result you can show beats an unfinished perfect one.
Nothing passes between the two agents except what you write down. That channel — you — is exactly the examined skill.
Use the Interview playbook and the two coaching helpers (Coach me · Suggest a question) in the portal. A vague question earns a vague answer; a precise one earns the truth. Then play it back and let them correct you.
You don't do the analysis — you direct it and judge what comes back. Same shape as every week.
Look at the model (vision is on): a low-confidence region or a shaky interface is obvious once you know to look. A finding that deflates the owner's plan is a win, not a failure.
The viewer is an interrogation instrument — build it so a sceptic (you first, the owner on Friday) can see whether the model supports the claim.
Everything through the portal Hand in button. Re-submit any time — we grade your latest. Submit even if you can't attend. A modest, honest result with a transcript that shows real checking is exactly what we're after.
Read the confidence that matches the claim, check the model is really their protein, and hand back the honest truth.