A person ran their MRI results through Claude and found it surfaced details their radiologist had mentioned but not fully explained. This is not a replacement for a doctor, but it is a preview of what it looks like when AI becomes the second reader you can actually afford.
Monday, June 29, 2026 · about a 2 minute read
Who's Watching the Watchers
Today's stories keep circling the same quiet question: when AI is grading your exam, reading your MRI, or writing your code, who is actually in charge of checking its work?
Get the calm version of AI news.
One email a day on what is actually happening in AI, in plain English. No hype, no doom.
Free. One calm email a day. No hype, no doom.
A Brown University professor flagged what appears to be widespread AI-written submissions on a single exam, enough to make it a public story. If you work in education, or manage people whose work you review, this is the moment where 'assume good faith' gets complicated.
A security-focused team at Semgrep ran their own benchmarks and found GLM 5.2, a Chinese open model, outperforming Claude on cybersecurity tasks. Benchmarks are always partial pictures, but this one got 875 upvotes because it came from a team with a specific, practical use case rather than a lab trying to sell something.
An open GitHub issue shows that OpenAI's Codex still has no reliable way to tell it which files to keep its hands off. If you are using an AI coding tool on a project with secrets, credentials, or private data sitting in the same folder, this is worth knowing before it becomes a problem.
Get this every morning.
Researchers tested whether AI models are better at judging answers than generating them, and the result is genuinely interesting: not always. Think of it like a student who can spot a bad essay but still writes bad essays. The AI tools that grade output, including their own output, are not some neutral referee sitting above the model. They are the same model wearing a different hat, and that matters every time a product uses 'AI as judge' to tell you whether another AI did a good job.
Jon Udell flipped the phrase 'human in the loop' to ' in the loop,' arguing the framing matters because it changes who we think of as being in charge. Small language shift, but worth sitting with if you are building workflows where AI takes actions on your behalf.
That's today. See you tomorrow.
Get this every morning.
One email a day on what is actually happening in AI, in plain English. No hype, no doom.
Free. One calm email a day. No hype, no doom.

The book behind this newsletter
Just Predicting Words
How ChatGPT, Claude, and Modern AI Actually Work
The trick is small. The world it built is not.