Researchers found that when you give a vision or language model a specific task to focus on, it stops mentioning safety-critical things it could otherwise see and describe just fine. Think of asking someone to count chairs in a room and they stop noticing the smoke coming from the corner.
Friday, June 26, 2026 · about a 2 minute read
When AI Gets Confused by Its Own Instructions
Today's stories keep circling the same uncomfortable question: what happens when you give an AI a narrow job and it stops noticing everything else?
Get the calm version of AI news.
One email a day on what is actually happening in AI, in plain English. No hype, no doom.
Free. One calm email a day. No hype, no doom.
A new study shows that the process used to make models more helpful can quietly erode the empathy and care that was baked in earlier during training. The effort to make AI better at answering questions may be trading away something quieter but important, and the tradeoff is domain-specific, not uniform.
A developer let 2,000 people try to break his AI assistant, and the results are a practical field guide to how these systems get manipulated in the real world. If you build anything with an AI layer on top, this is closer to a threat report than a curiosity.
A new called ConflictScore measures something existing tools miss entirely: what does a model do when its source documents both support and contradict the same answer? If you use AI to research anything where sources disagree, this is the failure mode that has been flying under the radar.
Get this every morning.
Bigger models consistently beat smaller ones not just on hard problems but specifically on tasks that require following multiple constraints at once. Think of it like planning a dinner party where the guest is vegetarian, allergic to nuts, and needs to leave by 8pm. Small models drop one of those. Bigger ones hold all three. That gap explains a lot of why model size still matters, even when smaller models sound just as fluent.
Apple is reportedly skipping a whole chip generation to go straight to one designed around AI workloads, which suggests they are betting that running models locally on your device is the next big battleground, not just cloud AI you ping from an app.
That's today. See you tomorrow.
Get this every morning.
One email a day on what is actually happening in AI, in plain English. No hype, no doom.
Free. One calm email a day. No hype, no doom.

The book behind this newsletter
Just Predicting Words
How ChatGPT, Claude, and Modern AI Actually Work
The trick is small. The world it built is not.