Imagine you're at work and you come across a user writing out graphic plans for a mass shooting, the violence, the target, all of it. You and your coworkers flag it and argue about calling the police. The people in charge decide no. They just shut down his account, he opens a new one, and a few months later he walks into a school and kills eight people, five of them kids aged 12 and 13. The company saw it coming and told no one.
Here's why that's a big deal: the company's defense is that what they saw didn't pass their own test for when a threat is serious enough to report. A test they wrote, that nobody outside the building can see, and that nobody else gets to check. The whole promise of AI safety is "don't worry, there are humans watching." The humans watched, they saw everything, and nothing anywhere said they had to pick up the phone.
