• 0 posts
  • 6 comments
Joined 8 days ago
Cake day: October 2nd, 2026
  • “The outputs of this opt-in vulnerability scanner will be fully model-generated, without human review or triage,” Anthropic explained. “This will enable faster and more frequent scanning, but means that it is possible reports will be incorrect or invalid.”

    When I saw the headline, I was wondering about this specifically. This may make this service not super useful.

    My experience with AI security reviews is that they’re fantastic at finding faults, but they always find a list of things to complain about. If there are no real/serious faults they’ll start finding things that kind of have the same shape as a security issue, but really aren’t if you dig into them. I’ve regularly had an LLM generate a list of 10-15 issues ranging in severity from “nits” to “critical” where none of them were actual issues.

    Periodic reviews seem like they could get annoying really quickly, becoming more of a maintenance burden than a help.

  • this is all done with custom models by individuals.

    This is really important for people to understand.

    Powerful open weight models are already out there. Anyone can download them and run them. There is no possibility for government control of AI content without North Korean style internet controls.

    Politicians and corporations will absolutely start trying to push for more controls. They will tell you to “think of the children” and they will use this to take power for themselves, knowing full-well that they can’t actually meaningfully prevent this from happening. We need people to understand that the push for control here has nothing to do with morality or protecting anyone.