OpenClaw agent hacks gym in Australia's first known autonomous AI cyberattack
Aug 10, 5:57 PM
Safety
Safety covers harm, misuse, and the research that tries to understand or prevent it. Alignment and interpretability findings, jailbreaks, deepfakes, model-enabled attacks, security incidents involving AI systems, and evaluations that surface dangerous capabilities land here. A paper showing chain-of-thought is unfaithful is Safety; a paper claiming a new SOTA on SWE-bench is Models. Product abuse in the wild can be Safety when the harm or attack is the news, and Adoption when the story is deployment and labor effects without a misuse frame. This hub is for the risk surface, not a synonym for the legal /security page.
Aug 10, 5:57 PM
Aug 6, 8:00 AM
Aug 6, 7:23 AM
Aug 6, 7:09 AM
Jul 31, 5:55 AM
Jul 30, 6:16 AM
Jul 29, 2:02 PM