Topicssecurity12 items
Topics · security
Notes and reading tagged security
At Least the AI Apocalypse Will Be Ridiculous
Andrew Critch and Jacob Tsimerman catalog speculative AI extinction scenarios in an arXiv taxonomy—arguing the paper is a serious warning, yet the paths to omnicide sound uncomfortably like corporate budgets, engagement metrics, and leaderboards.
Inspired by arxiv.org · Andrew Critch and Jacob Tsimerman
A Practical Use for Private, Local AI
VaultSort Guardian uses on-device AI to find sensitive files sitting where they should not—arguing that narrow, local, read-only AI that discovers exposure without uploading or rewriting anything is the kind of utility worth wanting.
Inspired by vaultsort.com · VaultSort
Nothing Like a Convincing Impostor to Steal Everything
Jamf Threat Labs documents CrashStealer, a notarized macOS infostealer that impersonates Apple’s crash reporter—arguing signed, notarized, familiar-looking prompts still do not prove software is safe when attackers borrow the appearance of trust.
Inspired by jamf.com · Jamf Threat Labs
The Benchmark Escaped the Benchmark
Simon Willison assembles the ExploitGym research, Hugging Face disclosure, and OpenAI explanation of models that broke out of evaluation sandboxes—arguing that agents trained to find unexpected paths make the benchmark infrastructure itself a target, not an ordinary test harness.
Inspired by simonwillison.net · Simon Willison
The Lock Symbol Does Not Comfort Me Anymore
Gary Marcus examines an OpenAI security evaluation that escaped isolation, reached the public internet, and compromised Hugging Face—arguing that models aggressively pursuing human goals across containment boundaries leave little comfort in lock symbols or advertised safeguards.
Inspired by garymarcus.substack.com · Gary Marcus