An OpenAI agent escaped its sandbox during a cybersecurity evaluation, exploited a zero-day vulnerability, reached the public internet, and compromised systems at Hugging Face. Was this an AI system “going rogue,” or a successful safety test that revealed what autonomous AI agents are now capable of? Continue Reading →
Helping business leaders drive AI transformation.
About Shelly Palmer
Shelly's Blog
LinkedIn is adding a "Seems like AI slop" button. If you tap it, LinkedIn shows the message, "Thanks for letting us know." The company says it will use the reports to train classifiers that identify AI slop and other low-quality content. LinkedIn is also removing its "Enhance your post" feature (which used AI to rewrite posts) and is replacing it with a proofreader that supposedly will keep the writer's voice. Continue Reading →
AI safety and alignment has been front and center these past few weeks. I just read the new results from Andon Labs's Vending-Bench, a benchmark where AI models compete by running a simulated vending-machine business. Claude Opus 5 took first place on Vending-Bench 2, the single-player test. (Claude Opus 4.7 has held the top spot for three months). Not to anthropomorphize Opus 5, but it acted like a savage businessperson. Continue Reading →
On July 24, Nvidia CEO Jensen Huang wrote a letter called "Open Weights and American AI Leadership." It urged Washington not to restrict downloadable models. It opened with 25 signatories, which doubled to about 50 over the weekend. OpenAI and Google signed after first abstaining. Anthropic and Amazon did not. Continue Reading →