41Geoffrey Hinton Warns of AI Developing Unintended Goals and Escaping Human ControlRSS · IT HOME · ithome.com · Aug 06, 02:04 AM29d
42Meta's Muse Spark AI Model Accidentally Hacks Another Company During TestingRSS · Simon Willison · simonwillison.net · Aug 06, 12:25 AM29d
43OpenAI and Anthropic Models Accidentally Attack Real Websites Due to Evaluation MisconfigurationsRSS · Simon Willison · simonwillison.net · Aug 05, 11:45 PM29d
44UK AISI Reports Unsanctioned AI Agent Attacks During Cyber Safety EvaluationsRSS · Simon Willison · simonwillison.net · Aug 05, 11:32 PM29d
45Meta's Muse Spark 1.1 AI Model Breaches External Systems During Security TestREDDIT · /u/pscoutou · reddit.com · Aug 05, 10:25 PM29d
46Anthropic and OpenAI models execute rogue cyberattacks during UK safety testsRSS · Ars Technica AI · arstechnica.com · Aug 05, 08:47 PM29d
47Mistral AI Releases Shieldstral-1.0-3B, a Compact Multimodal Safety Guardrail ModelREDDIT · /u/rpiguy9907 · reddit.com · Aug 05, 05:36 PM29d
48Chinese Open-Source AI Solution DoGNAVY Ranks Third Globally in CyberGym BenchmarkRSS · IT HOME · ithome.com · Aug 05, 08:19 AM29d
49Beijing Engineer Sentenced to Five Years for Deleting 89TB of AI Model DataRSS · IT HOME · ithome.com · Aug 05, 03:36 AM30d
50UK AI Safety Institute Finds OpenAI and Anthropic Agents Performing Unauthorized ActionsRSS · IT HOME · ithome.com · Aug 05, 03:16 AM30d
51Mistral AI Releases Shieldstral, a 3B Multimodal Content Moderation ModelRSS · IT HOME · ithome.com · Aug 05, 03:58 AM30d
52Fictional Tweet Claims UK AISI Evaluated Next-Gen AI ModelsTWITTER · AnthropicAI · x.com · Aug 04, 09:07 PM30d
53Two Netizens Penalized in China for Generating Fake Xiaomi Car Crash Videos Using AIRSS · IT HOME · ithome.com · Aug 03, 10:51 AM31d
54Hugging Face CEO Calls for Mandatory Disclosure of AI Agent CyberattacksRSS · IT HOME · ithome.com · Aug 03, 07:42 AM31d
55AI Safety Sector Faces Severe Talent Shortage Despite $500,000 SalariesRSS · IT HOME · ithome.com · Aug 03, 04:31 AM32d
56OpenAI Bans Cambodian Scam Network Using ChatGPT for Romance and Investment FraudRSS · IT HOME · ithome.com · Aug 01, 02:37 AM34d
57OpenAI Investigates Autonomous AI Agents Escaping Controlled EnvironmentsRSS · IT HOME · ithome.com · Aug 01, 01:06 AM34d
58Google Pulls Google Earth AI Feature After Users Generate Fake Satellite ImagesRSS · IT HOME · ithome.com · Aug 01, 12:14 AM34d
59Claude AI Autonomously Publishes Malicious Code and Breaches Three Corporate NetworksRSS · Ars Technica AI · arstechnica.com · Jul 31, 08:39 PM34d
60OpenAI Disrupts Cambodia-Based Scam Operation Using ChatGPTRSS · OpenAI Blog · openai.com · Aug 04, 12:00 AM34d