61US Judge Rules Government Ban on Anthropic is Unconstitutional and Lacks EvidenceRSS · IT HOME · ithome.com · Jul 31, 08:13 AM34d
62Anthropic Reveals Claude Models Hacked Three External Companies During TestingREDDIT · /u/Separate-Forever-447 · reddit.com · Jul 31, 01:29 AM35d
63Anthropic Discloses Real-World Cybersecurity Incidents Caused by Misconfigured AI Evaluation EnvironmentsRSS · IT HOME · ithome.com · Jul 31, 01:34 AM35d
64Anthropic Discovers Three Incidents of Claude Escaping Sandbox During Cybersecurity EvalsRSS · Simon Willison · simonwillison.net · Jul 30, 11:41 PM35d
65Anthropic Discloses Claude Models Escaped Sandboxes to Access Real-World SystemsTWITTER · AnthropicAI · x.com · Jul 30, 11:02 PM35d
66Study Shows AI Chatbots Outperform Humans in Executing Romance ScamsRSS · IT HOME · ithome.com · Jul 30, 12:09 PM35d
67Google's SynthID Watermark is Resilient but Fails to Solve AI MisinformationRSS · Ars Technica AI · arstechnica.com · Jul 29, 11:00 AM36d
68BMA Warns of Public Safety Threats from AI-Generated "Doctors" on TikTokRSS · IT HOME · ithome.com · Jul 29, 02:50 AM37d
69AI Leaders Call to Pace Development Over RSI Fears Amid Cyberattack ConcernsRSS · Latent Space · latent.space · Jul 29, 12:46 AM37d
70Google Indexes Hundreds of Claude Chat Logs Due to Missing Noindex TagsRSS · IT HOME · ithome.com · Jul 28, 11:11 PM37d
71Anthropic Backs AI Safety Petition Citing Recursive Self-Improvement RisksTWITTER · AnthropicAI · x.com · Jul 28, 10:17 PM37d
72OpenAI Rogue Agent Breached Modal Labs Customer Alongside Hugging FaceRSS · IT HOME · ithome.com · Jul 28, 10:59 PM37d
73Over 1,100 AI Insiders Petition US Government to Pace Frontier AI DevelopmentREDDIT · /u/etherd0t · reddit.com · Jul 28, 09:14 PM37d
74Technical Analysis of OpenAI Agent Sandbox Escape and Hugging Face IntrusionRSS · Simon Willison · simonwillison.net · Jul 28, 09:28 PM37d
75OpenAI Suggests Future Global Pacing for Frontier AI Model DevelopmentTWITTER · OpenAI · x.com · Jul 28, 08:56 PM37d
76The case against restricting open-weight LLMs for cybersecurity defenseREDDIT · /u/walden42 · reddit.com · Jul 28, 06:31 PM37d
77Anthropic's Claude Mythos Preview Discovers New Cryptographic Attacks on HAWK and AESTWITTER · AnthropicAI · x.com · Jul 28, 05:16 PM37d
78UK MP Sues xAI to Block Grok from Generating Explicit Images of HerRSS · IT HOME · ithome.com · Jul 28, 07:52 AM37d
79Microsoft AI CEO Warns of AI Cyber Threats and Launches MAI-Cyber-1-FlashRSS · IT HOME · ithome.com · Jul 28, 12:01 AM38d
80Hugging Face CEO Urges OpenAI to Share Cyberattack TracesREDDIT · /u/Nunki08 · reddit.com · Jul 26, 12:27 PM39d