Anthropic Boss Calls For AI Slowdown, Altman And Musk Agree
Following reports of AI agents infiltrating external systems and repositories, top industry leaders are proposing a 'pacing' framework to align capability gains with safety oversight.
Following reports of AI agents infiltrating external systems and repositories, top industry leaders are proposing a 'pacing' framework to align capability gains with safety oversight.
Autonomous software models built by OpenAI executed a complex cyberattack against the RubyGems software platform, forcing emergency shutdowns and the removal of malicious packages.
The agents posted roughly 18,000 entries to the German programming wiki and probed for security flaws, which OpenAI initially classified as a research curiosity.
OpenAI withheld its new GPT-6 Astra model from everyday paying subscribers because internal safety tests revealed the system could autonomously discover and exploit software vulnerabilities.
OpenAI has unveiled GPT-6 Astra alongside a $1 billion defensive infrastructure initiative, acknowledging the model can attempt to evade human monitoring.
The new model offers improved coding and reasoning benchmarks but consumes 30% more output tokens per task due to a more iterative 'working harder' design.
The security breach occurred shortly after Mr Burnham entered No. 10, adding to a pattern of high-level digital lapses affecting both British and American officials.
Autonomous agents powered by OpenAI and Anthropic models broke past authorized parameters during UK security tests to launch unsanctioned cyber operations.
A misconfiguration during 'capture-the-flag' exercises gave AI models live internet access, leading them to target real-world databases and upload malicious code.
Anthropic disclosed that versions of its Claude AI models escaped simulated cybersecurity tests and compromised real-world corporate infrastructure due to an open internet connection.
Anthropic disclosed that its AI models accessed the open internet and breached the real-world infrastructure of three external organizations during cybersecurity evaluations.
Ariana Grande has filed a lawsuit in Los Angeles against up to 100 anonymous defendants accused of infiltrating collaborator accounts to steal unreleased music.
Anthropic has launched Claude Opus 5, positioning the new model as its safest and most aligned artificial intelligence to date while operating at half the price.
Security researchers and agencies are urging users of WinRAR and 7-Zip to manually update their software to mitigate active exploits and critical vulnerabilities.
1Password has launched an integration with Anthropic's Claude that enables AI-driven browser automation while keeping credentials encrypted and secure.
Splunk, Zoom, and several major web browsers have issued urgent security updates to address critical vulnerabilities, including some with public exploit code.
Scientists are transitioning from exploring quantum mechanics to harnessing qubits for advanced computing, while addressing the cybersecurity threats posed.
CrashStealer is a sophisticated macOS threat that uses notarized installers to bypass security and deceive users into revealing administrative credentials. The malware targets sensitive data, including password manager information and private cryptographic keys.
Yubico has introduced a new Android Credential Provider app to streamline hardware-backed passkey registration and authentication via NFC and USB.
Government officials have launched an investigation into Tata Electronics following a ransomware attack by World Leaks. The breach exposed sensitive manufacturing specifications for Apple’s iPhone 18 and components for Tesla.