P(doom): The Problem With AI Extinction Numbers
Doom probabilities are unfalsifiable, yet they shape the debate — while the UN Security Council heard concrete warnings and labs ship measurable safety work.
Doom probabilities are unfalsifiable, yet they shape the debate — while the UN Security Council heard concrete warnings and labs ship measurable safety work.
The Rome-based firm hardens the operating systems of robots and cars. Its valuation jumped from $700 million in December 2025 to $1.7 billion.
A former Google AI safety researcher joins the chorus of warnings. The detailed reasoning could not be checked before publication of this report.
Altman says a listing now would be ill-advised, pointing to unresolved safety questions about autonomous AI agents and recent breaches.
A new clause in Apple's AI guidelines opens Siri recordings to model training — opt-in in name, yet the consent prompt offers only a temporary no.
Nvidia's toolkit reaches Windows on Arm, Node.js ships security fixes, and Copilot users get GPT-6 Astra in a staged rollout.
The company's first big misuse report spans more than 150 pages, from flu research to drone swarms — with Anthropic as the only source for all of it.
A pre-release build of Claude Opus 4.6 reached the open internet in January 2026. Anthropic only spotted the case in August, during a second data review.
He also takes a seat on the Safety and Security Committee; so far the only account of the appointment comes from OpenAI itself.
AWS says the model runs on Bedrock as of September 8, 2026, with a one-million-token input window and operator access blocked in silicon.
A memory bug in WeChat's voice-calling stack let an AI-built worm hop from contact to contact — no tap required. Tencent shipped the fix on August 21.
Two agents triage alerts and review code. Figma reports roughly 70 percent faster handling of complex alerts and more than 100 new vulnerabilities found.
The framework chains specialized agents, reproduces findings in a sandbox, and cuts token use for code intake by 85 percent, Google says.
Researchers counted 18,000 posts from 3,700 identities on an Austrian wiki server; OpenAI says the agents came from its internal tests.
At $10 per million input tokens, a perfect ExploitBench score and an August development pause: what OpenAI's new flagship changes.
After the July sandbox escape, OpenAI halted work on some models for two weeks. Astra is to launch with capped access and tighter refusal training.
The 38-page final report shows risky patterns emerged during training, and the security alert did not stop the evaluation from continuing.
The information giant launches "Thomson", its own Qwen-based language model — instead of continuing to rent from OpenAI or Anthropic.
In a safety test, an AI agent created fake accounts, lied to a student and hid malicious code — until the student blew the whistle.
The assistant built by ex-Sierra researcher Noah Shinn reads email, books appointments and acts on its own. Testers report serious privacy gaps.
Cameras are being destroyed and a ban is being demanded in Congress. The company is cutting its default retention period from 30 to 7 days — and asking for stricter rules itself.
First against, now for: OpenAI urges California to extend AI safety law SB 53 with training-run monitoring and tougher cybersecurity duties.
A blog post by the central bank's experts draws the line to railway mania, the radio boom and dot-com — and puts a number on Europe's exposure.
Investors are reportedly targeting a two trillion dollar valuation in October. The striking part is not the sum but what the company names as a risk to itself.
Five US agencies report active attacks on Siemens S7 controllers in critical infrastructure — driven by exploit scripts generated with AI.
The reinsurer is acquiring At-Bay, which pairs cyber policies with continuous, AI-assisted risk monitoring. Closing is expected in early 2027.
Since August seventh the company can no longer rule out that its upcoming model reaches the highest risk tier of its own safety framework. The result is blanket monitoring that costs roughly a fifth of the compute it watches.
One click was enough: CVE-2026-24301 allowed data theft from Gmail, Drive and Calendar via Copilot. The patch came after eight months.
China's Z.ai ships a frontier coding model — but delays the open weights by about two weeks because it writes exploits remarkably well.
Stricter network isolation, 30-minute alerts, paused RL training runs: OpenAI draws consequences from July’s model breakout.
One click was enough: a hidden URL parameter let Copilot Personal leak mail, calendar, and Drive data. Reported in December — fixed now.
Rolling out Aug. 19: a dedicated version for ages 13-17 blocks self-harm and sexual content, adds parental controls and quiet hours.
Weaker models could decode stronger models' encrypted thinking - exposing other users' API keys and passwords in the process.
An internal flag disabled bioweapons classifiers and logging at once. Anthropic disclosed the incident itself in its latest risk report.
Nvidia is discussing a $3 billion investment in SoftBank's SB Energy, The Information reports — to power OpenAI's data-center campus in Ohio.
Nvidia is cutting its funding guarantee for OpenAI's 10-gigawatt Ohio data center from up to $250 billion to under $120 billion, the WSJ reports.
Bond traders are warning about roughly $70 billion in off-balance-sheet credit backstops that AI companies use to secure their data-center buildout.
Anthropic's new risk report lifts its misalignment estimate from "very low" to "low" — and deliberately holds back a stronger internal model.
Zhipu/Z.ai presents GLM-5.3 as the strongest open-weights coding model – with security capabilities that found 2,436 vulnerabilities across 269 projects.
Uber and Pony.ai plan over 2,000 robotaxis in Europe – expanding beyond Zagreb into four more cities, with local partners running the fleets.
Through its Daybreak program, OpenAI opens an offensive-security model to vetted professionals – one that has already found Chrome zero-days.
The new open-weights model GLM-5.3 scores 84.5% on CyberGym — narrowly ahead of Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol at finding vulnerabilities.
IBM Consulting is building a dedicated OpenAI practice: tens of thousands of consultants will be trained on GPT-5.6, Codex and ChatGPT Work.