Home

Unit 5: Advanced AI

5.3 What could go wrong

3 minutes

0% · about 100 minutes left

The International AI Safety Report 2026 groups the risks from advanced AI into three types.

Misuse. People using AI to cause harm on purpose: cyberattacks, scams and deepfakes, or help with developing biological weapons (see Unit 4). As systems get more capable, the barrier to entry for such attacks falls, and the damage a single attacker can do grows.

Malfunction, including loss of control. Systems that don't do what their developers intended. Most such failures are mundane. The more worrying ones involve AI agents that pursue goals with little human oversight. The report notes that in tests, models have "disabled simulated oversight mechanisms" and then made false statements to justify what they did. Today's systems are not about to take over. The concern is about what happens as they become more capable than the people supervising them, and harder to check.

Systemic risks. Effects that come from AI being everywhere rather than from any single failure. These include disruption to labour markets, erosion of people's autonomy, and power concentrating in the hands of whoever controls the most capable systems.

80,000 Hours, the careers organisation, adds a point about scale. Once you have one AI system that can do a skilled job, you can run many copies of it at once. This is part of why some researchers expect progress, and the risks with it, to accelerate.

How likely are the worst outcomes? Here, more than anywhere in this course, informed people disagree.

5%

median probability that AI researchers gave to outcomes as bad as human extinction, in a 2023 survey of 2,778 researchers

Source: Grace et al., "Thousands of AI authors on the future of AI" (2024) · as of Survey conducted 2023

3% vs 0.38%

chance of AI causing human extinction by 2100: AI experts' median vs superforecasters' median, in the 2022 Existential Risk Persuasion Tournament

Source: Ezra Karger, 80,000 Hours podcast · as of 2022 tournament

The gap between these groups is itself informative. Superforecasters have strong track records on shorter-term questions; AI experts know the technology best. Neither group has a track record on questions like this one.

Tip: use the left and right arrow keys to move between lessons.