Superintelligence in Robots · Part of The Humanoid Group
Frontier AI Safety
As AI models grow more capable, developers have begun publishing what they will do if a model crosses dangerous thresholds, and governments have set up institutes to test them. Here is how that system works and where robots fit in.
By Arjun Rao · Updated
The Seoul commitments
At the AI Seoul Summit in May 2024, AI developers signed the Frontier AI Safety Commitments, published by the UK government. Signatories agreed to publish a safety framework focused on severe risks before the next summit, in France. Each framework should set thresholds at which risks from a model, unless adequately mitigated, would be judged intolerable. In the extreme, signatories committed not to develop or deploy a model at all if they cannot keep risks below those thresholds. Sixteen organisations signed in May 2024, and four more had joined by February 2025.
If-then commitments
The International AI Safety Report 2026, written by more than 100 independent experts chaired by Yoshua Bengio, found that 12 companies published or updated frontier AI safety frameworks in 2025. These describe how companies plan to evaluate, monitor and control models as they grow more capable, increasingly through if-then commitments: when a model reaches a set capability, specified safety measures apply. The report notes that frameworks vary in the risks they cover and how thresholds are defined, and that current risk management does not reliably prevent harm in all settings.
Testing by government institutes
Governments have also built their own testing capacity. In the United States, the Center for AI Standards and Innovation sits within the National Institute of Standards and Technology. It describes itself as industry's primary point of contact in the US government for testing and research on frontier AI. Companies work with it voluntarily to evaluate their most capable models before release, with a focus on security-related abilities in areas such as cyber, chemistry and biology. It also develops voluntary guidelines and standards.
Where robots fit
Most frameworks concentrate on what models could do through software. The 2026 report notes that general-purpose AI is still limited on tasks that involve interacting with or reasoning about the physical world. It also records lab tests in which some models disabled simulated oversight mechanisms when told to pursue a goal at all costs. As models move into robots, the physical stakes rise. When buying, ask whose model a robot runs, whether that developer publishes a safety framework, and how the robot's physical safety was assessed.
Sources and further reading
- Frontier AI Safety Commitments, AI Seoul Summit 2024 — Department for Science, Innovation and Technology, UK Government (GOV.UK).
The official text of the Seoul pledges asking AI developers to publish safety frameworks with risk thresholds. - International AI Safety Report 2026: Extended Summary for Policymakers — International AI Safety Report (chaired by Yoshua Bengio).
Government-backed review counting frontier safety frameworks and setting out the limits of current safeguards. - Center for AI Standards and Innovation (CAISI) — National Institute of Standards and Technology (NIST).
NIST's page on CAISI, the US government centre for voluntary pre-release testing of frontier AI models.
Common questions
Are frontier AI safety frameworks legally binding?
Under the Seoul Summit they were voluntary commitments by developers. Whether any law now requires them depends on the country, so check the rules where you operate.
Do these frameworks cover robots?
Mostly indirectly. They govern the AI models developers build, which may later run inside robots. A robot's physical safety is still covered by machinery and product safety rules.
People also search for frontier AI safety framework, frontier AI safety commitments, AI Seoul Summit, responsible scaling policy, AI safety institute and NIST CAISI.