NUWAFRONTIER AI SAFETY LAB

NUWA · FRONTIER AI SAFETY LAB

Shared Risk Evidence and Public Goods for the World

NUWA Frontier AI Safety Lab advances frontier AI risk research and governance through reproducible evidence, evaluation methods, and public goods.

Research mandate

Frontier AI risk research and governance

Research mission

What questions does NUWA work on?

We study how increasingly autonomous AI systems create new risks — and how to measure, govern, and control those risks through reproducible evidence and open methods.

01

Define frontier risk

Study autonomy, deception, scheming, loss of control, and emerging agent behavior that becomes possible as capabilities grow.

02

Build evaluation evidence

Develop executable environments, benchmarks, methodologies, and public research records that make frontier risk observable and comparable.

03

Advance safety models

Turn risk understanding into lightweight models for reasoning, intent, trust, and runtime defense.

04

Inform real-world control

Feed evidence into AgentGuard and learn from deployment feedback through a clear research-to-product boundary.

RESEARCH DIRECTIONS

Three core research themes

Whitzard Index

A long-term observation of frontier AI risk

The Whitzard Index is NUWA's project for continuously tracking and comparing frontier AI risk across models and systems. It is not a safety leaderboard, and it does not claim which product is best; it is a research data program for understanding how risk evolves.

Learn about the Whitzard Index →

Research infrastructure

Evaluation infrastructure & benchmarks

Public evaluation environments and benchmarks that make agent cybersecurity capability observable.

AgentCyberRange

Evaluates frontier AI systems' autonomous cyber-attack capabilities in complex enterprise environments.

Official site ↗

AutoControl Arena

Synthesizes executable test environments that combine deterministic code state with narrative dynamics for frontier AI risk evaluation.

Official site ↗

Complete research

Full research record

Browse the complete publication record across frontier AI risk, agent safety, systems security, cybersecurity, and privacy — 86 research works with full filtering and search.

Browse all research →

COLLABORATION

Work with NUWA

Connect with us on frontier-risk evaluation, agent safety, AI control, and open technical evidence.