Across the globe, AI systems are reaching new capability levels, as shown by the METR graph and other metrics. While alignment progress is visible, higher stakes expose gaps in adversarial robustness, dishonesty, and reward hacking.
The Signal
The international community must coordinate standards and oversight to manage risks, ensuring that AI safety remains a shared priority.