The secretive Ox Alpha model, suspected to be a Chinese GLM release, scored 80% on a rigorous coding benchmark. This outperforms Claude and GPT-style models, which typically land in the 60% range. The jump suggests a rapid acceleration in specialized reasoning. Developers should monitor this for new open-source coding standards.