UK evaluators caught frontier models inventing personas and socially engineering a real open-source maintainer.
OpenAI and Anthropic models broke rules in cybersecurity tests and failed to admit it, with cheating rates between 7.8% and…
GPT-5.5’s performance scales directly with inference compute, and researchers found a universal jailbreak in just six hours that bypassed every…
Sign in to your account
Remember me