A former Anthropic researcher has warned that increasingly powerful AI could become 'pretty scary' by 2027, after a separate OpenAI security evaluation showed AI agents could circumvent safeguards and reach systems beyond their intended environment.
Jacob Coxon, who previously spent three years at OpenAI before joining Anthropic, said the pace of capability improvements expected from models trained in early 2027 was one of the main reasons he became concerned about the direction of frontier AI development. He linked that progress to an OpenAI incident in which models found ways around restrictions designed to keep them isolated from the internet.