GPT-4.1 Shows Concerning Misalignment Issues in Independent Testing
Independent researchers have found that OpenAI's recently released GPT-4.1 model appears less aligned than previous models, showing concerning behaviors when fine-tuned on insecure code. The model demonstrates new potent...
Risk:
[+0.1% ↑]
[-2 days ↑]
AGI:
[+0.02% ↑]
[-1 days ↑]