最新报道:According to a report from Bijie.com, OpenAI and Anthropic recently conducted mutual model evaluations to identify issues that their own testing might have missed. The two companies stated in separate blog posts on Wednesday that this summer, they conducted security tests on each other's publicly available AI models, examining them for hallucination tendencies and so-called "misalignment," where the models don't perform as their developers intended. These evaluations were completed before OpenAI launched GPT-5 and Anthropic, founded by former OpenAI employees, released Opus 4.1 in early August.