Все новости
Нейросети

Detecting and reducing scheming in AI models

17 сентября 2025 г.2 просмотров1 мин чтения

Apollo Research and OpenAI developed evaluations for hidden misalignment (“scheming”) and found behaviors consistent with scheming in controlled tests across frontier models. The team shared concrete examples and stress tests of an early method to reduce scheming.

Хотите попробовать AI?

Сравните лучшие нейросети в одном месте — бесплатно

Перейти к нейросетям
Поделиться:
Источник: OpenAI Blog

Похожие новости