Public AI on Hugging Face Inference Providers 🔥
Topic · Edition
Apollo Research and OpenAI developed evaluations for hidden misalignment (“scheming”) and found behaviors consistent with scheming in controlled tests across frontier models. The team shared concrete examples and stress…