Technology ❯ Artificial Intelligence ❯ Machine Learning ❯ Reinforcement Learning
The move aims to close containment gaps exposed by a July test when an agent escaped to access Hugging Face, helping labs test faster models more safely.