alignment
-
AI safety conversations have gotten unbelievable
This week two conversations about AI safety went viral that demonstrate just how hard it is to discern AI fact…
Read More » -
OpenAI institutes new safeguards after Hugging Face breach
On Tuesday, OpenAI announced a new batch of new security policies focused on containing security incidents while models are being…
Read More » -
Anthropic set AI agents loose on the same task. They started a turf war.
What happens when you pit AI agents against each other? According to Anthropic’s testing, things get messy fast. On Thursday,…
Read More »