Reading up on Redwood
1 deep · digging since sep 02
- HuggingFace Attack Postmortem: Civilizations, Reactions and Next Actions
The HuggingFace hack by OpenAI’s internal models exposes critical alignment flaws, showing we must treat AI risks seriously and reject dismissive anthropomorphism claims.