← sladebyrd.com
AI Safety Portfolio
Work toward making AI go well.
- White-Box Attacks on the Best Open-Weight Model: CCP Bias vs Safety Training in Kimi K2.5
- How OpenAI's Models Broke Into Hugging Face
- Trajectory: A Monte Carlo Model of How AI Could Go
- AI Safety Literature Database
- AI Safety Orgs: Theory of Change for 120 Organizations
- People on AI: What 101 Strangers Think About Superhuman AI
- Responsible Scaling Policy Comparison: v1.0 vs v2.2 vs v3.0
- Mythos System Card vs Alignment Risk Report
- AI Safety Learning Tracker: A Completed 12-Week Curriculum