Project Moonshot is an open-source testing toolkit that assesses the safety and reliability of Large Language Models/Applications (LLM/apps) through benchmark testing and red teaming.
We’re very excited about Project Moonshot, it is one of the first tangible representations globally of what it means to approach AI Safety, in a way that is actionable for companies and AI teams.