Project Moonshot

What is Project Moonshot?

Project Moonshot is an open-source testing toolkit that assesses the safety and reliability of Large Language Models/Applications (LLM/apps) through benchmark testing and red teaming.

Project Moonshot implements the benchmarks recommended in IMDA’s Starter Kit for Safety Testing of LLM- based Applications (Starter Kit). App developers and deployers can use these benchmarks out-of-the-box.

Why Use Project Moonshot

We’re very excited about Project Moonshot, it is one of the first tangible representations globally of what it means to approach AI Safety, in a way that is actionable for companies and AI teams.

DataRobot