
Project Moonshot is an open-source testing toolkit that assesses the safety and reliability of Large Language Models/Applications (LLM/apps) through benchmark testing and red teaming.
Project Moonshot implements the benchmarks recommended in IMDA’s Starter Kit for Safety Testing of LLM- based Applications (Starter Kit). App developers and deployers can use these benchmarks out-of-the-box.

We’re very excited about Project Moonshot, it is one of the first tangible representations globally of what it means to approach AI Safety, in a way that is actionable for companies and AI teams.