In high-stakes use cases, testing AI systems for reliability and safety is essential. And the stakes don’t get much higher than in the public sector! So, for our August meetup, we brought 90+ practitioners from the public and private sector together.
We invited a few best-in-class GenAI dev and testing practitioners from the public sector to share their experiences:
Testing at scale across multiple use cases/teams by Benjamin Goh from National Group & GovTech Singapore
Ben shared about AI Guardian, a groundbreaking SaaS platform to enable GenAI testing across the public sector. Ben took us through the journey of implementing comprehensive AI testing and shared insights on creating standardised safety guardrails across teams through Litmus and Sentinel.
Embedding safety and reliability into GenAI for the legal sector by Eric Tan from IMDA Biztech
Eric spoke about how GPT Legal is transforming Singapore’s legal landscape. Eric revealed the intricate process of developing an LLM specifically for legal research, highlighting their partnership with Singapore Academy of Law and the implementation of robust safety measures throughout the development lifecycle.
Testing if your RAG application knows its limits by Jessica Foo & Shaun Khoo from GovTech Singapore
They shared exclusive insights into “KnowOrNot”, an innovative approach to detecting hallucinations in RAG applications, and how this open-source solution cleverly manipulates context to ensure AI remains grounded in reality – a game-changer for high-stakes applications.
Our speakers joined a collective, moderated Q&A session to take questions from the audience:
See you at our next meet up!