Jia Yi Goh

Jia Yi Goh

Strengthening Chatbot Safeguards with Red-Teaming Agents
The Lab

Strengthening Chatbot Safeguards with Red-Teaming Agents

How we used agentic AI to proactively test and strengthen chatbot safeguards for chat.gov.sg
AgenticSecurity
Decision Models for Guardrails: Exploring Jev and Kev for Moderation
The Lab

Decision Models for Guardrails: Exploring Jev and Kev for Moderation

A closer look at the new paradigm of frontier models, and how we explored Jev and adapted its open-source interpretation - Kev, for Singapore-context moderation.
Responsible AI
Building a Smart and Safe Whole-of-Government Chatbot for Singaporeans
The Studio

Building a Smart and Safe Whole-of-Government Chatbot for Singaporeans

How we build custom policy guardrails and balance safety with helpfulness for chat.gov.sg
Responsible AIEvals
Guardrails in the Wild: Closing the Retraining Loop for LionGuard 2
The Lab

Guardrails in the Wild: Closing the Retraining Loop for LionGuard 2

Operationalising MLOps with production feedback to continuously evaluate, retrain, and improve our guardrails.
Responsible AI
The Realities of Robot Deployment: What It Takes for Embodied AI to Succeed
The Studio

The Realities of Robot Deployment: What It Takes for Embodied AI to Succeed

The "hype" of robots ignore the unstructured environment problem.
Embodied AI

Does your LLM know when to say “I don’t know”?

Refusal by a model to answer may sometimes be more valuable.
Responsible AI