Responsible AI

A collection of 13 posts
Guardrails in the Wild: Closing the Retraining Loop for LionGuard 2
The Lab

Guardrails in the Wild: Closing the Retraining Loop for LionGuard 2

Operationalising MLOps with production feedback to continuously evaluate, retrain, and improve our guardrails.
Responsible AI
My 6 Months Interning Responsibly at AI Practice
The People

My 6 Months Interning Responsibly at AI Practice

Tell us about yourself- what did you expect from this internship? I’m Rohan, a penultimate year Business Analytics student at the National University of Singapore (NUS) who spent the first half of 2026 as an applied AI Data Scientist Intern at GovTech’s Responsible AI team.  Going in, I
Responsible AI
Building an automated Evals workflow that works (and open-sourcing it)
The Lab

Building an automated Evals workflow that works (and open-sourcing it)

How we built Kaleidoscope: A structured workflow for realistic, scalable, and human-aligned contextual AI evaluations.
Responsible AI
Yes, you’re absolutely right… Right? A mini survey on LLM sycophancy
The Lab

Yes, you’re absolutely right… Right? A mini survey on LLM sycophancy

Ever spoken to an AI and felt like it was responding with insincere praise?
Responsible AI
Benchmarking GPT-5 & GPT-OSS: A Responsible AI Approach
The Lab

Benchmarking GPT-5 & GPT-OSS: A Responsible AI Approach

Evaluating dimensions often overlooked by traditional benchmarks.
Responsible AIEvals

RabakBench: Multilingual AI Safety Evaluation Made Local

Global safety guardrails are often blind to local dialects and sensitivities.
Responsible AI

Does your LLM know when to say “I don’t know”?

Refusal by a model to answer may sometimes be more valuable.
Responsible AI

Fine-Tuning Language Models for Long-Context Data: Automated Stance Analysis of Citizen Discussions

Addressing technical challenges of processing high-volume public feedback for policy-making
Responsible AI

Securing Guardrails with Automated Red Teaming

Manual testing is no longer scalable.
Responsible AI
The Lab

(Part 2) LLM Safety Alignment for the Singapore Context using Supervised Fine-tuning and RLHF-based Methods

Safety must be "baked in".
Responsible AI
The Lab

(Part 1) LLM Safety Alignment for the Singapore Context using Supervised Fine-tuning and RLHF-based Methods

The process of "teaching" models to be safe
Responsible AI
The Lab

Eliciting Toxic Singlish from r1

A red-teaming exercise that proves even "reasoning" models can be coaxed.
Responsible AI