Research areas
Where the work is concentrated
R-01
Continuous Evaluation
Measuring bias, jailbreak resistance, grounding, and drift against live traffic rather than a frozen benchmark. Methodology and its limits published openly.
R-02
Policy Compilation
Translating written responsible-AI policy into executable guardrails, and studying where that translation loses fidelity.
R-03
Oversight Interfaces
What a human supervising an AI system needs to see, when, and in what form, for the oversight to be real rather than nominal.
R-04
Domain Model Training
Pre-training from scratch and fine-tuning toward AI governance, so the studio owns its pipeline end to end in a field where explainability is the product.
R-05
AI Literacy
How people build accurate mental models of what these systems do, and how interfaces help or hinder that.