Projects

Tools for making expert judgment measurable, reviewable, and easier to use.

Graphletter

Open-source GRC product that compares organizational evidence with control objectives across 81 supported frameworks.

LLM extraction
Secure Controls Framework
Evidence pipeline
Source-cited reasoning
+2 more

BarPlaybook

AI-graded bar-exam essay platform. Every scoring change is checked against a 32-essay human-graded golden set (MAE 2.58; 84.4% within 5 points).

Evaluation harness (MAE / within-5)
Golden set + calibration loop
LLM scoring with deterministic post-processing
Human-grounded evaluation
+1 more