Projects
Tools for making expert judgment measurable, reviewable, and easier to use.
Graphletter
Open-source GRC product that compares organizational evidence with control objectives across 81 supported frameworks.
LLM extraction
Secure Controls Framework
Evidence pipeline
Source-cited reasoning
+2 more
BarPlaybook
AI-graded bar-exam essay platform. Every scoring change is checked against a 32-essay human-graded golden set (MAE 2.58; 84.4% within 5 points).
Evaluation harness (MAE / within-5)
Golden set + calibration loop
LLM scoring with deterministic post-processing
Human-grounded evaluation
+1 more