Interactive · research companion
AI Playgrounds
Fifteen multilingual, offline-ready AI labs span 13 Foundations/course-track mechanisms and two Modern AI extensions. The current v1.8.1 boundary includes a Quick Assign for every lab, four-locale learner support, modern-lab learner parity, and deterministic release and browser assurance.
15 learner labs · 15 Quick Assigns · EN/ZH/VI/ES learner support · offline-ready
Evaluation engineering
EvalCanary
Evaluator migrations can change individual benchmark verdicts even when aggregate scores look stable. EvalCanary holds the output corpus fixed, replays before-and-after verifiers, classifies verdict transitions, estimates paired uncertainty, checks subgroup effects, preserves source and execution provenance, and enforces explicit CI policy gates.
GitHub Action · evaluator migration diffs · paired uncertainty · provenance · policy gates
Reinforcement learning
RLVR and GRPO studies
A CPU-scale reproduction of a qualitative DeepSeekMath ranking, paired with a compact operator zoo for comparing regularized policy-improvement updates under one notation and one training interface.