Where It All
Started
Born at the UT Southwestern Simulation Center, the MAPLES platform has graded over 7,000 clinical encounters in production. A 300-encounter retrospective multimodal preprint reported three-camera AI grading agreement of quadratic weighted kappa = 0.830 against a physician-adjudicated reference, compared with 0.732 for standard human evaluators against the same reference. Related work is published in NEJM AI and JMIR AI. Now, through the UT-REAL initiative, we're preparing that capability for governed validation across additional UT medical schools.