The distinguishing feature of this institute is not the hardware. It is the machinery for not fooling ourselves. Simulation and exhaustive search produce confident wrong answers easily, and most of the engineering effort here has gone into making that harder. What each instrument refuses is published. How it works is not.
Each program specifies its apparatus before any result exists, answering one question: what should the instrument be, specifically for this? Nothing is copied wholesale between programs. What counts as convergence, as completeness, or as a valid comparator is redefined per domain, and a rule that earned its place in one program is refused in another when the domain does not support it.
The measure of whether this works is not elegance. It is catches before results exist: in one program the apparatus surfaced four real defects, three of them before any result had been produced.
Every instrument failure is written up with the incident behind it and an applicable rule. The corpus runs to more than eighty entries across four programs and is held internally — it is the accumulated cost of the apparatus, and it is what the apparatus is for.
The test of whether an apparatus compounds is whether a lesson ever catches a failure without a human noticing first. Across programs, five have — and one of those was caught by a check that existed only because an earlier lesson demanded it. In one program a failure was prevented outright by a lesson recorded an hour earlier in another.
The characteristic failure this apparatus is built for does not crash. In one program, fourteen of sixteen recorded failures produced plausible, well-formed, entirely wrong output, and every one of them ran to completion successfully.