Sixteen promises and two
models that tried to break them — that's how the
announcement of skillmem 0.12 sounds. In version 0.11, the authors ran AI reviewers through the code for forty rounds and found many bugs. The problem was that each reviewer had their own idea of correctness. In 0.12 they started differently. What exactly changed