slow-stack is a small lab for work that has to survive its own audit: preregistered experiments, frozen evaluation sets, and SHA-256 lineage from corpus to weights.

What lives here

How to read the numbers

Thresholds and the aggregation rule were written down before the runs; three runs of the same configuration are reported by median, never picked. All data is synthetic — no real user data — and the evaluation set is asserted disjoint from the training corpus (leak audit 0/51). Code and claims pack: github.com/modusensus/laya.

Also here: ai-governance-compare — a claim-level EU↔China AI-regulation comparison with A/B/C evidence grading (conclusions may only rest on grade A).