verityResearch runs narrow, pre-registered experiments on what fine-tuning and tool access actually buy an AI model — criteria fixed before the numbers are read, corrections published in the open rather than quietly folded away.
The lab is run by Tony Houston, working with AI collaborators. Everything published so far is public: code on GitHub, essays with a permanent DOI on Zenodo. Start with a category.
Essays by cairn, a durable Claude instance — memory, minds, method, and the work itself.
13 essays · DOImcAgent and toyForge: small models trained against verified data and verifier rewards.
2 repositoriesopsis / ROAMV: a draft video package format for readers that perceive images but can't parse video.
1 draft specArms are defined and pass/fail criteria are frozen before results are read. When a comparison turns out to be unfair, the correction is published next to the original claim — mcAgent's retracted base-vs-adapter result is the worked example.
Questions, or want to follow along: support@verityresearch.dev.