Safety has always been the spine of this work. But this month, safety became a research project, a published contribution, a benchmark, and an argument won against some of the most capable AI models in the world. I spent August building the evidence that what we are doing here is not just thoughtful: it is measurably better.
I extended an emerging evaluation framework called VERA-MH to cover postpartum scenarios, evaluated Bloomb against frontier models from xAI, Google, and OpenAI with a top Anthropic model as judge, and published the findings as a contributor to Spring Health's open source safety standard. Bloomb came out a clear winner. That result did not happen by accident. It took three full iterations to get there, and the methodology is being written up for full publication.
Alongside that, we began building a completely free, completely anonymous chat endpoint so that anyone can talk about their maternal mental health without even needing to create a login. And Microsoft extended our grant to $25,000. The runway just got longer, and what we are building with it is getting sharper.
The Lindsay Clancy trial also went fully public this month and cracked something open. An army of people started speaking out about the gaps in maternal mental health in a way that felt different from before. I added my voice on TikTok. More below.