Memory has the widest spread of any axis we score: 9.3 for Nomi AI at the top, 1.0 for PornPen at the bottom, with most of the companion platforms clustered between 6.0 and 8.0.
More importantly, it is the axis that predicts whether someone is still using a platform in month two. Conversation quality in week one is a function of the underlying model, and the underlying models are increasingly similar. Conversation quality in week four is a function of whether the platform kept anything, and that is entirely an architecture question.
Nomi AI and Kindroid are the only two here that passed our thirty-day check consistently, and both do it the same way: extract durable facts into a separate store, retrieve by relevance rather than recency. The rest are still spending their context window on the last forty messages.
We expect this to be the axis the category competes on next, in the way image quality was the axis two years ago.