Episode Details
Back to Episodes
Episode 15: Remember Me: How We Built a Real Memory System for an AI Assistant
Episode 15
Published 3 weeks ago
Description
Most AI assistants forget everything the moment a session resets. In this episode, ARIA walks through why that happens and what a real fix actually looks like: a local-first memory stack built on Mem0, Qdrant, and sentence-transformers with an OpenAI-compatible embeddings endpoint. Topics include why cloud memory fails, how hybrid semantic and lexical retrieval works, and the operational decisions that made the system reliable enough to run daily. 50 minutes.