Key prompt templates used for evaluation and generation in the Mem0 paper
The Mem0 paper publicly releases three categories of key prompt templates. This page provides streamlined versions for reuse.
Uses another LLM to judge whether a generated answer is correct, outputting
Guides the LLM to answer questions based on memories from two speakers.
Extends the Mem0 template by additionally injecting graph memories.
Since ChatGPT has no external API to control its memory, evaluation manually injects via prompt:
LLM-as-a-Judge Scoring Template
Uses another LLM to judge whether a generated answer is correct, outputting CORRECT or WRONG.
J scores are mean ± standard deviation over 10 independent evaluations to avoid randomness from LLM scoring affecting conclusions.