
The longer the conversation, the bigger the gap.
Same user. Same messages. Same model.
By Session 3:
Full history: 6,070 input tokens/message
Belcore: 2,041
66% less context — while both answered every recall question correctly (4/4).
Long-term memory shouldn’t mean sending the entire conversation back to the model every time.
Top comments (0)