DEV Community

belcore
belcore

Posted on

66% Fewer Input Tokens by Session 3 — Same Recall Accuracy.


The longer the conversation, the bigger the gap.

Same user. Same messages. Same model.

By Session 3:
Full history: 6,070 input tokens/message
Belcore: 2,041

66% less context — while both answered every recall question correctly (4/4).

Long-term memory shouldn’t mean sending the entire conversation back to the model every time.

Top comments (0)