LLMs have a limited context window. When conversations grow too long this affects output quality, performance and cost.
New blog post from Earendil engineer
@vegardstikbakke on how compaction addresses this and how we’ve implemented it in Pi.
Read the full post below