jcoglan · 1w somehow I only recently learned that LLMs need the whole conversation fed back to them on each prompt so their i/o cost scales as O(n^2), something that would be considered completely unacceptable in ... abadidea @abadidea 1784818005 @nprofile1q... the main reason that the longer you use them, the crazier they get is that at some point it will resort to resummarizing the prompt to reduce how much of the context window it’s eating; a machine that destroys its own nuance to live