Damus
jcoglan profile picture
jcoglan
@jcoglan
somehow I only recently learned that LLMs need the whole conversation fed back to them on each prompt so their i/o cost scales as O(n^2), something that would be considered completely unacceptable in almost any other production network-accessible software
1
abadidea · 1w
nostr:nprofile1qy2hwumn8ghj7un9d3shjtnyd968gmewwp6kyqpqu0qy2ps70wkdvy9pnkljl0z85rj605jzmzzl0qvkdhc7fv0ztasqsyldna the main reason that the longer you use them, the crazier they get is that at some point it will resort to resummarizing the prompt to reduce how much of the context window it’s eating...