Original Reddit post

Well all I did was load termux and import a llama chat model in there but there’s a problem now it has a ceiling of like 4000 tokens or smth so a chat with it can’t go above that but I need it to remember the chat so what I can do it change the model to a better one but that’s a temporary solution BUT I can also make it save every single message in SQLITE basically a database which it will store every single message on and every fact that’s worth remembering ect the cons is that it will take memory like A LOT and it is still annoying working with only 4000 tokens while the responses are like 3tokens/second so it will be annoying changing the chat every once in a while butt I can also change the model whenever I want (prob when I get the laptop back 👀) or I should do it now idk I’ll make a poll and y’all decide View Poll submitted by /u/Fun-Operation7561

Originally posted by u/Fun-Operation7561 on r/ArtificialInteligence