logoalt Hacker News

indigo945today at 3:05 PM1 replyview on HN

There's no inherent, technological reason to feed the historic data into the prompt, though. You could train a network with a second input for model-states, and during inference, feed into that input historic model-states that occured for the same LLM when consuming input that is somehow associatively related to whatever input is currently being consumed. Can you definitely show that this would be meaningfully different from how humans "experience" memories?


Replies

mnewmetoday at 5:15 PM

That would still be circumventing. There are other architectures that could be more dynamic, LLMs are not.

It is completely different. Our brain rewires itself to store memories, there is no memory storage for humans.

show 1 reply