logoalt Hacker News

altcognitotoday at 2:41 PM0 repliesview on HN

I'm assuming you're referring to a harness that includes memory -- I generally think of the harness as anything beyond executing the generation loop, but I'm not an expert.

True as that may be, it may be better to optimize models for some amount of memory versus forcing some token count based on a reasoning level, right?