I am running 27B with Deepseek Harness these days and somehow just by using it, without any parameter changes, the model feels even more intelligent.
do LLMs tend to be homesick when not used in the same harness they sat in during some training phase?
do LLMs tend to be homesick when not used in the same harness they sat in during some training phase?