Running a local LLM model with llama-cli gets a 43% boost in tok/s running on CPU if I run llama inside a jail without anything much.
I am curious why. What would slow it down running on the host as is?
I am curious why. What would slow it down running on the host as is?