I stopped using my RTX GPU for every LLM.
Integrated graphics handled these local AI models better than expected.