I run a Qwen 3.8 27b at home on gaming hardware. Got it scraping and posting news articles and general stupid questions. Takes 5 minutes instead of seconds because it has a gemma model doing RAG in the back end.
Other than some coding, I really don’t have a reason to use a big model.
I run a Qwen 3.8 27b at home on gaming hardware. Got it scraping and posting news articles and general stupid questions. Takes 5 minutes instead of seconds because it has a gemma model doing RAG in the back end.
Other than some coding, I really don’t have a reason to use a big model.