this post was submitted on 04 Jan 2024
25 points (100.0% liked)
LocalLLaMA
2775 readers
25 users here now
Welcome to LocalLLama! This is a community to discuss local large language models such as LLama, Deepseek, Mistral, and Qwen.
Get support from the community! Ask questions, share prompts, discuss benchmarks, get hyped at the latest and greatest model releases! Enjoy talking about our awesome hobby.
As ambassadors of the self-hosting machine learning community, we strive to support eachother and share our enthusiasm in a positive constructive way.
founded 2 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
The 3060 is a nice cheap one for running okay sized models, but if you can find a way to stretch for a 3090 or a 7900 XTX you'll be able to run these 33B models with decent quant levels
I was hoping to avoid Nvidia's binary drivers although I don't know what the driver/support status of dedicated AI accelerators are like on Linux._
I run my Nvidia stuff in containers to not have to deal with all the stupid shenanigans