182
If open weight models are the future, U.S. AI companies are going to have a hard time
(www.fastcompany.com)
This is a most excellent place for technology news and articles.
Oh yeah, 16 GB of VRAM is a strange spot to be in. Most of the focus goes onto the models that fit in 24 GB cards.
Qwen 3.6 35b-a3b is pretty solid when you don't have enough VRAM for a dense model. Most of the weights can be left in RAM and it still runs really quickly.