Mistral NeMo 12B is the name of the new AI model, presented this week by Nvidia and Mistral. “We are fortunate to collaborate with the NVIDIA team, leveraging their top-tier hardware and software,” said Guillaume Lample, cofounder and chief scientist of Mistral AI. “Together, we have developed a model with unprecedented accuracy, flexibility, high-efficiency and enterprise-grade support and security thanks to NVIDIA AI Enterprise deployment.”
The promise of the new AI model is significant. Whereas previous LLMs were tied to datacenters, Mistral NeMo 12B moves to workstations. And it does this without sacrificing performance, or well, that’s the promise.
The best GPU to buy right now would be an Intel arc a770. You can get them for under $300 with 16gb vram.
You should also make sure that you have a motherboard and CPU that supports a feature called “resizeable bar”
https://game.intel.com/us/stories/wield-the-power-of-llms-on-intel-arc-gpus/
https://pcpartpicker.com/products/video-card/#P=17179869184,51539607552&sort=price&page=1
Isn’t there an AMD 16 GB card too, like a 7600 or something?
It’s in the list I linked on pcpartpicker they’re about $50 more.
There’s many 16gb AMD cards since at least the 6000 series. The 7600 XT is probably want you’re thinking of since the 7600 is only 8gb.
Just beware that like AMD, Intel GPUs suffer a performance hit when using LLMs because of the CUDA specific optimizations in frameworks like llama.cpp