this post was submitted on 22 Jul 2024
115 points (92.0% liked)

Technology

58091 readers
3064 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 1 year ago
MODERATORS
 

Mistral NeMo 12B is the name of the new AI model, presented this week by Nvidia and Mistral. “We are fortunate to collaborate with the NVIDIA team, leveraging their top-tier hardware and software,” said Guillaume Lample, cofounder and chief scientist of Mistral AI. “Together, we have developed a model with unprecedented accuracy, flexibility, high-efficiency and enterprise-grade support and security thanks to NVIDIA AI Enterprise deployment.”

The promise of the new AI model is significant. Whereas previous LLMs were tied to datacenters, Mistral NeMo 12B moves to workstations. And it does this without sacrificing performance, or well, that’s the promise.

you are viewing a single comment's thread
view the rest of the comments
[–] [email protected] 4 points 1 month ago* (last edited 1 month ago) (2 children)

The best GPU to buy right now would be an Intel arc a770. You can get them for under $300 with 16gb vram.

You should also make sure that you have a motherboard and CPU that supports a feature called "resizeable bar"

https://game.intel.com/us/stories/wield-the-power-of-llms-on-intel-arc-gpus/

https://pcpartpicker.com/products/video-card/#P=17179869184,51539607552&sort=price&page=1

[–] [email protected] 4 points 1 month ago

Just beware that like AMD, Intel GPUs suffer a performance hit when using LLMs because of the CUDA specific optimizations in frameworks like llama.cpp

[–] [email protected] 1 points 1 month ago (2 children)

Isn't there an AMD 16 GB card too, like a 7600 or something?

[–] [email protected] 2 points 1 month ago

It's in the list I linked on pcpartpicker they're about $50 more.

[–] [email protected] 2 points 1 month ago

There's many 16gb AMD cards since at least the 6000 series. The 7600 XT is probably want you're thinking of since the 7600 is only 8gb.