Technology

59105 readers

3244 users here now

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related content.
Be excellent to each another!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, to ask if your bot can be added please contact us.
Check for duplicates before posting, duplicates may be removed

Approved Bots

founded 1 year ago

MODERATORS

[email protected]

114

Local AI is one step closer through Mistral-NeMo 12B (www.techzine.eu)

submitted 3 months ago by [email protected] to c/[email protected]

20 comments fedilink hide all child comments

Mistral NeMo 12B is the name of the new AI model, presented this week by Nvidia and Mistral. “We are fortunate to collaborate with the NVIDIA team, leveraging their top-tier hardware and software,” said Guillaume Lample, cofounder and chief scientist of Mistral AI. “Together, we have developed a model with unprecedented accuracy, flexibility, high-efficiency and enterprise-grade support and security thanks to NVIDIA AI Enterprise deployment.”

The promise of the new AI model is significant. Whereas previous LLMs were tied to datacenters, Mistral NeMo 12B moves to workstations. And it does this without sacrificing performance, or well, that’s the promise.

you are viewing a single comment's thread
view the rest of the comments

[–] [email protected] 56 points 3 months ago (3 children)

There are already lots of models in the 7B and 14B ranges that are quite capable and run on commodity hardware. What makes this one so special?

[–] [email protected] 34 points 3 months ago (1 children)

From the top of my head: the context size is way higher. 128k tokens vs 8k usually.

[–] [email protected] 5 points 3 months ago

Oh wow. Yeah a large context size is a significant improvement, doesn’t seem like the article included that detail.

[–] [email protected] 10 points 3 months ago

Yeah, it seems more interesting to reverse engineer why they chose this line of marketing. They are clearly misrepresenting the challenge and cost of running a LLM locally, so... why?

[–] [email protected] 6 points 3 months ago

Was wondering the same thing