title

DarkThoughts , 2 hours ago to technology in Llama 3.1 AI Models Have Officially Released

Meta? lol

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

brucethemoose , 1 hour ago

Its really weird, but they’re kinda the heroes of open source AI now (as opposed to everything being locked behind a corporate API).

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

admin , 6 hours ago to technology in Llama 3.1 AI Models Have Officially Released

128k token context is pretty sweet. Mistral nemo also just launched with a similar context. Good times.

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

Throwaway4669332255 , 3 hours ago

How does the Nemo 12B compare to the Llama 3.1 8B?

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

admin , 2 hours ago

I haven’t given it a very thorough testing, and I’m by no means an expert, but from the few prompts I’ve ran so far, I’d have to hand it to Nemo concerning quality.

Using openrouter.ai, I’ve also given llama3.1 405B a shot, and that seems to be at least on par with (if not better than) Claude 3.5 Sonnet, whilst being a bit cheaper as well.

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

brucethemoose , 1 hour ago

Llama 70B is probably where its at, if you go the API route. It’s distilled from 405B, and its benchmarks are pretty close.

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

brucethemoose , 1 hour ago

At long context (close to the full 128K), Nemo is way better than llama 8B in my testing.

Turns out they are both very sensitive to quantization though.

TBH I didn’t know people here were running LLMs. Seems like most of Lemmy is very broadly anti AI?

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

admin , 1 hour ago

Yeah, there’s a massive negative circlejerk going on, but mostly with parroted arguments. Being able to locally run a model with this kind of context is huge. Can’t wait for the finetunes that will result from this (cough NeverSleep’s *-maid models come to mind).

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

brucethemoose , 39 minutes ago

I am looking into doing it on the 12B for myself, not so much for RP but novel style prose.

I am thinking literature + a fanfic dump as a dataset?

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

bilb , 1 hour ago

If forced to characterize the attitude of lemmy towards LLM/“AI,” I’d say people here are broadly interested in the tech but critical of the way it’s often used.

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

brucethemoose , 34 minutes ago

I dunno, with image models specifically it seems like they’re the devil because of the datasets they’re trained on, killing artists, and… that’s that. And LLMs to a lesser extent. There’s truth to all that, but there’s also a lot more.

I think most people don’t realize how much of an inflection point local running vs. corporate hosting could be, which is especially ironic on Lemmy.

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

General_Effort , 5 minutes ago

If by interested you mean willing to bullshit… Talking about AI here is like talking about evolution at bible camp in the deep south.

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

Halosheep , 1 hour ago

The loud minority is really loud.

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

ObsidianZed , 12 minutes ago

My impression is the general consensus is we don’t want huge corporations stealing data to train their AI models only to turn around and cram it down our throats anywhere they can with increasingly negative experiences. That being said, while I would generally agree with that, I still find it interesting and especially if I can host it myself.

Reply

Report

Activity

Open original URL

Copy original URL

Copy Mbin URL

Loading...

Federation

Status:

On | Off

/d/llama.meta.com

Threads

Comments

Domain

llama.meta.com

Active people

Random posts

Trish arrives home after a decade away. Patrick Cash reflects on what has changed and what remains the same in his short story Trish Malone....

16 hours ago to bookstodon

DATE: July 23, 2024 at 05:01AM...

13 hours ago to psychology

@chloroform_tea and I went into DEEP depths of discussion about the Arthur C Clarke Award shortlist! Read all our thoughts about octopus linguistics, fawn-spider hybrids, interiority and thoughts about mothers, fascism, criminal justice, social inequality, footnotes, and what a Clarkey type book actually is....

15 hours ago to bookstodon