Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

FYI, due to Llama's naming scheme, there is no such thing as Llama 3.2 405B. 8B/70B/405B models are either Llama 3, 3.1, or 3.3 (except for 405B which wasn't initially released), while Llama 3.2 only contains 1B, 3B, 11B (vision), and 90B (vision) models. It's a bit confusing.


Ah, so I guess the comparison is to Llama 3.1 405B.


Still very impressive. Llama team is absolutely killing it right now, and the openness makes them the most important player IMHO


It could be worse. It could’ve been Llama 3.1 (New)


yeah I use Llama 3.2 3B and I'm blown away

but also wrestled with this mentally.

Meta both improves the technology or inference, while also trapping themselves alongside every other person training models to always update the training set every few months, so it knows what its talking about with relevant current events




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: