# Llama 3.1 - 405B, 70B & 8B with multilinguality and long context

DevFeed: [Llama 3.1 - 405B, 70B & 8B with multilinguality and long context](<https://devfeed.tech/articles/llama-3-1-405b-70b-8b-with-multilinguality-and-long-context-7335.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/llama31>)

Author: Philipp Schmid; Omar Sanseviero; Alvaro Bartolome; Leandro von Werra; Daniel Vila; Vaibhav Srivastav; Marc Sun; Pedro Cuenca

Published: 2024-07-23T00:00:00Z

Content type: release

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [llama](<https://devfeed.tech/topics/llama.md>), [Large Language Model](<https://devfeed.tech/topics/llm.md>), [long-context](<https://devfeed.tech/topics/long-context.md>), [Meta](<https://devfeed.tech/topics/meta.md>), [Inference](<https://devfeed.tech/topics/inference.md>)

Tags: [community](<https://devfeed.tech/tags/community.md>), [inference](<https://devfeed.tech/tags/inference.md>), [llama](<https://devfeed.tech/tags/llama.md>), [llm](<https://devfeed.tech/tags/llm.md>), [long-context](<https://devfeed.tech/tags/long-context.md>), [meta](<https://devfeed.tech/tags/meta.md>), [nlp](<https://devfeed.tech/tags/nlp.md>), [research](<https://devfeed.tech/tags/research.md>)

## AI overview

Hugging Face describes the Llama 3.1 release: six base and instruction-tuned models in 8B, 70B, and 405B sizes, plus Llama Guard 3 and Prompt Guard. The models support 128K-token context lengths and eight languages, and the article covers integrations, quantization, fine-tuning, and synthetic-data workflows.

## Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.