An update to Meta Llama 3 8B Instruct that includes an expanded 128K context length, multilinguality and improved reasoning capabilities.
Model availability and access depend on your workspace and plan.
About the model
An update to Meta Llama 3 8B Instruct that includes an expanded 128K context length, multilinguality and improved reasoning capabilities.
What you get back. The response appears directly in your Relam chat, where you can continue the conversation.
Capabilities & limits
Ask questions, give instructions and continue with follow-up messages in a conversation.
The model can work with tools as part of a response. Relam determines which tools are available for your request and enforces their permissions.
Llama 3.1 8B Instruct has a 128,000-token context window. This is the model’s capacity for the text it considers in a request, including prompts and conversation context. Tokens are pieces of text, rather than a word or page count. The amount of context sent by your workspace can be lower.
The catalogue sets a maximum output of 8,192 tokens for Llama 3.1 8B Instruct. This is the response budget, distinct from its context window. Responses can be shorter depending on your request and the limits applied by your workspace.
Llama 3.1 8B Instruct is not marked as reasoning-capable in the current catalogue.
Questions & answers