Llama3.1-8B

Built with Llama. Unmodified mirror of Meta's Llama 3.1 8B base model. No fine-tuning, no quantization, no change to the weights.

architecture LlamaForCausalLM, 32 layers, vocabulary 128,256
precision bf16
context 131,072 tokens
formats safetensors (4 shards) plus original/consolidated.00.pth in Meta's checkpoint format

This is the base model, not Instruct. It ships no chat template. Use it for completion or as a fine-tuning starting point; for chat, use meta-llama/Llama-3.1-8B-Instruct.

Apple silicon conversions: 4-bit MLX, 8-bit MLX.

Full model card, benchmarks, training details and responsible-use guidance: meta-llama/Llama-3.1-8B.

License

Llama 3.1 is licensed under the Llama 3.1 Community License, Copyright (c) Meta Platforms, Inc. All Rights Reserved. The full agreement is in LICENSE and reproduced in the gate above. The acceptable use policy is in USE_POLICY.md.

Downloads last month
238
Safetensors
Model size
8B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for cyboghostginx/Llama3.1-8B

Finetuned
(1482)
this model
Quantizations
2 models