New state of the art 70B model. Llama 3.3 70B offers similar performance compared to the Llama 3.1 405B model.
4.1M Pulls 14 Tags Updated 1 year ago
Llama 3.1 is a new state-of-the-art model from Meta available in 8B, 70B and 405B parameter sizes.
119M Pulls 93 Tags Updated 1 year ago
New state of the art 70B model. Llama 3.3 70B offers similar performance compared to Llama 3.1 405B model.
56.8K Pulls 10 Tags Updated 1 year ago
9 Pulls 1 Tag Updated 1 year ago
New state of the art 70B model. Llama 3.3 70B offers similar performance compared to Llama 3.1 405B model. I-Quants models.
427 Pulls 9 Tags Updated 1 year ago
https://huggingface.co/bosonai/Higgs-Llama-3-70B
242 Pulls 1 Tag Updated 2 years ago
This model extends LLama-3 70B's context length from 8k to over 1m tokens. [I-Quants]
212 Pulls 4 Tags Updated 2 years ago
NousResearch/Hermes-3-Llama-3.1-70B Korean q4 model with CPT->SFT->DPO
98 Pulls 1 Tag Updated 1 year ago
NousResearch/Hermes-3-Llama-3.1-70B Korean q5 model with CPT->SFT->DPO
31 Pulls 1 Tag Updated 1 year ago
watt-tool-70B is a fine-tuned language model based on LLaMa-3.3-70B-Instruct, optimized for tool usage and multi-turn dialogue. It achieves state-of-the-art performance on the Berkeley Function-Calling Leaderboard (BFCL).
1,294 Pulls 1 Tag Updated 1 year ago
Llama-3.1-SauerkrautLM-70b-Instruct is a fine-tuned Model based on meta-llama/Meta-Llama-3.1-70B-Instruct.
54 Pulls 1 Tag Updated 1 year ago
llama-3-taiwan(70B-Instruct, 70B-Instruct-DPO, 70B-Instruct-128k, 8B-Instruct, 8B-Instruct-DPO) (FP16, Q8_0, Q6_K, Q5_1, Q5_0, Q5_K_M, Q5_K_S, Q4_1, Q4_0, Q4_K_M, Q3_K_L, Q4_K_S, Q3_K_M, Q3_K_S)
3,519 Pulls 70 Tags Updated 2 years ago
Llama 3.1 Swallow is a series of large language models (8B, 70B) that were built by continual pre-training on the Meta Llama 3.1 models.
2,339 Pulls 13 Tags Updated 1 year ago
Llama-3-Taiwan-70B is a 70B parameter model finetuned on a large corpus of Traditional Mandarin and English data using the Llama-3 architecture. It demonstrates state-of-the-art performance on various Traditional Mandarin NLP benchmarks.
427 Pulls 1 Tag Updated 2 years ago
The model used is a quantized version of `Llama-3-Taiwan-70B-Instruct`. More details can be found on the website (https://huggingface.co/yentinglin/Llama-3-Taiwan-70B-Instruct)
386 Pulls 11 Tags Updated 2 years ago
Hermes 2 Pro - Llama-3 70B (f16.q4 and .q5)
48 Pulls 2 Tags Updated 2 years ago
m42-health/Llama3-Med42-70B in Ollama
1,128 Pulls 1 Tag Updated 2 years ago
🦙🦙🦙 Llama3-70B-Chinese-Chat is an instruction-tuned language model for Chinese & English users with various abilities such as roleplaying & tool-using built upon the Meta-Llama-3-70B-Instruct model.
212 Pulls 1 Tag Updated 2 years ago
Llama3-70B-EnSecAI-Ru-Chat
24 Pulls 3 Tags Updated 1 year ago