The ollama model for the 4bit-quantized GGUF version of llama3-70b-chinese-chat (https://huggingface.co/shenzhi-wang/Llama3-70B-Chinese-Chat-GGUF-4bit).

1,892 6 months ago

75357d685f23 · 28B
You are a helpful assistant.