Download a model to run on your own hardware, or run it on Ollama’s cloud.
2 models
Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.