OpenReasoning-Nemotron-7B-GGUF is now available on Ollama
OpenReasoning-Nemotron-7B-GGUF is now available through local Ollama runtime. 128K context window listed. Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.
View source