Nemotron-Labs-3-Puzzle-75B-A9B-APEX-GGUF is now available on Ollama
Nemotron-Labs-3-Puzzle-75B-A9B-APEX-GGUF is now available through local Ollama runtime. 128K context window listed. Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.
View source