Plain-language provider view: what this company is strongest at, how many active models it has, what pricing we can verify, and which current models matter most.
27
Models
—
Top Rank
—
Avg Capability
784.0K
Downloads
795
Likes
27
Open Models
Pricing Posture
0 / 27 models have official company pricing
This tells you how often we can verify direct first-party pricing instead of only broker or router pricing.
Lowest Verified Entry
Free/M
The lowest reliable public price we could verify across this provider's active lineup.
Strategic Strength
LLMs
Qwen2.5-Coder-7B-Instruct leads current estimated value at ---.
Deployment Reach
14 / 27 models have verified deploy or runtime access
10 on your computer · 0 cloud servers you control · 4 hosted for you
Open weights do not always mean easy hosted access. For unsloth, they usually mean you bring the hardware yourself or rent a cloud GPU when the models are larger.
Most open models here can run on your own hardware.
15 easy personal-hardware fits · 9 desktop GPU fits · 2 cloud GPU fits · 1 high-memory cloud fits
Open Source
5
Recent launches, pricing moves, benchmark updates, API changes, and research signals linked to this provider.
Recent updates about new ways to use this provider's models, including self-host and official runtime options.
tinyllama is now available through local Ollama runtime. 2K context window listed. The TinyLlama project is an open endeavor to train a compact 1.1B Llama model on 3 trillion tokens.
| # | Model | Category | Params | Capability | Downloads | Cheapest Verified | Est. Value | Open |
|---|---|---|---|---|---|---|---|---|
| — | Qwen2.5-Coder-7B-Instruct | LLMs | 8B |
tinyllama is now available through local Ollama runtime. 2K context window listed. The TinyLlama project is an open endeavor to train a compact 1.1B Llama model on 3 trillion tokens.
SmolLM2-135M-Instruct-GGUF is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
View sourceSmolLM2-1.7B-Instruct-GGUF is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
View sourceSmolLM2-1.7B-Instruct-bnb-4bit is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
View sourceQwen3.8-Flash-Next is now available through local Ollama runtime. 256K context window listed. This experimental preview of the architecture that will underpin Qwen4.
View sourceSmolLM2-135M-Instruct-GGUF is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
View sourceSmolLM2-1.7B-Instruct-GGUF is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
View sourceSmolLM2-1.7B-Instruct-bnb-4bit is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
View source| — |
| 190.9K |
Price not verified View Plan |
| --- |
| — | Llama-3.2-3B-Instruct | LLMs | 3B | — | 164.4K | Price not verified View Plan | --- |
| — | GLM-4.7-Flash | LLMs | 31B | — | 117.8K | Price not verified View Plan | --- |
| — | Qwen2.5-1.5B-Instruct | LLMs | 2B | — | 49.7K | Price not verified View Plan | --- |
| — | Qwen2.5-0.5B-Instruct | LLMs | 494M | — | 30.2K | Price not verified View Plan | --- |
| — | gemma-4-E4B-it | Multimodal | 8B | — | 65.4K | Free Free | --- |
| — | gemma-3-270m-it | LLMs | 268M | — | 36.3K | Free Free | --- |
| — | DeepSeek-OCR-2 | Multimodal | 3B | — | 5.6K | Free Free | --- |
| — | LFM2.5-VL-3B | Multimodal | 3B | — | 3.4K | Price not verified View Plan | --- |
| — | Qwen3-1.7B-Base | LLMs | 2B | — | 2.6K | Price not verified View Plan | --- |
| — | DeepSeek-R1-Distill-Qwen-7B | LLMs | 8B | — | 1.7K | Free Free | --- |
| — | Phi-4-mini-reasoning | LLMs | 4B | — | 1.4K | Price not verified View Plan | --- |
| — | DeepSeek-V4-Flash | LLMs | Unknown | — | 1.1K | Price not verified Deploy | --- |
| — | GLM-5.3 | LLMs | 753B | — | 897 | Price not verified Deploy | --- |
| — | Qwen3-30B-A3B-Thinking-2507 | LLMs | 31B | — | 701 | Price not verified View Plan | --- |
| — | Kimi-K3 | Multimodal | Unknown | — | 303 | Price not verified Deploy | --- |
| — | DeepSeek-V4-Pro | LLMs | Unknown | — | 151 | Price not verified Deploy | --- |
| — | aya-vision-8b | Multimodal | 9B | — | 101 | Free Free | --- |
| — | embeddinggemma-300m | Embeddings | 303M | — | 70.8K | Free Free | --- |
| — | Qwen-Image-2.1 | Image Gen | 7B | — | 12.2K | Free Free | --- |
| — | csm-1b | Speech | 2B | — | 8.5K | Free Free | --- |
| — | Llama-OuteTTS-1.0-1B | Speech | 1B | — | 7.6K | Free Free | --- |
| — | Qwen3-Embedding-0.6B | Embeddings | 596M | — | 6.4K | Price not verified View Plan | --- |
| — | FLUX.2-VAE | Specialized | Unknown | — | 4.4K | Free Free | --- |
| — | Spark-TTS-0.5B | Speech | 500M | — | 926 | Free Free | --- |
| — | embeddinggemma-300m-qat-q8_0-unquantized | Embeddings | 303M | — | 345 | Free Free | --- |
| — | embeddinggemma-300m-qat-q4_0-unquantized | Embeddings | 303M | — | 166 | Free Free | --- |