Plain-language provider view: what this company is strongest at, how many active models it has, what pricing we can verify, and which current models matter most.
12
Models
—
Top Rank
—
Avg Capability
223
Downloads
19
Likes
12
Open Models
Pricing Posture
0 / 12 models have official company pricing
This tells you how often we can verify direct first-party pricing instead of only broker or router pricing.
Lowest Verified Entry
Free/M
The lowest reliable public price we could verify across this provider's active lineup.
Strategic Strength
Specialized
GLM-5.3-Flash-GGUF leads current estimated value at ---.
Deployment Reach
7 / 12 models have verified deploy or runtime access
7 on your computer · 0 cloud servers you control · 0 hosted for you
Open weights do not always mean easy hosted access. For backpack-run, they usually mean you bring the hardware yourself or rent a cloud GPU when the models are larger.
Most open models here can run on your own hardware.
11 easy personal-hardware fits · 0 desktop GPU fits · 0 cloud GPU fits · 1 high-memory cloud fits
Open Source
2
Recent launches, pricing moves, benchmark updates, API changes, and research signals linked to this provider.
Recent updates about new ways to use this provider's models, including self-host and official runtime options.
SmolLM2-1.7B-Instruct-GGUF is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
| # | Model | Category | Params | Capability | Downloads | Cheapest Verified | Est. Value | Open |
|---|---|---|---|---|---|---|---|---|
| — | GLM-5.3-Flash-GGUF | Multimodal |
SmolLM2-1.7B-Instruct-GGUF is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
SmolLM2-135M-Instruct-GGUF is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
View sourceSmolLM2-135M-Instruct-GGUF is now available through local Ollama runtime. 8K context window listed. SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
View source| Unknown |
| — |
| 0 |
Free Free |
| --- |
| — | Devstral-Small-2-24B-Instruct-2512-GGUF | Multimodal | 24B | — | 0 | Price not verified View Plan | --- |
| — | Kokoro-82M-Backpack-TTS | Speech | 82M | — | 181 | Free Free | --- |
| — | Qwen3-ASR-0.6B-Backpack-ASR | Speech | 938M | — | 42 | Price not verified View Plan | --- |
| — | Qwen3-Coder-Next-GGUF | Specialized | Unknown | — | 0 | Price not verified View Plan | --- |
| — | Qwen2.5-0.5B-Instruct-GGUF | Specialized | 500M | — | 0 | Price not verified View Plan | --- |
| — | SmolLM2-1.7B-Instruct-GGUF | Specialized | 2B | — | 0 | Price not verified View Plan | --- |
| — | SmolLM2-135M-Instruct-GGUF | Specialized | 135M | — | 0 | Price not verified View Plan | --- |
| — | Z-Image-Turbo-Backpack-Image | Image Gen | Unknown | — | 0 | Free Free | --- |
| — | Wan2.2-TI2V-5B-Backpack-Video | Video | 5B | — | 0 | Free Free | --- |
| — | Qwen3-Coder-30B-A3B-Instruct-GGUF | Specialized | 30B | — | 0 | Price not verified View Plan | --- |
| — | whisper-large-v3-turbo-Backpack-ASR | Speech | Unknown | — | 0 | Free Free | --- |