Models · 25 tracked
Pick a model by what you have and what you want it for.
Every model here can be downloaded and run on your own hardware. The number that matters most is the memory it needs, so the list is grouped by it.
Run locally12 of 25
Needs 22 GB or less at 8K context. Ordered by how many reference systems run it well. Show all models
Most laptops
up to 6 GBRuns on most laptops, even without a graphics card.
- Compact general-purpose model for coding and tool use.Qwen Team (Alibaba Cloud) · Runs well on most reference systems~6 GB estimateto runPermissive
- Compact general-purpose model for coding and tool use.Meta · Runs well on most reference systems~6 GB estimateto runPermissive
- ~6 GB estimateto runPermissive
A 16 GB graphics card
6–15 GBOr a Mac with 24 GB of memory.
- General-purpose model for coding and reasoning.Microsoft · Runs well on most reference systems~10 GB estimateto runPermissive
- General-purpose model for coding and tool use.Qwen Team (Alibaba Cloud) · Runs well on most reference systems~10 GB estimateto runPermissive
- ~9 GB estimateto runRestricted
- ~15 GB estimateto runPermissive
A 24 GB graphics card
15–22 GBOr a Mac with 32 GB of memory.
- Model for coding. Among the strongest open models at coding.Qwen Team (Alibaba Cloud) · Runs well on many reference systems~21 GB estimateto runPermissive
- Efficient general-purpose model for coding, reasoning and tool use.Qwen Team (Alibaba Cloud) · Runs well on many reference systems~18 GB estimateto runPermissive
- General-purpose model for coding, reasoning and tool use. Among the strongest open models at reasoning.Qwen Team (Alibaba Cloud) · Runs well on many reference systems~21 GB estimateto runPermissive
- Chat model that also understands images. Among the strongest open models at following instructions.Google DeepMind · Runs well on many reference systems~20 GB estimateto runRestricted
- ~20 GB estimateto runRestricted
Parameters, context length, capability profiles and side-by-side comparison are in the Technical view.