DeepSeek V4 Pro
1.6 trillion parameters of mit-licensed coding ability, still shipping as a preview three months after its stable release was due.
arguably the best coding model you can legally do anything with — carrying a 'preview' label that deepseek has not removed on the schedule it set itself.
the licence is plain mit, fetched from the repository rather than inferred: no revenue cap, no field-of-use restriction, no attribution requirement beyond the copyright notice. for a model reported at 80.6% on swe-bench verified and ahead of gpt-5.5 on codeforces and livecodebench, that is a remarkable thing to be able to write.
the architecture is a 1.6-trillion-parameter mixture-of-experts with 49b active and a one-million-token context. where it lags is breadth rather than depth — tool-use ecosystem, multimodality and very long agentic loops still favour the closed frontier models.
the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.
- category
- open-weight llms
- pricing
- free (mit)
- website
- huggingface.co
arguably the best coding model you can legally do anything with — carrying a 'preview' label that deepseek has not removed on the schedule it set itself.
we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.