verifier.org

Jan

gui, server and inference engine all apache 2.0 — the only tool here where every layer is genuinely open.

free#anyone-who-wants-ollama-

the cleanest answer in the category since ollama's gui went closed — smaller and less polished, and the only one with nothing to explain.

gui, local openai-compatible server and the cortex inference engine underneath are all apache 2.0 in the same repository. of the nine tools here that is a claim only jan, llama.cpp, gpt4all and localai can make, and jan is the only one of those four that is also a finished desktop application.

practically it does what you want: cuda, metal and vulkan backends, gguf models, an openai-compatible endpoint so existing code points at it unchanged, and installers for macos, windows and linux.

the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.

pricing
free
website
github.com
our verdict

the cleanest answer in the category since ollama's gui went closed — smaller and less polished, and the only one with nothing to explain.

we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.

more local llm tools

Text Generation WebUI

oobabooga's gradio interface for running local models and loaders

#open-source#gradio

vLLM

high-throughput serving engine you can run on your own gpu

#serving#throughput

MLX LM

apple's framework for running and fine-tuning models on apple silicon

#apple-silicon

llama.cpp

we checked thisfree

mit, and the engine most of this list is quietly built on top of.

#anyone-comfortable-at-a-#anyone-who-wants-the-new