Skip to content
Categories

Local AI.

Run language models on a laptop or a home GPU. Chat, complete, and serve weights without sending prompts to a vendor.

  1. Rank 1. OllamaLocal model runner with a one-line CLI and a REST API.GitHub stars: 180.7k
  2. Rank 2. Open WebUISelf-hosted chat UI that sits on Ollama, OpenAI, or a local runner.GitHub stars: 151.8k
  3. Rank 3. exoLink everyday devices into one cluster and run models too big for any single machine.GitHub stars: 47.4k
  4. Rank 4. llama.cppC/C++ runtime that made local GGUF models practical.GitHub stars: 128k
  5. Rank 5. JanDesktop ChatGPT-style app that runs models locally.GitHub stars: 44.4k
  6. Rank 6. LocalAIOpenAI-compatible API that can sit in front of several local backends.GitHub stars: 49.1k
  7. Rank 7. MLXApple’s framework for running and training models on Mac unified memory.GitHub stars: 28.4k
  8. Rank 8. GPT4AllNomic’s local chat app and model catalog for laptops without a GPU.GitHub stars: 77.4k
  9. Rank 9. KoboldCppOne-file llama.cpp build with a built-in chat UI aimed at offline storytelling.GitHub stars: 11.7k
  10. Rank 10. text-generation-webuiGradio UI for local models with loaders, LoRA, and extensions.GitHub stars: 47.7k
  11. Rank 11. PocketPal AIRun small GGUF chat models fully offline on iOS and Android.GitHub stars: 8.3k
  12. Rank 12. llamafileSingle-file LLM runtime from Mozilla. Download, chmod, run.GitHub stars: 26k
  13. Rank 13. shell_gptCLI that runs an LLM as a shell command for scripts, refactors, and answers.GitHub stars: 12.3k