Skip to content
Is there an AI for this?

Alternatives

Self-hosted alternatives to DeepSeek API

Projects in the registry that can run on your own hardware and serve at least one of the same use cases as DeepSeek API.

Source
Registry relations over fetched repository records — the evidence sits on each repository page
Verified
Evidence not verified
Confidence
High

01How this list was built

Compared against
DeepSeek API
Delivery
Cloud platforms
Shared use cases
Choosing a model, Local LLM, AI coding assistant, Translation
Selection rule
open source, or self-hostable, and shares a use case
Candidates found
13

02Alternatives

13
  • AnythingLLM

    Hybrid — hosted or self-hosted

    Desktop and server application that turns a document set into a chat workspace, with per-workspace embeddings, multiple model back ends and a built-in vector store.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Local llm

  • Continue

    Hybrid — hosted or self-hosted

    Open-source IDE extension for VS Code and JetBrains that connects completion and chat to any model back end, including a local server, under a checked-in configuration file.

    Stars
    35,622
    Licence
    Apache-2.0
    In stacks
    1

    Shares: Ai coding assistant · Local llm

  • Jan

    Hybrid — hosted or self-hosted

    Open-source desktop assistant that runs open-weight models locally and can also call remote endpoints. The open alternative to LM Studio for a single-machine deployment.

    Repository
    janhq/jan
    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Local llm

  • LiteLLM

    Cloud platforms

    Gateway that presents one OpenAI-compatible API in front of many providers and local servers, with per-team keys, budgets, rate limits, fallbacks and request logging.

    Repository
    BerriAI/litellm
    Stars
    57,216
    Licence
    Other
    In stacks
    3

    Shares: Local llm · Ai coding assistant

  • llama.cpp

    Libraries

    C++ inference engine for quantised models on CPU, Apple Silicon and GPUs, with a bundled HTTP server. The engine underneath many desktop runtimes and the GGUF quantisation format.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Local llm

  • LMDeploy

    Model servers

    Serving toolkit from the InternLM team with its own inference engine and quantisation tooling, and strong coverage of Chinese open-weight model families.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Local llm · Translation

  • LM Studio

    Hybrid — hosted or self-hosted

    Desktop application for downloading and running open-weight models locally, with a chat interface and a local OpenAI-compatible server. Windows, macOS and Linux.

    Repository
    none linked
    Stars
    Licence
    unknown
    In stacks
    1

    Shares: Local llm

  • LocalAI

    Model servers

    Drop-in OpenAI-compatible API server that runs text, embedding, image and audio models locally across several back ends, including CPU-only deployments.

    Repository
    mudler/LocalAI
    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Local llm

  • MLflow

    Cloud platforms

    Experiment tracking, a model registry and deployment packaging. The component that lets a prediction be traced back to the model version and data that produced it.

    Repository
    mlflow/mlflow
    Stars
    Licence
    unknown
    In stacks
    1

    Shares: Model selection

  • NVIDIA NIM

    Model servers

    Prebuilt inference containers with an OpenAI-compatible API, deployable into your own cluster including air-gapped sites. Licensed through NVIDIA AI Enterprise rather than open source.

    Repository
    none linked
    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Local llm

  • Compiles models into optimised engines for NVIDIA GPUs. The build step is the trade: peak throughput on a fixed configuration, in exchange for recompiling when it changes.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Local llm

  • Ollama

    Model servers

    Local model runtime with a one-command install, a model library and an OpenAI-compatible API. The usual starting point for running open-weight models on a workstation or small server.

    Repository
    ollama/ollama
    Stars
    179,403
    Licence
    MIT
    In stacks
    4

    Shares: Local llm · Ai coding assistant

  • Open WebUI

    Self-hosted projects

    Self-hosted chat interface for local and hosted models, with user accounts, groups, document upload and built-in retrieval. Runs in Docker against Ollama, vLLM or any OpenAI-compatible endpoint.

    Stars
    149,870
    Licence
    Other
    In stacks
    5

    Shares: Local llm · Translation


03Repository facts

RepositoryStarsHealth
Mintplex-Labs/anything-llmHealth not measuredunknown
continuedev/continueopen-source coding agent36kMaintenance 100%, Releases 60%, Contributors 50%, Issues 74%, Community 91%80
janhq/janHealth not measuredunknown
BerriAI/litellmThe fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]57kMaintenance 100%, Releases 100%, Contributors 50%, Issues 14%, Community 95%83
ggml-org/llama.cppHealth not measuredunknown
InternLM/lmdeployHealth not measuredunknown
mudler/LocalAIHealth not measuredunknown
mlflow/mlflowHealth not measuredunknown
nvidia/tensorrt-llmHealth not measuredunknown
ollama/ollamaGet up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.179kMaintenance 100%, Releases 80%, Contributors 50%, Issues 79%, Community 100%86
open-webui/open-webuiUser-friendly AI Interface (Supports Ollama, OpenAI API, ...)150kMaintenance 100%, Releases 60%, Contributors 50%, Issues 99%, Community 100%84

04Ask about your own case

Which alternative is right depends on how many people will use it, what the data is and what you can operate. Ask and the answer is researched for your organisation.

  1. 01What is a self-hosted alternative to DeepSeek API that we can run ourselves?