Skip to content
Is there an AI for this?

Alternatives

Self-hosted alternatives to LibreChat

Projects in the registry that can run on your own hardware and serve at least one of the same use cases as LibreChat.

Source
Registry relations over fetched repository records — the evidence sits on each repository page
Verified
Evidence not verified
Confidence
High

01How this list was built

Compared against
LibreChat
Delivery
Self-hosted projects
Shared use cases
Private company ChatGPT, Private LLM, Email drafting, Document Q&A
Selection rule
open source, or self-hostable, and shares a use case
Candidates found
28

02Alternatives

28
  • AnythingLLM

    Hybrid — hosted or self-hosted

    Desktop and server application that turns a document set into a chat workspace, with per-workspace embeddings, multiple model back ends and a built-in vector store.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Private company chatgpt · Document qa

  • authentik

    Hybrid — hosted or self-hosted

    Identity provider with OpenID Connect, SAML and proxy-based authentication, application-level policies and a forward-auth outpost for services that have no login of their own.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Private company chatgpt

  • Chroma

    Hybrid — hosted or self-hosted

    Embedded and server-mode vector store with a small API surface, often used for prototypes and single-node retrieval before a larger engine is justified.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Document qa

  • Dify

    Hybrid — hosted or self-hosted

    Platform for building LLM applications: visual workflow editor, retrieval pipelines, agent tools, prompt management and an API layer. Available self-hosted or as a managed cloud service.

    Repository
    langgenius/dify
    Stars
    153,448
    Licence
    Other
    In stacks
    1

    Shares: Private company chatgpt · Document qa

  • Docling

    Libraries

    Document conversion toolkit that parses PDF, Office and image files into structured Markdown or JSON, preserving reading order, tables and figures for downstream retrieval.

    Stars
    65,519
    Licence
    MIT
    In stacks
    4

    Shares: Document qa

  • Flowise

    Hybrid — hosted or self-hosted

    Visual builder for LLM chains and agents, with retrieval nodes, tool calling and an API or embeddable chat widget for the finished flow.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Document qa

  • Keycloak

    Self-hosted projects

    Identity and access management server providing OpenID Connect and SAML single sign-on, user federation, groups and roles for self-hosted applications.

    Stars
    Licence
    unknown
    In stacks
    3

    Shares: Private company chatgpt

  • Kotaemon

    Self-hosted projects

    Document question-answering application with a configurable retrieval pipeline, inline citations that highlight the source passage, and support for local or hosted models.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Document qa

  • LangChain

    Frameworks

    Library for composing model calls, retrieval and tools into applications, with adapters for most providers and vector stores. A building block, not a deployable product.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Document qa

  • Langflow

    Hybrid — hosted or self-hosted

    Visual environment for composing LLM pipelines and agents from components, with export to a Python application or an API endpoint.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Document qa

  • Langfuse

    Hybrid — hosted or self-hosted

    Tracing, evaluation and prompt management for LLM applications, self-hostable so prompts and responses can stay inside the network being audited.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Private company chatgpt

  • LiteLLM

    Cloud platforms

    Gateway that presents one OpenAI-compatible API in front of many providers and local servers, with per-team keys, budgets, rate limits, fallbacks and request logging.

    Repository
    BerriAI/litellm
    Stars
    57,216
    Licence
    Other
    In stacks
    2

    Shares: Private llm · Private company chatgpt

  • llama.cpp

    Libraries

    C++ inference engine for quantised models on CPU, Apple Silicon and GPUs, with a bundled HTTP server. The engine underneath many desktop runtimes and the GGUF quantisation format.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Private llm

  • LlamaIndex

    Frameworks

    Data framework for retrieval applications: loaders for many document types, indexing and query pipelines, and evaluation helpers for retrieval quality.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Document qa

  • LM Studio

    Hybrid — hosted or self-hosted

    Desktop application for downloading and running open-weight models locally, with a chat interface and a local OpenAI-compatible server. Windows, macOS and Linux.

    Repository
    none linked
    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Private llm · Document qa

  • LocalAI

    Model servers

    Drop-in OpenAI-compatible API server that runs text, embedding, image and audio models locally across several back ends, including CPU-only deployments.

    Repository
    mudler/LocalAI
    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Private llm

  • Milvus

    Hybrid — hosted or self-hosted

    Distributed vector database designed for large collections, with several index types, GPU indexing options and a separated storage and compute architecture.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Document qa

  • Ollama

    Model servers

    Local model runtime with a one-command install, a model library and an OpenAI-compatible API. The usual starting point for running open-weight models on a workstation or small server.

    Repository
    ollama/ollama
    Stars
    179,403
    Licence
    MIT
    In stacks
    3

    Shares: Private llm · Private company chatgpt

  • Onyx

    Hybrid — hosted or self-hosted

    Open-source enterprise search and chat over company systems, with connectors to common SaaS tools, permission-aware indexing and a self-hosted deployment path.

    Stars
    31,754
    Licence
    Other
    In stacks
    0

    Shares: Document qa

  • Open WebUI

    Self-hosted projects

    Self-hosted chat interface for local and hosted models, with user accounts, groups, document upload and built-in retrieval. Runs in Docker against Ollama, vLLM or any OpenAI-compatible endpoint.

    Stars
    149,870
    Licence
    Other
    In stacks
    4

    Shares: Private company chatgpt · Document qa · Private llm

  • pgvector

    Libraries

    PostgreSQL extension adding vector types and HNSW or IVFFlat indexes, so embeddings live in the same database and the same backup as the rest of the application data.

    Stars
    22,735
    Licence
    Other
    In stacks
    5

    Shares: Document qa

  • PostgreSQL

    Self-hosted projects

    Relational database used here as the system of record for documents, metadata and job queues, and — with pgvector — for embeddings as well.

    Repository
    none linked
    Stars
    Licence
    unknown
    In stacks
    3

    Shares: Document qa

  • Qdrant

    Hybrid — hosted or self-hosted

    Vector search engine written in Rust with payload filtering, hybrid search, quantisation and snapshots. Runs as a single container or a cluster, or as a managed cloud service.

    Repository
    qdrant/qdrant
    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Document qa · Private llm

  • RAGFlow

    Self-hosted projects

    Retrieval engine built around deep document parsing: layout-aware chunking of PDFs, tables and scans, citation-backed answers, and a visual pipeline for building knowledge bases.

    Stars
    Licence
    unknown
    In stacks
    1

    Shares: Document qa

  • SGLang

    Model servers

    Serving framework for large models with prefix caching and structured-output support, aimed at high-throughput deployments and multi-GPU serving.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Private llm

  • Unstructured

    Hybrid — hosted or self-hosted

    Library and hosted API that partition documents of many formats into typed elements for indexing, with connectors to common storage systems and vector databases.

    Stars
    Licence
    unknown
    In stacks
    1

    Shares: Document qa

  • vLLM

    Model servers

    High-throughput GPU inference server using paged attention and continuous batching, with an OpenAI-compatible API. Built for many concurrent users rather than single-session use.

    Stars
    89,956
    Licence
    Apache-2.0
    In stacks
    7

    Shares: Private llm · Private company chatgpt

  • Weaviate

    Hybrid — hosted or self-hosted

    Vector database with a schema model, hybrid keyword and vector search, and optional built-in vectorisation modules. Self-hosted or managed, single-tenant or multi-tenant collections.

    Stars
    Licence
    unknown
    In stacks
    0

    Shares: Document qa


03Repository facts

RepositoryStarsHealth
Mintplex-Labs/anything-llmHealth not measuredunknown
goauthentik/authentikHealth not measuredunknown
chroma-core/chromaHealth not measuredunknown
langgenius/difyBuild Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without rebuilding the stack.153kMaintenance 100%, Releases 80%, Contributors 50%, Issues 94%, Community 100%88
docling-project/doclingGet your documents ready for gen AI66kMaintenance 100%, Releases 100%, Contributors 50%, Issues 85%, Community 96%90
FlowiseAI/FlowiseHealth not measuredunknown
keycloak/keycloakHealth not measuredunknown
Cinnamon/kotaemonHealth not measuredunknown
langchain-ai/langchainHealth not measuredunknown
langflow-ai/langflowHealth not measuredunknown
langfuse/langfuseHealth not measuredunknown
BerriAI/litellmThe fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]57kMaintenance 100%, Releases 100%, Contributors 50%, Issues 14%, Community 95%83
ggml-org/llama.cppHealth not measuredunknown
run-llama/llama_indexHealth not measuredunknown
mudler/LocalAIHealth not measuredunknown
milvus-io/milvusHealth not measuredunknown
ollama/ollamaGet up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.179kMaintenance 100%, Releases 100%, Contributors 50%, Issues 79%, Community 100%90
onyx-dot-app/onyxOpen Source AI Platform - AI Chat with advanced features that works with every LLM32kMaintenance 100%, Releases 100%, Contributors 50%, Issues 87%, Community 90%89
open-webui/open-webuiUser-friendly AI Interface (Supports Ollama, OpenAI API, ...)150kMaintenance 100%, Releases 80%, Contributors 50%, Issues 99%, Community 100%88
pgvector/pgvectorOpen-source vector similarity search for Postgres23kMaintenance 100%, Releases 30%, Contributors 50%, Issues 99%, Community 87%75
qdrant/qdrantHealth not measuredunknown
infiniflow/ragflowHealth not measuredunknown
sgl-project/sglangHealth not measuredunknown
Unstructured-IO/unstructuredHealth not measuredunknown
vllm-project/vllmA high-throughput and memory-efficient inference and serving engine for LLMs90kMaintenance 100%, Releases 80%, Contributors 50%, Issues 22%, Community 99%81
weaviate/weaviateHealth not measuredunknown

04Ask about your own case

Which alternative is right depends on how many people will use it, what the data is and what you can operate. Ask and the answer is researched for your organisation.

  1. 01What is a self-hosted alternative to LibreChat that we can run ourselves?