modelship

modelship

Self-hosted, multi-model AI inference server. Run LLMs, TTS, STT, embeddings, and image generation with an OpenAI-compatible API.

37 4
Apache-2.0
last commit 2026-06-19
Source
Share:

About

Self-hosted, OpenAI-compatible inference for the agentic era: reasoning LLMs, universal tool calling, and the Responses API alongside embeddings, speech, and image models — many models sharing your GPUs, one gateway. Powered by Ray Serve.

Languages

Contributors2

No features listed.

Comments Theme
slug: modelship