
xyntetik-runner
Single-binary GGUF model runtime in C — CPU/CUDA/Metal, OpenAI-compatible. Serves, scores, and trains LoRA directly through the quantized weights it deploys, with byte-reproducible adapters. Tool calls survive the token limit; sparse MoE and schema-constrained decoding included.
About
Languages
Contributors2
No features listed.
Comments Theme
Platforms
Hosting
Self-hosted
Install
Docker
Build from source




