xyntetik-runner

xyntetik-runner

Single-binary GGUF model runtime in C — CPU/CUDA/Metal, OpenAI-compatible. Serves, scores, and trains LoRA directly through the quantized weights it deploys, with byte-reproducible adapters. Tool calls survive the token limit; sparse MoE and schema-constrained decoding included.

16 0
Apache-2.0
last commit 2026-08-25
Website Source
Share:

About

Local inference you can prove. A single-binary GGUF engine in C: serves, scores, and trains LoRA through the quantized weights it deploys, byte-reproducibly. Tool calls survive the token limit; every run replays as a signed receipt. CPU/CUDA/Metal, OpenAI- and Anthropic-compatible.

xyntetik-runner website preview

Languages

Contributors2

No features listed.

Comments Theme
slug: xyntetik-runner