Top repos they contribute to (2)
Ranked by their commits
MTPLXPython2.24x decode TPS increase On Qwen 3.6 27B @ temp 0.6 | Native MTP Speculative Decoding On Apple Silicon With No External Drafter.
1.1k2commits
club-3090PythonCommunity recipes for serving LLMs on RTX 3090/4090/5090 CUDA gpus. Multi-engine (vLLM, llama.cpp, ik_llama) and model-agnostic. Currently shipping Qwen3.6-27B Qwen3.6 35B Gemma 4 26B Gemma 4 31B configs for 1× and 2× cards.
1.8k1commits
Related contributors
People who work on the same repos- noonghunna1shared repo1.4kcommitsclub-3090
- youssofal1shared repo271commitsMTPLX
- github-actions[bot]1shared repo28commitsclub-3090
- dependabot[bot]1shared repo13commitsMTPLX
- daniel-farina1shared repo9commitsMTPLX
- mmmugh1shared repo8commitsMTPLX
- steamEngineer1shared repo7commitsclub-3090
- PhilipJohnBasile1shared repo4commitsMTPLX
- easel1shared repo4commitsclub-3090
- claude1shared repo3commitsMTPLX
- danbedford1shared repo3commitsclub-3090
- SuperMarioYL1shared repo2commitsMTPLX




