
| ÍNDICE | GitHub - carloslfu/slotstream: Run Qwen3.8-Flash-Next (125B MoE, 104 GB at 4-bit) on Macs https://github.com/carloslfu/slotstream carloslfu /slotstreamPublic Go to file Code 90 Commits
About Run Qwen3.8-Flash-Next (125B MoE, 104 GB at 4-bit) on Macs with a fraction of that RAM by streaming experts from SSD. MLX + Swift, Ollama-compatible API. apple-siliconllmllm-inferencelocal-llmmacosmixture-of-expertsmlxollamaqwenswift 235 stars 1 watching Releases Packages Contributors Languages You can’t perform that action at this time. |