Pinned Loading
-
qwen3.8-flash-next-exllamav3-3x5070ti-recipe
qwen3.8-flash-next-exllamav3-3x5070ti-recipe PublicReproducible ExLlamaV3 recipe and optimization log for Qwen3.8-Flash-Next on 3x RTX 5070 Ti 16GB
Python
-
surface-lenovo-active3pen-mapping
surface-lenovo-active3pen-mapping Publicsurface-lenovo-active3pen-mapping
C++
-
qwen3.8-flash-next-strata-gpu-per-lane-recipe
qwen3.8-flash-next-strata-gpu-per-lane-recipe PublicMeasured 3x RTX 5070 Ti recipe for Qwen3.8-Flash-Next on the Strata multi-GPU fork: 3x262K lanes, shared expert arena, and real Hermes concurrency results.
TeX
-
Niko1221/Strata
Niko1221/Strata PublicQwen3.8-Flash-Next on any consumer hardware: one-click install for Windows / Linux. Strata inference engine, OpenAI/Anthropic API on localhost, optional image input.
-
Strata-Lanes
Strata-Lanes PublicForked from Niko1221/Strata
Qwen3.8-Flash-Next (125B MoE) on a 12-24 GB NVIDIA GPU + 64 GB RAM: one-click install for Windows / Linux. Strata inference engine, OpenAI/Anthropic API on localhost, optional image input.
C++
If the problem persists, check the GitHub status page or contact support.


