Skip to content
View rhgo1749's full-sized avatar
🤔
🤔

Block or report rhgo1749

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. qwen3.8-flash-next-exllamav3-3x5070ti-recipe qwen3.8-flash-next-exllamav3-3x5070ti-recipe Public

    Reproducible ExLlamaV3 recipe and optimization log for Qwen3.8-Flash-Next on 3x RTX 5070 Ti 16GB

    Python

  2. surface-lenovo-active3pen-mapping surface-lenovo-active3pen-mapping Public

    surface-lenovo-active3pen-mapping

    C++

  3. qwen3.8-flash-next-strata-gpu-per-lane-recipe qwen3.8-flash-next-strata-gpu-per-lane-recipe Public

    Measured 3x RTX 5070 Ti recipe for Qwen3.8-Flash-Next on the Strata multi-GPU fork: 3x262K lanes, shared expert arena, and real Hermes concurrency results.

    TeX

  4. Niko1221/Strata Niko1221/Strata Public

    Qwen3.8-Flash-Next on any consumer hardware: one-click install for Windows / Linux. Strata inference engine, OpenAI/Anthropic API on localhost, optional image input.

    C++ 2.9k 289

  5. Strata-Lanes Strata-Lanes Public

    Forked from Niko1221/Strata

    Qwen3.8-Flash-Next (125B MoE) on a 12-24 GB NVIDIA GPU + 64 GB RAM: one-click install for Windows / Linux. Strata inference engine, OpenAI/Anthropic API on localhost, optional image input.

    C++