Pinned Loading
-
llama-moe-cache
llama-moe-cache PublicExpert cache + predictive prefetch for MoE inference in llama.cpp. A 12GB GPU can run a 120GB model at native speed. At 3% sparsity, a single workstation could theoretically run a 1T parameter mode…
-
VU-Campusnet
VU-Campusnet PublicA simple script to connect VU-Campusnet on linux if the official script is frustrating for you too
Shell 1
-
-
terminal-search
terminal-search Publica web search cli tool assisted with ai for quick research in between commands
Shell 2
-
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.


