Pinned Loading
-
llama-cpp-pascal-cuda-windows
llama-cpp-pascal-cuda-windows PublicCUDA compile guides for llama-cpp-python on legacy NVIDIA Pascal GPUs (GTX 1050/1060/1070/1080) — Windows, VS2019/CUDA 12.1 and VS2022/CUDA 12.6 paths
-
local-whisper-transcription-pipeline
local-whisper-transcription-pipeline PublicLocal video/audio to subtitles pipeline. Zero cloud, VRAM-aware model selection, runs locally on any NVIDIA GPU. Built on faster-whisper and ffmpeg. Designed for readable orchestration logic, not d…
Python
-
sdxl-lora-gguf-blueprint
sdxl-lora-gguf-blueprint PublicBlueprint for training custom SDXL LoRAs and running local GGUF inference on 8GB VRAM — no cloud, no ComfyUI
-
sovereign-swat-blueprint
sovereign-swat-blueprint PublicDescription: Solo-built local multi-agent LLM system — 3 pipelines with evidence-gated reasoning, memory governance, and zero cloud dependencies
-
SovereignKernel
SovereignKernel PublicDescription: Custom C++ LLM inference runtime built from scratch — tokenizer, attention, RoPE, KV cache, GGUF/Q4_0 loading, no framework dependencies. CUDA backend in progress.
C++
-
turbovec-quantization-benchmark
turbovec-quantization-benchmark PublicDoes TurboVec quantization hold up on exact-fact retrieval, or only on semantic similarity? Tests recall@k across bit_width 4/2 vs full precision on a synthetic report with 104 ground-truth facts. …
Python
If the problem persists, check the GitHub status page or contact support.