Skip to content
#

rtx-3060

Here are 3 public repositories matching this topic...

Language: All
Filter by language

Knowledge distillation from GPT-5.5-xhigh into Qwen2.5-1.5B-Instruct: bf16 LoRA trained on a 6GB RTX 3060, served as a containerized OpenAI-compatible API (FastAPI + llama.cpp) with a streaming React chat UI. 570-prompt dataset across 10 categories, held-out perplexity evaluation, GGUF quantized deployment.

  • Updated Jul 28, 2026
  • Python

Improve this page

Add a description, image, and links to the rtx-3060 topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the rtx-3060 topic, visit your repo's landing page and select "manage topics."

Learn more