Skip to content
#

rtx-3060

Here are 11 public repositories matching this topic...

Unofficial FreeToken fork: on one RTX 3060 12 GB, a 35B MoE at 250k of context or gpt-oss-120b; Flash-Next 125B on two. Half the RAM, image input. Runs on Turing: RTX 2060, RTX 20 series, sm_75.

  • Updated Sep 10, 2026
  • Python

Knowledge distillation from GPT-5.5-xhigh into Qwen2.5-1.5B-Instruct: bf16 LoRA trained on a 6GB RTX 3060, served as a containerized OpenAI-compatible API (FastAPI + llama.cpp) with a streaming React chat UI. 570-prompt dataset across 10 categories, held-out perplexity evaluation, GGUF quantized deployment.

  • Updated Aug 10, 2026
  • Python

Add this topic to your repo

To associate your repository with the rtx-3060 topic, visit your repo's landing page and select "manage topics."

Learn more