Skip to content
#

sm120a

Here are 4 public repositories matching this topic...

Language: All
Filter by language

memra is a Rust + CUDA inference engine built for NVIDIA RTX PRO 6000 Blackwell and RTX 5090. It serves GGUF models over an OpenAI-compatible API, and every default it ships was measured on those two cards — both Blackwell sm_120a, with a separately compile-gated Hopper/H100 (sm_90a) lane alongside.

  • Updated Aug 15, 2026
  • Rust

Improve this page

Add a description, image, and links to the sm120a topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the sm120a topic, visit your repo's landing page and select "manage topics."

Learn more