Skip to content
#

b200

Here are 15 public repositories matching this topic...

memra

Rust + CUDA LLM inference engine for Blackwell (Tuned specifically on RTX PRO 6000, RTX 5090, B200): OpenAI-compatible (+converse and ant) serving, per-model X hardware exactness gates. NVFP4/mixed (fp8 hybrid, 4o6, etc - correctness, performance, hardware specific adapted) main quant support.

  • Updated Sep 6, 2026
  • OpenEdge ABL

Spheron — independent third-party profile of a public API surface, by API Evangelist. Spheron Network is a decentralized GPU and cloud compute marketplace that aggregates enterprise-grade NVIDIA GPU capacity from certified Tier 3/4 data centers worldwide and exposes it through a single on-demand, per-minute billed interface.

  • Updated Sep 4, 2026

Add this topic to your repo

To associate your repository with the b200 topic, visit your repo's landing page and select "manage topics."

Learn more