Skip to content
#

ternary-quantization

Here are 22 public repositories matching this topic...

Deploy Bonsai-2-27B (ternary 27B, 5.95GB) on a single 8GB consumer GPU via an AI agent. 用 AI 部署 Bonsai-2-27B:一条提示词让 AI 装驱动、编 CUDA、下模型并接入 Claude Code。实测 51 tok/s @4k / 39 tok/s @100k,含按机器调参指南与完整测评报告。

  • Updated Oct 6, 2026
  • Shell

Native low-bit (NLT) diffusion models — the weights live in the quantized space from step 0 instead of being compressed after fp16 pre-training. Home of the AquariusImage series: Aquarius Terimage (ternary text-to-image, released) and Aquarius Binimage (planned).

  • Updated Oct 4, 2026
  • Python

Add this topic to your repo

To associate your repository with the ternary-quantization topic, visit your repo's landing page and select "manage topics."

Learn more