You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
HIP/ROCm fork optimized for AMD RDNA2 (gfx1030) with PrismML Q1_0_G128 1-bit quant support, RotorQuant, TurboQuant, EAGLE3 and P-EAGLE speculative decoding, and full Wave32 kernel optimizations.
Precompiled PrismML/llama.cpp backend for LM Studio on Windows. Run PrismML Bonsai 1-bit/ternary GGUF models natively in LM Studio with NVIDIA CUDA. No compiling required.
Windows x64 installer for Bonsai 2 / PrismML GGUF models in LM Studio. Isolates PrismML server binaries, preserves native libraries, verifies downloads, and backs up runtime manifests.
GGUFly — Interactive TUI launcher and model manager for GGUF models with llama.cpp, PrismML/Bonsai, and Mirai S runtimes. Built on Omarchy / Arch Linux.
Local chat app for running Bonsai GGUF models with streaming responses, built-in Prism runtime management, model switching, runtime diagnostics, and saved conversation history.