You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
HIP/ROCm fork optimized for AMD RDNA2 (gfx1030) with PrismML Q1_0_G128 1-bit quant support, RotorQuant, TurboQuant, EAGLE3 and P-EAGLE speculative decoding, and full Wave32 kernel optimizations.
Precompiled PrismML/llama.cpp backend for LM Studio on Windows. Run PrismML Bonsai 1-bit/ternary GGUF models natively in LM Studio with NVIDIA CUDA. No compiling required.
Local chat app for running Bonsai GGUF models with streaming responses, built-in Prism runtime management, model switching, runtime diagnostics, and saved conversation history.
GGUFly — Interactive TUI launcher and model manager for GGUF models with llama.cpp, PrismML/Bonsai, and Mirai S runtimes. Built on Omarchy / Arch Linux.