Model-to-NPU pipelines for Qualcomm Snapdragon: QNN/ONNX/Android runtimes for on-device image and video generation.
-
Updated
Jul 23, 2026 - Python
Model-to-NPU pipelines for Qualcomm Snapdragon: QNN/ONNX/Android runtimes for on-device image and video generation.
llama.cpp fork with an experimental QNN backend for the Hexagon NPU on Windows on Snapdragon
Snapdragon X Elite (Hexagon NPU) LLM: Genie server exposing Anthropic + OpenAI APIs for on-device inference
Claude Code skills for Qualcomm QAIRT/QNN model deployment — ONNX to Hexagon HTP, AIMET quantization, eSDK cross-compilation
Keyboard-driven .NET terminal UI to find, run, chat with, and serve GGUF/LLM models on the Qualcomm Snapdragon NPU via GenieX — with auto CPU/NPU/GPU compute selection and one-key offload to external/NAS drives.
To associate your repository with the qairt topic, visit your repo's landing page and select "manage topics."