This is your work, valued
A C# developer
rk-llama.cpp. Llama.cpp with the Rockchip NPU integration as a GGML backend.
llama.cpp. LLM inference in C/C++