3 comments

[ 1.6 ms ] story [ 15.2 ms ] thread
and use the original llama.cpp directly. Its infinitely more easy to setup and use now