Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

It's literally the third paragraph:

> Deploying a Python inference stack means resolving a dependency tree at install time, on the target machine, against whatever CUDA and glibc that machine has. Deploying a ggml port means copying a shared library and a GGUF file.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: