skip to content

llama.cpp

0 questions

llama.cpp is the ggml C/C++ engine under Ollama and LM Studio, shipping convert_hf_to_gguf.py, llama-quantize and llama-server. Interviewers use it to test local inference below the wrapper.

questions

no questions here yet

this part of the tree is still being written

>
0 questions in this topic or below it

explore