llama.cpp
0 questions
llama.cpp is the ggml C/C++ engine under Ollama and LM Studio, shipping convert_hf_to_gguf.py, llama-quantize and llama-server. Interviewers use it to test local inference below the wrapper.
questions
no questions here yet
this part of the tree is still being written
>
0 questions in this topic or below it