IntermediateVocabulary#llama.cpp#GGUF#quantization#local LLM

llama.cpp Inference Exercises

llama.cpp enables efficient local inference of large language models on consumer hardware. These exercises cover the GGUF file format, quantization levels and tradeoffs, GPU layer offloading with -ngl, context size configuration, and platform-specific backends including Metal for Apple Silicon.

0 / 5 completed
1 / 5
What is the GGUF file format used by llama.cpp?

Frequently Asked Questions

What does the "llama.cpp Inference Exercises" vocabulary exercise cover?

This exercise tests real IT vocabulary related to llama.cpp inference exercises through 5 multiple-choice questions, each built from realistic workplace sentences rather than abstract definitions.

Is this vocabulary exercise free to use?

Yes. Every exercise on CoderSlingo, including this one, is completely free — no account, sign-up, or payment required.

How many questions does this exercise have?

This exercise has 5 questions. Each one shows a real-world sentence or scenario with multiple-choice options and an explanation once you answer.