explainx.ai0k
TrendingAI News TodayPathwaysSkills
Pricing
explainx.ai

Upskill in AI — 16 free pathways, live workshops & bootcamps, and 50+ courses from practitioners. Plus the skills, tools, and MCP servers to practice on.

follow us

follow on google

Add explainx.ai as a preferred source

corporate training

support@explainx.ai

get started

Find your pathTake Free Evaluation

community

Join the community

learn

mind: share how you thinkpathways — start freeworkshopsbootcampscoursescompare Explainxcertificationsmock testsexplainx universitycorporate traininglearn skills & mcp

discover

skillsmcp serversexplainx mcptoolsmdx readeragentsllmsdesignsdictionarypeopleagi trackerfelony benchranks

company

aboutvisionmissionteaminstructorsteach on explainxpartnershipscommunityhackathonscareers

content

daily AI newsstate of AI — live resultsblogreleasespromptsgeneratorsresource libraryfor LLMsexplainx.ai kids

solutions

all solutionsdeveloper upskillingmarketing upskillingproduct manager upskillingleadership upskilling

newsletter · weekly

Get AI news, tools, and insights in your inbox.

supportcontactprivacytermsdata rightshow we create contentsubmission guidelines

© 2026 AISOLO Technologies Pvt Ltd

explainx.ai

  1. Home
  2. /
  3. Dictionary
  4. /
  5. llama.cpp
Inference & Deploymentaka llamacpp

llama.cpp

A C/C++ inference engine for running LLMs on CPUs and consumer GPUs with extensive quantization support.

Ask Melo about this← all terms

A C/C++ inference engine for running LLMs on CPUs and consumer GPUs, supporting extensive quantization formats and enabling local inference without Python or CUDA dependencies.

Related terms

On-Device InferenceOllamaModel Quantization (Inference)Inference EngineMLXModel Weights

Where llama.cpp comes up

  • Hugging Face Transformers Now Matches llama.cpp on GGUF Performance
  • Tencent Hy3 GGUF — 1-Bit and 4-Bit Quants for Single-GPU llama.cpp
  • What Is llama.cpp? Install, Run GGUF Models, and Serve OpenAI-Compatible APIs
  • Qwen 3.6 27B Local Dev Guide: llama.cpp, OpenCode, and Why Dense Beats MoE