Technology · LLMs → Serving · Open Source · wiki:stub

llama.cpp

This pack is still a wiki stub. Opening text below if present; full article bands appear after deepen (`Wiki-status: deep`).

llama.cpp runs LLMs efficiently on consumer and edge hardware via quantized C/C++ inference. Underpins many local model deployments. Enables private/on-prem workers without GPU clouds — trade throughput for data locality. LLM inference in C/C++ ggml / ops / maintainer PRs%20sort%3Aupdated-desc) / dev stats / lib llama API / llama-server REST API - Visit https://llama.app and follow the instructions - Run with Docker - see our Docker documentation - Download pre-built binaries from the releases page - Build from source by cloning this repository.

Research inventory

9 tags · 10 out · 11 in · 3 artifacts · 1 gaps · 14 corpus docs

Catalog tags

landscape.layer
LLMs
landscape.subcategory
Serving
license_tag
Open Source
maps.dm
absent
maps.features
5
one_liner
LLMs
review.depth
science
slug
llama-cpp
title
llama.cpp

Artifacts

  • dm_map · absent
  • features_map · present · technologies/llama-cpp/features.md
  • readme · present · technologies/llama-cpp/README.md

Out · alternative_to

Out · maps_to

In · alternative_to

In · in_stack

In · maps_to

Corpus tags

category
LLMs → Serving
dedication
open-source
feature
bring-your-own-model
cost-latency-ops
deployment-data-residency
model-flexibility-routing
self-host
kind
map_edge
tech_features
tech_quote
tech_readme
needs_deepen
false
quality
ok
slug
llama-cpp
source_id
ggml-note
llama-cpp-github
vllm-contrast
technology
llama-cpp

Corpus documents (14)

map_edge · 5

  • llama-cpp → bring-your-own-model
  • llama-cpp → cost-latency-ops
  • llama-cpp → deployment-data-residency
  • llama-cpp → model-flexibility-routing
  • llama-cpp → self-host

tech_features · 1

  • llama.cpp · features

tech_quote · 7

  • llama.cpp · ggml-note
  • llama.cpp · ggml-note
  • llama.cpp · ggml-note
  • llama.cpp · llama-cpp-github
  • llama.cpp · llama-cpp-github
  • llama.cpp · vllm-contrast
  • llama.cpp · vllm-contrast

tech_readme · 1

  • llama.cpp