# llama.cpp mit Homebrew, dnf, Nix, pacman, apk, MacPorts, winget installieren

Prüfe Installationswege, Executables, Metadaten und Sicherheitshinweise für llama.cpp in AI-Agent-Workflows.

## Installation

```sh
sudo av install brew:llama.cpp
```

Weitere Installationsbefehle:

### macOS

- Homebrew (100%):

```sh
brew install llama.cpp
```

  Evidenz: local Homebrew formula metadata

- MacPorts (94%):

```sh
sudo port install llama.cpp
```

  Evidenz: MacPorts ports tree: llm/llama.cpp/Portfile from https://api.github.com/repos/macports/macports-ports/git/trees/master?recursive=1

### Linux

- dnf (92%):

```sh
sudo dnf install llama-cpp
```

  Evidenz: Fedora Rawhide package metadata: llama-cpp from https://dl.fedoraproject.org/pub/fedora/linux/development/rawhide/Everything/x86_64/os/repodata/07190dc5ae9f35ae73866675fed6d95fe6e8d9fe22c9d7cdf85862cb2ed24a4c-primary.xml.zst

- Nix (92%):

```sh
nix profile install nixpkgs#llama-cpp
```

  Evidenz: nixpkgs package indexes: pkgs/by-name/ll/llama-cpp/package.nix from https://api.github.com/repos/NixOS/nixpkgs/git/trees/master?recursive=1

- pacman (92%):

```sh
sudo pacman -S llama-cpp
```

  Evidenz: Arch Linux sync databases: llama-cpp from https://geo.mirror.pkgbuild.com/extra/os/x86_64/extra.db.tar.gz

- apk (92%):

```sh
sudo apk add llama-server
```

  Evidenz: Alpine Linux edge package indexes: llama-server from https://dl-cdn.alpinelinux.org/alpine/edge/community/x86_64/APKINDEX.tar.gz

### Windows

- winget (92%):

```sh
winget install --id ggml.llamacpp -e
```

  Evidenz: Windows Package Manager source index: ggml.llamacpp from https://cdn.winget.microsoft.com/cache/source.msix

## Paketfakten

- **Paketschlüssel:** brew:llama.cpp
- **Paketmanager:** Homebrew
- **Version:** 10210
- **Quellzusammenfassung:** LLM inference in C/C++
- **Homepage:** <https://llama.app>
- **Repository:** <https://github.com/ggml-org/llama.cpp>
- **Zuletzt aktualisiert:** 2026-07-31T19:04:02Z
- **Generiert:** 2026-08-03T19:37:03+00:00

## Executables

- llama (Alias)
- llama-batched (Alias)
- llama-batched-bench (Alias)
- llama-bench (Alias)
- llama-cli (Alias)
- llama-completion (Alias)
- llama-debug (Alias)
- llama-debug-template-parser (Alias)
- llama-diffusion-cli (Alias)
- llama-embedding (Alias)
- llama-eval-callback (Alias)
- llama-finetune (Alias)
- llama-fit-params (Alias)
- llama-gen-docs (Alias)
- llama-gguf (Alias)
- llama-gguf-hash (Alias)
- llama-gguf-split (Alias)
- llama-idle (Alias)
- llama-imatrix (Alias)
- llama-lookahead (Alias)
- llama-lookup (Alias)
- llama-lookup-create (Alias)
- llama-lookup-merge (Alias)
- llama-lookup-stats (Alias)
- llama-mtmd-cli (Alias)
- llama-parallel (Alias)
- llama-passkey (Alias)
- llama-perplexity (Alias)
- llama-quantize (Alias)
- llama-results (Alias)
- llama-retrieval (Alias)
- llama-server (Alias)
- llama-simple (Alias)
- llama-simple-chat (Alias)
- llama-speculative (Alias)
- llama-speculative-simple (Alias)
- llama-template-analysis (Alias)
- llama-tokenize (Alias)
- llama-tts (Alias)

## Installationsverhalten

- Bottle: nicht verfügbar

## Version und Aktualität

- Seite generiert: 2026-08-03
- Manager-Version: 10210
## Projektgeschichte und Nutzung

llama.cpp is one of the defining packages of the local-LLM era: a C/C++ inference stack that made it practical to run quantized transformer models on laptops, desktops, servers, and small devices without a heavyweight Python runtime.

### Projektgeschichte

The repository was created on GitHub on March 10, 2023, shortly after Meta's LLaMA model release changed the center of gravity for local language-model experimentation. The README states the project goal as LLM inference with minimal setup and strong performance across local and cloud hardware.

The project is closely tied to ggml. Its README describes llama.cpp as the main playground for developing new ggml features, and the implementation grew around plain C/C++, integer quantization, CPU backends, and hardware accelerators such as Metal, CUDA, Vulkan, SYCL, HIP, and related GPU paths.

As model support broadened beyond the original LLaMA family, llama.cpp became a runtime and tooling umbrella: converters, quantizers, benchmarking tools, embedding tools, an OpenAI-compatible server, multimodal support, and many model-family loaders are represented in the command set and documentation.

### Adoptionsgeschichte

Package adoption spread because llama.cpp lowered the cost of trying local inference: build from source, install from Homebrew, Nix, winget, conda-forge, Docker, or download release binaries, then run a model file or fetch one from Hugging Face-oriented workflows.

The README's bindings list shows the surrounding ecosystem that formed around the C/C++ core, including Python, Go, Node.js, Ruby, browser/Wasm, editor-completion plugins, and server clients. That ecosystem made llama.cpp both an end-user CLI and a library/runtime target for other packages.

Its high-frequency build-tag release pattern reflects active downstream pressure: package managers, bindings, model hubs, and local-AI applications all depend on fast propagation of backend, quantization, and model-format changes.

### Wie es verwendet wird

Users run llama-cli for local prompts, llama-server for an OpenAI-compatible HTTP API, llama-bench for performance testing, llama-quantize for smaller model files, and many auxiliary tools for embeddings, perplexity, retrieval, tokenization, and model-file manipulation.

The package is especially useful when a developer wants a self-contained inference engine: compile once, point it at a model, and choose a CPU/GPU backend without adopting a full ML framework stack.

### Warum Paket-Nerds sich dafür interessieren

For package maintainers, llama.cpp is unusually dynamic: hardware backend flags, model-format transitions, CLI renames, bundled tools, and release cadence all matter. It turned local AI into something package managers had to treat like a fast-moving systems tool rather than a single Python application.

It is also a packaging bridge between model hubs and Unix tooling. The same project can be installed as a formula, used as a server daemon, linked by language bindings, wrapped by desktop apps, or embedded in another inference product.

### Zeitleiste

- 2023: ggml-org/llama.cpp repository created on March 10.
- 2023: The project manifesto discussion documented goals and direction for the early community.
- 2024: llama.cpp issue trackers began maintaining separate API changelogs for libllama and llama-server.
- 2025: Multimodal support reached llama-server through the upstream pull request and documentation referenced from the README.
- 2026: Homebrew, Nix, winget, conda-forge, Docker, and release binaries were documented installation paths in the upstream README.

### Related projects

- ggml is the closest related project, because llama.cpp is the main development playground for that tensor library.
- Important downstream and adjacent projects include llama-cpp-python, node-llama-cpp, go-llama.cpp, wllama, llama.vscode, llama.vim, and applications that wrap llama-server's OpenAI-compatible API.

### Quellen

- <https://api.github.com/repos/ggml-org/llama.cpp>
- <https://formulae.brew.sh/formula/llama.cpp>
- <https://github.com/ggml-org/ggml>
- <https://github.com/ggml-org/llama.cpp>
- <https://github.com/ggml-org/llama.cpp/discussions/205>
- <https://github.com/ggml-org/llama.cpp/issues/9289>
- <https://github.com/ggml-org/llama.cpp/issues/9291>
- <https://github.com/ggml-org/llama.cpp/pull/12898>
- <https://github.com/ggml-org/llama.cpp/releases>
- <https://raw.githubusercontent.com/ggml-org/llama.cpp/master/README.md>


## Sicherheitshinweise

Für llama.cpp wurde kein passendes lokales Secret-Handling-Manifest gefunden. Nucleus-Paketmetadaten bleiben hier veröffentlicht, damit künftige Abdeckung eine stabile Paket-URL hat.


## Andere Paketmanager-Einträge

- Nix - llama-cpp: normalized package name match | nixpkgs package indexes: pkgs/by-name/ll/llama-cpp/package.nix from https://api.github.com/repos/NixOS/nixpkgs/git/trees/master?recursive=1
- apk - llama-server - 0.0.9672-r0: normalized package name match | Alpine Linux edge package indexes: llama-server from https://dl-cdn.alpinelinux.org/alpine/edge/community/x86_64/APKINDEX.tar.gz | llama.cpp server | https://github.com/ggml-org/llama.cpp
- apk - llama-server-openrc - 0.0.9672-r0: normalized package name match | Alpine Linux edge package indexes: llama-server-openrc from https://dl-cdn.alpinelinux.org/alpine/edge/community/x86_64/APKINDEX.tar.gz | llama.cpp server (OpenRC init scripts) | https://github.com/ggml-org/llama.cpp
- apk - llama.cpp - 0.0.9672-r0: normalized package name match | Alpine Linux edge package indexes: llama.cpp from https://dl-cdn.alpinelinux.org/alpine/edge/community/x86_64/APKINDEX.tar.gz | LLM inference in C/C++ (with Vulkan GPU acceleration) | https://github.com/ggml-org/llama.cpp
- apk - llama.cpp-cpu - 0.0.9672-r0: normalized package name match | Alpine Linux edge package indexes: llama.cpp-cpu from https://dl-cdn.alpinelinux.org/alpine/edge/community/x86_64/APKINDEX.tar.gz | LLM inference in C/C++ (with Vulkan GPU acceleration) | https://github.com/ggml-org/llama.cpp
- apk - llama.cpp-dev - 0.0.9672-r0: normalized package name match | Alpine Linux edge package indexes: llama.cpp-dev from https://dl-cdn.alpinelinux.org/alpine/edge/community/x86_64/APKINDEX.tar.gz | LLM inference in C/C++ (with Vulkan GPU acceleration) (development files) | https://github.com/ggml-org/llama.cpp
- apk - llama.cpp-extras - 0.0.9672-r0: normalized package name match | Alpine Linux edge package indexes: llama.cpp-extras from https://dl-cdn.alpinelinux.org/alpine/edge/community/x86_64/APKINDEX.tar.gz | llama.cpp additional binaries | https://github.com/ggml-org/llama.cpp
- apk - llama.cpp-libs - 0.0.9672-r0: normalized package name match | Alpine Linux edge package indexes: llama.cpp-libs from https://dl-cdn.alpinelinux.org/alpine/edge/community/x86_64/APKINDEX.tar.gz | LLM inference in C/C++ (with Vulkan GPU acceleration) (shared libraries) | https://github.com/ggml-org/llama.cpp
- apk - llama.cpp-vulkan - 0.0.9672-r0: normalized package name match | Alpine Linux edge package indexes: llama.cpp-vulkan from https://dl-cdn.alpinelinux.org/alpine/edge/community/x86_64/APKINDEX.tar.gz | LLM inference in C/C++ (with Vulkan GPU acceleration) | https://github.com/ggml-org/llama.cpp
- dnf - llama-cpp - b9840-2.fc45: normalized package name match | Fedora Rawhide package metadata: llama-cpp from https://dl.fedoraproject.org/pub/fedora/linux/development/rawhide/Everything/x86_64/os/repodata/07190dc5ae9f35ae73866675fed6d95fe6e8d9fe22c9d7cdf85862cb2ed24a4c-primary.xml.zst | Port of Facebook's LLaMA model in C/C++ | https://github.com/ggerganov/llama.cpp
- dnf - llama-cpp-devel - b9840-2.fc45: normalized package name match | Fedora Rawhide package metadata: llama-cpp-devel from https://dl.fedoraproject.org/pub/fedora/linux/development/rawhide/Everything/x86_64/os/repodata/07190dc5ae9f35ae73866675fed6d95fe6e8d9fe22c9d7cdf85862cb2ed24a4c-primary.xml.zst | Port of Facebook's LLaMA model in C/C++ | https://github.com/ggerganov/llama.cpp
- pacman - llama-cpp - b10221-1: normalized package name match | Arch Linux sync databases: llama-cpp from https://geo.mirror.pkgbuild.com/extra/os/x86_64/extra.db.tar.gz | LLM inference in C/C++ | https://github.com/ggerganov/llama.cpp
- MacPorts - llama.cpp: normalized package name match | MacPorts ports tree: llm/llama.cpp/Portfile from https://api.github.com/repos/macports/macports-ports/git/trees/master?recursive=1
- winget - ggml.llamacpp: normalized package name match | Windows Package Manager source index: ggml.llamacpp from https://cdn.winget.microsoft.com/cache/source.msix


## Combined YAML source

View the package source record on GitHub. [combined/llama.cpp.yml](https://github.com/mxcl/pkgdb/blob/main/combined/llama.cpp.yml)


## Quellen

- pkg.so package database
- Geiger risk classifier
- curated package history
- pkgdb category and tag curation
- external package-manager database matches
- cross-ecosystem install command graph
