sci-misc/llama-cpp (stuff)

Search

Install

Install this package:

emerge -a sci-misc/llama-cpp

Package Information

Description:
Port of Facebook's LLaMA model in C/C++
Homepage:
https://github.com/ggml-org/llama.cpp
License:
MIT

Versions

Version EAPI Keywords Slot
9999 8 ~amd64 0
0_pre10818 8 ~amd64 ~arm64 0
0_pre10784 8 ~amd64 ~arm64 0
0.4.0 8 ~amd64 ~arm64 0
0.3.0 8 ~amd64 ~arm64 0
0.2.0 8 ~amd64 ~arm64 0
0.1.2 8 ~amd64 ~arm64 0

Metadata

Maintainers

Upstream

Raw Metadata XML
<pkgmetadata>
	<maintainer type="person">
		<email>iohann.s.titov@gmail.com</email>
		<name>Ivan S. Titov</name>
	</maintainer>
	<use>
		<flag name="blis">Build a BLIS backend</flag>
		<flag name="flexiblas">Build a FlexiBLAS backend</flag>
		<flag name="rocm">Build a HIP (ROCm) backend</flag>
		<flag name="openblas">Build an OpenBLAS backend</flag>
		<flag name="opencl">Build an OpenCL backend, so far only works on Adreno and Intel GPUs</flag>
		<flag name="openssl">Use openssl to support HTTPS</flag>
		<flag name="sycl">Build an Intel SYCL backend (Arc GPU, Intel CPU via
    oneAPI). Requires a -fsycl-capable compiler (Intel icpx or clang++
    with SYCL patches) installed separately.</flag>
		<flag name="webui">Build the embedded llama-server web UI. Fetches
    prebuilt assets from the upstream Hugging Face bucket at configure
    time; disable for a server binary with only the HTTP API.</flag>
	</use>
	<upstream>
		<remote-id type="github">ggml-org/llama.cpp</remote-id>
	</upstream>
</pkgmetadata>

Lint Warnings

USE Flags

Manage flags for this package: euse -i <flag> -p sci-misc/llama-cpp | euse -E <flag> -p sci-misc/llama-cpp | euse -D <flag> -p sci-misc/llama-cpp

Flag Description 9999 0_pre10818 0_pre10784 0.4.0 0.3.0 0.2.0 0.1.2
( ⚠️
) ⚠️
avx ⚠️
avx2 ⚠️
avx512f ⚠️
avx512vbmi ⚠️
blis Build a BLIS backend
bmi2 ⚠️
cuda Build the CUDA (cuda_v13) llama-server GPU backend via <pkg>dev-util/nvidia-cuda-toolkit</pkg> ⚠️
examples Build and install the example programs ⚠️
f16c ⚠️
flexiblas Build a FlexiBLAS backend
fma3 ⚠️
openblas Build an OpenBLAS backend
opencl Build an OpenCL backend, so far only works on Adreno and Intel GPUs
openmp Build the host backend with OpenMP threading ⚠️
openssl Use openssl to support HTTPS
rocm Build a HIP (ROCm) backend
sse4_2 ⚠️
sycl Build an Intel SYCL backend (Arc GPU, Intel CPU via oneAPI). Requires a -fsycl-capable compiler (Intel icpx or clang++ with SYCL patches) installed separately.
vulkan Build the Vulkan llama-server GPU backend ⚠️
webui Build the embedded llama-server web UI. Fetches prebuilt assets from the upstream Hugging Face bucket at configure time; disable for a server binary with only the HTTP API.

Manifest

Type File Size Versions
DIST ggml-org_models_tinyllamas_stories15M-q4_0-99dd1a73db5a37100bd4ae633f4cfce6560e1567.gguf 19077344 bytes 9999, 0_pre10818, 0_pre10784, 0.4.0, 0.3.0, 0.2.0, 0.1.2
DIST llama-cpp-0.1.2.tar.gz 36847636 bytes 0.1.2
DIST llama-cpp-0.2.0.tar.gz 36879187 bytes 0.2.0
DIST llama-cpp-0.3.0.tar.gz 36951021 bytes 0.3.0
DIST llama-cpp-0.4.0.tar.gz 37288249 bytes 0.4.0
DIST llama-cpp-0_pre10784.tar.gz 37145388 bytes 0_pre10784
DIST llama-cpp-0_pre10818.tar.gz 37316806 bytes 0_pre10818
Unmatched Entries
Type File Size