sci-ml/llama-cpp (gentoo-zh)

Search

Install

Install this package:

emerge -a sci-ml/llama-cpp

Package Information

Description:
LLM inference in C/C++
Homepage:
https://github.com/ggml-org/llama.cpp
License:
MIT

Versions

Version EAPI Keywords Slot
9999 8 ~amd64 ~riscv 0
0.6.0 8 ~amd64 ~riscv 0

Metadata

Maintainers

Upstream

Raw Metadata XML
<pkgmetadata>
	<maintainer type="person">
		<email>me@puqns67.icu</email>
		<name>Puqns67</name>
	</maintainer>
	<use>
		<flag name="blis">Build a BLIS backend</flag>
		<flag name="flexiblas">Build a FlexiBLAS backend</flag>
		<flag name="openblas">Build an OpenBLAS backend</flag>
		<flag name="rocm">Build a HIP (ROCm) backend</flag>
		<flag name="wmma">Use rocWMMA to enhance flash attention performance</flag>
		<flag name="opencl">Build an OpenCL backend, so far only works on Adreno and Intel GPUs</flag>
		<flag name="rpc">Build with rpc-server</flag>
		<flag name="server">Build with example server</flag>
		<flag name="webui">Build server with embedded WebUI</flag>
	</use>
	<upstream>
		<remote-id type="github">ggml-org/llama.cpp</remote-id>
	</upstream>
</pkgmetadata>

Lint Warnings

USE Flags

Manage flags for this package: euse -i <flag> -p sci-ml/llama-cpp | euse -E <flag> -p sci-ml/llama-cpp | euse -D <flag> -p sci-ml/llama-cpp

Flag Description 9999 0.6.0
"( ⚠️ ✓ ✓
( ⚠️ ✓ ✓
) ⚠️ ✓ ✓
)" ⚠️ ✓ ✓
amx_bf16 ⚠️ ✓ ✓
amx_int8 ⚠️ ✓ ✓
amx_tile ⚠️ ✓ ✓
avx ⚠️ ✓ ✓
avx2 ⚠️ ✓ ✓
avx512_bf16 ⚠️ ✓ ✓
avx512_vnni ⚠️ ✓ ✓
avx512bw ⚠️ ✓ ✓
avx512cd ⚠️ ✓ ✓
avx512dq ⚠️ ✓ ✓
avx512f ⚠️ ✓ ✓
avx512vbmi ⚠️ ✓ ✓
avx512vl ⚠️ ✓ ✓
avx_vnni ⚠️ ✓ ✓
blis Build a BLIS backend ✓ ✓
bmi2 ⚠️ ✓ ✓
cuda ⚠️ ✓ ✓
examples ⚠️ ✓ ✓
f16c ⚠️ ✓ ✓
flexiblas Build a FlexiBLAS backend ✓ ✓
fma3 ⚠️ ✓ ✓
openblas Build an OpenBLAS backend ✓ ✓
opencl Build an OpenCL backend, so far only works on Adreno and Intel GPUs ✓ ✓
openmp ⚠️ ⊕ ⊕
rocm Build a HIP (ROCm) backend ✓ ✓
rpc Build with rpc-server ✓ ✓
server Build with example server ⊕ ⊕
sse4_2 ⚠️ ✓ ✓
v ⚠️ ✓ ✓
vulkan ⚠️ ✓ ✓
webui Build server with embedded WebUI ✓ ✓
wmma Use rocWMMA to enhance flash attention performance ✓ ✓
xtheadvector ⚠️ ✓ ✓
zba ⚠️ ✓ ✓
zfh ⚠️ ✓ ✓
zicbop ⚠️ ✓ ✓
zihintpause ⚠️ ✓ ✓
zvfh ⚠️ ✓ ✓

Manifest

Type File Size Versions
DIST ggml-org_models_tinyllamas_stories15M-q4_0-99dd1a73db5a37100bd4ae633f4cfce6560e1567.gguf 19077344 bytes 9999, 0.6.0
DIST llama-cpp-0.6.0-ui.tar.gz 3099802 bytes 0.6.0
DIST llama-cpp-0.6.0.tar.gz 37874349 bytes 0.6.0
Unmatched Entries
Type File Size