Install this package:
emerge -a dev-python/tokenspeed-mla-bin
| Version | EAPI | Keywords | Slot |
|---|---|---|---|
| 0.2.16 | 8 | ~amd64 | 0 |
| 0.2.15 | 8 | ~amd64 | 0 |
| 0.2.14 | 8 | ~amd64 | 0 |
| 0.1.8 | 8 | ~amd64 | 0 |
<pkgmetadata> <maintainer type="person"> <email>iohann.s.titov@gmail.com</email> <name>Ivan S. Titov</name> </maintainer> <longdescription>TokenSpeed MLA provides precompiled and just-in-time CUDA kernels for multi-head latent attention on NVIDIA Blackwell GPUs.</longdescription> <upstream> <remote-id type="pypi">tokenspeed-mla</remote-id> </upstream> </pkgmetadata>
| Type | File | Size | Versions |
|---|---|---|---|
| DIST | tokenspeed_mla-0.1.8-py3-none-manylinux_2_28_x86_64.whl | 755827 bytes | 0.1.8 |
| DIST | tokenspeed_mla-0.2.14-py3-none-any.whl | 151474 bytes | 0.2.14 |
| DIST | tokenspeed_mla-0.2.15-py3-none-any.whl | 151477 bytes | 0.2.15 |
| DIST | tokenspeed_mla-0.2.16-py3-none-any.whl | 154122 bytes | 0.2.16 |
| Type | File | Size |
|---|