back

by EwanG·2y ago·view on hn ↗
The important stuff from the readme (if you're not looking to tinker with it directly):

We have tested PowerInfer on the following platforms:

x86-64 CPU (with AVX2 instructions) on Linux

x86-64 CPU and NVIDIA GPU on Linux

Apple M Chips on macOS (As we do not optimize for Mac, the performance improvement is not significant now.)

And new features coming soon:

Mistral-7B model

Metal backend for sparse inference on macOS

1 comments
Also worth mentioning the downloadable llama2 models, and the convert.py file.