This sorta reminds me of the lie that was pushed when the Snapdragon X laptops w...

wmf · 2025-05-03T21:37:52 1746308272

AFAIK Windows 11 does use the NPU to run Phi Silica language models and this is available to any app through some API. The models are quite small as you said though.

eyeris · 2025-05-04T16:34:07 1746376447

Agreed. Same with intel’s NPUs. I’ve been testing with my intel core evo 155x. The npu only runs int8 as well. At least, Intel has put in a decent amount of effort into the ecosystem

There are a couple ways to interface — DirectML by MS and Intel’s native api (they provide OpenVINO model conversion to convert normal Python ml models) I’ve tried ONNXRuntime conversions for both backends to little success. Additionally the OpenVINO model conversion seems to break the model if the model small enough.

OpenVINO model server seems pretty polished and has openapi compatible endpoints.

nullpoint420 · 2025-05-03T17:02:15 1746291735

If you don’t mind me asking, what OS do you use on it?

cowmix · 2025-05-03T17:35:16 1746293716

I use Windows 11. Podman/WLS2 works way better than I thought it would. And when Gitbash was finally ported (officially) - that filled other gaps I was missing in my workflow. Windows ARM Python still is lacking all sorts of stuff, but overall I'm pretty productive on it.

I pre/ordered the Snapdragon X Dev kit from Qualcomm - but they ended up delivering a few units -- to only cancel the whole program. The whole thing turned out to be a hot-mess express saga. THAT computer was going to be my Debian rig.

captainregex · 2025-05-03T17:21:46 1746292906

AnythingLLM uses NPU

mikaraento · 2025-05-03T18:25:33 1746296733

Could you provide a pointer to docs for this? It wasn't obvious from an initial read of their docs.

tough · 2025-05-03T19:38:42 1746301122

could find some about it on the 1.7.2 changelog here https://docs.anythingllm.com/changelog/v1.7.2