In short
Lemonade provides local model tools for several backends, including supported Ryzen AI routes. Selecting Lemonade does not automatically mean the NPU is active: the model and backend determine execution.
Requirements: compatible Windows hardware, Ryzen AI drivers and a model supported by the intended backend. This is a documentation-based procedure, not a hardware validation claim.
What is Lemonade Server?
Lemonade organises model execution behind an application and service interface. A model using llama.cpp is different from one prepared for a Ryzen AI NPU backend, even when both appear in the same application.
Install Lemonade Server
Download and run the current official Windows MSI installer. Update Ryzen AI drivers as directed by the compatible AMD guide. Open the application and check the detected hardware and backends. Use the current installer rather than an executable name from an older tutorial.
Run a model on the NPU
Select a catalogue model explicitly supported by the detected Ryzen AI backend, download it and start a chat. Inspect backend status and logs. A response proves that a model ran, not that it used the NPU.
NPU vs iGPU in Lemonade
Choose the route according to model support. The current command interface can be checked with:
lemonade --help
lemonade status
lemonade backendsA hybrid path can use multiple processors and should be labelled accordingly. Speed comparisons need matching model conditions.
Frequently asked questions
Does Lemonade Server use the NPU?
It can with compatible hardware, drivers and a supported model/backend. Other models use GPU or CPU routes.