Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.
Cost / License
- Free Personal
- Open Source (MIT)
Application typeApplication types
Platforms
- Windows
- Online
- Self-Hosted
RWKV Runner is described as 'This project aims to eliminate the barriers of using large language models by automating everything for you. All you need is a lightweight executable program of just a few megabytes. Additionally, this project provides an interface compatible with the OpenAI API, which means' and is a large language model (llm) tool in the ai tools & services category. There are more than 10 alternatives to RWKV Runner for a variety of platforms, including Windows, Linux, Mac, Self-Hosted and Android apps. The best RWKV Runner alternative is Ollama, which is both free and Open Source. Other great apps like RWKV Runner are Jan.ai, GPT4ALL, AnythingLLM and LM Studio.
Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.
This application provides a full suite of generative AI features for chat, code assistance, document search, image analysis, image and video generation. All features run offline and are powered by your PC’s Intel® Core™ Ultra with built-in Intel Arc GPU or Intel Arc™ dGPU...


Learn how to add AI with local models and APIs to Windows apps. Discover AI scenarios and models such as Phi, Mistral, Stable Diffusion, Whisper, and many more to delight your users. The AI Dev Gallery is an open-source app designed to help Windows developers integrate AI...







Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs.




AI00 RWKV Server is an inference API server for the RWKV language model based upon the web-rwkv inference engine.
A modern web interface for managing and interacting with vLLM servers (www.github.com/vllm-project/vllm). Supports both GPU and CPU modes, with special optimizations for macOS Apple Silicon and enterprise deployment on OpenShift/Kubernetes.




MLC LLM is a machine learning compiler and high-performance deployment engine for large language models. The mission of this project is to enable everyone to develop, optimize, and deploy AI models natively on everyone’s platforms.



📱 The first fully functional, standalone AI assistant for mobile devices with powerful tool-calling capabilities 📱



A local-first Windows AI ecosystem: a multi-module workspace, a persistent AI agent, and a voice assistant. Runs offline. One-time purchase, no subscription.
