Run AI models locally on your device. Foundry Local provides on-device inference with complete data privacy, no Azure subscription required.
Cost / License
- Free
- Open Source
Platforms
- Windows
- Mac
Ollama is described as 'Facilitates local deployment of Llama 3, Code Llama, and other language models, enabling customization and offline AI development. Perfect for creating personalized AI chatbots and writing tools' and is a very popular large language model (llm) tool in the ai tools & services category. There are more than 100 alternatives to Ollama for a variety of platforms, including Mac, Windows, Linux, Web-based and Android apps. The best Ollama alternative is Jan.ai, which is both free and Open Source. Other great apps like Ollama are Ensu, AnythingLLM, LM Studio and Alpaca - Ollama Client.
Run AI models locally on your device. Foundry Local provides on-device inference with complete data privacy, no Azure subscription required.
This application provides a full suite of generative AI features for chat, code assistance, document search, image analysis, image and video generation. All features run offline and are powered by your PC’s Intel® Core™ Ultra with built-in Intel Arc GPU or Intel Arc™ dGPU...



AI00 RWKV Server is an inference API server for the RWKV language model based upon the web-rwkv inference engine.
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs.




Learn how to add AI with local models and APIs to Windows apps. Discover AI scenarios and models such as Phi, Mistral, Stable Diffusion, Whisper, and many more to delight your users. The AI Dev Gallery is an open-source app designed to help Windows developers integrate AI...







This project aims to eliminate the barriers of using large language models by automating everything for you. All you need is a lightweight executable program of just a few megabytes. Additionally, this project provides an interface compatible with the OpenAI API, which means...




A modern web interface for managing and interacting with vLLM servers (www.github.com/vllm-project/vllm). Supports both GPU and CPU modes, with special optimizations for macOS Apple Silicon and enterprise deployment on OpenShift/Kubernetes.




LM-Kit.NET is a versatile SDK for integrating Large Language Models (LLM) into C# applications. It provides advanced AI features like text generation, Natural Language Processing (NLP), data retrieval, content enhancement, and translation, enabling diverse industry use cases with ease.




Automate Unresolved Go-to-Market Challenges and Empower Your CRM, Marketing, and Sales Teams to Achieve Desired Results with Multi-agent Framework, B2B Database Creation, Agentic Workflows, and Integrations.




Inferencer lets you run, host and deeply control the latest SOTA AI models (OSS, DeepSeek, Qwen, Kimi, GLM, MiniMax and more) from your own computer.



