vMLX provides functions no other MLX inferencing app does, including LM Studio, from KV Cache Quantization (save 2-4x the RAM), Prefix Caching, and full VL support.
Cost / License
- Free
- Proprietary
Application type
Platforms
- Mac
United States


Ollama is described as 'Facilitates local deployment of Llama 3, Code Llama, and other language models, enabling customization and offline AI development. Perfect for creating personalized AI chatbots and writing tools' and is a very popular large language model (llm) tool in the ai tools & services category. There are more than 100 alternatives to Ollama for a variety of platforms, including Mac, Windows, Linux, Web-based and Android apps. The best Ollama alternative is Jan.ai, which is both free and Open Source. Other great apps like Ollama are Ensu, AnythingLLM, LM Studio and Alpaca - Ollama Client.
vMLX provides functions no other MLX inferencing app does, including LM Studio, from KV Cache Quantization (save 2-4x the RAM), Prefix Caching, and full VL support.


SmolChat allows you to download and run popular LLMs on your Android device, locally, without needing an internet connection. Customize the model used for each chat, tune settings like temperature and min-p, and pin your favourite chats on the home-screen with shortcuts.



📱 The first fully functional, standalone AI assistant for mobile devices with powerful tool-calling capabilities 📱



A local-first AI workspace that runs open-weight models (GGUF via llama.cpp and MLX) directly on consumer hardware with integrated MCP tool support.




Cortex is the open-source brain for robots: vision, speech, language, tabular, and action -- the cloud is optional.


DS AI Chat is your powerful AI assistant on your Windows PC. It includes the latest LLMs (large language models) like DeepSeek and Qwen. You don’t need any coding skills. Just follow 3 easy steps to install it on your PC and start chatting with AI.




HugstonOne Enterprise Edition — Awesome AI App with Code editor and Live Preview Create games, dashboards, maps, tables, charts, webpages, data analysis, converters etc in seconds.




Greenative Studio lets you run local LLMs with your files, tools, and coding agents - without sending sensitive data or source code to cloud AI services.



Nativ is a native macOS workspace for running AI models locally on Apple silicon. It bundles an mlx-vlm server, finds compatible models in your Hugging Face cache (honoring HF_HUB_CACHE and HF_HOME), and wraps the whole experience in a polished SwiftUI app.


An offline AI assistant that runs entirely from a USB drive on Windows and macOS, with no internet connection, account, or subscription.



