Local AI agents on your Mac. Private, offline, free. Open models run on Apple Silicon; add ChatGPT, Claude or Gemini only when you choose.
Cost / License
- Free
- Open Source (MIT)
Application typeApplication types
Platforms
- Mac
United States



Local AI agents on your Mac. Private, offline, free. Open models run on Apple Silicon; add ChatGPT, Claude or Gemini only when you choose.



Merlin AI is a free Chrome extension that unifies top AI models to help you instantly write, summarize, translate, code, and manage tasks directly in your browser.




Experience the power of RWKV models directly on your device. Completely offline, privacy-first, and efficient. No internet required.



Transform institutional knowledge into frontier-grade LLMs—without infrastructure burden or cloud lock-in.

Boost your app with AI: In-app chatbot capabilities and intelligent text generation, all open-sourced for React devs.


OfflineLLM is unlimited, private, offline, 24/7, free access to AI. Augment your day-to-day life by using this Ai chatbot for a multiplicity of applications.



AI Sparks Studio is a user interface that allows you to efficiently utilize your own API access to state-of-the-art AI models like ChatGPT, GPT-4, Whisper or ElevenLabs.




Crush is an open-source, terminal-based AI coding agent built in Go by Charm (charmbracelet), designed to bring agentic coding directly into your shell with a polished TUI experience.

Connect your local or hosted instance using your URL and API key. Once connected, all available models are loaded automatically. You can switch between them at any time, start new chats, or continue existing ones in a clean and focused interface.




Run LLMs on device or connect to various commercial or open source APIs. ChatterUI aims to provide a mobile-friendly interface with fine-grained control over chat structuring.



Hermes Agent is a mobile-first AI agent app from Nous Research for running local models and practical Android workflows on your phone.



Google AI Studio is the fastest way to start building with Gemini, our next generation family of multimodal generative AI models.




Use your locally running AI models to assist you in your web browsing.




1-bit Bonsai 8B implements a proprietary 1-bit model design across the entire network: embeddings, attention layers, MLP layers, and the LM head are all 1-bit. There are no higher-precision escape hatches. It is a true 1-bit model, end to end, across 8.2 billion parameters.

Embeddings databases are a union of vector indexes (sparse and dense), graph networks and relational databases. This enables vector search with SQL, topic modeling, retrieval augmented generation and more.




Project referring to gpt-oss-120b and gpt-oss-20b, two open-source (weight) language models by OpenAI.

Digital Life Project 2 (DLP3D) is an open-source real-time framework that brings Large Language Models (LLMs) to life through expressive 3D avatars. Users converse naturally by voice, while characters respond on demand with unified audio, whole-body animation, and physics...




Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.
Echo combines specialist open-weight models into one system, allocating compute where it improves the result.


Run Llama, Gemma, Qwen, DeepSeek, and more locally on your iPhone, iPad, and Mac. Offline. Private. No login. Optimized for Apple Silicon.



🌟 An AI desktop pet with long-term memory, expressive character sprites, computer control, and voice features—perfect for Galgame-style characters 🌟

This application provides a full suite of generative AI features for chat, code assistance, document search, image analysis, image and video generation. All features run offline and are powered by your PC’s Intel® Core™ Ultra with built-in Intel Arc GPU or Intel Arc™ dGPU...

