Ask questions to your documents without an internet connection, using the power of LLMs. 100% private, no data leaves your execution environment at any point. You can ingest documents and ask questions without an internet connection!

MLC LLM is described as 'Machine learning compiler and high-performance deployment engine for large language models. The mission of this project is to enable everyone to develop, optimize, and deploy AI models natively on everyone’s platforms' and is a AI Chatbot in the ai tools & services category. There are more than 50 alternatives to MLC LLM for a variety of platforms, including Mac, Windows, iPhone, Android and Linux apps. The best MLC LLM alternative is Lumo by Proton, which is both free and Open Source. Other great apps like MLC LLM are ChatGPT, Ollama, Google Gemini and Claude.
Ask questions to your documents without an internet connection, using the power of LLMs. 100% private, no data leaves your execution environment at any point. You can ingest documents and ask questions without an internet connection!

LlamaGPT is a chatbot that provides a ChatGPT-like experience, with no data leaving your device.



Empowering LLM researchers and hobbyists with seamless control over self-hosted models. Connect remotely, customize prompts, manage chats, and fine-tune configurations. All in one intuitive app.




Drop-In OpenAI replacement, On-device, local-first, Generate text/image/speech/music/etc... Backend Agnostic: (llama.cpp, diffusers, bark.cpp, etc...), Optional Distributed Inference(P2P/Federated).




NanoGPT is revolutionizing AI access with a core mission: democratizing state-of-the-art models like ChatGPT, Claude, and more for everyone, globally. We believe cutting-edge AI should be accessible, not expensive or complex.




Comprehensive AI suite for Galaxy devices includes photo, video, and audio editing, generative and side-by-side tools, background noise removal, writing and note-taking assistance, AI-powered briefings, real-time info, content organization, and seamless device integration.




Run LLMs on AMD Ryzen™ AI NPUs in minutes. Just like Ollama - but purpose-built and deeply optimized for the AMD NPUs.
OptiQ is an MLX-native toolkit for running large language models on Apple Silicon. No PyTorch, no CUDA, no cloud.




Chat with AI models, generate images, convert books into audiobooks, and run text-to-speech – all locally on your device. No cloud. No compromises.




A modern web interface for managing and interacting with vLLM servers (www.github.com/vllm-project/vllm). Supports both GPU and CPU modes, with special optimizations for macOS Apple Silicon and enterprise deployment on OpenShift/Kubernetes.




Run frontier LLMs and VLMs with day-0 model support across GPU, NPU, and CPU, with comprehensive runtime coverage for PC (Python/C++), mobile (Android & iOS), and Linux/IoT (Arm64 & x86 Docker). Supporting OpenAI GPT-OSS, IBM Granite-4, Qwen-3-VL, Gemma-3n, Ministral-3, and more.




Privacy-first AI runs on-device for iPhone and iPad, manages files, calendar, contacts, and health offline, analyzes PDFs, enables secure chats, custom briefings, and supports 13 languages, with optional cloud models using your OpenAI API key for flexibility.



