Powered by Google's latest Gemma technology, Google AI Edge Eloquent is an advanced dictation app engineered to bridge the gap between natural speech and professional, ready-to-use text.




Powered by Google's latest Gemma technology, Google AI Edge Eloquent is an advanced dictation app engineered to bridge the gap between natural speech and professional, ready-to-use text.




Records and analyzes screen content locally with AI, supports semantic/keyword search, chat-based recall, multiple privacy safeguards, and robust integrations.



Custom Swift and Metal runtime enabling 26B instruction-tuned model inference in around 2 GB RAM by streaming weights from SSD, optimized for Apple Silicon Macs.

Chat with AI models, generate images, convert books into audiobooks, and run text-to-speech – all locally on your device. No cloud. No compromises.




Local AI app and inference engine for agents. Run open-weight LLMs locally — private, 100% offline on your computer.





OfflineLLM is the fastest large language model (LLM) engine designed specifically for Apple devices, including iPhone, iPad, Mac and Vision Pro. With OfflineLLM, users can engage in private conversations with AI chatbots without the need for an internet connection, ensuring that...




Lyra rewrites what you type in every Mac app — Mail, Slack, WhatsApp, browsers. Runs entirely on your Mac. Learns how you write to each person.




AURA is an open-source Android app for AI face reading (physiognomy) and palm reading. The key idea is privacy: your photos never leave your device. All analysis runs locally using Google Gemma (on-device LLM) and MediaPipe for detection, so nothing is uploaded to a server.




Run and fine-tune generative AI models with easy-to-use APIs and highly scalable infrastructure. Train and deploy models at scale on our AI Acceleration Cloud and scalable GPU clusters. Optimize performance and cost.

ShieldGemma is a set of instruction tuned models for evaluating the safety of text and images against a set of defined safety policies. You can use this model as part of a larger implementation of a generative AI application to help evaluate and prevent generative AI...

The desktop app for Ollama. Chat, code, and plan with your local models — and scale to faster cloud models when you need them. Free local use, no account required.

WebBrain is a free, open-source browser extension that brings AI agent capabilities to your browser. Read pages, extract data, and automate web tasks — powered by your choice of LLM. The self-hostable alternative to proprietary browser AI plugins.



