OptiQ is an MLX-native toolkit for running large language models on Apple Silicon. No PyTorch, no CUDA, no cloud.




Cloudflare Workers AI is described as 'A serverless AI platform that runs models on Cloudflare's network, offering over 50 open-source models and a comprehensive suite for global application deployment' and is an app. There are more than 50 alternatives to Cloudflare Workers AI for a variety of platforms, including Mac, Windows, Linux, Web-based and Self-Hosted apps. The best Cloudflare Workers AI alternative is Ollama, which is both free and Open Source. Other great apps like Cloudflare Workers AI are Jan.ai, GPT4ALL, Open WebUI and AnythingLLM.
OptiQ is an MLX-native toolkit for running large language models on Apple Silicon. No PyTorch, no CUDA, no cloud.




Advanced Slack bot integrating OpenAI's ChatGPT-4 and DALL-E-3 for interactive AI conversations and image generation.

KoboldCpp is an easy-to-use AI text-generation software for GGML models. It's a single self contained distributable from Concedo, that builds off llama.cpp, and adds a versatile Kobold API endpoint, additional format support, backward compatibility, as well as a fancy UI...




Run open-source LLMs on your computer. You can download and make custom characters for your models. Works offline. The new name is Back yard though it is the same.



A Gradio web UI for Large Language Models. Supports transformers, GPTQ, llama.cpp (GGUF), Llama models.

AI edge infrastructure for macOS. Run local or cloud models, share tools across apps via MCP, and power AI workflows with a native, always-on runtime.

Transform institutional knowledge into frontier-grade LLMs—without infrastructure burden or cloud lock-in.


RustGPT is my latest experiment in cloning the abilities of OpenAI's ChatGPT. It represents the fourth iteration in a series of clones, each built with different tech stacks to evaluate their functionality in creating a ChatGPT-like application.

Harness state-of-the-art open-source LLMs and image models at blazing speeds with Fireworks AI. Utilize rapid deployment, fine-tuning without extra costs, FireAttention for model efficiency, and FireFunction for complex AI applications including automation and domain-expert copilots.




OfflineLLM is unlimited, private, offline, 24/7, free access to AI. Augment your day-to-day life by using this Ai chatbot for a multiplicity of applications.



LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar.

