Ollaya icon
Ollaya icon

Ollaya

Run open decision models locally: pull and serve Laya, decider, NLI and GLiClass behind a TypeSafe-compatible API. Ollama for decision models.

Ollaya screenshot 1

Cost / License

Platforms

  • Mac
  • Linux
  • Windows
  • Docker
1like
0articles
Save

Features

Properties

  1.  Privacy focused

Features

  1.  Ad-free
  2.  Works Offline
  3.  No registration required
  4.  No Tracking
  5.  Local AI

Ollaya News & Activities

Highlights All activities

Recent activities

Ollaya information

GitHub repository

  •  820 Stars
  •  38 Forks
  •  4 Open Issues
  •   Updated  
View on GitHub

Popular alternatives

View all
Ollaya was added to AlternativeTo by Paul on and this page was last updated .

No comments or reviews, maybe you want to be first?

Official Links

What is Ollaya?

Run open decision models locally, the way Ollama runs LLMs.

A decision model reads a state (a message, an email, a ticket, any JSON) plus typed questions (choice, score, noul) and returns calibrated probabilities in a single forward pass, in milliseconds. It never generates text. Ollaya pulls these models by name, serves them from a local daemon, and speaks TypeSafe's /v1/systemone wire format, so existing Jev clients work by changing one environment variable.

Features:

  • One binary. ollaya serve runs the daemon; ollaya run, pull, list, ps, show, rm, cp, stop and create work the way they do in Ollama. If the daemon isn't running, the CLI starts it.
  • TypeSafe-compatible. POST /v1/systemone, /v1/decisions and GET /v1/models are wire-identical to TypeSafe. The official SDK works unchanged when you set TYPESAFE_BASE_URL=http://localhost:11435.
  • Native API. /api/decide adds routing information and timings. /api/pull streams NDJSON progress, and there are /api/tags, /api/show, /api/ps and more. See docs/api.md.
  • Weights come from their authors. Ollaya publishes only small ONNX graphs, about 3 MB each. These graphs read the original weight files (usually model.safetensors) from the author's Hugging Face repository, pinned to a commit and verified by sha256. Models whose authors publish GGUF files (winnow, jevk5) run that file itself on llama.cpp. Ollaya never re-hosts weights.
  • For agents. ollaya mcp serves the models to Claude Code, Claude Desktop, Cursor and other MCP clients (claude mcp add ollaya -- ollaya mcp), and the ollaya-decisions skill teaches agents when and how to use them (npx skills add ollaya-dev/ollaya --skill ollaya-decisions).
  • Routers. laya detects the script and language of each request, then answers with laya:en or laya:multilingual.
  • Modelfiles. You can bake a question set into your own model.