Skip to content
TriageTriage
Esc
↑↓navigate↵open⌘Jpreview
On this page

Choices

Host triage and run its models wherever you like, local or hosted, with no vendor lock-in.

Triage doesn’t tie you to any vendor. Every part can run on your own hardware or on a service you pick, and none of them is the intended one with the rest as fallbacks. Mix them however suits you: the server on one machine, decision models on another, and a hosted language model, or everything on one box with nothing leaving your network.

See Privacy for exactly what each of them is sent, and Configuration for the settings.

Hosting the server

Option Where data goes Cost
Arch Linux user service Your machine Free
Container, with Docker Compose Your machine, plus Cloudflare if you use a Cloudflare Tunnel Free
Home Assistant app Your Home Assistant Free
Cloudflare (planned) Your Cloudflare account Cloudflare Workers pricing

Hosts and workers only need the server’s URL, so you can move it later without changing anything else.

Decision models

Decision models decide which issues are worth fixing. Set TRIAGE_DECISION_PROVIDER to typesafe for any TypeSafe System One API, or cloudflare for Clef.

Option Where data goes Cost
Ollaya, such as laya or winnow Your machine Free
Ollama 0.35 or later, such as nimble or tev1 Your machine Free
TypeSafe’s API TypeSafe TypeSafe’s pricing
Jev on OpenCode Zen OpenCode OpenCode Zen’s pricing
Clef on Cloudflare Workers AI, clef or clef-flash Your Cloudflare account Workers AI’s free 10,000 Neurons a day, then its pricing

Language models

Language models suggest fixes. Set TRIAGE_LLM_PROVIDER to openai for any OpenAI-compatible API, anthropic for any Anthropic-compatible one, or cloudflare for Workers AI.

Option Where data goes Cost
Ollama, LM Studio or llama.cpp Your machine Free
OpenAI OpenAI OpenAI’s pricing
Anthropic Anthropic Anthropic’s pricing
OpenRouter OpenRouter and the model’s provider Each model’s price on OpenRouter
OpenCode Zen, through either API OpenCode OpenCode Zen’s pricing
GitHub Models GitHub GitHub’s free allowance, then its pricing
Workers AI Your Cloudflare account Workers AI’s free 10,000 Neurons a day, then its pricing

One Ollama can serve both the decision model and the language model.

Planned:

  • A coding agent such as OpenCode, Pi, Cursor, Claude Code, Codex, Copilot or Gemini, which reaches whatever providers it’s set up with, such as a Copilot subscription you already have.
  • Home Assistant’s AI Task action, which uses whichever AI provider your Home Assistant is set up with.

Limits

Decide and suggest are off until you turn them on, and stop at their daily limits (TRIAGE_DECIDE_DAILY and TRIAGE_SUGGEST_DAILY), shared by the server and all its workers. Each suggestion is capped at 4,096 tokens, so its cost stays predictable on any paid API.

Last updated on