Roadmap¶
In progress
- Native in Ollama. opendecider-small-td retrained on Ollama's own
/v1/systemoneprompt as well as ours: 0.793 on typed-decisions through Ollama's endpoint, up from 0.719, and still 0.794 through the opendecider package. It goes on ollama.com once Ollama 0.35.1 (the first release that accepts third-party decision models) is out.
Next
- More than 26 options: shortlist-then-letters for the Qwen-based models, measured on every benchmark before it ships (0.3).
- Rule-labelled evaluation: every model on tasksource/procedural-typed-decisions, whose answers are computed exactly from rules, so it measures correctness rather than agreement with a teacher model.
- Fine-tune nano on your own labels: a script and a guide for adapting opendecider-nano to your decisions.
- Agent frameworks: LangChain / LangGraph and LlamaIndex tools.
- An MCP server, so AI assistants and coding agents can call OpenDecider as a tool.
- A TypeScript client for
opendecider serve. - ONNX export of opendecider-nano for edge and in-browser use.
- Multilingual evaluation.
Want to help with one of these? See where to help and open an issue to agree on the approach first.