Files
interactive-story/planning/DECISIONS/002-ollama-only-v1.md
T

867 B

ADR 002 — Ollama Is the v1 Model Backend

Status: Accepted

Decision

v1 will target local Ollama inference.

Context

The intended deployment already has a local Ollama inference engine. The project prioritizes local control, privacy, and predictable integration.

Alternatives Considered

  • multiple cloud providers,
  • LM Studio,
  • llama.cpp direct integration,
  • arbitrary OpenAI-compatible endpoints,
  • Ollama.

Reason

Ollama is already available locally, provides a simple local API, supports both text-generation and embedding models, and avoids requiring external inference services.

Consequences

  • candidate forks supporting multiple cloud providers should be simplified or hardened,
  • candidate projects using another local API need an adapter,
  • future backend abstraction may be added, but v1 should not be delayed to support it.