# ADR 002 — Ollama Is the v1 Model Backend **Status:** Accepted ## Decision v1 will target local Ollama inference. ## Context The intended deployment already has a local Ollama inference engine. The project prioritizes local control, privacy, and predictable integration. ## Alternatives Considered - multiple cloud providers, - LM Studio, - llama.cpp direct integration, - arbitrary OpenAI-compatible endpoints, - Ollama. ## Reason Ollama is already available locally, provides a simple local API, supports both text-generation and embedding models, and avoids requiring external inference services. ## Consequences - candidate forks supporting multiple cloud providers should be simplified or hardened, - candidate projects using another local API need an adapter, - future backend abstraction may be added, but v1 should not be delayed to support it.