30 lines
867 B
Markdown
30 lines
867 B
Markdown
# ADR 002 — Ollama Is the v1 Model Backend
|
|
|
|
**Status:** Accepted
|
|
|
|
## Decision
|
|
|
|
v1 will target local Ollama inference.
|
|
|
|
## Context
|
|
|
|
The intended deployment already has a local Ollama inference engine. The project prioritizes local control, privacy, and predictable integration.
|
|
|
|
## Alternatives Considered
|
|
|
|
- multiple cloud providers,
|
|
- LM Studio,
|
|
- llama.cpp direct integration,
|
|
- arbitrary OpenAI-compatible endpoints,
|
|
- Ollama.
|
|
|
|
## Reason
|
|
|
|
Ollama is already available locally, provides a simple local API, supports both text-generation and embedding models, and avoids requiring external inference services.
|
|
|
|
## Consequences
|
|
|
|
- candidate forks supporting multiple cloud providers should be simplified or hardened,
|
|
- candidate projects using another local API need an adapter,
|
|
- future backend abstraction may be added, but v1 should not be delayed to support it.
|