Add initial planning files from ChatGPT research here
This commit is contained in:
@@ -0,0 +1,29 @@
|
||||
# ADR 002 — Ollama Is the v1 Model Backend
|
||||
|
||||
**Status:** Accepted
|
||||
|
||||
## Decision
|
||||
|
||||
v1 will target local Ollama inference.
|
||||
|
||||
## Context
|
||||
|
||||
The intended deployment already has a local Ollama inference engine. The project prioritizes local control, privacy, and predictable integration.
|
||||
|
||||
## Alternatives Considered
|
||||
|
||||
- multiple cloud providers,
|
||||
- LM Studio,
|
||||
- llama.cpp direct integration,
|
||||
- arbitrary OpenAI-compatible endpoints,
|
||||
- Ollama.
|
||||
|
||||
## Reason
|
||||
|
||||
Ollama is already available locally, provides a simple local API, supports both text-generation and embedding models, and avoids requiring external inference services.
|
||||
|
||||
## Consequences
|
||||
|
||||
- candidate forks supporting multiple cloud providers should be simplified or hardened,
|
||||
- candidate projects using another local API need an adapter,
|
||||
- future backend abstraction may be added, but v1 should not be delayed to support it.
|
||||
Reference in New Issue
Block a user