What
Connect the side panel to a local Ollama/llamafile server. Implement model selection, streaming responses, and graceful offline fallback.
Model Selection & Download
Auto-Download (Current Sprint 2)
- Add setup modal: "Download qwen2.5 (~400MB)?"
- Download + launch llamafile binary on first-use
- Seamless for users, zero manual steps
- Setup wizard on first-use: "Connect to local AI model"
- Model selector populated from
GET /api/tags endpoint
Acceptance Criteria
What
Connect the side panel to a local Ollama/llamafile server. Implement model selection, streaming responses, and graceful offline fallback.
Model Selection & Download
Auto-Download (Current Sprint 2)
GET /api/tagsendpointAcceptance Criteria
GET localhost:11434/api/tagslocalhost:11434/api/generateai_suggestioneventai_interactionwithacceptance: 'fully_accepted'ai_interactionwithacceptance: 'rejected'localhost:11434/api/generate)