Files

54 lines
1.6 KiB
Markdown

# Ollama Project Setup
## Project Overview
This repository contains the Ollama project configuration for local development. The project uses the [Ollama framework](https://ollama.ai) to run large language models (LLMs) locally, along with [OpenWebUI](https://openwebui.ai) (web interface) and [SearXNG](https://searxng.net) (search engine).
## Environment Details
- **OS**: Linux (6.8.0-117-generic)
- **Working Directory**: /opt/ollama
- **Model**: qwen2.5-coder:7b-instruct-q4_K_M-64k
- **Context Limit**: 40960 tokens (64k context window)
## Project Features
- Local model serving with Ollama
- Custom OpenAI-compatible provider
- 7B parameter model with 64k context window
- Web interface with OpenWebUI
- Search capabilities with SearXNG
## How to Run
### 1. Start the model
```bash
curl http://localhost:11434/api/generate -d '{
"model": "qwen2.5-coder:7b-instruct-q4_K_M-64k",
"prompt": "Hello, world!"
}'"
```
### 2. Check model list
```bash
curl http://localhost:11434/api/models
```
### 3. Start OpenWebUI (web interface)
```bash
docker run -p 3000:3000 -v /opt/ollama/openwebui:/app/data -d --name openwebui openwebui/openwebui:latest
```
### 4. Start SearXNG (search engine)
```bash
docker run -p 8888:8888 -v /opt/ollama/searxng:/app/data -d --name searxng searxng/searxng:latest
```
## Notes
- The model has a 64k context window limit
- No Docker required for Ollama - uses native API
- OpenWebUI accessible at http://localhost:3000
- SearXNG accessible at http://localhost:8888
- All operations are performed in the /opt/ollama directory
- For GPU acceleration, ensure CUDA is properly configured