skip to main content

Installation harvey

harvey is a terminal coding agent that runs language models locally. No cloud account or API key required.

Requirements

Hardware requirements

Harvey runs entirely on local hardware, so the practical limits are RAM (to hold the model) and storage (for the model files themselves), not just raw compute.

Compile from source

git clone https://github.com/rsdoiel/harvey
cd harvey
make
make test
make install

make install copies the binaries into $HOME/bin/ by default. Override with make install prefix=/usr/local.

Model backend

Install Ollama from https://ollama.com/download, then pull a model:

ollama pull qwen2.5-coder:7b

Harvey detects Ollama automatically on startup.

Llamafile (self-contained, no install)

Download a pre-built llamafile from: https://docs.mozilla.ai/llamafile/getting-started/pre-built-llamafiles

Recommended: - Qwen2.5-Coder-7B-Q5_K_S.llamafile (~5 GB, good for most hardware) - Phi-3.5-mini-instruct-Q4_K_M.llamafile (~2 GB, low-VRAM / CPU)

Place it in ~/Models/ and make it executable (Linux / macOS):

chmod +x ~/Models/Qwen2.5-Coder-7B-Q5_K_S.llamafile

Harvey finds the llamafile automatically and connects.

Running harvey

cd ~/myproject
harvey