Ollama logo
AI · self-hosted

Ollama in one click.

Ollama is a local llm runner with openai-compatible api you can run on your own machine. Launch it in Even Stacks in one click, serve it over trusted HTTPS, and let your AI agents operate it.

What is Ollama?

Ollama is a tool for running large language models locally. It supports models like Llama, Mistral, Gemma, and Qwen, and exposes an OpenAI-compatible REST API on port 11434.

CategoryAI
Default port11434
Container imageollama/ollama:latest
Official siteollama.com

Run Ollama in Even

Even Stacks launches Ollama as a managed container on your own machine. No compose files, no manual setup.

  1. Open the Even Stacks panel in Even. The container engine starts on demand.
  2. Find Ollama in the one-click services and click Launch.
  3. Even pulls the image, starts it on port 11434, and serves it over trusted HTTPS at https://ollama.localhost.

Drive Ollama with your AI agents

Even ships an MCP, so any agent you run inside Even (Claude Code, Codex, and more) can operate Ollama directly, at both the UI level and the container level.

Through the Even MCP an agent can pull a model, call the local endpoint, and wire it into a tool, all on your own hardware. It runs commands inside the container, reads its logs, and for web apps opens and clicks through the interface in Even's own browser pane. Ask once, for example "set up Ollama and get it ready", and the agent handles it end to end.

How agents reach it: Pull models with `docker exec <container> ollama pull <model>` and run inference via POST to `http://localhost:11434/api/generate` with `{"model":"<name>","prompt":"..."}`. Use `/api/chat` for multi-turn. Models are stored in `/root/.ollama`.

More AI services

Other one-click services you can launch in Even Stacks.