Skip to the page
Conch
DocsGitHub

Providers

On this computer

Private, free, and works offline

An open model that runs right here, through Ollama. Your chats aren’t sent to any AI company, it costs nothing, and it keeps working when the internet doesn’t. Conch lends it your integrations.

  • Private
  • Free
  • Works offline
Connects with
A program on this computer
Your apps
Conch hands it their tools
To start
Free
  • NoYour files and commands
  • YesSaves memories itself
  • YesAsks before each step
  • YesWorks offline

Worth knowing first

  • Slower and less capable than the big cloud models: best for everyday questions, drafts and quick jobs.
  • Tool-capable models can use Conch’s files and connected apps. Commands need the OS sandbox and run without network access; chat-only models cannot take actions.

Set it up

One button does all of it.

  1. Open Settings → Providers → On this computer.
  2. Press Get on the model Conch suggests for this computer. It shows the size and how long the download should take.
  3. If Ollama (the program that runs the model) isn't here, the same press installs it first, then carries on to the model. On Linux, Conch links to Ollama's installer instead, and notices by itself when it's there.

You can pause the download and pick it up later. When it finishes, the model is in the picker like any other.

Which model

Conch only offers models that can use tools, and suggests one by how much memory this computer has: about 2 GB of download on a small laptop, up to about 23 GB on a workstation.

It never offers one that won't fit. If memory or disk is short, it says by how much, in one sentence, and there's no button to press.

What it's for

  • Private. Your chats go to no AI company. The model only ever talks to this computer.
  • Free. Nothing to pay, and no limit.
  • Offline. It answers on a train, and takes over when the internet drops. See Offline and at a limit.

It is slower and less capable than the big cloud models. Use it for everyday questions, drafts and quick jobs, and a cloud provider for hard ones. A chat can move between them.

Good to know

  • Thinking is off by default, so small models answer straight away. Turn it up in the model picker.
  • It reads as much as this computer allows. Conch works out how much of the chat the model can hold from the model itself and this computer's memory, between 16,000 and 64,000 tokens.
  • A small model gets a lean setup. When the model reads little at once, Conch gives it short instructions and hands it tools as it needs them, so the chat itself has room. It's automatic.
  • Conch keeps Ollama running. If it has stopped when you need it, Conch starts it quietly and leaves a note under Fixed on its own.