Back to dollama.net

What works today

Dollama is an experimental developer tool. This page lists what is available, what is available but experimental, what is in testing and not yet available, and what is only planned.

Availability and limitations

Available
Usable in the released product.
Available · experimental
Usable, with the limitations stated in the notes.
In testing · not yet available
Being evaluated; not enabled for users.
Planned
Intended; not available.

Last verified 1 October 2026 · applies to CLI v0.72.0 and relay v0.64.0

FeatureStatusNotes
Local endpoints for coding tools and applicationsAvailableAnthropic Messages, OpenAI Chat Completions and Responses
Private, Group and Open Network modesAvailableYour own machines are always tried first
Worker, planner and helper rolesAvailableAssigned to each machine from its benchmark
Recovery: retries, continuation, tool-call repairAvailableMeasured per model on the Models page
Web search and page fetch for the modelAvailableRuns from your machine
Contribution-based priorityAvailableOrders the queue when the network is busy
Speech-to-textAvailableRuns on CPU or GPU. See Capabilities
EmbeddingsAvailable · experimentalLightly benchmarked so far
Personal hosting of static sitesAvailable · experimentalSee the limits on the Hosting page
Direct links between your own machinesAvailable · experimentalOpt in with dollama peer enable; falls back to the relay
Models outside Ollama (Bonsai 27B on llama.cpp)Available · experimentalYour own machines only; slow to first token on a 12 GB GPU
Evaluating a candidate model and submitting the resultAvailable · experimentalResults feed the list of models the network accepts
Several capabilities on one machineAvailable · experimentalFor example speech-to-text next to a language model
Decision model for routing choices (Julia-1)In testing · not yet availableBenchmarked for deciding when a request needs planning; not used in the product yet
Offline reference search (Kiwix)In testing · not yet availableLocal-only lookup of offline reference collections; not enabled yet
Client-to-client encryptionPlannedWould encrypt each request from your machine to the machine that runs it. Today each hop is encrypted in transit, but the relay can read requests. The serving machine always has to read it.
Sending less of each request to the relayPlannedBuilt, but off by default
Self-service account deletionPlannedBy email until then
Separate settings for where your requests run and who may use your hardwarePlannedToday one mode sets both
Public source code and licencePlannedNo date yet. See Source release plans

Source release plans

  1. The source is not public yet. You can read the installation scripts: /install.sh and /install.ps1.
  2. No licence has been published yet, so nothing is licensed for reuse yet.
  3. An open-source release is intended. There is no announced date.
  4. It will be announced on the Releases page and on this page.

Roadmap

The aim is that models, relays and applications stay replaceable, so your access does not depend on any one provider, including Dollama.

Next

Work that already exists in some form and is being finished.

  • Easier personal-cluster setup
  • Self-service account deletion
  • Sending less of each request to the relay
  • Separate settings for where your requests run and who may use your hardware
  • Public source code under an open-source licence

Longer term

Intended directions. None has a date.

  • Client-to-client encryption, so the relay cannot read requests
  • Automated review and correction of completed work
  • More direct connections between your machines
  • Relays run by other people, using a documented protocol
  • Running models on more runtimes
  • Persistent agents and personal data hosted at the edge
  • A network that keeps working without Dollama-operated infrastructure