The local proxy does not serve these. It has no /v1/audio/transcriptions or /v1/embeddings endpoint; a request to /v1/embeddings returns a 404. There is no OpenAI-compatible route for either.
Today the supported way is to call the CLI from your application and read its JSON output. Both commands print JSON to standard output, or to a file with --output.
# Speech-to-text: prints {"text","language","words","duration_s"}
dollama transcribe recording.wav --output transcript.json
# Embeddings: prints {"model","dims","embeddings","usage"}
dollama embed "first chunk" "second chunk" > vectors.json
You cannot call the relay for these with a plain HTTP client either. Every request needs a live content tunnel from your machine, and audio is fetched over it, so the CLI is the route to use.
Limits. Speech-to-text and embeddings are served only by machines in Open Network mode. By default, dollama transcribe sends audio only to machines on your own account, not to group members' machines, so one of yours must be running dollama network. Pass --public to also allow public volunteers' machines. In CLI v0.72.0, dollama embed always uses the Open Network, whatever mode the app is in. See known gaps. Flags, output and routing are in Capabilities.