RIME_API_KEY on the server and call Rime over HTTPS with fetch. No Rime-specific npm package is required; Express is the only dependency.
Prerequisites
- A Rime API key from the API Tokens page, exported as
RIME_API_KEY - Node.js 20.11+ and
npm install express
1. Server: Express app with a TTS route
Createserver.mjs:
server.mjs
test.mp3 should be playable audio. A 401 means RIME_API_KEY isn’t visible to the server process.
2. Client: mic in, Rime audio out
Createindex.html next to server.mjs. It uses the browser’s built-in SpeechRecognition for input (Chrome/Edge/Safari), a stub respond() function as the agent brain, and your /api/tts route for the voice:
index.html
http://localhost:3000, select Speak, say something, and the agent answers in Rime’s astra voice.
Add streaming when latency matters
For incremental synthesis, attach a WebSocket bridge to the same HTTP server: browser WS ↔ Express server ↔wss://users-ws.rime.ai/ws3 with header authentication. Audio can then start while later sentences are still generating. Browser WebSockets cannot send the required Authorization header, so the bridge stays server-side. The Next.js guide provides the complete bridge; it works with Express’s HTTP server through const server = app.listen(3000). The Coda WebSocket reference defines the /ws3 schema.
Production building blocks
WebSocket API overview
Endpoints, word-level timestamps, context IDs, and interruption handling.
Voices
Swap
astra for any Coda voice. Coda covers eight languages, and each voice serves one of them.Streaming formats
Choose between Opus, MP3, WAV, PCM, and μ-law for your latency budget.
Plain Node (no framework)
Build the same starter with only
node:http and built-in fetch.
