Skip to main content

Cartesia STT

Streaming speech-to-text using Cartesia’s ink-whisper model. Uses ws, no vendor SDK required.

Install

Usage

Plug it into an agent:
Supported sample rates: 8000, 16000, 24000, 44100, 48000 Hz.