Skip to main content
Modulate’s models are grouped into families by the kind of output they produce. Within a family, endpoints differ by language coverage, latency, and whether they take a file or a live stream. Each capability page below carries the full parameter, response, and audio format detail for its endpoints.

Model families

No Detection model returns a transcript.

Endpoints that overlap

Three models return a transcript

Multilingual Transcription, PII/PHI Redaction, and Velma Triage all return a transcript. Pairing any of the last two with a transcription call returns the same text twice.

Four signals have two paths

Each is available as a flag on Multilingual Transcription, attached per utterance, or as a dedicated endpoint that produces no transcript. English Fast and Multilingual Fast Transcription carry none of them.

Common scenarios