Intelligence

We don't guess who was talking

Because we record in dual channel, the rep and the prospect are already on separate audio tracks before transcription starts.

Proof Points

  • AssemblyAI Universal-3.5 Pro, multichannel
  • Speaker separation from the recording, not from a model's guess
  • Click-to-jump timestamps
  • Colour-coded speakers
  • Talk-time percentage per speaker
  • Automatic retry, up to 3 attempts over a 30-day window
  • $0.02 per minute

How It Works

  1. Record the call
  2. Transcription runs automatically
  3. Read it with speakers colour-coded
  4. Click any line to jump the audio

Benefits

Separation before transcription, not after

Twilio records dual channel. Channel one is the rep's leg, channel two is the dialed leg. AssemblyAI's multichannel model transcribes each independently. There is no inference step to get wrong.

Click a line, hear it

Every line in the transcript is timestamped. Click and the audio jumps there.

It fixes itself

When a transcription fails — a provider hiccup, a quota limit — a sweeper retries it automatically, up to three attempts, waiting long enough that a live call is never transcribed twice.

Frequently asked questions

Why does dual channel matter so much?

Diarization guesses who spoke. Dual channel knows. On a call where both people talk over each other, that's the difference between a usable transcript and a confusing one.

What languages are supported?

English today.

What if transcription fails?

It retries automatically.

Explore