Making a Recording

Choose Language

By default, XSpeak is configured to transcribe the English language. If you need to record in another language, press the Conversation Settings button in the top toolbar and choose the appropriate language:

Language chooser in meeting settings
Language chooser

Supported Languages

  • French
  • Korean
  • Portuguese
  • German
  • Italian
  • Chinese
  • Spanish
  • English
  • Cantonese
  • Japanese

Choose Conversation Context

You can also choose the context in meeting settings. It affects AI analysis and changes the prompt passed to the model for AI actions, chat, and real-time insights.

Supported Contexts

  • Neutral
  • Personal
  • Work

Choose Participants

If you know who will be participating in the conversation and you want to improve speaker identification quality, you can select the participants of the conversation in Conversation Settings. Remove the “All Speakers” badge and the “Other” badge and add only those who are participating. It’ll help the speaker diarization engine properly recognize speakers:

Selected participants in meeting settings
Selected participants

Starting Recording

Press the recording button (at the top toolbar). You’ll see a recording consent disclaimer. Please obtain the consent of all recorded parties before starting the recording if that’s required by law in your region. After you obtain consent, press “Agree and Start” and recording will start:

Recording consent disclaimer
Recording

Recording consists of individual statements, each labeled with the speaker.

Conversation Naming

Once the recording has enough context, it’ll be named automatically by the chosen AI model. You can rename it manually if you need to.

Assigning Speakers

XSpeak learns speakers’ voices when you assign the correct speakers to statements. Initially, it’ll assign “Other” to all unknown voices. To change the speaker and teach XSpeak other people’s voices, press the Speaker label:

Speaker assignment popup
Speaker assignment

Choose “Add Speaker…” to create a new speaker right in place, or choose a speaker you created before.

The more you assign speakers, the better speaker identification becomes. In a perfect case, the app should have samples of different lengths, intonations, and volumes of each speaker.