Guide

Speaker Diarization on iPhone: Identify Who Said What

Loro 7 min

Learn how LoroNote automatically separates speakers in meetings and interviews, then rename speakers and correct individual or multiple segments on iPhone and iPad.

A transcript tells you what was said. In a meeting, interview, or group discussion, that is only half the job—you also need to know who said it.

Speaker diarization solves that problem by dividing a recording into segments and assigning each segment to a speaker. LoroNote performs this analysis directly on your iPhone or iPad, then presents the transcript with color-coded speaker labels and timestamps.

This guide walks through the complete workflow: detecting speakers, reviewing the result, replacing generic labels with names, and correcting any segment assigned to the wrong person.

The quick answer

Open a completed recording, switch to Show Segments, and use the speaker button to identify participants. You can let LoroNote detect the number automatically or choose a known count from 2 to 8.

What you want to doWhere to do it
Detect speakers automaticallySpeaker menu → Auto Detect
Set a known number of participantsSpeaker menu → choose 2–8 Speakers
Rename, add, or merge speaker labelsSpeakers identifiedEdit
Fix one segmentUse the edit control beside that transcript segment
Move several segments to another personEdit → select segments → Change Speaker

Before identifying speakers

You can use speaker diarization with a recording made in LoroNote or with an imported audio file. For the cleanest result:

  • Place the iPhone or iPad where every participant can be heard clearly.
  • Avoid covering the microphone or putting the device in a pocket.
  • Ask participants not to speak over one another when possible.
  • If you already know the exact number of speakers, use that number instead of automatic detection.

Speaker diarization works best when each person has enough uninterrupted speech for the model to learn the differences between their voices.

1. Choose automatic detection or a speaker count

LoroNote toolbar with the speaker diarization icon between refresh and more

Open the recording and tap the speaker icon in the top toolbar. The Identify Speakers menu offers two approaches:

  • Auto Detect is the easiest option when you do not know how many people are in the recording.
  • 2–8 Speakers tells LoroNote the exact number to look for. Use this for a scheduled interview, panel, or meeting with a known attendance count.

Choosing a count can reduce unnecessary labels when background voices or short interruptions might otherwise be mistaken for another participant.

LoroNote Identify Speakers menu with automatic detection and options for two through eight speakers
Let LoroNote choose automatically, or provide a known speaker count from 2 to 8.

2. Review the color-coded transcript

After analysis, the Speakers identified summary shows how many voices were found. In Show Segments, each speaker receives a color and every section includes its speaker label and timestamp.

Tap a segment to jump to that exact moment in the audio. This is the fastest way to verify a boundary: listen to the surrounding few seconds and confirm that the label changes when the voice changes.

A LoroNote transcript divided into blue and purple speaker segments with timestamps
Speaker colors and timestamps make a long conversation easier to scan and replay.

3. Replace Speaker 1 with real names

Generic labels are useful during detection, but names make a transcript much easier to read and share.

Tap Edit beside Speakers identified to open the speaker list. From here you can:

  1. Tap a speaker name and replace it with the participant’s real name or role.
  2. Add a missing participant with Add Speaker.
  3. Remove an unnecessary speaker. Removing one merges its segments into the previous speaker, so review the affected sections afterward.
  4. Tap Done to save the list.

For recurring meetings, descriptive labels such as “Interviewer,” “Client,” or “Presenter” can be clearer than names alone.

LoroNote Edit Speaker Names sheet with three speakers and an Add Speaker button
Rename participants, add a missing speaker, or merge an unnecessary label.

4. Correct the speaker assigned to a segment

Even a strong diarization result may need a quick correction around interruptions, laughter, or overlapping speech.

For a single mistake, use the edit control beside that transcript segment and assign the correct speaker. Play the audio from its timestamp if you need to confirm the voice first.

To correct several segments at once:

  1. Switch from Read to Edit.
  2. Select every segment that belongs to the same person.
  3. Tap Change Speaker.
  4. Choose the correct speaker from the list.

Bulk correction is especially useful when one participant was split into two labels throughout a long recording.

Individual transcript segments in LoroNote with timestamp and edit controls
Edit a single segment when only one label is wrong.
Two selected transcript segments and the Change Speaker button in LoroNote
Select multiple segments and move them to another speaker in one action.

Tips for more reliable speaker separation

  • Prioritize clean audio. Distance and room noise affect speaker separation as well as transcription accuracy.
  • Use the known count when possible. A fixed count gives the model a useful constraint.
  • Check transitions first. Mistakes are most likely where one person interrupts another or two voices overlap.
  • Listen from the timestamp. A few seconds of audio is usually enough to confirm the correct speaker.
  • Name speakers before sharing. “Alex” and “Product Lead” are more useful to readers than “Speaker 1.”
  • Use bulk editing for repeated errors. Do not correct the same split speaker one segment at a time.

What speaker diarization can—and cannot—do

LoroNote can distinguish different voices and group their speech, but it does not automatically know anyone’s real-world identity. Detection begins with neutral labels such as Speaker 1 and Speaker 2; you decide what names to assign.

Overlapping speech remains the hardest case for any diarization system. If two people talk at exactly the same time, the transcript may need a small manual correction. That is why timestamps, direct audio playback, and both single and bulk reassignment are part of the workflow.

As with transcription, speaker analysis is performed on-device. Your recording does not need to be uploaded to a transcription server, and the feature remains useful for private meetings and interviews.

Frequently asked questions

Can LoroNote identify speakers in an imported audio file? Yes. Import the file, open its transcript, and run speaker identification in the same way as a recording made in LoroNote.

Should I use Auto Detect or enter the number myself? Use a fixed number when you know exactly how many people spoke. Use Auto Detect when the count is uncertain or changes during the recording.

How many speakers can I choose manually? The speaker menu provides fixed choices from 2 through 8, plus automatic detection.

Can I rename speakers after detection? Yes. Open the speaker list with Edit and replace the generic labels with names or roles at any time.

What if several segments have the wrong speaker? Switch to Edit, select all affected segments, and use Change Speaker to reassign them together.

Does speaker identification require the internet? No. LoroNote processes the audio on your device, just like its offline transcription workflow.

Turn your voice into text — offline

LoroNote transcribes meetings, lectures, and interviews right on your iPhone — private, accurate, and unlimited.

Download on the App Store