Summary
Imported recordings with multiple speakers can produce significant speaker misattribution even when the source does not contain isolated channels that could be lost during downmixing.
Steps to reproduce
- Import a plain stereo recording containing two speakers.
- Let import transcription and diarization complete.
- Compare the assigned speaker IDs with the recording.
Actual behavior
Speech is attributed to the wrong speaker. A one-speaker control imported correctly, while a two-speaker plain-stereo file reproduced the problem. A separate four-person import folded much of one person's speech into the other identities.
Expected behavior
File-import diarization should distinguish speakers with accuracy comparable to live-recorded audio.
Notes
This is separate from the confirmed channel-collapse issue: #5. The two-speaker reproduction had no isolated per-speaker tracks to lose, so unconditional downmixing does not fully explain it.
Reported by Joey Virrueta.
Summary
Imported recordings with multiple speakers can produce significant speaker misattribution even when the source does not contain isolated channels that could be lost during downmixing.
Steps to reproduce
Actual behavior
Speech is attributed to the wrong speaker. A one-speaker control imported correctly, while a two-speaker plain-stereo file reproduced the problem. A separate four-person import folded much of one person's speech into the other identities.
Expected behavior
File-import diarization should distinguish speakers with accuracy comparable to live-recorded audio.
Notes
This is separate from the confirmed channel-collapse issue: #5. The two-speaker reproduction had no isolated per-speaker tracks to lose, so unconditional downmixing does not fully explain it.
Reported by Joey Virrueta.