Speaker recognition that never leaves your device. Runs AntResearch/AntSpeaker (MECT-B2) in your browser.
Record three short clips of the same person, about four seconds each. Any words work.
Finds where people speak, fingerprints every 1.5 s of speech and groups the fingerprints by voice. Known voices from Voice ID get named automatically.
Each person needs one clip of a few seconds. Adding more clips to the same name sharpens their voiceprint.