Tuesday, February 9, 2010

Voxeo’s Announcement

Today’s announcement that IVR provider Voxeo will partner with four voice biometrics vendors (CSIdentity, Persay, TradeHarbor and Vocalect) has promise. It offers 100,000 telephony/Web developers, who use Voxeo’s IVR developer portal, the opportunity to develop creative voice biometrics applications for their clients. This move has the potential to mainstream the use of voice biometrics in organizations large and small. The one thing I’m unclear about is how the developers are going to differentiate between the voice biometrics applications offered by the four vendors. Time will tell.

Monday, February 8, 2010

Speech Recognition vs. Speaker Recognition

People often get confused between speech recognition and speaker recognition. Let’s explore the difference.

Simply stated, speech recognition is the ability to recognize what a person is saying. Speaker recognition on the other hand is the ability to identify the person speaking.


Speech recognition has been in use for decades. Today, many organizations have replaced live agents with automated interactive voice response (IVR) systems. The IVRs may prompt users to speak their entries. For example, when calling directory assistance, users are prompted to speak the name of the individual or business that they are inquiring about. The system converts the spoken phrases to text, recites them back to the user for verification, and searches the database for a match.


Speaker recognition is a relatively new technology. Using voice biometrics, a speaker recognition system is able to identify, or verify the identity, of a person who has previously enrolled (i.e., provided a base voiceprint) in the system. A voiceprint is a stored set of measurable characteristics of a human voice that are unique to each individual. A sample voiceprint can be compared to a base voiceprint to identify, or verify the identity, of a person. Speaker recognition is an ideal way to verify the identity of a person remotely (e.g., performing a financial transaction via telephone).

Sunday, February 7, 2010

Speaker Verification vs. Speaker Identification

Voice biometrics applications can be classified as speaker verification or speaker identification. Let’s examine the difference.

Speaker verification is a one-to-one process of confirming the identity of an enrolled person. The person provides a sample voiceprint which is compared to their base voiceprint in the database. If the voiceprints match, the person’s identity is verified (accepted), if they don’t match, the person’s identity is not verified (rejected). Most commercial applications are classified as speaker verification.


Speaker identification is a one-to-many process of determining the identity of an unknown person. The person provides a sample voiceprint which is compared to all of the base voiceprints in the database. If there is a match, the person is identified; if there is no match, the person is not identified. Speaker identification has been successfully used in law enforcement and intelligence applications.