官术网_书友最值得收藏!

Speech

Adding one of the Speech APIs allows your application to hear and speak to your users. The APIs can filter noise and identify speakers. Based on the recognized intent, they can drive further actions in your application.

The speech domain contains three APIs that are outlined in the following sections.

Bing Speech

Adding the Bing Speech API to your application allows you to convert speech to text and vice versa. You can convert spoken audio to text either by utilizing a microphone or other sources in real time or by converting audio from files. The API also offers speech intent recognition, which is trained by the Language Understanding Intelligent Service (LUIS) to understand the intent.

Speaker recognition

The speaker recognition API gives your application the ability to know who is talking. By using this API, you can verify that the person that is speaking is who they claim to be. You can also determine who an unknown speaker is based on a group of selected speakers.

Translator speech API

The translator speech API is a cloud-based automatic translation service for spoken audio. Using this API, you can add end-to-end translation across web apps, mobile apps, and desktop applications. Depending on your use cases, it can provide you with partial translations, full translations, and transcripts of the translations cover all speech-related APIs in Chapter 5, Speak with Your Application.

主站蜘蛛池模板: 涟水县| 泊头市| 聂拉木县| 资溪县| 盈江县| 临沭县| 册亨县| 衡阳市| 沐川县| 古蔺县| 宁城县| 无极县| 丹巴县| 分宜县| 玉田县| 玛纳斯县| 云南省| 从江县| 岑巩县| 凯里市| 天全县| 靖远县| 郓城县| 新营市| 泰和县| 甘德县| 北宁市| 伊川县| 西乡县| 巴马| 高密市| 新晃| 泉州市| 龙口市| 天长市| 白银市| 城市| 腾冲县| 天水市| 谷城县| 禹城市|