白丝美女被狂躁免费视频网站,500av导航大全精品,yw.193.cnc爆乳尤物未满,97se亚洲综合色区,аⅴ天堂中文在线网官网

Systems and methods for voice-assisted media content selection

專利號
US11175880B2
公開日期
2021-11-16
申請人
Sonos, Inc.(US CA Santa Barbara)
發(fā)明人
Sherwin Liu; Paul Bates
IPC分類
G10L15/30; G06F3/16; G10L15/22
技術領域
playback,vas,mps,voice,zone,media,may,content,in,or
地域: CA CA Santa Barbara

摘要

Systems and methods for media playback via a media playback system include (i) capturing a voice input comprising a request for media content, (ii) receiving information derived at least from the request for media content, (iii) requesting and receiving information from at least one remote computing device associated with a first media content service and at least one remote computing device associated with a second media content service, wherein (a) the information identifies first media content available via the first media content service for playback and identifies second media content available via the second media content service for playback, and (b) the first and second media content are related to the requested media content, and (iv) after receiving at least one of the first information and the second information, (a) selecting the first media content instead of the second media content, and (b) playing back the first media content.

說明書

A network microphone device further includes components for detecting and facilitating capture of voice input. For example, the network microphone device 503 shown in FIG. 5A includes beam former components 551, acoustic echo cancellation (AEC) components 552, voice activity detector components 553, and/or wake word detector components 554. In various embodiments, one or more of the components 551-556 may be a subcomponent of the processor 512. The beamforming and AEC components 551 and 552 are configured to detect an audio signal and determine aspects of voice input within the detect audio, such as the direction, amplitude, frequency spectrum, etc. For example, the beamforming and AEC components 551 and 552 may be used in a process to determine an approximate distance between a network microphone device and a user speaking to the network microphone device. In another example, a network microphone device may detective a relative proximity of a user to another network microphone device in a media playback system.

The voice activity detector activity components 553 are configured to work closely with the beamforming and AEC components 551 and 552 to capture sound from directions where voice activity is detected. Potential speech directions can be identified by monitoring metrics which distinguish speech from other sounds. Such metrics can include, for example, energy within the speech band relative to background noise and entropy within the speech band, which is measure of spectral structure. Speech typically has a lower entropy than most common background noise.

權利要求

1
微信群二維碼
意見反饋