Example 35: A system comprising a first playback device and a second playback device. The first playback device may comprise one or more processors, a microphone array, and a first computer-readable medium storing instructions that, when executed by the one or more processors, cause the first playback device to perform first operations. The first operations may comprise: detecting sound via the microphone array; transmitting data associated with the detected sound to a second playback device over a local area network. The second playback device may comprise one or more processors and a second computer-readable medium storing instructions that, when executed by the one or more processors, cause the second playback device to perform second operations. The second operations may comprise analyzing, via a wake word engine of the second playback device, the transmitted data associated with the detected sound from the first playback device for identification of a wake word; identifying that the detected sound contains the wake word based on the analysis via the wake word engine; based on the identification, transmitting sound data corresponding to the detected sound to a remote computing device over a wide area network, wherein the remote computing device is associated with a particular voice assistant service; receiving a response from the remote computing device, wherein the response is based on the detected sound; and transmitting a message to the first playback device over the local area network, wherein the message is based on the response from the remote computing device and includes instructions to perform an action. The first computer-readable medium of the first playback device may cause the first playback device to perform the action from the instructions received from the second playback device. Example 36: the system of Example 35, wherein the action is a first action and the second operations further comprise performing a second action via the second playback device, where the second action is based on the response from the remote computing device. Example 37: the system of Example 35 or Example 36, wherein the second operations may further comprise disabling a wake word engine of the first playback device in response to the identification of the wake word via the wake word engine of the second playback device. Example 38: the system of any one of Examples 35 to 37, wherein the second operations may further comprise enabling the wake word engine of the first playback device after the second playback device receives the response from the remote computing device. Example 39: the system of any one of Examples 35 to 38, wherein the first playback device may be configured to communicate with the remote computing device associated with the particular voice assistant service. Example 40: the system of any one of Examples 35 to 39, wherein the remote computing device is a first remote computing device and the voice assistant service is a first voice assistant service, and wherein the first playback device is configured to detect a wake word associated with a second voice assistant service different than the first voice assistant service.