US 12,394,414 B2
Enterprise type models for voice interfaces
Austin Chang Ming Liu, Burnaby (CA); Gonzalo Alvarez Barrio, Seattle, WA (US); Gregory James Wade, Coquitlam (CA); Harsh Agarwal, Seattle, WA (US); Sam Anthony Sullivan, Lions Bay (CA); Chieh Chien, Bellevue, WA (US); and Cameron L Chinn, Seattle, WA (US)
Assigned to Amazon Technologies, Inc., Seattle, WA (US)
Filed by Amazon Technologies, Inc., Seattle, WA (US)
Filed on Dec. 16, 2022, as Appl. No. 18/082,786.
Prior Publication US 2024/0203412 A1, Jun. 20, 2024
Int. Cl. G06N 20/00 (2019.01); G10L 15/06 (2013.01); G10L 15/22 (2006.01)
CPC G10L 15/22 (2013.01) [G10L 15/063 (2013.01)] 20 Claims
OG exemplary drawing
 
1. A system, comprising:
one or more processors; and
non-transitory computer-readable media storing computer-executable instructions that, when executed by the one or more processors, cause the one or more processors to perform operations comprising:
receiving, from a device associated with an enterprise, first input data indicating an enterprise type of the enterprise, the enterprise type indicating an industry of the enterprise;
receiving, from a voice interface device, audio data representing a user utterance including a request that an action be performed, the voice interface device associated with account data identifying the voice interface device as being associated with the enterprise;
determining that the account data is associated with an enterprise type model configured to identify a plurality of enterprise applications to service user requests, wherein the enterprise type model is generated based on the enterprise type of the enterprise;
retrieving the enterprise type model from a model repository;
applying the enterprise type model to the user utterance to identify a first enterprise application from the plurality of enterprise applications;
ranking the first enterprise application with other candidate applications to service the request;
selecting the first enterprise application to service the request;
generating, utilizing the first enterprise application, a speech processing result associated with the action to be performed; and
sending, to the voice interface device, a directive to cause the action to be performed.