{"id":2606,"date":"2018-05-25T17:00:37","date_gmt":"2018-05-25T17:00:37","guid":{"rendered":"http:\/\/digitalscientists.com\/?p=2606"},"modified":"2023-03-06T21:06:01","modified_gmt":"2023-03-06T21:06:01","slug":"can-you-hear-me-now","status":"publish","type":"post","link":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/","title":{"rendered":"Can you hear me now?"},"content":{"rendered":"\n<h3 class=\"wp-block-heading\">The limitations of voice<\/h3>\n\n\n\n<p>Finally, after decades of watching characters on science fiction movies and television programs tell computers what to do, we have voice-recognition technology. Devices like the Amazon Echo, Google Home, and Apple HomePod allow users to command them to perform a variety of functions, simply by telling them what to do.<\/p>\n\n\n\n<p>It\u2019s game-changing technology, but it\u2019s young.<\/p>\n\n\n\n<p>With any young technology, you\u2019re going to have limitations. Still in its infancy, voice recognition and operation is no exception. Many of these limitations came into sharp focus during a project we recently completed for a client. While this wasn\u2019t our first voice user interface design project (we\u2019ve been doing them for years), it was filled with many new learnings as the technology continues to change.<\/p>\n\n\n\n<p>The product is an<strong><a href=\"https:\/\/digitalscientists.com\/mobile-app-development\/\">&nbsp;iPad app<\/a><\/strong>&nbsp;that serves as a digital store room attendant for facility managers at hotels and apartment complexes. It features a voice-activated inventory management and ordering interface that allows maintenance technicians to ask if a part is available, just as they would to a human attendant.<\/p>\n\n\n\n<p>If the part is available, the app allows the technician to verbally \u201ccheck out\u201d the part and automatically updates the inventory. Read about our work to develop this app&nbsp;<strong><a href=\"https:\/\/digitalscientists.com\/case-studies\/hd-supply\/\">here<\/a>.<\/strong><\/p>\n\n\n\n<p>The project is a proprietary, voice-activated digital assistant. Because we built it from the ground up, without relying on existing voice-activated technology like Alexa (the product was built as a defense against Amazon), we learned a great deal about the limitations of voice technology.<\/p>\n\n\n\n<p>Through the process of developing this app, we learned firsthand some of the key challenges related to voice operation.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">New technology<\/h3>\n\n\n\n<p>The first has to do with the immaturity of the technology and its ability to recognize certain vocal commands. Human speech patterns have nearly endless variation, certainly between different countries and languages, as well as different cities and regions in the United States. Even individuals from the same place can speak different ways.<\/p>\n\n\n\n<p>At this point, there simply isn\u2019t enough information on speech patterns available to computers to be able to process commands from different people. That\u2019s why, for iPhone users, you have to repeat \u201cHey Siri\u201d when you set up your phone. The phone has to \u201clearn\u201d how you talk so it can recognize your voice and speech patterns.<\/p>\n\n\n\n<p>That same setup process is necessary for any voice-controlled device or application. In the case of our store room app, it requires any and all technicians who will use it to input their voice so their commands will be recognized.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">Determining intent<\/h3>\n\n\n\n<p>The second challenge with voice technology is determining intent. When you interact with a software application on a desktop or a touchscreen, intent is simple. You indicate you want the software to perform a function by clicking a button or a link. It\u2019s very clear.<\/p>\n\n\n\n<p>With voice commands, it\u2019s far less clear, in large part because of the speech pattern recognition problem mentioned above. It\u2019s not easy to figure out what the user wants, because people can say the same thing different ways.<\/p>\n\n\n\n<p>This can mean having to go through multiple steps to confirm the user\u2019s desires. In the case of the store room attendant app, a technician may inquire about the available inventory of a part, then vocally request their desired quantity, then confirm they are taking it. That\u2019s a minimum of three voice commands, compared to one click on a screen.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\">Speed<\/h3>\n\n\n\n<p>Those steps to determine intent lead to the third challenge with voice, speed. The amount of time it takes to proceed through multiple layers to confirm intent is an obvious detriment to the speed at which a computer can perform a function.<\/p>\n\n\n\n<p>But processing the commands themselves is also slower. Clicking a button or a link sends a network request through the software, and the computer begins performing the function within nanoseconds. But a voice command can add several seconds to the time it takes for the network request to occur. That\u2019s essentially a lifetime.<\/p>\n\n\n\n<p>The main reason for that again goes back to being able to recognize speech patterns and intent. The computer first has to recognize that it\u2019s being addressed (which is the reason voice assistants are given names, like Alexa and Siri). Then it has to recognize when the command is complete, as indicated by a second or two of silence.<\/p>\n\n\n\n<p>What takes almost no time at all with a screen interface can take five, ten, even 20 seconds or more to complete with a voice assistant.<\/p>\n\n\n\n<p>Clearly,&nbsp;<strong><a href=\"https:\/\/digitalscientists.com\/blog\/how-to-use-voice-now\/\">voice technology<\/a>&nbsp;<\/strong>has a ways to go. But that doesn\u2019t mean companies you should ignore it. Instead, start to learn what potential it has for you.<\/p>\n\n\n\n<p>One of the benefits of today\u2019s technology is it allows you to run tests and learn how it adds value and how your customers might accept it. This means you can experiment with immature technologies, like voice, without deploying it.<\/p>\n\n\n\n<p>When the technology matures, it will happen fast. And you\u2019ll be ready.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Still in its infancy, voice recognition and operation is prone to limitations. Many of these limitations came into sharp focus during a project we recently completed for a client.<\/p>\n","protected":false},"author":4,"featured_media":2278,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"_acf_changed":false,"om_disable_all_campaigns":false,"inline_featured_image":false,"_price":"","_stock":"","_tribe_ticket_header":"","_tribe_default_ticket_provider":"","_tribe_ticket_capacity":"0","_ticket_start_date":"","_ticket_end_date":"","_tribe_ticket_show_description":"","_tribe_ticket_show_not_going":false,"_tribe_ticket_use_global_stock":"","_tribe_ticket_global_stock_level":"","_global_stock_mode":"","_global_stock_cap":"","_tribe_rsvp_for_event":"","_tribe_ticket_going_count":"","_tribe_ticket_not_going_count":"","_tribe_tickets_list":"[]","_tribe_ticket_has_attendee_info_fields":false,"footnotes":""},"categories":[42,45,43],"tags":[],"coauthors":[40],"class_list":["post-2606","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-archived","category-development","category-product"],"acf":[],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v23.5 - https:\/\/yoast.com\/wordpress\/plugins\/seo\/ -->\n<title>Can you hear me now? - Digital Scientists<\/title>\n<meta name=\"description\" content=\"Still in its infancy, voice recognition is prone to limitations. Many of these limitations came into sharp focus during a project we recently completed for a client.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"Can you hear me now? - Digital Scientists\" \/>\n<meta property=\"og:description\" content=\"Still in its infancy, voice recognition is prone to limitations. Many of these limitations came into sharp focus during a project we recently completed for a client.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/\" \/>\n<meta property=\"og:site_name\" content=\"Digital Scientists\" \/>\n<meta property=\"article:published_time\" content=\"2018-05-25T17:00:37+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2023-03-06T21:06:01+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg\" \/>\n\t<meta property=\"og:image:width\" content=\"2560\" \/>\n\t<meta property=\"og:image:height\" content=\"1707\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/jpeg\" \/>\n<meta name=\"author\" content=\"Katie Walters\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Katie Walters\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"4 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\/\/schema.org\",\"@graph\":[{\"@type\":[\"Article\",\"BlogPosting\"],\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#article\",\"isPartOf\":{\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/\"},\"author\":{\"name\":\"Katie Walters\",\"@id\":\"https:\/\/digitalscientists.com\/#\/schema\/person\/b5cc0d863bf3c9792ce6954cf821f97d\"},\"headline\":\"Can you hear me now?\",\"datePublished\":\"2018-05-25T17:00:37+00:00\",\"dateModified\":\"2023-03-06T21:06:01+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/\"},\"wordCount\":852,\"publisher\":{\"@id\":\"https:\/\/digitalscientists.com\/#organization\"},\"image\":{\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#primaryimage\"},\"thumbnailUrl\":\"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg\",\"articleSection\":[\"archived\",\"development\",\"product\"],\"inLanguage\":\"en-US\"},{\"@type\":\"WebPage\",\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/\",\"url\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/\",\"name\":\"Can you hear me now? - Digital Scientists\",\"isPartOf\":{\"@id\":\"https:\/\/digitalscientists.com\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#primaryimage\"},\"image\":{\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#primaryimage\"},\"thumbnailUrl\":\"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg\",\"datePublished\":\"2018-05-25T17:00:37+00:00\",\"dateModified\":\"2023-03-06T21:06:01+00:00\",\"description\":\"Still in its infancy, voice recognition is prone to limitations. Many of these limitations came into sharp focus during a project we recently completed for a client.\",\"breadcrumb\":{\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#primaryimage\",\"url\":\"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg\",\"contentUrl\":\"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg\",\"width\":2560,\"height\":1707,\"caption\":\"Machine Learning Echo Speaker\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\/\/digitalscientists.com\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"Can you hear me now?\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\/\/digitalscientists.com\/#website\",\"url\":\"https:\/\/digitalscientists.com\/\",\"name\":\"Digital Scientists\",\"description\":\"\",\"publisher\":{\"@id\":\"https:\/\/digitalscientists.com\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\/\/digitalscientists.com\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Organization\",\"@id\":\"https:\/\/digitalscientists.com\/#organization\",\"name\":\"Digital Scientists\",\"url\":\"https:\/\/digitalscientists.com\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/digitalscientists.com\/#\/schema\/logo\/image\/\",\"url\":\"https:\/\/digitalscientists.com\/wp-content\/uploads\/2022\/11\/header-logo-light.svg\",\"contentUrl\":\"https:\/\/digitalscientists.com\/wp-content\/uploads\/2022\/11\/header-logo-light.svg\",\"width\":165,\"height\":48,\"caption\":\"Digital Scientists\"},\"image\":{\"@id\":\"https:\/\/digitalscientists.com\/#\/schema\/logo\/image\/\"}},{\"@type\":\"Person\",\"@id\":\"https:\/\/digitalscientists.com\/#\/schema\/person\/b5cc0d863bf3c9792ce6954cf821f97d\",\"name\":\"Katie Walters\",\"url\":\"https:\/\/digitalscientists.com\/blog\/author\/katiewalters\/\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"Can you hear me now? - Digital Scientists","description":"Still in its infancy, voice recognition is prone to limitations. Many of these limitations came into sharp focus during a project we recently completed for a client.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/","og_locale":"en_US","og_type":"article","og_title":"Can you hear me now? - Digital Scientists","og_description":"Still in its infancy, voice recognition is prone to limitations. Many of these limitations came into sharp focus during a project we recently completed for a client.","og_url":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/","og_site_name":"Digital Scientists","article_published_time":"2018-05-25T17:00:37+00:00","article_modified_time":"2023-03-06T21:06:01+00:00","og_image":[{"width":2560,"height":1707,"url":"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg","type":"image\/jpeg"}],"author":"Katie Walters","twitter_card":"summary_large_image","twitter_misc":{"Written by":"Katie Walters","Est. reading time":"4 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":["Article","BlogPosting"],"@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#article","isPartOf":{"@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/"},"author":{"name":"Katie Walters","@id":"https:\/\/digitalscientists.com\/#\/schema\/person\/b5cc0d863bf3c9792ce6954cf821f97d"},"headline":"Can you hear me now?","datePublished":"2018-05-25T17:00:37+00:00","dateModified":"2023-03-06T21:06:01+00:00","mainEntityOfPage":{"@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/"},"wordCount":852,"publisher":{"@id":"https:\/\/digitalscientists.com\/#organization"},"image":{"@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#primaryimage"},"thumbnailUrl":"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg","articleSection":["archived","development","product"],"inLanguage":"en-US"},{"@type":"WebPage","@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/","url":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/","name":"Can you hear me now? - Digital Scientists","isPartOf":{"@id":"https:\/\/digitalscientists.com\/#website"},"primaryImageOfPage":{"@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#primaryimage"},"image":{"@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#primaryimage"},"thumbnailUrl":"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg","datePublished":"2018-05-25T17:00:37+00:00","dateModified":"2023-03-06T21:06:01+00:00","description":"Still in its infancy, voice recognition is prone to limitations. Many of these limitations came into sharp focus during a project we recently completed for a client.","breadcrumb":{"@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#primaryimage","url":"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg","contentUrl":"https:\/\/digitalscientists.com\/wp-content\/uploads\/2020\/09\/ML3-scaled-1.jpg","width":2560,"height":1707,"caption":"Machine Learning Echo Speaker"},{"@type":"BreadcrumbList","@id":"https:\/\/digitalscientists.com\/blog\/can-you-hear-me-now\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/digitalscientists.com\/"},{"@type":"ListItem","position":2,"name":"Can you hear me now?"}]},{"@type":"WebSite","@id":"https:\/\/digitalscientists.com\/#website","url":"https:\/\/digitalscientists.com\/","name":"Digital Scientists","description":"","publisher":{"@id":"https:\/\/digitalscientists.com\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/digitalscientists.com\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/digitalscientists.com\/#organization","name":"Digital Scientists","url":"https:\/\/digitalscientists.com\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/digitalscientists.com\/#\/schema\/logo\/image\/","url":"https:\/\/digitalscientists.com\/wp-content\/uploads\/2022\/11\/header-logo-light.svg","contentUrl":"https:\/\/digitalscientists.com\/wp-content\/uploads\/2022\/11\/header-logo-light.svg","width":165,"height":48,"caption":"Digital Scientists"},"image":{"@id":"https:\/\/digitalscientists.com\/#\/schema\/logo\/image\/"}},{"@type":"Person","@id":"https:\/\/digitalscientists.com\/#\/schema\/person\/b5cc0d863bf3c9792ce6954cf821f97d","name":"Katie Walters","url":"https:\/\/digitalscientists.com\/blog\/author\/katiewalters\/"}]}},"ticketed":false,"_links":{"self":[{"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/posts\/2606"}],"collection":[{"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/users\/4"}],"replies":[{"embeddable":true,"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/comments?post=2606"}],"version-history":[{"count":0,"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/posts\/2606\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/media\/2278"}],"wp:attachment":[{"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/media?parent=2606"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/categories?post=2606"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/tags?post=2606"},{"taxonomy":"author","embeddable":true,"href":"https:\/\/digitalscientists.com\/wp-json\/wp\/v2\/coauthors?post=2606"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}