MARATTO

book chapter

Bridging the Gap

Abstract

Human-computer interaction (HCI) has evolved significantly with speech processing technologies, yet a substantial gap remains between natural human communication and machine interaction capabilities. This paper examines speech processing's role in enhancing HCI by analyzing acoustic signal processing, natural language understanding, and user interface design. The study investigates key technologies including automatic speech recognition (ASR), text-to-speech synthesis, and speaker identification across diverse applications such as voice assistants, accessibility tools, gaming environments, and interactive voice response systems. Our analysis reveals that while speech processing has successfully transformed HCI applications, critical issues including acoustic noise sensitivity, limited inclusivity for diverse speaker populations, and the “uncanny valley” effect continue impeding optimal user experience.

Read the original research

This page summarises published work. The authoritative version sits with the publisher.

DOI: 10.4018/979-8-3373-3048-8.ch001

Is something wrong with this record? Report it or request removal.

Discussion

Discuss this research

Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.

No discussion yet. Open the first thread.