Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “speech communication”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Preliminary Design of an 'Autonomous Medical Response Agent' Interface Prototype for Long Duration Spaceflight

Major challenges for astronauts in future long-duration exploration missions (LDEMs) will be that crewmembers are not expected to be medical professionals, may be under high workload and stress, are facing physiological challenges caused by spaceflight, and will have limited, delayed voice communications with medical support from Earth. An autonomous medical response agent (AMRA) is envisioned to help astronauts address medical complaints, develop a differential diagnosis, and guide self-treatment until a healthy state is restored. AMRA develops a process of personalized diagnosis and treatment through a Bayesian predictive control system that recommends therapeutic control actions including diagnostic tests and treatments to crewmembers (Menon, 2020). The Human Computer Interaction (HCI) lab from NASA Ames Research Center’s Human Systems Integration Division (Code TH) has collaborated with Nahlia Inc in human-centered design augmentation research for AMRA. The project, titled Design of ‘Autonomous Medical Response Agent Interface Prototype for Long Duration Spaceflight, has been funded by the Translational Research Institute for Space Health (TRISH) and introduces an interactive user-interface prototype that guides astronauts through self-diagnosis, treatment, and rehabilitation while communicating with remote specialists in ground support (most notably a patient’s flight surgeon). Our project develops the interaction design for the crewmember using AMRA through user research, iterative design, and usability testing to evaluate the user interface and workflow designed. The interface design deliverable for this project, titled AMRA Aggregate Information Display (AMRA AID) is an integrated information display system for comprehensive autonomous medical guidance, diagnosis, and treatment of in-flight medical conditions experienced by crewmembers. AMRA AID demonstrates how we might ensure crew autonomy, increase the crew’s medical capabilities, and decrease cognitive burden within a front-end user interface. AMRA AID refrains from relying on input from ground or mission control for self-treatment of medical issues—though ground awareness and communication with ground is maintained as a means of ensuring trust between mission control and crew. AMRA AID demonstrates how the crew’s on-board medical system might integrate with information from vehicle monitoring and crew schedule, without assuming causal relationships. AMRA AID’s comprehensive view enables efficient information access for both crew and ground support, reducing cognitive burden in the event of an unplanned or emergency medical incident and enabling informed analytical decisions to be made based on both crew and vehicle health. Human-centered design augmentation advanced within the prototype included: enhanced workflow and treatment guidance for two medical scenarios for a non-specialist user base with various levels of medical training, interaction design which considered speech (conversational user interface) elements and on-screen interactions to be developed in future iterations of the project, communication design and functional requirements relevant to self-care versus caring for another astronaut, as well as user testing of the prototype with an international space medical community. This project arrives at critical findings regarding usability needs, communication requirements, and integrated information requirements for a future technology interface functioning to increase confidence between ground support and LDEM crewmembers.

TRISH↗

Audio problems in space.

Audio signal processing techniques in future space exploration, discussing channel capacity, speech processing, bandwidth narrowing, etc

SPACE EXPLORATION↗

Towards Understanding Data Requirements for Developing Automatic Speech Recognition Systems for Air Traffic Control

In recent years, the application of automatic speech recognition has gained popularity across diverse industries, including aviation. Given the many applications focusing on transcribing air traffic control and management communication, this paper explores the training of OpenAI's Whisper model across multiple existing public and private air traffic control voice datasets in an effort to improve robustness. Combining roughly 60+ hours of various air traffic datasets, our goal is to train a unified Whisper model and expect an average word error rate reduction across testing datasets. Furthermore, this work aims to understand the data quantity requirements for achieving state-of-the art results by comprehensively training Whisper on varying dataset sizes. This work has the potential to improve automatic speech recognition performance across the domain, improve understanding of the quantity of data required by an aviation speech recognition system, and lastly provide metrics to compare and improve upon in future research.

ATC↗

Linguistic methodology for the analysis of aviation accidents

A linguistic method for the analysis of small group discourse, was developed and the use of this method on transcripts of commercial air transpot accidents is demonstrated. The method identifies the discourse types that occur and determine their linguistic structure; it identifies significant linguistic variables based upon these structures or other linguistic concepts such as speech act and topic; it tests hypotheses that support significance and reliability of these variables; and it indicates the implications of the validated hypotheses. These implications fall into three categories: (1) to train crews to use more nearly optimal communication patterns; (2) to use linguistic variables as indices for aspects of crew performance such as attention; and (3) to provide guidelines for the design of aviation procedures and equipment, especially those that involve speech.

Goguen, J. A.↗

Hypermedia = hypercommunication

New hardware and software technology gave application designers the freedom to use new realism in human computer interaction. High-quality images, motion video, stereo sound and music, speech, touch, gesture provide richer data channels between the person and the machine. Ultimately, this will lead to richer communication between people with the computer as an intermediary. The whole point of hyper-books, hyper-newspapers, virtual worlds, is to transfer the concept and relationships, the 'data structure', from the mind of creator to that of user. Some of the characteristics of this rich information channel are discussed, and some examples are presented.

Laff, Mark R.↗

Integrate oral communication with technical writing: Towards a rationale

Integrating oral communication and technical writing instruction, to give students the opportunity to learn and practice interpersonal skills, is proposed. By linking speech and writing the importance of small-group interaction in developing transferrable ideas is acknowledged. Three reasons for integration are examined: workday activities, application of role-taking to writing, and conflict resolution. Four advantages of integration are stated.

Skelton, T.↗

The MSAT-X MARECS B2 satellite experiment - Ground segment results

The results of a recently completed satellite experiment employing the JPL MSAT-X developed land-mobile satellite communication terminal are described. In this experiment, a full duplex 4800-b/s digital data and voice communication link was established through the INMARSAT Marecs B2 satellite between Atlantic City, New Jersey, and Southbury, Connecticut. A series of experiments was performed to characterize the terminal performance over this link. The basic experimental setup and the preliminary results of the speech and data experiments are presented. The satellite environment proved to be near to what was expected, and as a result the experimental results were very close to theory/simulation/laboratory experiments. It was found that the ground-to-ground communication links were more benign links than the ground-to-air and air-to-ground links, and this is reflected in the improved margins for the ground-to-ground links (approximately 5 dB versus 3.2 dB for the aeronautical links).

Jedrey, Thomas C.↗

Considerations for Implementing Voice-Controlled Spacecraft Systems Through a Human-Centered Design Approach

As computational power and speech recognition algorithms improve, the consumer market will see better-performing speech recognition applications. The cell phone and Internet-related service industry have further enhanced speech recognition applications using artificial intelligence and statistical data-mining techniques. These improvements to speech recognition technology (SRT) may one day help astronauts on future deep space human missions that require control of complex spacecraft systems or spacesuit applications by voice. Though SRT and more advanced speech recognition techniques show promise, use of this technology for a space application such as vehicle/habitat/spacesuit requires careful considerations. There are still recognition challenges to overcome such as background noise, human speech variability, and task loading. However, implemented correctly, a voice-controlled spacecraft system (VCSS) can provide a useful, natural and efficient form of human-machine communications during complex tasks such as when the hands and eyes are busy or as an aid for vehicle situational awareness inquiries or collaborative human-robot tasks. This document provides considerations and guidance for the use of SRT in VCSS applications for space missions, specifically in command-and-control (C2) applications where the commanding is user-initiated. First, current SRT limitations as known at the time of this report are given. Then, highlights of SRT used in the space program provide the reader with a history of some of the human spaceflight applications and research. Next, an overview of the speech production process and the intrinsic variations of speech are provided. Finally, general guidance and considerations are given for development of a VCSS using a human-centered design approach for space applications that includes vocabulary selection and performance testing, as well as VCSS considerations for C2 dialogue management design, feedback, error handling, and evaluation/usability testing. Available from

Salazar, George A.↗

DAISY-DAMP: A distributed AI system for the dynamic allocation and management of power

One of the critical parameters that must be addressed when designing a loosely coupled Distributed AI SYstem (DAISY) has to do with the degree to which authority is centralized or decentralized. The decision to implement the Dynamic Allocation and Management of Power (DAMP) system as a network of cooperating agents mandated this study. The DAISY-DAMP problem is described; the component agents of the system are characterized; and the communication protocols system elucidated. The motivations and advantages in designing the system with authority decentralized is discussed. Progress in the area of Speech Act theory is proposed as playing a role in constructing decentralized systems.

Hall, Steven B.↗

Communication variations related to leader personality

While communication and captain personality type have been separately shown to relate to overall crew performance, this study attempts to establish a link between communication and personality among 12 crews whose captains represent three pre-selected personality profiles (IE+, I-, Ec-). Results from analyzing transcribed speech from one leg of a full-mission simulation study, so far indicate several discriminating patterns involving ratios of total initiating speech (captain to crew members); in particular commands, questions and observations.

Kanki, Barbara G.↗

Are You Talking to Me? Dialogue Systems Supporting Mixed Teams of Humans and Robots

This position paper describes an approach to building spoken dialogue systems for environments containing multiple human speakers and hearers, and multiple robotic speakers and hearers. We address the issue, for robotic hearers, of whether the speech they hear is intended for them, or more likely to be intended for some other hearer. We will describe data collected during a series of experiments involving teams of multiple human and robots (and other software participants), and some preliminary results for distinguishing robot-directed speech from human-directed speech. The domain of these experiments is Mars-analogue planetary exploration. These Mars-analogue field studies involve two subjects in simulated planetary space suits doing geological exploration with the help of 1-2 robots, supporting software agents, a habitat communicator and links to a remote science team. The two subjects are performing a task (geological exploration) which requires them to speak with each other while also speaking with their assistants. The technique used here is to use a probabilistic context-free grammar language model in the speech recognizer that is trained on prior robot-directed speech. Intuitively, the recognizer will give higher confidence to an utterance if it is similar to utterances that have been directed to the robot in the past.

Dowding, John↗

Speech Recognition Interfaces Improve Flight Safety

"Alpha, Golf, November, Echo, Zulu." "Sierra, Alpha, Golf, Echo, Sierra." "Lima, Hotel, Yankee." It looks like some strange word game, but the combinations of words above actually communicate the first three points of a flight plan from Albany, New York to Florence, South Carolina. Spoken by air traffic controllers and pilots, the aviation industry s standard International Civil Aviation Organization phonetic alphabet uses words to represent letters. The first letter of each word in the series is combined to spell waypoints, or reference points, used in flight navigation. The first waypoint above is AGNEZ (alpha for A, golf for G, etc.). The second is SAGES, and the third is LHY. For pilots of general aviation aircraft, the traditional method of entering the letters of each waypoint into a GPS device is a time-consuming process. For each of the 16 waypoints required for the complete flight plan from Albany to Florence, the pilot uses a knob to scroll through each letter of the alphabet. It takes approximately 5 minutes of the pilot s focused attention to complete this particular plan. Entering such a long flight plan into a GPS can pose a safety hazard because it can take the pilot s attention from other critical tasks like scanning gauges or avoiding other aircraft. For more than five decades, NASA has supported research and development in aviation safety, including through its Vehicle Systems Safety Technology (VSST) program, which works to advance safer and more capable flight decks (cockpits) in aircraft. Randy Bailey, a lead aerospace engineer in the VSST program at Langley Research Center, says the technology in cockpits is directly related to flight safety. For example, "GPS navigation systems are wonderful as far as improving a pilot s ability to navigate, but if you can find ways to reduce the draw of the pilot s attention into the cockpit while using the GPS, it could potentially improve safety," he says.

Source record↗

Checklist interruption and resumption: A linguistic study

This study forms part of a project investigating the relationships among the formal structure of aviation procedures, the ways in which the crew members are taught to execute them, and the ways in which thet are actually performed in flight. Specifically, this report examines the interactions between the performance of checklists and interruptions, considering both interruptions by radio communications and by other crew members. The data consists of 14 crews' performance of a full mission simulation of a higher ratio of checklist speech acts to all speech acts within the span of the performance of the checklist. Further, it is not number of interruptions but length of interruptions which is associated with crew performance quality. Use of explicit holds is also associated with crew performance.

Linde, Charlotte↗

Brahms Mobile Agents: Architecture and Field Tests

We have developed a model-based, distributed architecture that integrates diverse components in a system designed for lunar and planetary surface operations: an astronaut's space suit, cameras, rover/All-Terrain Vehicle (ATV), robotic assistant, other personnel in a local habitat, and a remote mission support team (with time delay). Software processes, called agents, implemented in the Brahms language, run on multiple, mobile platforms. These mobile agents interpret and transform available data to help people and robotic systems coordinate their actions to make operations more safe and efficient. The Brahms-based mobile agent architecture (MAA) uses a novel combination of agent types so the software agents may understand and facilitate communications between people and between system components. A state-of-the-art spoken dialogue interface is integrated with Brahms models, supporting a speech-driven field observation record and rover command system (e.g., return here later and bring this back to the habitat ). This combination of agents, rover, and model-based spoken dialogue interface constitutes a personal assistant. An important aspect of the methodology involves first simulating the entire system in Brahms, then configuring the agents into a run-time system.

Clancey, William J.↗

Advantages of Brahms for Specifying and Implementing a Multiagent Human-Robotic Exploration System

We have developed a model-based, distributed architecture that integrates diverse components in a system designed for lunar and planetary surface operations: an astronaut's space suit, cameras, all-terrain vehicles, robotic assistant, crew in a local habitat, and mission support team. Software processes ('agents') implemented in the Brahms language, run on multiple, mobile platforms. These mobile agents interpret and transform available data to help people and robotic systems coordinate their actions to make operations more safe and efficient. The Brahms-based mobile agent architecture (MAA) uses a novel combination of agent types so the software agents may understand and facilitate communications between people and between system components. A state-of-the-art spoken dialogue interface is integrated with Brahms models, supporting a speech-driven field observation record and rover command system. An important aspect of the methodology involves first simulating the entire system in Brahms, then configuring the agents into a runtime system Thus, Brahms provides a language, engine, and system builder's toolkit for specifying and implementing multiagent systems.

Clancey, William J.↗

Automatic speech recognition in air-ground data link

In the present air traffic system, information presented to the transport aircraft cockpit crew may originate from a variety of sources and may be presented to the crew in visual or aural form, either through cockpit instrument displays or, most often, through voice communication. Voice radio communications are the most error prone method for air-ground data link. Voice messages can be misstated or misunderstood and radio frequency congestion can delay or obscure important messages. To prevent proliferation, a multiplexed data link display can be designed to present information from multiple data link sources on a shared cockpit display unit (CDU) or multi-function display (MFD) or some future combination of flight management and data link information. An aural data link which incorporates an automatic speech recognition (ASR) system for crew response offers several advantages over visual displays. The possibility of applying ASR to the air-ground data link was investigated. The first step was to review current efforts in ASR applications in the cockpit and in air traffic control and evaluated their possible data line application. Next, a series of preliminary research questions is to be developed for possible future collaboration.

Armstrong, Herbert B.↗

Proceedings of the Second International Mobile Satellite Conference (IMSC 1990)

Presented here are the proceedings of the Second International Mobile Satellite Conference (IMSC), held June 17-20, 1990 in Ottawa, Canada. Topics covered include future mobile satellite communications concepts, aeronautical applications, modulation and coding, propagation and experimental systems, mobile terminal equipment, network architecture and control, regulatory and policy considerations, vehicle antennas, and speech compression.

Huck, R. W.↗

Talking Wheelchair

Communication is made possible for disabled individuals by means of an electronic system, developed at Stanford University's School of Medicine, which produces highly intelligible synthesized speech. Familiarly known as the "talking wheelchair" and formally as the Versatile Portable Speech Prosthesis (VPSP). Wheelchair mounted system consists of a word processor, a video screen, a voice synthesizer and a computer program which instructs the synthesizer how to produce intelligible sounds in response to user commands. Computer's memory contains 925 words plus a number of common phrases and questions. Memory can also store several thousand other words of the user's choice. Message units are selected by operating a simple switch, joystick or keyboard. Completed message appears on the video screen, then user activates speech synthesizer, which generates a voice with a somewhat mechanical tone. With the keyboard, an experienced user can construct messages as rapidly as 30 words per minute.

Source record↗