Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Speech Intelligibility”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Acoustic Issues in Human Spaceflight

NASA is concerned about acute effect of sound on crew performance on International Space Station (ISS), and is developing strategies to assess and reduce acute, chronic, and delayed effects of sound. High noise levels can cause headaches, irritation, fatigue, impaired sleep, headache, and tinnitus and have resulted in an inability to hear alarms. Speech intelligibility may be more impaired for crew understanding non-native language in a noisy environment. No hearing loss occurred, but significant effects on crew performance and communication occurred. Permanent Threshold Shifts (PTS) have not been observed in the US shuttle program. Russian specification for noise in spacecraft is 60 dBA (awake) and 50 dBA (asleep) while the U.S. noise specification on ISS is NC 50 (awake) and NC 40 (asleep) with a 85 dBA hazard limit. Background noise levels of ISS modules have measured 56-69 dBA. Treadmill exercise operations measure 77 dBA. Alarms are required to be 20 dBA above ambient. Hearing protection is recommended when noise exceeds 60 dB 24 hour Leq. Countermeasures include hearing protection and design/ engineering controls. Advanced composite materials with excellent low frequency attenuation properties could be applied as a barrier protection around noisy equipment, or used on personal protective equipment worn by the crew. Hearing protection countermeasures include foam ear inserts, passive muff headsets, and active noise reduction headsets. Oto-acoustic emissions (OAE) could be used to monitor effectiveness of hearing protection countermeasures and tailor hearing protection countermeasures to individual crewmembers. Micro-gravity, vibration, toxic fumes, air quality/composition, stress, temperature, physical exertion or some combination of the above factors may have interacted with moderate long-term noise exposure to cause significant hearing loss. Longitudinal studies will need to address what co-morbidity factors, such as radiation, toxicology, microgravity effects (fluid shift), aging, are involved with hearing loss.

Clark, Jonathan B.↗

Cabin Noise Studies for the Orion Spacecraft Crew Module

Controlling cabin acoustic noise levels in the Crew Module (CM) of the Orion spacecraft is critical for adequate speech intelligibility, to avoid fatigue and to prevent any possibility of temporary and permanent hearing loss. A vibroacoustic model of the Orion CM cabin has been developed using Statistical Energy Analysis (SEA) to assess compliance with acoustic Constellation Human Systems Integration Requirements (HSIR) for the on-orbit mission phase. Cabin noise in the Orion CM needs to be analyzed at the vehicle-level to assess the cumulative acoustic effect of various Orion systems at the crewmember's ear. The SEA model includes all major structural and acoustic subsystems inside the CM including the Environmental Control and Life Support System (ECLSS), which is the primary noise contributor in the cabin during the on-orbit phase. The ECLSS noise sources used to excite the vehicle acoustic model were derived using a combination of established empirical predictions and fan development acoustic testing. Baseline noise predictions were compared against acoustic HSIR requirements. Key noise offenders and paths were identified and ranked using noise transfer path analysis. Parametric studies were conducted with various acoustic treatment packages in the cabin to reduce the noise levels and define vehicle-level mass impacts. An acoustic test mockup of the CM cabin has also been developed and noise treatment optimization tests were conducted to validate the results of the analyses.

Dandaroy, Indranil↗

Analysis of Noise Exposure Measurements Made Onboard the International Space Station

The International Space Station (ISS) is a unique workplace environment for U.S. astronauts and Russian cosmonauts to conduct research and live for a period of six months or more. Noise has been an enduring environmental physical hazard that has been a challenge for the U.S. space program since before the Apollo era. Noise exposure in ISS poses significant risks to the crewmembers, such as; hearing loss (temporary or permanent), possible disruptions of crew sleep, interference with speech intelligibility and communication, possible interference with crew task performance, and possible reduction in alarm audibility. Acoustic measurements are made aboard ISS and compared to requirements in order to assess the acoustic environment to which the crewmembers are exposed. The purpose of this paper is to describe in detail the noise exposure monitoring program as well as an assessment of the acoustic dosimeter data collected to date. The hardware currently being used for monitoring the noise exposure onboard ISS will be discussed. Acoustic data onboard ISS has been collected since the beginning of ISS (Increment 1, November 2000). Noise exposure data analysis will include acoustic dosimetry logged data from crew-worn during work and sleep periods and also fixed-location measurements from Increment 1 to present day. Noise exposure levels (8-, 16- and 24-hr), LEQ, will also be provided and discussed in this paper. Discussions related to hearing protection will also be included. Future directions and recommendations for the noise exposure monitoring program will be highlighted. This acoustic data is used to ensure a safe and healthy working and living environment for the crewmembers aboard the ISS.

Limardo, Jose G.↗

Analysis of Noise Exposure Measurements Acquired Onboard the International Space Station

The International Space Station (ISS) is a unique workplace environment for U.S. astronauts and Russian cosmonauts to conduct research and live for a period of six months or more. Noise has been an enduring environmental physical hazard that has been a challenge for the U.S. space program since before the Apollo era. Noise exposure in ISS poses significant risks to the crewmembers, such as; hearing loss (temporary or permanent), possible disruptions of crew sleep, interference with speech intelligibility and communication, possible interference with crew task performance, and possible reduction in alarm audibility. Acoustic measurements were made onboard ISS and compared to requirements in order to assess the acoustic environment to which the crewmembers are exposed. The purpose of this paper is to describe in detail the noise exposure monitoring program as well as an assessment of the acoustic dosimeter data collected to date. The hardware currently being used for monitoring the noise exposure onboard ISS will be discussed. Acoustic data onboard ISS has been collected since the beginning of ISS (Increment 1, November 2001). Noise exposure data analysis will include acoustic dosimetry logged data from crew-worn dosimeters during work and sleep periods and also fixed-location measurements from Increment 1 to present day. Noise exposure levels (8-, 16- and 24-hr), LEQ, will also be provided and discussed in this paper. Future directions and recommendations for the noise exposure monitoring program will be highlighted. This acoustic data is used to ensure a safe and healthy working and living environment for the crewmembers onboard the ISS.

Limardo, Jose G.↗

International Space Station Noise Constraints Flight Rule Process

Crewmembers onboard the International Space Station (ISS) live in a unique workplace environment for as long as 6 ‐12 months. During these long‐duration ISS missions, noise exposures from onboard equipment are posing concerns for human factors and crewmember health risks, such as possible reductions in hearing sensitivity, disruptions of crew sleep, interference with speech intelligibility and voice communications, interference with crew task performance, and reduced alarm audibility. The purpose of this poster is to describe how a recently‐updated noise constraints flight rule is being used to implement a NASA‐created Noise Exposure Estimation Tool and Noise Hazard Inventory to predict crew noise exposures and recommend when hearing protection devices are needed.

Limardo, Jose G.↗

Status: Crewmember Noise Exposures on the International Space Station

The International Space Station (ISS) provides a unique environment where crewmembers from the US and our international partners work and live for as long as 6 to 12 consecutive months. During these long-durations ISS missions, noise exposures from onboard equipment are posing concerns for human factors and crewmember health risks, such as possible reductions in hearing sensitivity, disruptions of crew sleep, interference with speech intelligibility and voice communications, interference with crew task performance, and reduced alarm audibility. It is crucial to control acoustical noise aboard ISS to acceptable noise exposure levels during the work-time period, and to also provide a restful sleep environment during the sleep-time period. Acoustic dosimeter measurements, obtained when the crewmember wears the dosimeter for 24-hour periods, are conducted onboard ISS every 60 days and compared to ISS flight rules. NASA personnel then assess the acoustic environment to which the crewmembers are exposed, and provide recommendations for hearing protection device usage. The purpose of this paper is to provide an update on the status of ISS noise exposure monitoring and hearing conservation strategies, as well as to summarize assessments of acoustic dosimeter data collected since the Increment 36 mission (April 2013). A description of the updated noise level constraints flight rule, as well as the Noise Exposure Estimation Tool and the Noise Hazard Inventory implementation for predicting crew noise exposures and recommending to ISS crewmembers when hearing protection devices are required, will be described.

Limardo-Rodriguez, Jose G.↗

Acoustically Tailored Composite Rotorcraft Fuselage Panels

A rotorcraft roof sandwich panel has been redesigned to optimize sound power transmission loss (TL) and minimize structure-borne sound for frequencies between 1 and 4 kHz where gear meshing noise from the transmission has the most impact on speech intelligibility. The roof section, framed by a grid of ribs, was originally constructed of a single honeycomb core/composite face sheet panel. The original panel has coincidence frequencies near 700 Hz, leading to poor TL across the frequency range of 1 to 4 kHz. To quiet the panel, the cross section was split into two thinner sandwich subpanels separated by an air gap. The air gap was sized to target the fundamental mass-spring-mass resonance of the double panel system to less than 500 Hz. The panels were designed to withstand structural loading from normal rotorcraft operation, as well as 'man-on-the-roof' static loads experienced during maintenance operations. Thin layers of VHB 9469 viscoelastomer from 3M were also included in the face sheet ply layups, increasing panel damping loss factors from about 0.01 to 0.05. Measurements in the NASA SALT facility show the optimized panel provides 6-11 dB of acoustic transmission loss improvement, and 6-15 dB of structure-borne sound reduction at critical rotorcraft transmission tonal frequencies. Analytic panel TL theory simulates the measured performance quite well. Detailed finite element/boundary element modeling of the baseline panel simulates TL slightly more accurately, and also simulates structure-borne sound well.

Hambric, Stephen↗

Development of Acoustic Mufflers for Cabin Noise Reduction in Orion Spacecraft

Controlling cabin acoustic noise levels in the Crew Module (CM) of the Orion spacecraft is critical to ensure adequate speech intelligibility, to avoid fatigue, and prevent any possibility of temporary and permanent hearing loss to the crew. The primary source of cabin noise for the on-orbit phase of the mission is from the Environmental Control and Life Support System (ECLSS) which recycles and conditions breathing air and maintains cabin pressurization through its duct network and components. Unfortunately, as a side effect, noise from the ECLSS fans propagates through theses ducts and emanate into the cabin habitable volume via the ECLSS inlet and outlets. To mitigate excessive duct-borne noise, two ECLSS mufflers have been designed to provide significant acoustic transmission loss (TL) so that the cabin noise requirements can be met. Each muffler is meant to be installed in the ducting of the ECLSS air inlet and outlet sides, respectively. Packaging constraints and tight volume requirements necessitated the mufflers to be of complex geometry and compatible with the bends of the ECLSS duct layout. To design and characterize the acoustic performance of the inlet and outlet mufflers, computational acoustic models were developed using the Finite Element Method (FEM) with 𝑤𝑎𝑣𝑒଺ vibroacoustic software. Characterization of the acoustic material and perforations in the mufflers were addressed with poroelastic theory. Once the mufflers were designed on paper and its TL predicted, prototypes of these mufflers were created using additive manufacturing. The muffler prototypes were subsequently tested for acoustic TL in the laboratory with various con-figurations of acoustic materials. Comparing the analytical predictions to the test performance yielded excellent correlation for acoustic TL and demonstrated significant broadband noise attenuation. The ECLSS mufflers are currently scheduled to be installed on the Artemis II Crew Module (CM) of the Orion spacecraft and will provide significant cabin comfort to crew during the mission.

Indranil Dandaroy↗

Talking Wheelchair

Communication is made possible for disabled individuals by means of an electronic system, developed at Stanford University's School of Medicine, which produces highly intelligible synthesized speech. Familiarly known as the "talking wheelchair" and formally as the Versatile Portable Speech Prosthesis (VPSP). Wheelchair mounted system consists of a word processor, a video screen, a voice synthesizer and a computer program which instructs the synthesizer how to produce intelligible sounds in response to user commands. Computer's memory contains 925 words plus a number of common phrases and questions. Memory can also store several thousand other words of the user's choice. Message units are selected by operating a simple switch, joystick or keyboard. Completed message appears on the video screen, then user activates speech synthesizer, which generates a voice with a somewhat mechanical tone. With the keyboard, an experienced user can construct messages as rapidly as 30 words per minute.

Source record↗

Space-Station-Interior Noise-Analysis Program

Intelligibility of speech evaluated for specified acoustical environments. Program makes systematic prediction of noise and vibration environment of craft defined by user and evaluates relative acceptability of predicted environment for effective communication by speech. Written in MicroSoft FORTRAN Version 3.3.

Stusnick, Eric↗

Synthesized speech rate and pitch effects on intelligibility of warning messages for pilots

In civilian and military operations, a future threat-warning system with a voice display could warn pilots of other traffic, obstacles in the flight path, and/or terrain during low-altitude helicopter flights. The present study was conducted to learn whether speech rate and voice pitch of phoneme-synthesized speech affects pilot accuracy and response time to typical threat-warning messages. Helicopter pilots engaged in an attention-demanding flying task and listened for voice threat warnings presented in a background of simulated helicopter cockpit noise. Performance was measured by flying-task performance, threat-warning intelligibility, and response time. Pilot ratings were elicited for the different voice pitches and speech rates. Significant effects were obtained only for response time and for pilot ratings, both as a function of speech rate. For the few cases when pilots forgot to respond to a voice message, they remembered 90 percent of the messages accurately when queried for their response 8 to 10 sec later.

Simpson, C. A.↗

Speech communications in noise

The physical characteristics of speech, the methods of speech masking measurement, and the effects of noise on speech communication are investigated. Topics include the speech signal and intelligibility, the effects of noise on intelligibility, the articulation index, and various devices for evaluating speech systems.

Source record↗

Laboratory and in-flight experiments to evaluate 3-D audio display technology

Laboratory and in-flight experiments were conducted to evaluate 3-D audio display technology for cockpit applications. A 3-D audio display generator was developed which digitally encodes naturally occurring direction information onto any audio signal and presents the binaural sound over headphones. The acoustic image is stabilized for head movement by use of an electromagnetic head-tracking device. In the laboratory, a 3-D audio display generator was used to spatially separate competing speech messages to improve the intelligibility of each message. Up to a 25 percent improvement in intelligibility was measured for spatially separated speech at high ambient noise levels (115 dB SPL). During the in-flight experiments, pilots reported that spatial separation of speech communications provided a noticeable improvement in intelligibility. The use of 3-D audio for target acquisition was also investigated. In the laboratory, 3-D audio enabled the acquisition of visual targets in about two seconds average response time at 17 degrees accuracy. During the in-flight experiments, pilots correctly identified ground targets 50, 75, and 100 percent of the time at separation angles of 12, 20, and 35 degrees, respectively. In general, pilot performance in the field with the 3-D audio display generator was as expected, based on data from laboratory experiments.

Ericson, Mark↗

Voice technology and BBN

The following research was discussed: (1) speech signal processing; (2) automatic speech recognition; (3) continuous speech understanding; (4) speaker recognition; (5) speech compression; (6) subjective and objective evaluation of speech communication system; (7) measurement of the intelligibility and quality of speech when degraded by noise or other masking stimuli; (8) speech synthesis; (9) instructional aids for second-language learning and for training of the deaf; and (10) investigation of speech correlates of psychological stress. Experimental psychology, control systems, and human factors engineering, which are often relevant to the proper design and operation of speech systems are described.

Wolf, Jared J.↗

The design of a device for hearer and feeler differentiation, part A

A speech modulated white noise device is reported that gives the rhythmic characteristics of a speech signal for intelligible reception by deaf persons. The signal is composed of random amplitudes and frequencies as modulated by the speech envelope characteristics of rhythm and stress. Time intensity parameters of speech are conveyed through the vibro-tactile sensation stimuli.

Creecy, R.↗

Machine Learning for Predicting Team Functioning in HERA Missions

Team functioning is integral to success in future long term space exploration missions. Proactively detecting declines in team functioning can mitigate conflict and ensure mission success. This project developed a speech-based artificial intelligence (AI) system that unobtrusively predicts degradation in team functioning, including performance and cohesion, in the Human Exploration Research Analog (HERA) Campaigns 4 and 5. The AI system conducted automated analysis of the prosodic (tone of voice) and linguistic (language content) components of speech, modeling interpersonal dynamics at both the turn-taking and day-wide levels. We investigated team functioning via observing structured interactions (i.e., multi-mission space exploration vehicle-extra vehicular activity [MMSEV-EVA], team interaction battery [TIB]) and unstructured interactions before the MMSEV-EVA task. We developed machine learning models to predict team functioning (objective task accuracy, self reported team efficacy and self reported team cohesion) by analyzing OpenSmile acoustic features, linguistic descriptors extracted via the linguistic inquiry and word count (LIWC) dictionary, and semantic embeddings. In the TIB, static models using logistic regression and random forests were not able to predict task accuracy, but predicted team efficacy and cohesion during both the decision making and relational tasks to a moderate level (60-70%). Majority voting on the individual turns to predict day long team efficacy further increased accuracies (70-80%). Finally, long short-term memory (LSTM) models showed the best performance across all variables (80-91%), including task performance. In the MMSEV-EVA, static models achieved an accuracy of 60% with majority voting, which increased to 80% through the incorporation of mission day as a variable, accounting for the learning effect. A key finding across both tasks was the "team-dependent" nature of these interactions; models achieved much higher accuracy when trained on prior days of the same team's data rather than attempting to generalize across entirely different teams, with even 1-2 days of prior data per team achieving 5-15% improvement over team-independent models. In addition, the incorporation of pre-task data from the same team also improves model performance, e.g., incorporating data from the decision-making task of the TIB, which preceded the relational task, improved the prediction of team efficacy and cohesion during the latter. We compared model performance when trained on machine-generated data compared to data that had been further corrected by human annotators. Overall, models trained on human-corrected data exhibited a modest improvement in performance, particularly when acoustic features were used. We found no significant correlation between word error rate (WER) and model accuracy (r(55) = -0.08, p = 0.51), but model’s accuracy was significantly higher for medium/high quality transcription (0.74 (SD = 0.48)) compared to the low-quality group (0.64 (SD = 0.36)) (t(63)=2.82, p = 0.006). Based on these, several design recommendation emerge, that could inform Standards at NASA. Models predicting team functioning should incorporate at least one to two days of historical interaction data, include brief pre-task discussions, and explicitly model temporal learning effects, especially for longer operational tasks. Minimum quality standards for automated speech-processing pipelines are needed, given the performance gains observed with manually corrected acoustic data. Finally, systems should leverage both acoustic features and language embeddings in complementary ways, with modality choices and fusion strategies tailored to mission context, task demands, and data quality requirements.

Shrivatsa Mishra↗

Predicting Team Functioning in Long Term Space Missions Using Acoustic and Linguistic Measures

Maintaining optimal team functioning is critical for long-duration space exploration missions, yet traditional monitoring methods, such as self-reports and wearable sensors, often impose operational burdens or suffer from bias. This paper investigates a non-intrusive speech-based artificial intelligence (AI) framework to predict degradations in team functioning using data from the Human Exploration Research Analog (HERA) of the U.S. National Aeronautics and Space Administration (NASA). Using acoustic features, linguistic descriptors, and semantic embeddings, we evaluate static non-linear and temporal machine learning models to predict both objective (task accuracy) and subjective (self-reported efficacy and cohesion) team functioning outcomes. Results indicate that temporal models outperform static approaches, with prediction of objective task accuracy in Team Interaction Battery (TIB) improving from near chance to 71%. Self-reported outcomes, including team efficacy and cohesion, are predicted more reliably than task performance, achieving balanced accuracies of up to 85.56% and 78.12%, respectively, and are found to be most strongly associated with acoustic features. In a second interdependent task, the MMSEV–EVA, accuracies of up to 78% are achieved using temporal models with acoustic features. Furthermore, incorporating just 1–2 days of team-specific historical data systematically improved performance, and acoustic markers from informal pre-task interactions provided modest predictive gains. Finally, while automated preprocessing yielded viable accuracy, humancorrected data provided moderate performance gains, though transcription error rates did not significantly correlate with model performance. These findings highlight the potential of speech as a passive, high-fidelity monitoring tool for autonomous habitats.

Temporal modeling↗

Beyond the sterile cockpit

Consideration is given to some of the negative aspects of the trend toward increased automation of aircraft flight decks. The history of automated devices for navigation, communications and detection on board aircraft is reviewed. Instances of automatic system failure are identified which have led to accidents, and the events surrounding the downing of Korean Airlines Flight 747 are reexamined within the context of a computer-based system failure. Finally, new software and interactive systems to reduce navigational error due to inadequate computer-assisted flight instruction (CAI) are described, with emphasis given to speech processing and intelligent CAI systems.

Wiener, E. L.↗