Attention:The NSF Public Access Repository (PAR) system and access will be unavailable from 5:00 PM ET until 8:00 PM ET on Friday, September 11 due to maintenance. We apologize for the inconvenience.


Title: Non-invasive identification of swallows via deep learning in high resolution cervical auscultation recordings
Abstract High resolution cervical auscultation is a very promising noninvasive method for dysphagia screening and aspiration detection, as it does not involve the use of harmful ionizing radiation approaches. Automatic extraction of swallowing events in cervical auscultation is a key step for swallowing analysis to be clinically effective. Using time-varying spectral estimation of swallowing signals and deep feed forward neural networks, we propose an automatic segmentation algorithm for swallowing accelerometry and sounds that works directly on the raw swallowing signals in an online fashion. The algorithm was validated qualitatively and quantitatively using the swallowing data collected from 248 patients, yielding over 3000 swallows manually labeled by experienced speech language pathologists. With a detection accuracy that exceeded 95%, the algorithm has shown superior performance in comparison to the existing algorithms and demonstrated its generalizability when tested over 76 completely unseen swallows from a different population. The proposed method is not only of great importance to any subsequent swallowing signal analysis steps, but also provides an evidence that such signals can capture the physiological signature of the swallowing process.  more » « less
Award ID(s):
1652203
PAR ID:
10341517
Author(s) / Creator(s):
; ;
Date Published:
Journal Name:
Scientific Reports
Volume:
10
Issue:
1
ISSN:
2045-2322
Format(s):
Medium: X
Sponsoring Org:
National Science Foundation
More Like this
  1. null (Ed.)
    Purpose Safe swallowing requires adequate protection of the airway to prevent swallowed materials from entering the trachea or lungs (i.e., aspiration). Laryngeal vestibule closure (LVC) is the first line of defense against swallowed materials entering the airway. Absent LVC or mistimed/shortened closure duration can lead to aspiration, adverse medical consequences, and even death. LVC mechanisms can be judged commonly through the videofluoroscopic swallowing study; however, this type of instrumentation exposes patients to radiation and is not available or acceptable to all patients. There is growing interest in noninvasive methods to assess/monitor swallow physiology. In this study, we hypothesized that our noninvasive sensor-based system, which has been shown to accurately track hyoid displacement and upper esophageal sphincter opening duration during swallowing, could predict laryngeal vestibule status, including the onset of LVC and the onset of laryngeal vestibule reopening, in real time and estimate the closure duration with a comparable degree of accuracy as trained human raters. Method The sensor-based system used in this study is high-resolution cervical auscultation (HRCA). Advanced machine learning techniques enable HRCA signal analysis through feature extraction and complex algorithms. A deep learning model was developed with a data set of 588 swallows from 120 patients with suspected dysphagia and further tested on 45 swallows from 16 healthy participants. Results The new technique achieved an overall mean accuracy of 74.90% and 75.48% for the two data sets, respectively, in distinguishing LVC status. Closure duration ratios between automated and gold-standard human judgment of LVC duration were 1.13 for the patient data set and 0.93 for the healthy participant data set. Conclusions This study found that HRCA signal analysis using advanced machine learning techniques can effectively predict laryngeal vestibule status (closure or opening) and further estimate LVC duration. HRCA is potentially a noninvasive tool to estimate LVC duration for diagnostic and biofeedback purposes without X-ray imaging. 
    more » « less
  2. Aspiration is the most serious complication of dysphagia, which may lead to pneumonia. Detection of aspiration is limited by the presence of its signs like coughing and choking, which may be absent in many cases. High resolution cervical auscultations (HRCA) represent a promising non-invasive method intended for the detection of swallowing disorders. In this study, we investigate the potential of HRCA in detection of penetration-aspiration in patients suspected of dysphagia. A variety of features were extracted from HRCA in both time and frequency domains and they were tested for association with the presence of penetration-aspiration. Multiple classifiers were implemented also for aspiration detection using the extracted signal features. The results showed the presence of strong association between some HRCA signal features and penetration-aspiration, furthermore, they direct towards future directions to enhance prediction capability of aspiration using HRCA signals. 
    more » « less
  3. null (Ed.)
    High-resolution cervical auscultation (HRCA) is an evolving clinical method for noninvasive screening of dysphagia that relies on data science, machine learning, and wearable sensors to investigate the characteristics of disordered swallowing function in people with dysphagia. HRCA has shown promising results in categorizing normal and disordered swallowing (i.e., screening) independent of human input, identifying a variety of swallowing physiological events as accurately as trained human judges. The system has been developed through a collaboration of data scientists, computer–electrical engineers, and speech-language pathologists. Its potential to automate dysphagia screening and contribute to evaluation lies in its noninvasive nature (wearable electronic sensors) and its growing ability to accurately replicate human judgments of swallowing data typically formed on the basis of videofluoroscopic imaging data. Potential contributions of HRCA when videofluoroscopic swallowing study may be unavailable, undesired, or not feasible for many patients in various settings are discussed, along with the development and capabilities of HRCA. The use of technological advances and wearable devices can extend the dysphagia clinician's reach and reinforce top-of-license practice for patients with swallowing disorders. 
    more » « less
  4. Abstract In‐field visual inspections have inherent challenges associated with humans such as low accuracy, excessive cost and time, and safety. To overcome these barriers, researchers and industry leaders have developed image‐based methods for automatic structural crack detection. More recently, researchers have proposed using augmented reality (AR) to interface human visual inspection with automatic image‐based crack detection. However, to date, AR crack detection is limited because: (1) it is not available in real time and (2) it requires an external processing device. This paper describes a new AR methodology that addresses both problems enabling a standalone real‐time crack detection system for field inspection. A Canny algorithm is transformed into the single‐dimensional mathematical environment of the AR headset digital platform. Then, the algorithm is simplified based on the limited headset processing capacity toward lower processing time. The test of the AR crack‐detection method eliminates AR image‐processing dependence on external processors and has practical real‐time image‐processing. 
    more » « less
  5. Year-round recordings of bearded seal calls were collected in the northeastern edge of the Chukchi Continental Slope (Alaska, within the Arctic Circle) in 2016–2017, 2018–2019, and 2019–2020. While the underwater vocalizations of bearded seals are often analyzed manually or using automatic detections manually validated, in this article, a detection and classification system (DCS) based on the You Only Look Once Version 5 (YOLOV5) algorithm is proposed. With YOLOV5, the network learns how to detect and classify these marine mammals’ calls using the principle of computer vision for object detection in images where bounding boxes enclose the objects of interest. During training, validation, and testing, YOLOV5 achieved an accuracy of 96.54%, 93.36%, and 93.87%, respectively. The DCS was applied to the three-yearlong dataset, and an analysis of the vocal behavior of the bearded seals showed that there exists a geographical dependence where this species prefers shallower water depths in the Chukchi Continental Slope. Another advantage of using YOLOV5 over other typical DCS is that the predicted bounding boxes have embedded statistical information about the vocalization, such as the duration, bandwidth, and center frequency of the signals. This additional information equips biologists with statistical data that facilitate the analysis of animal vocal behavior. 
    more » « less