Journal of Iranian Medical Council

Journal of Iranian Medical Council

Word-In-Noise Perception Test in Adults

Document Type : Original article

Authors
1 Department of Audiology, School of Rehabilitation Sciences, Shahid Beheshti University of Medical Sciences, Tehran, Iran
2 Department of Pediatrics, School of Medicine, Besat Hospital, Hamadan University of Medical Sciences, Hamadan, Iran
3 Research Center for Health Sciences, Hamadan University of Medical Sciences, Hamadan, Iran
4 Department of Audiology, School of Rehabilitation Sciences, Hamadan University of Medical Sciences, Hamadan, Iran
Abstract
Background: It is essential to have a test that creates the best competitive conditions for evaluating meaning perception. Therefore, this study aimed to investigate the psychometric properties of the Word-In-Noise Perception test (WINP) using Homotonic-Monosyllabic Words (HMWs) and Spectrum Speech Noise (SSN) for Persian speakers aged 18 to 22 years.
Method: This study was a test-development type that was conducted in a cross-sectional-comparative way. Assessments included 110 young Persian speakers. The evaluations included checking the state of general health, sleep and mental states, basic audiological evaluations, dichotic digit test and WINP. Mann-Whitney and Friedman tests were used to compare group scores.
Results: The validity and reliability of WINP using HMWs and SSN were determined. The mean Conversion Rate (CVR) and mean Content Validity Index (CVI) of the HMWs were 0.99 and 0.95, respectively. The minimum acceptable reliability (ICC) was also calculated for single and mean measurements. The results showed that the obtained scores are stable and without measurement errors. Normal values of WINP were obtained using HMWs and SSN. Gender was not an effective factor in causing differences in mean WINP scores. There was no significant difference between WINP mean scores in Signal-to-Noise Ratios (SNR) in different noises between the right and left ears. Also, there was a significant difference between mean scores of WINP in SNRs from -5 to +15 for left and right ears.
Conclusion: The psychometric properties of the WINP using HMWs and SNN have been confirmed for Persian speaking adults.
Keywords
Subjects

Introduction
Since everyday conversations and social communication are conducted in competitive situations, the most important human need for the sense of hearing is speech-in-noise perception (1). Word-In-Noise Perception test (WINP) using Homotonic-Monosyllabic Words (HMWs) along with the use white noise was recently developed (2-5). WINP consists of 6 lists, each list containing 25 HMWs, whose structural pattern is consonant/vowel/consonant and the vowel of each list is the same. Since there are 6 vowels in Persian, there are 6 HMW lists in this test (Appendix 1). However, considering that the signal strength of white noise is evenly distributed over time, white noise cannot be a suitable masker for the speech signal (2). For this reason, the use of Speech Spectrum Noise (SSN) prevents the excessive masking of the speech signal and increase the accuracy of the WINP (2,4).
Speech processing is carried out with the participation of the auditory brain and the brainstem. The brainstem is sensitive to the fundamental frequency of sound (f0), which is the psychoacoustical equivalent of speech vowels, and can be observed in the frequency following response. This is why musicians who are exposed to more f0 are better at understanding speech in noise than normal people. 
In general, the brainstem is most sensitive in detecting musical rhythm and speech pitch, which is done based on the tonotopic map. The auditory brain also monitors and controls the neural signals received from the brainstem as a correction and coordination center, and its parts even participate in pitch processing based on the tonotopic map. There is another processing in the auditory brain called the phonemotopic mapping, which is not sensitive to f0 in this function. 
In fact, the auditory brain is sensitive to the frequency difference between the first and second formant of vowels for processing the phonemes and detecting the consonants, and it is the f1-f2 connection processor that understands the meaning and concepts. As a result, a sentence uttered by different speakers (woman, man and child) conveys the same meaning to the listener (1-6). Accordingly, if HMWs are presented to the individual in order and sequence, they cause the brainstem to have the least participation in speech processing and the neural function to focus on recognizing speech consonants by the auditory brain (3-6). It seems that WINP using SSN and HMWs can be a valid measure in the fields of prevention, diagnosis and rehabilitation of hearing disorders. For this reason, this study aimed to investigate the psychometric properties of the WINP using HMWs and SSN for Persian speakers aged 18 to 22 years.

Materials and Methods
Setting 
This was a cross-sectional, comparative, experimental study. The location of the sound recording was the acoustic recording studio of the Hamadan University of Medical Sciences, the Fourier analysis of HMWs and psychoacoustics measurement of the recorded audio file were performed and the practical work was carried out at the Department of Audiology of Hamadan University of Medical Sciences.
Inclusion criteria included high school level, age less than 22 years and more than 18 years. Normal conditions in general health, nocturnal sleep, hearing thresholds, middle ear function, acoustic reflexes, word-in-noise discrimination, dichotic digit and pitch pattern scores, and no history of the following: head injury, ototoxicity, metabolic and cardiovascular diseases, anemia, diabetes, hypothyroidism, hyperthyroidism, medication use psychedelic, nervous system diseases and addiction. 
Exclusion criteria comprised unwillingness to continue participating in the research, inappropriate cooperation, and insufficient attention from participants.Participants included 110 Persian-speaking young people (61 males and 49 females) with a mean age of 20 (0.56) years. The ethical considerations of this research were met by maintaining privacy, respecting the safety of participants, and ensuring that the assessments were free of charge.
The main outcome variables were obtained after explaining the practical work steps and obtaining written consent to participate in the study, a 28-item general health questionnaire, the Petersburg Sleep Quality Index, and a brief psychological status examination among the participants (131 people). Those who had normal results in this stage (n=118) underwent audiological tests, which included acoustic immittance measurement (by clarinet middle ear immittance device), pure tone audiometry, speech reception threshold and word-in-noise discrimination (by Interacoustic AC33 audiometer), dichotic digit and pitch pattern tests (with the audio files recorded on a laptop). Of the 118 participants, 110 had normal results on the mindfulness-based assessments, which were considered our core sample size and were assessed by the WINP using HMWs and SSN. 
In order to assess the content validity, HMWs lists were provided to 20 Persian linguists. After reviewing the responses, the levels of Content Validity Ratio (CVR) and Content Validity Index (CVI) were determined for each of the items related to specificity, simplicity and clarity for the word list. Words with CVR<0.7 and or CVR<0.42 were excluded. Psychoacoustics measurement of MMWs were performed using Fourier analysis. Then, in a studio equipped with the highest quality audio equipment, HMWs were recorded in silence with the voice of a female Persian speaker without accent. HMWs with similar frequency bandwidth and energy level were selected and non-HMWs were eliminated. The generated SSN was speech-filtered white noise with a flat spectrum in the frequency range of 250 to 1000 Hz and an energy reduction of 6 dB/octave from 1000 to 6000 Hz (7). Finally, SSN was added to the recorded voice at Signal-to-Noise Ratios (SNRs) of -5, 0, +5, +10, and +15 dB. 
The WINP practical task was conducted using HMWs and SSN at the participants’ comfortable listening level (3). The recorded sound was presented to each ear separately through a high-quality headphone connected to the computer. Participants were asked to respond quickly, with a gap of 2000 ms between words. If there was no response, the next trial automatically began 2000 ms after the last word was played (4). 
The interval between trials of each 25 HMW list was 1000 ms. In other words, there was a 1000 ms pause after the last word in each list was presented and then the first word from the second list was presented after 3000 ms. To calculate the WINP norm, the number of correct answers was multiplied by 4, resulting in a percentage score (4,5). 
WINP reliability using HMWs and SSN was calculated in terms of repeatability across assessment times. Repeatability during testing was assessed using Intra-class Correlation Coefficient (ICC). WINP was assessed twice for participants, with a two-week interval between them, to avoid participants’ memory loss, and all six lists of HMWs were presented to participants at each assessment. The median of all 6 lists in each ear was considered as the norm of that ear, and the median of the entire right and left ear was considered as the test norm. The values of the calculated indices in the two repetitions were completely identical and the correlation value was equal to one. 

Data analysis 
The analysis of the findings was performed using Stata17 software from StataCorp. Quantitative variables included median, Interquartile Range (IQR), mean, and standard deviation. Descriptive statistics were Kolmogorov-Smirnov and Shapiro-Wilk tests were used to examine the normality of data distribution. The Mann-Whitney and Friedman tests were used to compare the differences between scores. All statistical analyses were performed at a significance level of 0.05.
 
Results
In this study, the validity and reliability of WINP test using HMWs and SSN were determined in 110 Persian speakers. The mean CVR and mean CVI of the remaining words were equal to 0.99 and to 0.95, respectively. ICC was also calculated on individual and average measurements. An ICC close to 1 indicates greater validity of WPS and better consistency of HMWs across lists. The results showed that the obtained scores are stable and free from measurement errors. Normal WINP values were obtained using HMWs and SSN (Table 1).
The results of the Mann-Whitney test showed that there was no significant difference between the mean WINP at different SNRs between the right and left ears, given that the p-value of the test was greater than 0.05. Also, the results of Friedman’s test showed that considering that the p-value of the test was less than 0.001 and there was a significant difference between the mean WINP for SNRs of -5 to +15 in the left and right ears (Table 2). 
Since there were significant differences between SNRs, pairwise comparisons of SNRs were performed by Friedman’s supplementary test (Table 3). Also, no significant difference was observed between the mean WINP scores by gender (p=0.989). 

 Appendix 1. Lists of Phonetically balanced monosyllabic words of Iranian Persian language with the arrangement of homotonic-monosyllabic words for word-in-noise perception test, based on the Iranian Persian international transliteration alphabet *

Vowel

/ uː / 

Vowel

/ æ / 

Vowel

/ /

Vowel

/ iː /

Vowel

/ o /

Vowel

/ e /

n

ɡuːʃ

Sær

Kɒːr

Siːb

Ʃol

Ʃen

1

Muːʃ

Dær

Bɒːr

Ʃiːb

Pol

Sen

2

Duːʃ

Kær

Mɒːr

Dʒiːb

ɡol

Dʒen

3

Nuːʃ

Ʃær

Qɒːr

Siːr

Kol

ɡel

4

Huːʃ

Pær

Zɒːr

Piːr

Khol

Del

5

Dʒuːʃ

Sær

Jɒːr

Diːr

Qom

Hel

6

Tʃuːb

Xær

Hɒːr

Ʃiːr

ɡom

Sel

7

Muːr

Nær

Dɒːr

Qiːr

Xom

Vel

8

Ʃuːr

Tær

Nɒːr

Ziːr

Dom

Ʒel

9

Suːr

ɡær

Sɒːr

Dʒiːr

Som

Beh

10

Duːr

Sæm

Xɒːl

Miːr

Ʃok

Meh

11

Zuːr

Næm

Sɒːl

Tiːr

Nok

Deh

12

Kuːr

m

Mɒːl

Miːʃ

Ʃod

Leh

13

Dʒuːr

Gæm

Bɒːl

Niːʃ

Xod

Deq

14

Nuːr

Dæm

Zɒːl

Piːʃ

Por

Neq

15

Buːr

Bæm

Kɒːl

Kiːʃ

Sor

Nej

16

Tuːr

Kæm

Ʃɒːl

Biːʃ

Ʃor

Dej

17

Duːd

Qæm

Tʃɒːl

Fiːl

Lor

Keʃ

18

Suːd

Jæx

Nɒːm

Miːl

Lop

Ʃeʃ

19

Ruːd

Ʃæb

Ʃɒːm

Biːl

Boz

Fer

20

Kuːd

Tæb

Dɒːm

Siːx

Hoz

Qer

21

Buːd

Læb

Dʒɒːm

Miːx

Moz

Kez

22

Zuːd

Ʃæk

Kɒːm

Niːm

Motʃ

Vez

23

Buːq

Tæk

Vɒːm

Biːm

Xoʃ

Mes

24

Duːq

Sæɡ

Rɒːm

Siːm

Ʃoʃ

Hes

25

* Emami SF. The use of homotonic monosyllabic words in the Persian language for the word-in-noise perception test. Aud Vestib Res. 2024. 33(1):28–33. doi.org/10.18502/avr.v33i1.14271.

Table 1. Reliability of the word-in-noise perception test (WINP) using homotonic monosyllabic words (HMWs) and speech spectrum noise (SSN) for Persian speakers aged 18 to 22 (n=110)

Second test score

First test score

 

*ICC2

 

Intraclass correlation coefficient limit (0.95)

 

*ICC1

 

Intraclass correlation coefficient limit (0.95)

SNR

SD

Mean

M

IQR

SD

Mean

M

IQR

Lower limit

Upper limit

Lower limit

Upper limit

1.03

55

42

24.45

1.93

53

40

25.41

0.96

0.80

1.00

0.92

0.83

0.98

-5

2.89

68

54

26.89

2.22

69

52

28.36

0.92

0.78

0.99

0.90

0.78

0.99

0

3.74

83

64

22.68

2.43

82

66

23.18

0.93

0.79

0.98

0.95

0.79

0.98

+5

2.63

87

74

26.34

1.50

90

72

27.65

0.92

0.81

0.96

0.94

0.86

0.96

+10

1.08

95

97

24.96

1.44

95

99

26.47

0.95

0.85

1.00

0.96

0.88

1.00

+15

SNR; signal-to-noise ratio, Interquartile range: IQR, Median or norm values: M, Standard deviation: SD, *ICC1: In mean of measurements, *ICC2: In one measurement.

Table 2. Comparing of mean scores of the word-in-noise perception test (WINP) using homotonic monosyllabic words (HMWs) and speech spectrum noise (SSN) for Persian speakers aged 18 to 22 years at different signal-to-noise ratio (SNR) by ear

** p

SNR

Ear

+15

+10

+5

0

-5

M

IQR

M±SD

M

IQR

M±SD

M

IQR

M±SD

M

IQR

M±SD

M

IQR

M±SD

<0.001

98

25.13

94±5.08

71

27.91

90±6.63

62

21.98

82±5.74

55

27.91

69±4.89

40

23.12

54±6.03

Right

<0.001

95

23.62

96±3.90

75

25.87

89±7.09

65

23.80

83±4.88

53

25.34

68±5.61

43

26.19

53±6.24

Left

 

0.063

0.830

0.080

0.433

0.346

* p

*Mann-Whitney test, ** Friedman tests, SNR: signal-to-noise ratio.

Table 3. pairwise Comparing of mean scores of the word-in-noise perception test (WINP) using homotonic monosyllabic words (HMWs) and speech spectrum noise (SSN) for Persian speakers aged 18 to 22 years at different signal-to-noise ratio (SNR)

SNR

Right ear

10-15

5-10

0-5

-5-0

<0.001

<0.001

<0.001

<0.001

p-value

10-15

5-10

0-5

-5-0

Left ear

<0.001

<0.001

<0.001

<0.001

p-value

SNR: signal-to-noise ratio.

Discussion 
By comparing the mean and standard deviation values of WINP using SSN for Persian speakers aged 18 to 22 years (at the SNRs of -5 dB=53±1.93,  0 dB=69±2.22,  +5 dB=82±2.43, +10dB=90±1.50 and +15 dB=95±1.44) with the values obtained from the WINP using white noise for Persian speakers aged 18 to 22 years (at the SNRs of 0 dB=53±14.80, +5 dB=68.15±13.20, and +10 dB=88±12.60) in a recent study (4), it can be seen that the mean WINP using SSN is higher in both ears and the standard deviations are lower. 
Previous researchers have reported that the high standard deviation of WINP using white noise in their studies (3-5) may be due to the occurrence of Central Auditory Processing Disorder (CAPD) in young adults (4), elderly (3) and patients with renal failure (5). In this study, it can be assumed that the lower limit of the mean WINP scores using SSN (in SNRs of -5 dB=48%, 0 dB=44%, +5 dB=64%, +10 dB=68% and +15 dB=84% probably occurs as a hidden impairment that cannot be detected.
Audiological tests used to diagnose CAPD are unable to localize lesions due to the simultaneous involvement of central auditory centers in various auditory processing functions (1,2,6). For example, word-in-noise discrimination, dichotic digit, and pitch pattern tests that can detect abnormalities of the auditory brain and brainstem (2). Various tests are usually used to diagnose CAPD, and since the WINP test minimizes the participation of the brainstem in vowel recognition, it can be one of the appropriate tests to assessment the function of the auditory brain in recognizing of speech consonants and lexical-semantic processing (3-5). 
In silent conditions, the left hemisphere of the auditory brain prefers verbal stimuli (6,8). The left frontal pole, left dorsolateral prefrontal cortex, and right posterior parietal lobe are activated in any function related to attentional mechanisms, especially speech perception (9,10). In the presence of noise, each hemisphere works with the participation of all cortical areas (11,12). Two broad lexical-semantic processing streams have been described for speech-in-noise perception: Bilaterally and largely symmetric streams, organized in superior temporal gyrus. The dorsal stream or posterior part of the superior temporal gyrus is for sensory-motor interaction, which is left-dominant and involves structures at the parietal-temporal-frontal junction (13). 
In the process of speech-in-noise perception, the intraparietal sulcus is activated and the planum temporale also acts as a computational center and plays an important role in distinguishing two simultaneous speech signals. The left insula cortex processes fundamentals frequency sounds containing semantic information, and the right insula cortex processes fundamentals frequency sounds containing phonological information. The ventral premotor cortex also enhances the ability to recognize vowels, especially when the signal-to-noise ratio is reduced. Mirror neurons in the medial frontal cortex, which are active in imitative functions, are also involved in speech-in-noise perception (2,14,15).
When SNR decreases, and speech-in-noise perception in becomes difficult, brain areas associated with semantic processing and speech production are recruited and techniques that improve listeners’ understanding are employed, such as speaking louder (2,10). Wernicke’s area’s principal function is speech perception, but it also participates in production. The main role of Broca’s area’s is in speech production, but it also plays a role in perception (2,9). When someone reads a text aloud, word imagery is created in the occipital cortex. Each word is formed in a spatial format, that is different for each person. Then, visual signals generated from the occipital cortex are transmitted to the angular gyrus and on to Wernicke’s area (16).   
Regardless of whether CAPD and other auditory impairments are unilateral or bilateral, speech-in-noise perception impairments appear to be due to damages to the superior temporal lobe of the language-dominant hemisphere, suggesting that speech inputs are processed asymmetrically at early stages (2,8). Bilateral superior temporal lobe dysfunction causes word deafness, a condition in which pure tone hearing thresholds are within normal limits. But the ability to speech perception is virtually nil. Damage to posterior temporal lobe regions is the strongest predictive of speech perception deficits (in aphasia) (12).
Damage to the posterior portion of superior temporal gyrus (in certain forms of aphasia) causes speech perception deficits. However, these conditions involve only mild phonological impairments (13). Anterior temporal lobe regions have also been implicated in sentence-level lexical-semantic processing. Patients with semantic dementia show bilateral anterior temporal lobe atrophy along with deficits in lexical tasks such as naming, semantic association, and single-word comprehension (14). The anterior and inferior temporal lobes appear to be more involved in the process of meaning-making, while the anterior temporal lobe is involved in integrating specific forms of semantic knowledge across different dimensions (17).  
Damage to posterior temporal lobes, particularly along the middle temporal gyrus has been associated with speech perception deficits (14). It is well-established that auditory inputs can have rapid and automatic effects on speech production. For example, delayed auditory feedback of one’s own voice disrupts speech fluency. Adult-onset deafness is associated with decreased articulation, suggesting that auditory feedback is important in maintaining articulation inflection (15,17).  
Damage to the left dorsal posterior superior temporal gyrus is associated with production disorders (12). In particular, such damage is associated with conduction aphasia, a syndrome usually caused by stroke and characterized by good speech perception, but frequent phonological errors in production, naming difficulties, and difficulty with word-for-word repetition (16). 
Finally, since patients with CAPD and various types of peripheral hearing loss have difficulty speech-in-noise perception (2,10), it is essential to perform WINP using SSN. 

Conclusion 
Psychometric properties of the WINP have been confirmed using HMWs and SNN for Persian-speaking adults.

Ethics approval and consent to participate 
The study was approved by the research ethics committee of Hamadan University of Medical Sciences (Code: IR. IR.UMSHA.REC.1402.018). Informed written consent to participate in the study was provided by all participants.

Funding 
This study was financially supported by Hamadan University of Medical Sciences (Grant Number: 140205173928).

Acknowledgement
The authors would like to thank the participants for their cooperation in this study.

Conflict of Interest 
The authors declare that they have no conflicts of interests.

1. Emami SF. Is all human hearing cochlear? ScientificWorldJournal 2013 Dec 19;2013:147160.  https://pubmed.ncbi.nlm.nih.gov/24453793/
2. Emami SF, Shariatpanahi E, Gohari N, Mehrabifard M. Word-in-noise perception test in children. The Egyptian Journal of Otolaryngology 2024 Jun 25;40(1):64.
3. Emami SF, Shariatpanahi E, Gohari N, Mehrabifard M. Aging and Speech-in-Noise Perception. Indian J Otolaryngol Head Neck Surg 2023 Sep;75(3):1579-1585. https://pubmed.ncbi.nlm.nih.gov/37636642/
4. Emami SF. The use of homotonic monosyllabic words in the Persian language for the word-in-noise perception test. Aud Vestib Res 2024;33(1):28-33.
5. Emami SF, Momtaz HE, Mehrabifard M. Central Auditory Processing Impairment in Renal Failure. Indian J Otolaryngol Head Neck Surg 2024 Feb;76(1):1010-1013. https://pubmed.ncbi.nlm.nih.gov/38440591/
6. Emami SF. Acoustic sensitivity of the saccule and daf music. Iran J Otorhinolaryngol 2014 Apr;26(75):105-10. https://pubmed.ncbi.nlm.nih.gov/24744999/
7. Ebrahimi A, Mahdavi ME, Jalilvand H. Auditory recognition of Persian digits in presence of speech-spectrum noise and multi-talker babble: a validation study. Auditory and Vestibular Research 2020;29(1):39-47.
8. Chowsilpa S, Bamiou DE, Koohi N. Effectiveness of the Auditory Temporal Ordering and Resolution Tests to Detect Central Auditory Processing Disorder in Adults With Evidence of Brain Pathology: A Systematic Review and Meta-Analysis. Front Neurol 2021 Jun 2;12:656117. https://pubmed.ncbi.nlm.nih.gov/34149594/
9. Moulin A. Ear Asymmetry and Contextual Influences on Speech Perception in Hearing-Impaired Patients. Front Neurosci 2022 Mar 18;16:801699. https://pubmed.ncbi.nlm.nih.gov/35368258/
10. Scott SK, Sinex DG, 2010. Speech. In: Rees A, Palmer A, editors. The Oxford handbook of auditory science: the auditory brain. 1st ed. New York: Oxford university press: 193-214.
11. Hickok G, Okada K, Barr W, Pa J, Rogalsky C, Donnelly K, et al. Bilateral capacity for speech sound processing in auditory comprehension: evidence from Wada procedures. Brain Lang 2008 Dec;107(3):179-84. https://pubmed.ncbi.nlm.nih.gov/18976806/
12. Vaden KI Jr, Kuchinsky SE, Ahlstrom JB, Dubno JR, Eckert MA. Cortical activity predicts which older adults recognize speech in noise and when. J Neurosci 2015 Mar 4;35(9):3929-37. https://pubmed.ncbi.nlm.nih.gov/25740521/
13. Tremblay P, Perron M, Deschamps I, Kennedy-Higgins D, Houde JC, Dick AS, et al. The role of the arcuate and middle longitudinal fasciculi in speech perception in noise in adulthood. Hum Brain Mapp 2019 Jan;40(1):226-241. https://pubmed.ncbi.nlm.nih.gov/30277622/
14. Humphries C, Willard K, Buchsbaum B, Hickok G. Role of anterior temporal cortex in auditory sentence comprehension: an fMRI study. Neuroreport 2001 Jun 13;12(8):1749-52.
15. Tremblay P, Brisson V, Deschamps I. Brain aging and speech perception: Effects of background noise and talker variability. Neuroimage 2021 Feb 15;227:117675. https://pubmed.ncbi.nlm.nih.gov/33359849/
16. Hickok G. The functional neuroanatomy of language. Phys Life Rev 2009 Sep;6(3):121-43. https://pmc.ncbi.nlm.nih.gov/articles/PMC2747108/
17. Hickok G, Poeppel D. The cortical organization of speech processing. Nature Reviews Neuroscience 2007 May;8(5):393-402.