3. Voice Is
p. 35-38
Texte intégral
‘αίαι, ay, ay!…stutterer Demosthenes
gob full of pebbles outshouting seas—’Them and [µz]
Tony Harrison, poet
(Harrison, 1978, p. 64)

1Demosthenes, the greatest of the Greek orators, was celebrated for the content, form, and feeling which he brought to his oratory. But when he first began to speak in public, his declamations were mocked. His voice was far from perfect. He had to overcome a speech defect which, it has been deduced, was either a stammer or an inability to correctly pronounce the /r/ sound. He was said to have done this, according to Plutarch (1919, p. 27), by filling his mouth with pebbles whilst practising his vocal delivery. There is another take on this: according to Cicero (Pearson, 1975, p. 214), Demosthenes put stones into his mouth to improve his breathing control. He could not breathe deeply for fear of swallowing the stones. The more well-known story concerning Demosthenes is that he developed the power of his voice in stormy weather by reciting speeches on the seashore, pitting his voice against the noise of the waves.1
2Unlike Demosthenes, most second language learners do not aspire to be great orators, and we think it highly unlikely that many would fill their mouths with stones to improve their vocal quality. However, like the Greek orator, many of them are less than satisfied with the sound which they produce when speaking. This is a dissatisfaction which may have a negative effect on their learning. Even without the inhibiting effect of a specific vocal limitation, many people have an uneasy relationship with their own voice in their native tongue. We all have a problem reconciling ourselves with our voice, which we hear in a recording, and the voice which we hear when we are talking. This is because we are used to hearing much of our own voice resonating through our skull, rather than hearing it through the ear; it is the difference between bone-conducted and air-conducted sound. The recorded voice may sound like the voice of a stranger, particularly with regard to its pitch, which is perceived as being higher in a recording. Many people are not happy with their voice because they feel that it does not serve them well in all that they wish to express. In the same way that they can feel frustrated because they cannot find the right words to say in a certain situation, they may also feel that their voice does not sound right in expressing a certain sentiment or emotion. This discontentment is aggravated by a feeling of resignation; the voice they have is part of their lot in life, and there is not much to be done about it. They are most unlikely to explore radical solutions in the way that Demosthenes did.
3This difficult relationship between the speaker and their voice can become even more strained when trying to reproduce the sounds of a second language. The importance of the effect of the sound of their own voice on the speaker is greatly underestimated in language learning, and this is, we believe, a consequential omission. The more comfortable that a speaker feels with their own voice, the more they will be prepared to express themselves, whether it be in their mother tongue or in another language.
4The reasons for wanting, or needing, to learn a second language will vary from person to person. However, nearly everyone learning another language has in mind the ultimate goal of being able to communicate in that language in the same way that they already do in their native language. Anything less inevitably induces a sense of frustration.
5Aside from cloistered monks and study-bound writers of epic tomes, what people do most in terms of communication in their native language is speak. So it is quite right that the focus of language teaching should be on enabling the learner to speak in the manner of the native speaker. Now that observation, in this age of English as a lingua franca (ELF), may shock some teachers; they may argue that the purpose of ELF teaching is not to allow communication with native or first-language English speakers but to facilitate communication between second-language speakers. That, of course, is true, but a focus on grammar and syntax at the expense of oral communication does not meet the learners’ needs. Too often, oral communication is rated simply in terms of the quality of the vocabulary and grammar employed when speaking. Many English language teachers will freely admit that they are frustrated by the quality of the oral production of their learners, even when the standard of vocabulary and grammar is good, and they confess that they are often at a loss as to how to improve it. They have discovered that simple exhortations like ‘no, say it like this’, ‘speak slower’, ‘speak louder’, ‘articulate more clearly’ are vain utterances in the majority of cases. It is one thing to enhance the grammar and syntax skills of a learner, but quite another to work on the sound of their voice.
6But what do we mean when we use the word ‘voice’? It is important to be precise, particularly when it is used with other words sharing a semantic kinship. So, before we move on, let us establish what we are talking about when referring to voice, speech and language.
Voice
7This is the sound produced when air from the lungs passes through the vocal folds in the larynx. Voice is not always speech. Laughter, giggling, and crying are all vocalised sounds. The children use voice in the form of babbling from very early on, long before they use speech. This babbling period has been identified as being essential in the development of speech in the young child. It is in this process that the baby discovers control over the organs of speech. Proud parents might like to convince themselves that, right from the first sounds being produced, the baby is attempting to imitate the parents. However, although after a few months the baby’s babbling might be initiated or stimulated by the sound of the parent’s voice, the baby is in fact in the process of discovering the mechanism of making sounds, and not all of those sounds will later be incorporated into speech.
8There has always been some debate over the role of babbling. It was questioned whether it was part of a linguistic process or simply an expression of the development of motor skills like chewing or crying, merely ‘mouth-flapping’. The latter would lend credence to the idea that language was almost accidental in its origins. We would like to propose that, ideally, second language learning would also begin with babbling. And with pleasure! The learner would listen to the second language and imitate the language through making sounds. Of course, what was produced would risk being dismissed as nonsense. Perhaps teachers are too focused on getting learners to make sense in a language right from the beginning, rather than exploring the range of its new sounds. As we have already said, there is a strong case for the use of nonsense for pedagogical purposes, which we will be arguing for later.
Speech
9This is the means by which humans express ideas, thoughts, and feelings, and is regarded as one of the most important features of homo loquens in differentiating the species from the rest of the animal kingdom. It requires control over precisely-coordinated muscle actions in the abdomen, chest, neck and head, synchronically with cognitive development, and it takes several years before it is mastered by the growing child. An essential aspect of the process of acquiring speech in the young infant is that of assimilating the phonemes of the language. These phonemes (which are the smallest units of speech sounds, as opposed to voice sounds) have to be integrated into the child’s consciousness long before the child can start to combine the phonemes to make words. These have to be assimilated during a long period of reception before their reproduction can take place. This period of reception, which in the case of second-language learners is named by linguist and educational researcher Steven Krashen as the Silent Period (Krashen, 1982), is a fundamental part of the learning process. The child then experiences for the first time the pleasure of exercising some sort of control over their world through combining these units of speech to make language.
Language
10Language is a formal system of signs governed by grammatical rules of combination to communicate meaning. These are rules of combination of vocabulary, grammar and syntax. In language teaching, it is this formal system, in both its written and oral form, which is evaluated in order to assess the learner.
Evaluating voice
11An evaluation of voice rarely takes place in the language classroom, and its role in determining the quality of speech receives little attention. Any judgement that does take place focuses on an evaluation of pronunciation, that is to say, the way in which phonemes, the sounds of different words, are made. This is referred to linguistically as the segmental feature of the language. However, pronunciation work in English is complicated by the wide range of national and regional dialects, which have to be taken into account in determining the correctness of the sound produced. This means that work on the pronunciation of someone learning American English will be different from that of someone learning British English or Received Pronunciation.
12The poem Them and [µz] (Harrison, 1978) is a wonderful description of how his voice and language were mocked by his English teacher in his Yorkshire school. Harrison went on to read classics at university, and he obtained a degree in linguistics before becoming, as well as a poet, a translator and dramatist. His references in the poem to the voice of his teacher, the language of Shakespeare, Keats and Wordsworth, and the differences in rhyme over time and between cultures illustrate the challenges in evaluating voice.
13English teachers should understand that voice work, which they carry out with their students, will have its benefits in the learner’s first language. If learners are made aware of this, it will give them the incentive to work on their voice, and there will be a knock-on effect on the learner’s overall confidence level in vocal expression.
14We referred above to the pleasure that a young child demonstrates when playing with the sounds of the language during the period of speech acquisition and the notion of comfort related to voice for the language learner. It is this emphasis on the pleasure and comfort aspects which are sometimes missing when it comes to learning a second language. Pleasure in the production is more likely to be experienced by the speaker when they feel that their vocal quality bears some resemblance to that generally recognised as being a feature of the language they are learning. In the case of French learners of English, the complaint most often heard is that their oral production “doesn’t sound English”. Here, they are clearly talking about prosody, the suprasegmental feature of speech, which is often confused with pronunciation. However, prosody, pronunciation and voice quality are obviously closely intertwined. As Stephan Wilhelm (2019) writes:
… voice quality should be included among suprasegmental elements in so far as phonatory and articulatory settings, when conceived of as components of voice quality, fulfil a function that is superposed to that of the phonatory and articulatory components of any acoustic realisation of the English phonemes. (p. 27)
Phonatory and articulatory settings take centre stage early on in THEMPPO workshops, with a focus on breathing, beginning with posture, and exercises on the organs of articulation.
Sound
15It may seem strange to conclude a chapter which has focused on voice by making reference to sound. Surely sound should precede voice? But we bring up sound at this point in order to talk about its absence. The idiom ‘silence is golden’ is commonplace in usage, although the full form of the idiom, ‘speech is silver, silence is golden’, is less heard. But the banality of the expression should not detract from the pointedness of the observation it expresses. Susan Sontag wrote that ‘silence remains, inescapably, a form of speech’ (Sontag, 1969, p. 17). In the same vein, it is well-recognised that in musical composition the silences between notes are as important as the notes themselves. If we bring the two observations together, it would be reasonable to affirm that the silence between words is key to all spoken language. Silence, therefore, merits further discussion and will be the focus of Chapter 6.

Notes de bas de page
1See the painting in the preface to this book.
Le texte seul est utilisable sous licence Creative Commons - Attribution - Pas d'Utilisation Commerciale - Pas de Modification 4.0 International - CC BY-NC-ND 4.0. Les autres éléments (illustrations, fichiers annexes importés) sont « Tous droits réservés », sauf mention contraire.
Méthode Parm : PArler et Raconter en Maternelle
Démarche et outils de la PS à la GS
Laurence Buson, Isabelle Rousset et Solange Rossato
2026
Vocal Revelations: Body and Voice in the Language Classroom
Pedagogical Guide
Marieke de Koning, Chris Mitchell et Rebecca Guy
2026
Outils MIMNA : Médiation de l'Information à destination des Mineurs Non Accompagnés
Guide d'utilisation pour les professionnels
Isabelle Estève et Guillaume Coron
2026
