# Re-thinking Labov's Pronunciation Drift Data

**URL:** <http://discourse.iapct.org/t/re-thinking-labovs-pronunciation-drift-data/16022>\
**Category:** Collective phenomena\
**Created:** [February 7, 2023, 6:22pm UTC](http://discourse.iapct.org/t/re-thinking-labovs-pronunciation-drift-data/16022 "2023-02-07T18:22:40Z")\
**Posts on this page:** 1\
**Showing post:** 3

<div class="post-metadata">

**Author:** ![rsmarken](http://discourse.iapct.org/user_avatar/discourse.iapct.org/rsmarken/32/3553_2.png) [@rsmarken](http://discourse.iapct.org/u/rsmarken)\
**Post date:** [November 14, 2023, 4:18am UTC](http://discourse.iapct.org/t/re-thinking-labovs-pronunciation-drift-data/16022/3 "2023-11-14T04:18:38Z")

</div>

> [@bnhpct](#):
>
> > [@rsmarken](#):
> >
> > On re-reading Labov’s paper I detected evidence that his model of the results is essentially equivalent to mine inasmuch as it assumes that similarities in pronunciation result from _imitation_. Here’s a quote from Labov’s paper that suggests that this is the case:
> > 
> > > Only when social meaning is assigned to such variations will they be _imitated_ and begin to play a role in the language. [emphasis mine-RM]
> > 
> > So Labov’s model seems to be that people tend to imitate pronunciations that have a preferred social meaning.
> 
> What is a ‘social meaning’? What perceptions are controlled by ‘preferring’ one? Where in your model do preferences for social meanings appear?

That was Labov’s phrase so I don’t know know what he meant by it. The only point I was making by quoting Labov was to show that he, like me, believes imitation is the basis of the observed phonemic drift.

> [@bnhpct](#):
>
> CI is a convenient abbreviation for a narrow range of variation in two perceptual variables, the center frequencies in the lowest two bands of amplified harmonics in the speech sound, called the the first formant and second formant. ‘Centralization index’ is not a common term of art in the field. Labov appears to have invented it for the sake of a compact representation in his tables.

I think CI is a good hypothesis about a possible controlled variable, which is a function of the two other variables that happen to be more commonly used in the field.

> [@bnhpct](#):
>
> These variables, F1 and F2, are co-varied over much wider ranges to produce and perceive the full range of vowel and consonant sounds. It is possible but very unlikely that speakers in this population developed a unique CI perceptual input function, and combining them into one higher-level ‘amount of centralization’ variable is not necessary to account for control of sounding like or unlike others.

I also think it’s unlikely that CI is a controlled variable but the actual controlled variable is probably very close to it. Evidence for how close CI is to being the correct definition of a controlled variable would have been the size of the variance of the CI values for each group; the smaller the variance, the more likely it is that CI is a CV, It would have been nice if Labov had reported the variance data.

> [@bnhpct](#):
>
> As a confound, greater or lesser centralization of all vowels is a function of gain (careful pronunciation vs. lax, unstressed pronunciation).
> 
> What is involved here is a change in the articulatory/acoustic reference values for the onset of the two diphthongs. Reference values may be achieved in careful, high-gain pronunciation, but typically are not completely attained, and control is typically disturbed by control of what comes immediately before and after (‘co-articulation effects’). Individuals displaying social identification with high gain control the socially differentiated reference value with high gain. The observed phenomenon of hypercorrection resulted.

That’s a nice point about gain (careful pronunciation vs. lax, unstressed pronunciation) being a possible confounding variable. One way to see if it was a confound would be to look at the variance of the CI values in the different groups. If gain, and not CI references, were what is different across groups then an indication of that would be low CI variance in the high gain groups and high CI variance in the low gain groups. And the average CI value in the high gain groups should be nearly the same since they are all keeping the CI value at the “true” reference.

> [@bnhpct](#):
>
> The model asserts that individuals approximate to one another’s pronunciation at each encounter. It is an obvious truism in linguistics that people often do this, but together with the contrary observation that individuals may differentiate features of their own pronunciation from values of those features as perceived in others, or they may exaggerate the observed values or extend them beyond the range in which they can be observed in others. Labov provides data on both of these phenomena, and both are important for the questions that he was investigating and for the explanation that he discovered. Those questions were, why did an extended period of gradual approximation come to an end and then reverse the direction of change, and why was the reversal generalized from the environment before a high back vowel (as in _bout_) to the environment before a high front vowel (as in _bite_).

Don’t these questions require longitudinal data – data showing variations in pronunciation over time? I don’t think Labov collected such data. Oh, and I don’t understand the second question. What does it mean for a reversal to be generalized from the environment before a high back vowel to the environment before a high front vowel? What is the evidence that a reversal has been generalized from or to an environment?

> [@bnhpct](#):
>
> The investigation disclosed that the generalization took place in a subpopulation, adolescents, but only in that further subpopulation who intended not to move away. It is clear that this was due to their controlling higher level variables, and Labov clearly said so, although he did not use control theoretic terms.

My model doesn’t rule out control of “higher level variables” as an explanation of Labov’s results. Indeed, the large pronunciation difference for people with a positive, neutral or negative orientation toward Martha’s Vineyard probably reflects control of a higher level variable that might be called “amount of socializing with people who share your attitude”.

> [@bnhpct](#):
>
> If the individual’s pronunciation changes during or due to an encounter, it is because a higher level of control changes the reference values for the variables involved.

Yes, that’s the way my model currently works, with the higher level controlled variable being “degree of imitation”.

> [@What is collective control?](http://discourse.iapct.org/t/what-is-collective-control/16087/16):
>
> > [@What is collective control?](http://discourse.iapct.org/t/what-is-collective-control/16087/16):
> >
> > The CI averages can be seen to quickly diverge but eventually settle down to different, fairly stable values. But this stability is rather fragile as we see the average CI of Group 2 suddenly diverging to a new value after a long period of stability.
> 
> There is no such fragility in the data, including data from many other such investigations in New York and Philadelphia, and elsewhere by others.
> 
> This defect in the model might be remedied by modeling higher levels of control adjusting reference values for formants, rather than making those changes an uncontrolled consequence of proximity or ‘encounter’.

This isn’t a defect of the model unless the other studies that you mention kept track of average pronunciation **over long periods of time** and found no sudden spontaneous changes. The sudden, spontaneous changes in pronunciation seen in the behavior of the model were a surprise and represent a prediction of the model. Now what is needed is a study of variation in pronunciation over many years. If no spontaneous shifts occur in that time – or if they occur for obvious outside reasons, such as a large change in who people interact with – it would be a count against the model.

> [@bnhpct](#):
>
> What is it for a group to have an identity, and for members of it to **control perception of that identity with high gain**? (I hope you have no objection to that PCT translation of “ **the group fights to maintain its identity**.”)

I think a group identity exists when each member of a group is controlling for what all agree is essentially the same system concept perception. The people who do it with high gain are the “priests” or “leaders” of the group.

> [@bnhpct](#):
>
> > This gradual transition to dependency on, and outright ownership by the summer people has produced reactions varying from a fiercely defensive contempt for outsiders to enthusiastic plans for furthering the tourist economy. A study of the data shows that high centralization of /ai/ and /au/ is closely correlated with expressions of strong resistance to the incursions of the summer people.
> 
> Does this support a model of imitating those who you most frequently encounter?

Maybe, maybe not. It is not necessarily inconsistent with it. If the people who express strong resistance to summer people tend to interact with people who feel the same way then my imitation model gets some support. If they don’t, then the model will probably have to be revised somewhat. But I think the model is on the right track. What is needed is the right kind of data to test it.

Labov’s data is a good start and all the data he reports can be accounted for by my imitation model. But to really test the model we need more data; variance measures, frequency of interaction between people in different groups, time variations in pronunciation, and so on. What we don’t need is “seat of the pants” theorizing based on what people deeply steeped in the conventional view of behavior have to say about what the data imply.

---

_[View the full topic](http://discourse.iapct.org/t/re-thinking-labovs-pronunciation-drift-data/16022)._
