Results from second Republican debate

Regular readers will know that, especially in a crowded marketplace, politicians try to stand out and attract votes by presenting themselves in the best possible light that they can. This is a form of deception, and carries the word-use signals associated with deception, so it can be measured using some straightforward linguistic analysis.

Generally speaking, the candidate who achieves the highest level of this persona deception wins, so candidates try as hard as they can. There are, however, a number of countervailing forces. First, different candidates have quite different levels of ability to put on this kind of persona (Bill Clinton excelled at it). Second, it seems to be quite exhausting, so that candidates have trouble maintaining it from day to day. Third, the difficulty depends on the magnitude of the difference between the previous role and the new one that is the target of a campaign: if a vice-president runs for president, he is necessarily lumbered with the persona that’s been on view in the previous job; if not, it’s easier to present a new persona and make it seem compelling (e.g. Obama in 2008). Outsiders therefore have a greater opportunity to re-invent themselves. Fourth, it depends on the content of what is said: a speech that’s about pie in the sky can easily present a new persona, while one that talks about a candidate’s track record cannot, because it drags the previous persona into at least the candidate’s mind.

Some kinds of preparation can help to improve the persona being presented — a good actor has to be able to do this. But politicians aren’t usually actors manqué so the levels of persona deception that they achieve from day to day emerge from their subconscious and so provide fine-grained insights into how they’re perceiving themselves.

The results from the second round of debates are shown in the figure:


The red and green points represent artificial debate participants who use all of the words of the deception model at high frequency and low frequency respectively.

Most of the candidates fall into the band between these two extremes, with Rand Paul with the lowest level of persona deception (which is what you might expect). The highest levels of deception are Christie and Fiorina, who had obviously prepped extensively and were regarded as having done well; and Jindal, who is roughly at the same level, but via completely different word use.

Comparing these to the results from the first round of debates, there are two obvious changes: Trump has moved from being at the low end of the spectrum to being in the upper-middle; and Carson has moved from having very different language patterns from all of the other candidates to being quite similar to most of them. This suggests that both of them are learning to be better politicians (or being sucked into the political machine, depending on your point of view).

The candidates in the early debate have clustered together on the left hand side of the figure, showing that there was a different dynamic in the two different debates. This is an interesting datum about the strength of verbal mimicry.

Presidential speech word patterns

In the continuing saga of presidential campaign speech language, I’ve been analyzing parts of speech that don’t get much attention such as verbs, adverbs, and adjectives. Looking at the way in which each candidate uses such words over time turns up some interesting patterns. I don’t understand their deep significance, but there’s some work suggesting that variability in writing is a sign of health; and Ashby’s Law of requisite variety can be interpreted to mean that the actor in a system with the most available options tends to control the system.

Here are the plots of adjective use (in a common framework) for the 2008 and 2012 candidates (up to the time that Santorum dropped out of the race).

It’s striking how much the patterns over time form a kind of spiral, moving from one particular combination of adjectives to another and another and eventually back to the original pattern. The exception is Obama who displays a much more radial structure, with an adjective combination that he uses a lot, and occasional deviations to something else, but a rapid return to his “home ground”.

You can see (the extremal set of) adjectives and their relationships in this figure:

You can see that they form 3 poles: on the left, adjectives associated with energy policy; at the bottom, adjectives associated with patriotism; and on the right, adjectives associated with defence [yes, it is spelled that way]. This figure can be overlaid on those of the candidates to get a sense of which poles they are visiting. For example, Obama’s “home ground” is largely associated with the energy-related adjectives.

Comparing content in the US presidential campaign 2008 vs 2012

I posted about the content in the 2012 presidential campaign speeches. It’s still relatively early in the campaign so comparisons aren’t necessarily going to reveal a lot, but I went back and looked at the speeches in 2008 by Hillary Clinton, McCain, and Obama; and compared them to the four remaining Republican contenders and President Obama so far this year.

Here’s the result of looking just at the nouns:

The key is:   Clinton — magenta circles; Obama 2008 — red circles, McCain — light blue stars;

Gingrich — green circles; Paul — yellow circles; Romney — blue circles; Santorum — black circles; Obama 2012 — red squares.

Recall that the way to interpret these plots is that points far from the origin are more interesting speeches (in the sense that they use more variable word patterns) while different directions represent different “themes” in the words used.

The most obvious difference is that the topics talked about were much more wide-ranging in 2008 than they have been this year. This may be partly because of the early stage of the campaign, the long Republican primary season keeping those candidates focused on a narrow range of topics aimed at the base, or a change in the world that has focused our collective attention on different, and fewer, topics.

This can be teased out a bit by looking at the words that are associated with each direction and distance. The next figure shows the nouns that were actually used (only those that are substantially above the median level of interestingness are labelled):

You can see that there are four “poles” or topics that differentiate the speech content. To the right are words associated with the economy, but from a consumer perspective. At the bottom are words associated with energy. To the left are actually two groups of words, although they interleave a little. At the lower end are words associated with terrorism and the associated wars and threats. At the upper end are words associated with the human side of war and patriotism.

These two figures can be lined up with each other to get a sense of which candidates are talking about which topics. The 2012 speeches and Obama’s 2008 speeches all lean heavily towards the economic words. In 2008, McCain and Clinton largely talked about the war/security issues, with a slight bias by Clinton towards the patriotism cluster.

Obama’s 2012 speeches tend towards the energy cluster but, at this point, quite weakly given the overall constellation of topics and candidates.

The other thing that is noticeable is how similar the topics for some of the Republican contenders are: their speeches cluster quite tightly.

Negative words in the campaign

Yesterday we looked at the use of positive words in the campaign. Today, I want to present the use of negative words.

We saw the President Obama is much better at using positive words than the Republican contenders; but they are all about the same at using negative words. Note that these two flavors of words are not necessarily opposites; someone can use both positive and negative words at high rates (although that itself might be interesting).

Here are the speeches according to their patterns of negative word use:

Again, distance from the origin indicates intensity of negative word use, and direction indicates different words being used.

Romney has the strongest use of negative words (and the associated words are ones like “disappointments” and “worrying”). Ron Paul also has quite strong use of negative words. His word choices are quite different from those of the other candidates, though; they include “bankrupt”, “flawed” and “inconvenient”.

President Obama and Gingrich have moderate levels of negative word use; the most popular word for both of them is “problem”, followed by “challenge”.

Santorum has the lowest levels of negative word use of all five of them.

The differences are interesting because they shed some light on how each candidate views those aspects of the situation that are not favorable to them. Obama and Gingrich have a more proactive view: negatives to them are problems. The other candidates have a more outward focus on the source of difficulties and, at the same time, a more negative inward focus, that is they use negative words that reflect how they feel about themselves.

I also ran an experiment weighting the positive words positively and the negative words negatively, to see if there is any ranking from, as it were, most positive person to most negative person. It turns out that there isn’t such a ranking. All of them use mixtures of positive and negative words, different mixtures for each, but all of about the same ratio of positivity to negativity.

Positive words in the campaign

Yesterday I posted about the content of the speeches of the campaigners for the 2012 presidential election cycle: the Republican contenders and President Obama. Today I have similar results for the use of positive words.

Here are the speeches:

The figure should be interpreted like this:  distance from the origin indicates intensity of positive word use; direction indicates the use of a different set of positive words. So President Obama is much more positive than the Republican contenders, of which Gingrich is noticeably more positive than the rest. These are only based on the use of positive words so a placement close to the origin should be interpreted as the absence of positive words, not any kind of negativity (stay tuned). In other words, speeches near the origin are not positive (they could be either neutral or negative but this analysis can’t differentiate).

Some of the positive words associated with President Obama are: “profitable”, “creative”, “efficiency” and “outstanding”.

Some of the positive words associated with Gingrich are: “tremendous”, “optimistic”, “gains”, “happiness”, and “positive” itself.

You can see why the Republican approval numbers are dropping — people pick up on the tone of speeches, and they are attracted to positive language — which they aren’t getting. Even Gingrich’s positive words are mostly about the improvement (perceived) in his chances, not in the wider US situation.

Content in Presidential Campaign Speeches

Last week I posted details of the level of “persona deception” among the Republican presidential candidates and President Obama. Persona deception measures how much a candidate is trying to present himself as “better” in some way than he really is. This is the essence of campaigning — we don’t elect politicians based on the quality of their proposals; and we don’t fail to elect them because they tell us factual lies. Almost everything is based on our assessment of character which we get from appearance and behavior, and also from language.

Today I’ll post a description of the different content of the speeches so far in 2012. This is less informative than levels of deception, but it does give some insight into what candidates are thinking is of interest or importance to the voters they are currently targeting. Here is an overview of the topic space:

You can see that most of the Republican candidates are talking about very similar things. In fact, the speeches in the upper right-hand corner are associated strongly with words such as “greatness”, “freedom”, “opportunity”, “principles” and “prosperity” — all very abstract nouns without much content that could come back to haunt them.

Gingrich’s speeches towards the bottom of the figure are quite different, although still associated with quite abstract words: “bureaucracy”, “media”, “pipeline”, “elite”, “establishment”. These are almost all things that he is against — stay tuned for an analysis of negative word use later in the week.

Obama’s speeches, on the left-hand side, are heavily oriented to manufacturing associated with words such as: “cars”, “hi-tech”, “plant”, “oil”, “demand”, “prices”.

What a candidate chooses to talk about seems to be a mix of his personal hobbyhorses (at the time) and some judgement of what issues are of interest to the general public, or at least which can create daylight between one candidate’s position and the others. From this perspective, Gingrich separates himself from the other Republicans quite well. Somewhat surprisingly, Ron Paul’s content is not very different from that of Romney and Santorum. Probably this can be accounted for as a function of the three of them all trying to appeal to a very similar segment of the base. Whether Gingrich is consciously trying to address different issues, or whether his history or personality compel him to is not clear.

2012 US Election, Republicans plus President Obama

Yesterday I posted details about the levels of persona deception in the speeches by the Republican candidates since the beginning of 2012. In striking contrast to the 2008 cycle, the speeches fall along a single axis, indicating widespread commonalities in the way that they use words, particularly the words of the deception model.

Today I’ve included President Obama’s speeches this year in the mix. I’ve tried to select only those speeches where there was an audience. Of course, for a sitting president, the distinction between an ordinary speech and a campaign speech is difficult to draw. Almost all of these are labelled as campaign events at

Here is the plots of the persona deception levels, with Obama’s speeches added in magenta.

Generally speaking, Obama’s levels of persona deception (see yesterday’s post to be clear on what this means) are in the low range compared to the Republican presidential candidates. This is quite different from what happened in the 2008 cycle, where his levels were almost always well above those of McCain and Clinton. It’s not altogether surprising, though. First, he can no longer be the mirror in which voters see what they want to see since he has a substantial and visible track record. Second, he doesn’t have to try as hard to project a persona (at least at this stage of the campaign) since he has no competitor. I expect that his values will climb as the campaign progresses, particularly after the Republican nominee becomes an actual person and not a potential one.

The interesting point is the outlier at the top left of the figure. This is Obama’s speech to AIPAC. Clearly this is not really a campaign speech, so the language might be expected to be different. On the other hand, if it were projected onto the single-factor line formed by the other speeches, it would be much more towards the deceptive end of that axis. Since the underlying model detects all kinds of deception, not just that associated with persona deception in campaigns, this may be revealing of the attitude of the administration to the content expressed in this speech.

Republican presidential candidates — first analysis of persona deception

Regular readers of this blog will know that I carried out extensive analysis of the speeches of the contenders in the 2008 US presidential election cycle (see earlier postings). I’m now beginning similar analysis for the 2012 cycle, concentrating on the Republican contenders for now.

You will recall that Pennebaker’s deception model enables a set of documents to be ranked in order of their deceptiveness, detected via changes in the frequency of occurrence of 86 words in four categories: first-person singular pronouns, exclusive words, negative-emotion words, and action verbs. Words in the first two categories decrease in the presence of deception, while those in the last two categories increase. The model only allows for ranking, rather than true/false determination, because “increase” and “decrease” are always relative to some norm for the set of documents being considered.

How does this apply to politics? First of all, the point isn’t to detect when a politician is lying (Cynical joke: Q: How do you tell when a politician is lying? A: His lips are moving). Politicians tell factual lies, but this seems to have no impact on how voters perceive them, perhaps because we’re come to expect it. Rather, the kind of deception that is interesting is the kind where a politician is trying to present him/herself as a much better person (smarter, wiser, more competent) than they really are. This is what politicians do all the time.

Why should we care? There are two reasons. The first is that it works — typically the politician who is able to deliver the highest level of what we call “persona deception” gets elected. Voters have to decide on the basis of something, and this kind of presentation as a great individual seems to play more of a role than, say, actual plans for action.

Second, though, watching the changes in the levels of persona deception gives us a window into how each candidate (and campaign) is perceiving themselves (and, it turns out, their rivals) from day to day. Constructing and maintaining an artificial persona is difficult and expensive. Levels of persona deception tend to drop sharply when a candidate becomes confident that they’re doing well; and when some issue surfaces about which they don’t really have a persona opinion because, apparently, it takes time to construct the new piece.

So, with that preliminary, on to some results.

The figure shows the speeches in a space where speeches with greater person deception (spin) are further to the right, and those with less persona deception are further to the left. Ron Paul shows the lowest level of persona deception which is not surprising — nobody has ever accused him of trying to be what he is not. In contrast, Romney shows the highest level of persona deception — again not surprising as he has had to try hardest to make himself appealing to voters. Note that this also predicts that he will do well. Both Gingrich and Santorum occupy the middle ground; both are running on a very overt track record and are not trying as hard to make themselves seem different from who they are. Indeed, candidates with a strong history tend to have lower levels of persona deception simply because it’s very difficult to construct a new, more attractive persona when you already have a strong one. (The two points vertically separated from the rest are the result of a sudden burst of using “I’d” in these two speeches.)

The following figures break out the temporal patterns for the four candidates:

What’s striking about Romney is how much the level of persona deception changes from speech to speech. In the last election cycle, this wasn’t associated with audience type or recent success but seemed to be much more internally driven. This zig-zag pattern is much more the norm than a constant level of persona deception — some mystery remains.