Saturday, 29 December 2018

Word of the Week: Foundationalism


Meet Susan Haack - a new foundationalist.

Are there secure foundations to knowledge? 'Foundationalism' is a theory of knowledge, or how we justify our beliefs. Foundationalists think all beliefs ultimately get their justification from a set of very basic beliefs or sensory states which are the 'foundations' of knowledge. Otherwise, there'd be no foundation, and to fully justify any claim we'd have to say 'A is the case, because B, because C, because D', and so on until infinity! Foundationalists think infinite justification is absurd, so they prefer to base knowledge claims upon a secure foundation - a bit like building a house on rock, not on sand...

So there ought to be secure foundations for knowledge. But it is difficult to see what these foundations really are. The general form of foundationalism seems to be acceptable, but it is difficult to point to the content of any particular theory. There are secure foundations out there, but we don't really know what they are! Perhaps a moderate solution is needed: we might admit that Susan Haack is right in embracing 'foundherentism', which combines the 'foundationalist' insight with the 'coherentist' view that beliefs can justify each other without resting solely on a bedrock of 'basic' beliefs which don't need further justification. Foundationalists think knowledge is like a pyramid. Coherentists think it's like a web: not only does basic belief B justify derivative belief D, but D also can justify B and vice versa - within webs of reinforcing beliefs. Perhaps by 'secure' foundations we don't mean an indubitable foundation but some kind of sound basis for knowledge which can be strengthened by coherence relations and, perhaps, reliable processes too. Contrary to Descartes' assertion that we can all be dead certain of certain beliefs - e.g., 'I exist' and 'God exists' - maybe knowledge has somewhat weaker, but still significant, foundations. The alternative is the scepticism which results from rejecting basic beliefs.

To start with, let's consider epistemic foundationalism - the view that there must be some foundations to knowledge, whatever they are. I embrace epistemic foundationalism because the alternatives are pretty bad. The argument for epistemic foundationalism is the so-called 'regress argument'. Richard Fumerton uses the Principle of Inferential Justification, which states that: S's belief that P on basis E is justified if (1) S's belief that E is justified and (2) S's belief that E makes probable P is itself justified. Let's take (1) first, and return to (2) later. 

For S's belief that E to be justified, S must belief some other proposition - E1, say - which must also be justified. And so on. So we can see that part (1) of Fumerton's simple principle leads us to regress: the justified belief that E1 depends on another proposition E2, then another proposition E3, and so on until infinity. According to the regress argument for the view that knowledge has foundations, a contradiction is generated. On the one hand, regress leads us into there being an infinity of prior justifications for any given belief. On the other hand, it seems implausible that there can exist such an infinite regress! How can finite minds ever comprehend infinite chains of reasons? What's more, if the infinite chain were completed, it would no longer be infinite! The obvious way of resolving this dilemma is by embracing the claim that some knowledge is non-inferential - i.e., not inferred from any other belief. This non-inferential knowledge is therefore foundational. It either justifies itself or doesn't need justification: these are the foundations you're looking for! We don't know exactly what they are, mind you. All that Fumerton's argument proves is that there must be foundations of some kind - otherwise knowledge as we understand it is impossible.

But infinitists such as Peter Klein think infinite regress is perfectly reasonable: we can have infinite chains of justification that enable us to really have 'knowledge'. He rejects one alternative - 'coherentism' - on the grounds that it leads to circularity. Coherentists say that S's belief that P is justified if S's whole system of beliefs is justified by the coherence of the relations between the beliefs in that system - like a spider's web. But this suggests that it's possible for P to justify itself - albeit indirectly - by a chain of reasons which loops back on P. But Klein thinks we should all accept the Principle of Avoiding Circularity: circular reasoning is unreasonable. So we should reject pure coherentism. (I'll return to Haack's modified version - 'foundherentism' - later.) Also, Klein rejects foundationalism, because he thinks we should all accept the Principle of Avoiding Arbitrariness, which stops knowledge from having arbitrary foundations that have no justification. In Klein's view, to argue for the view that knowledge has secure foundations is like arguing for the view that one can simply believe something without having any good reason for believing it - so one foundationalist cannot challenge another if their arbitrary foundations differ! For instance, foundationalist Bertrand Russell thinks sense data - or sensory states, like a perception of a chair - are basic reasons for believing propositions such as 'There's a chair over there!'. But religious foundationalists like Alvin Plantinga think that a religious believer's belief is God is foundational, or basic. But both these foundations might be arbitrary - after all, a religious believer's so-called 'basic' belief that God exists is not likely to convince an atheist who has different 'foundations' to their knowledge. So if you want to avoid arbitrariness and circularity, Klein wants you to embrace infinitism!

Infinitism, however, isn't really that convincing. Look back at condition (2) of Fumerton's Principle of Inferential Justification: S's first belief that P on basis E is justified IF S's further belief, 'E makes probable P', is itself justified. Like (1), this leads to regress. If one's belief that it will rain tomorrow (P) is justified by your belief that the Met Office said so (E) and your further belief that what the Met Office says is reliable (i.e., E makes probable P), one would also need a reason for thinking that reliability itself is a good reason for believing reports!  You would need the further belief, which also requires justification, such as 'The Met Office's reliability is a justification for my believing its predictions'. But what is this the case? And so on, and so on. So (1) leads to regress, (2) leads to regress, and (2) also leads to an infinite number of infinitely long chains of reasons! Even if infinitism is not incoherent, it's surely absurd. As Carl Ginet points out, '[i]nference cannot originate justification, it can only transfer it' from one belief to the next. And as Jonathan Dancy argues, '[j]ustification by inference is conditional justification only'. To argue on the basis of inference from one belief to another presupposes that the inferential relation itself is justified - which at some point must require some kind of foundation, such as the basic belief that the inductive principle (inferring generalisations from observations) is valid. If one doesn't have any foundations, one doesn't have justification - infinitism cannot originate justification, it can only transfer it

So what are the foundations? Laurence BonJour thinks logical beliefs like 'anything deduced from a true proposition is true' (Bertrand Russell's example) are foundational. They are the basis for inference. This seems reasonable - without certain basic logical beliefs, in maths or in scientific principles, we can't really go about making justifiable inferences about, well, anything. But what makes these beliefs justified? Well, if they're basic, they don't require further justification! But perhaps what we mean is not 'they don't require further justification at all', but 'they don't require further justification of the usual kind'. They require special justification, which might still allow them to be foundational - or not requiring normal justification. They might be arrived at by reliable processes. As Carruthers (1992) argues, the means by which 'one's belief comes to be innate is reliable' - or generated by a process which tends to generate true beliefs. Our cognitive processes must be reliable for epistemic foundations to be secure. So it seems even self-evidence principles like 'The rules of addition are valid' need to be based on reliable cognitive processes. They are still strictly foundational, but they're not as secure as we once thought. Their security depends on processes outside of them. Beliefs derive their full warrant from external (cognitive) processes, not just internal justifications. Underneath the superficial security of foundational beliefs lie important reliable processes

Another possible foundation is sense-data: e.g., the sense-datum of a table. We base our knowledge on such sense-data, which don't need justifying. They're just experiences which form the basis of our justified beliefs. But as BonJour notes, this kind of foundationalism is problematic. Either sense-data assert that something exists or they don't. If they don't, then how could they justify beliefs which do make such assertions? On the other hand, if a sense-datum of a table can assert that the table exists, then what makes this assertion justified? The assertive representational content itself - not the sense-data - is doing the justification! Sense-data can't justify beliefs, because either they don't have the necessary conceptual content to do so, or they implicitly contain beliefs about existence which are the real sources of justification. These implicit beliefs (e.g., the belief that 'There's a table there' whenever we see, or have a sense-datum of, a table) surely need further justification. But this justification needs yet another justification, and so on, until we get to something basic. The alternative is infinite regress. What's certain is that sense data are not the secure foundations we're looking for.

Susan Haack wants us to embrace a synthesis. Foundations can't be indubitable - or beyond doubt. But they can be secure enough to provide reasons for believing non-basic beliefs. We may not have indubitable knowledge in sense-datum-like 'S-states'. But we can have some kind of knowledge in such states, mutually supported by cohering 'C-evidence' which ascribes S-states to a subject. Basic beliefs can be justified not just by themselves but at least in part by their coherence relations with other beliefs. Justification goes up (from basic beliefs to non-basic beliefs) and back the way down again (as non-basic beliefs cohere with one another, and support the basic beliefs on which they rely). Basic beliefs justify non-basic beliefs, but then non-basic beliefs justify each other through their coherence (consistency or mutual derivability). These non-basic beliefs can then lend further support to the basic beliefs - in a virtuous circle of foundherentist justification! A problem with Haack's account is how basic beliefs can derive any justification at all from coherence with non-basic beliefs independently of the justification originated from basic beliefs. If basic beliefs have already done the justifying of non-basic beliefs, isn't non-basic beliefs' supporting basic beliefs just circular? Maybe, but maybe not: as Peirce once argued, these special chains of justification might be relations of 'mutual support', not simple 'circularity'. But as an alternative to the extremes of logical and sense-datum foundationalism, Haack's foundherentism is at least a promising alternative. As Haack puts it: foundherentism, 'without sacrificing objectivity, acknowledges something of how complex and confusing evidence can be'. Foundherentism might someday solve the dilemma of foundationalism's good form but poor content.

So it seems there must be some foundations for knowledge if our beliefs are to be justified. But these foundations might not be indubitable - they might rely on non-basic beliefs (as Haack claims) or on reliable cognitive processes (as Carruthers thinks). We can't definitely point to particular foundations, and we can't claim these foundations are totally secure. We can't be certain that knowledge has secure foundations. But in order to have justification at all, it seems there must be some kind of foundation - however secure that foundation may be.

Image sourced from Creative Commons.

Tuesday, 4 December 2018

Word of the Day: Time

A moment in time (Creative Commons)

Political times are different from other times. For example, thermodynamic time is irreversible: entropy (roughly: a measure of energy dispersal or disorderliness) never decreases in a 'closed' system (e.g., the universe, or a very good thermos flask). In these systems, entropy goes up and up, but never down. So thermodynamic time isn't that different from psychological time - it just keeps moving on, never stopping, never reversing. You can't turn back the universe's clock.

But political times are often different from this version of time. (Perhaps not all times, but I'm no physicist - and certainly no quantum mechanic!). Most people see time as a linear progression, or at least as an irreversible trajectory ('past, the point of no return...', with apologies to Andrew Lloyd Weber). But for others, political time goes up, down, and back round again... Or should I say political times?

In this post, I take us through three different versions of political time: linear, cyclical, and plural. The first is a bit like thermodynamic time, but the second and third are different. I think the pluralist interpretation of time is best for our diverse world.

'Political time is a line'
For many political thinkers, admittedly, time is a line. And it's irreversible. A bit like thermodynamic time. There is a 'right' and 'wrong' side of history - like a ladder rising into heaven, with anything low being 'backward' and anything high 'advanced' and 'developed'. Time is progress. Both Marxists and liberals often think this way. Their common ancestor is Hegel, for whom history was a progression of ideas ('theses') being challenged by other ones ('antitheses'), from which improved 'syntheses' were formed which drew on the advantages, while casting away the drawbacks, of both thesis and antithesis. History culminated in an Absolute Idea, superior to all other ideas thanks to its constituting the synthesis of a long dialectical process. Hegel construed this 'end of history' as Napoleonic good government. So history ended with the invasion of Jena by Napoleon's forces in 1806. But Karl Marx thought Hegel neglected the continuing economic struggles after Jena between proletarians and bourgeoisie, which would lead to the challenging of capitalism from class antitheses. This new dialectic would lead to a different end of history: communism (after taking socialism as a further stepping-stone). For political scientist Francis Fukuyama, however, Hegel was more accurate than Marx: history was the culmination not of a dialectic of material struggles (as Marx suggested) but of a dialectic of ideational clashes (as Hegel contended). The final clash was the Cold War, which ended in the replacement of communism in eastern Europe with some form of democratic capitalism. Democratic capitalism, for Fukuyama, was the 'final form of human government' - the end of history, if 'history' is taken to be the progression of ideas until they reach a utopian end-point.

But despite their disagreements, Hegel, Marx and Fukuyama concur on this: history is a progression of contradictory ideas, and the end of history will eliminate these contradictions for good in a worldly utopia - be it Napoleonic-style rule, communism, or democratic capitalism. Time is a line - or, more specifically, a ladder - rising from the dark depths of the First Man of prehistory to the Last Man of modernity. Life for the First Man is 'nasty, brutish and short' (in Hobbes's words), but life for the Last Man is prosperous, free, and long. Time is a line - and it's pointed towards the sky.

'Political time is a loop'
But for many political scientists, political time doesn't go up and up. It goes down and comes back round again. Time is cyclical, not linear. For James Madison, the American republic was less of an 'end of history' than a modest return to the republican virtues of antiquity. For Donald Trump, the ideal America is to be sought some time in the past, which we must rise back up to meet, in order to 'make America great again'. For sociologist Saskia Sassen, we must return to the 'logic of inclusion' under the Keynesian forms of democratic capitalism practised between 1945 and 1978, rejecting the 'logic of exclusion' and 'expulsion' which has dominated since then. Geographer David Harvey similarly scorns the 'neoliberal' phase of capitalism. So we might need to seek out alternatives in the past, from the movements of indigenous people in Ecuador (and their ecological 'buen vivir' thinking) to the social democracy of the post-war era. Finally, for Alexandria Ocasio-Cortéz, the immediate post-war phase of regulated capitalism was better than the twenty-first century's phase of deregulated capitalism. For all these thinkers, Marx and Hegel are wrong to think that political time doesn't 'reverse'. Because time is not a line. It's a loop. Make {{insert a given period in the past two centuries}} great again!

'Political times have many shapes'
Both the linear and cyclical interpretations of time are, in my judgment, misled. Political time, of course, is itself a construct of our imaginations - but this is not my grounds for criticism. Religion, nation, and human rights are all constructs of our imagination - but that doesn't mean we should reject religion, nation, and human rights! Political time, similarly, is a construct we can use to look at how our political theories view historical change through time. Most political projects have some moralised version of time - as a line, as a loop, or as some other shape (a zig-zag, perhaps?). I do not propose we reject the notion of political times. But I do propose that we reject the singular notion of one political time. Because there are, I think, many political times. 

This doesn't just mean that different people have different opinions on whether things are getting 'better' or 'worse' (see the debate between the optimistic psychologist Steven Pinker and pessimistic intellectual John Gray for more details). I mean that different elements of society - from violence and poverty to life-expectancy and political systems - each have distinct political times. While violence between humans seems to have decreased as a proportion of the total population (as Steven Pinker has meticulously demonstrated), violence against nature has increased exponentially during the 'Anthropocene' (see novelist Amitav Ghosh, journalist Naomi Klein, or theorists Bonneuil and Fressoz for more details). Absolute poverty has decreased, but relative poverty and inequality have increased. We live in a 'high risk, high opportunity' society (as sociologist Anthony Giddens contends). Some things seem to progress towards a brave new world, while others go back to the 'good' or 'bad' old worlds. Some political times seem to be travelling on a trajectory towards a Hegelian Utopia - the political time of life expectancy, for instance, propped up by stabler political and economic systems. But the political times of democracy, socialism, and alternative ideologies are looping, rising, falling, and zig-zagging in myriad ways. There are many political times, depending on your vantage point.

This is not to say that political time is unimportant. It can help us to imagine where we've come from, and where we might be going. But political time is plural. There is no one 'political time', both due to different views and due to different trajectories seen in different aspects in the world today. There are many times, and they have many shapes. The upshot? Politics is many-sided, so we should adapt to each problem in a different way, rather than imposing a one-size-fits-all model to humanity's disparate issues. Political times have many shapes - so should our political choices.

Monday, 3 December 2018

Word of the Day: Food-sharing




Food-sharing: then and now... (Creative Commons)

What made us human? In 1978, archaeologist Glynn Isaac suggested one possible answer: food-sharing. Early hominins, Isaac argued, 'made tools and carried food to a home base', where they would share the fruits of their hunting, scavenging, and foraging. Hominins wanted to avoid confronting menacing grassland predators. And a division of labour was developing between hunting and foraging. So perhaps it became advantageous for hunter-gatherers to take surpluses from their hunting and gathering back to a 'home base' where they could share their food with other kin relations. At least, that is Isaac's opinion. The problem is that he does not make consistently accurate predictions about the distribution of bones and artefacts in early hominid sites. Luckily, however, his hypothesis does make broadly accurate predictions about the direction of behavioural and biological evolution. The hypothesis does not agree with all existing observations, but at least it is simple and powerful. I think it is 'empirically ambiguous', as it's unclear whether it explains all observations, but 'theoretically parsimonious', as it's simple and powerful. We can only hope that further excavations of early hominin sites, dating from nearly 2 million years ago, help us to get to the bottom of these questions: (1) 'what explains the artefact and stone distributions at these ancient sites?', and (2) 'what made us human?'. 


The first food-sharer? Meet Australopithecus boisei, an early hominid. (Creative Commons)

Firstly, then, does the food-sharing hypothesis explain the distribution of bones and artefacts in early hominin sites? Isaac thinks that some archaeological sites in East Africa, where a high concentration of bones and artefacts are found, represent some of the first 'home bases', where hominins engaged in food-sharing. Isaac claims that ‘protohuman hominids’ (or 'hominins' in today's terminology) engaged in ‘food-sharing behaviour’ around home-bases. He details artefact and animal bone concentrations at East African Oldowan sites such as Kay Behrensmeyer. There, a concentration of stones and bones lends weight to the argument that protohumans ‘made and discarded their tools here’ and ‘were also responsible for the bone accumulation because they met here to share their food’ (Isaac, 1978:100). But perhaps water flow, not hominin activity, led to these artefact and bone concentrations (Potts, 2017:11). This is quite plausible, given that the Olduvai Gorge sites surround ancient water sources (Schick, 2017:802). As Binford (1983:70) contends, ‘considerable quantities of bone can be expected to occur around water sources’ regardless of hominin activity. This suggests that natural fluvial depositional processes may be responsible for the dense bone concentrations in the Olduvai beds. But only 7% of Olduvai Bed I specimens were affected by abrasion, and weathering at the 'FLK Zinj' site is even rarer (Klein, 2009; Potts, 1984). Water flow alone cannot account for these concentrations. Perhaps hominin activity can.

But Isaac’s theory also implies continuous occupation of a single site. This is implausible, for two reasons. Firstly, the concentration of Olduvai sites around water-holes makes continuous occupation by hominins seem very unlikely. ‘Actualistic studies’ of modern water holes indicate ungulate presence during the day but predation during the night, making water-holes dangerous places indeed to rest when the sun goes down over the savannah (Binford, 1983:66-67). This is the reason modern hunter-gatherers (e.g., the Botswanan San of the Kalahari Desert) rarely locate their camps next to water sources (ibid.:68). This makes it more likely that the water sources were simply locations of ‘natural deaths’ of animals, ‘predatory kills’, ‘hyenas gnawing’, and fluvial depositional factors (ibid.:68-70). Perhaps a combination of these ‘natural’ (non-hominin) factors led to the observed concentrations in Olduvai and more recent (dating to 200-400 ka) waterhole sites like Elandsfontein. Also, computer simulations indicate the likelihood of bases being used as ad hoc energy-saving strategies rather than as permanent eating and sleeping locations (Potts, 1984:345). Early hominins might have deposited stones there for future usage, but it is unlikely they used these sites as 'home bases' for sleeping. So it is unlikely that these sites were Isaac's continuously occupied home bases. Food-sharing alone cannot explain the concentrations at these sites. Isaac's hypothesis, then, is ambiguous when it comes to explaining bone/artefact concentrations at early hominin sites.

Secondly, does the food-sharing hypothesis (partly) explain behavioural and biological evolution? I think it does, for a couple of reasons. First, the biological origins of modern humans may be partly sought in the possible dawn of food-sharing nearly 2 million years ago. Meat-eating, perhaps around home bases, may have played a role in leading to increasing brain size (Aiello & Wheeler 1995). The cut marks and fracture patterns of ‘large mammal carcasses’ reinforce the notion that home bases were locations for systematic butchery (Bunn, 1986). This led to our ancestors' consuming high-nutrient animal fat and protein which, over time, simplified our digestive tracks and freed up metabolic energy for enlarging the brain (Toth & Schick, 2018:65). Perhaps our collective meat-eating around home bases led to 'physical selection pressures' which promoted brain enlargement (Isaac, 1978:108). Food-sharing contributed to certain long-term biological changes, through collective meat-eating, socialising, and stone-manufacture around home bases.

Food-sharing also explains other human 'social' characteristics. Early hominin sociality is particularly well explained by Jones’s (1984) concept of ‘tolerated theft’. When a latecomer arrives to a hominin group devouring a carcass, early hominins may have begun to 'tolerate' this latecomer's 'theft' of their scavenged find.  They allowed latecomers to seize some of the catch for themselves. Then, when they were 'late' to a kill or scavenge, they too demanded some share in the food. So a system of tolerated theft and reciprocal altruism developed - presumably around a home base. This reciprocity was a dynamo of behavioural evolution (Trivers, 1971) and modern human reciprocity (Fukuyama, 2014:88-89; Mauss, 1925). What does it mean to be human? It means to reciprocate: 'you scratch my back, I'll scratch yours', but in a way that means you have to think about future debts and obligations. This need to ‘calculate’ future ‘contingencies’ of food-procurement would select for more ‘logical’ cognitive capabilities (Isaac, 1978:106). Food-sharing contributed to more complex modes of reciprocity than is found in other primates. We think about how much food we need to feed the entire home base - not just ourselves. So food-sharing might have developed both more complex reciprocity and enhanced cognition. The food-sharing hypothesis helps to explain many ‘human’ characteristics, from brain enlargement to reciprocal gift-giving. This helps to explain human phenomena from Albert Einstein to feasting and gift-reciprocating at Christmas time.

Overall, the food-sharing hypothesis doesn't fully explain the bone/artefact distributions in East and South African sites. But Isaac’s hypothesis provides a simple and powerful explanation of some long-term trends in human biological and behavioural evolution. But we might need to wait for a while until further excavations, observation of hunter-gatherer activity, and dating methods rigorously test Isaac's and others' hypotheses. The food-sharing hypothesis has not so far proved as empirically watertight as it is theoretically parsimonious. Only time will tell whether the data of pre-history is better explained by competing hypotheses. As of yet, the food-sharing hypothesis has fared surprisingly well in relation to its competitors—particularly in the light of its empirical shortfalls. At least Isaac has got us thinking about one of the greatest questions of all: what made us human?

Tuesday, 7 August 2018

Peace For Our Time? Explaining War’s Decline



Edward Hicks, ‘Peaceable Kingdom’

The following post is taken from my recent INK article, which details the possible causes of the decline in inter-state war since 1945...



“The tide of war is receding,” President Obama proclaimed in 2011. Which is ironic, because Obama was referring to the seemingly intractable war in Afghanistan, which President Trump is set on continuing. But ultimately, Obama was right. Since the Second World War, the tide of war has receded—worldwide. Annual battle deaths have fallen 90% since 1950. War between countries has declined, the 2003 Iraq invasion being the only conflict this century meeting the criteria of a deadly inter-state war (though there remain a fair few civil wars with outside involvement). But why has full war between states become virtually extinct?

One explanation is democratisation. Between the 76-odd democratic states, peace prevails for two reasons, according to proponents of democratic peace theory such as Yale’s Bruce Russett. Firstly, politicians could simply be voted out of power if they use force against the public’s will, so the political cost of using force is greater in democracies than in non-democracies. Secondly, democracies expect to resolve domestic conflicts by compromise, leading them to “externalise” these internal norms and behave peacefully with other democracies. The last decade of relative peace may therefore be due in part to the sheer number of democracies. But inter-state war is uncommon worldwide, not just in the democratic sphere. To understand this, we need a more truly global explanation.

Such an explanation might be found in the process of globalisation, or the growing interdependence of the peoples of planet Earth. Economic interdependence has expanded apace this century, creating trade interests which do not favour war, argues Columbia University’s Michael Doyle. Maybe ‘perpetual peace’, in Kant’s famous phrase, requires commercial ties and free trade. So capitalism may be ‘a more powerful force for peace than democracy’, as academic Michael Mousseau put it.

For example, in 2014 the non-democratic regimes of China and Russia agreed a $400bn thirty-year contract to supply China with Russian gas. This economic interdependence should help to pacify relations, since commerce between countries reduces the incentive to go to war. But sometimes economic interdependence can exacerbate tensions: for instance, the EU fears that Russia will use Europe’s dependence on Russian energy to blackmail Europe, according to one recent analysis. Nevertheless, the fact that no EU member state has fought a war with Russia demonstrates the ability of some other mechanism to pacify relations.

This mechanism is liberalisation, or the creation of a liberal world order. Liberal institutions such as the EU, the OSCE and the UN employ soft power to make aggression seem more costly to a state’s reputation than pacifism. As Kant predicted, a “pacific union” is forming, established by various treaties of international law, softening normative conflicts between hostile states. These institutions also provide a home for a broad, constant conversation among nations. The development of the human rights regime has also bolstered cooperation and reduced the probability of inter-state war. The liberal world order, propped up for 70 years by the democratic, liberal-minded US, has contributed to the decline in violence between countries.

But cracks are showing in this liberal world order. Mirroring the US’s recent critiques of liberal institutions such as the EU, the UN and NATO, in 2016 China condemned a UN-mandated court’s rejection of Chinese claims to the South China Sea. From the rise of populism in Europe and the Trump-led US to the demagoguery of Putin, Erdogan and Duterte, identity-driven tribal politics within states is fuelling aggressive power politics between them. This threatens all three of the forces for peace we have discussed. The US’s democracy can hardly prevent the possibility of conflict with a rapidly growing China; in any case, the US is one of 89 countries in which the quality of democracy declined in 2017, according to The Economist. Globalisation is threatened by understandable though often misguided popular protests. And the US can hardly prop up a liberal world order which its leader views as unfair and anti-American. As the liberal empire of the post-war era retreats, so the spectre of war might advance.

But the collapse of peace is not unstoppable. Just as democratisation, globalisation and liberalisation progressively reduced conflict in past decades, they could do so in decades to come. Only by tackling inequality within and between states can we reduce the populist pressures that threaten the considerable achievements of the liberal world order. And only with awareness of the virtues of democracy, interdependence and liberal values can we make the future resemble the peaceful present rather than the conflicted past.


Image used under Creative Commons license.

Friday, 9 June 2017

Great powers, superpowers and America


Is America a superpower, or just a great power? It certainly dominates other states in its naval might, as this image of the USS Carl Vinson aircraft carrier demonstrates. The UK, by contrast, has 0 aircraft carriers.

A great power is a political entity which exercises significant international influence and military strength. Within a given system of polities, great powers collectively possess a near monopoly of politico-military muscle. Traditionally, the great powers of the European states system set the rules by which international relations were to be managed. For instance, following the Thirty Years War a series of treaties signed by great powers in Westphalia set down the rules of the road for great and not-so-great powers alike. Similarly, the War of the Seventh Coalition concluded in 1815 with the Congress of Vienna, which arranged a new set of rules by which international society was to be organised. After 1919, the victorious great powers – France, Britain and the United States – drew up another set of rules. Therefore, great powers, as well as possessing a monopoly over hard power, which can be used militarily, also possess a substantial amount of soft power, or the economic, diplomatic and persuasive ability to influence the behaviour of other states. Soft power may be exercised through, amongst other things, drafting and signing important treaties.

A superpower, however, does not tend to share power so diversely. Whilst a given states system may have up to 20 great powers (as is the case with the G20 in the modern states system), there will only tend to be one or two superpowers. Generally, one superpower will tend to dominate a system – if, that is, superpowers exist at all. This single superpower may incorporate other states into its own imperium or, at the very least, ‘sphere of influence’. However, as was the case during the Cold War, there may be two superpowers. A superpower, therefore, is a transcendent power on the world stage.

Nevertheless, just as different great powers may have different areas of expertise, so may different superpowers have different areas in which they exercise hegemony, or dominance. For instance, it is arguable that China is currently just an economic – rather than a politico-military – superpower. It is the largest emitter of carbon dioxide and has the second largest Gross Domestic Product (GDP) for a single-state entity, standing at $11.2. Although this lags behind the European Union’s GDP of $16.4T, the EU is a multistate entity, and thus does not have as much political clout as China, which has a permanent seat at the UN Security Council. Nevertheless, both China and the EU have small economies in comparison with the US, which has a GDP of $18.6T.

The US, moreover, is perhaps the only all-round superpower in 2017. Whilst the Cold War featured a clash between two superpowers  - the US and the USSR – today only the US controls its region of the world. Therefore, the US is, to use John Mearsheimer’s term, the only ‘regional hegemon’ on Earth. Whilst the Cold War was bipolar, to some extent America’s superpower status has shifted the world towards an era of unipolarity, where only one power is dominant. But America’s superpower status should not be exaggerated. Admittedly, in 2014 its economy amounted to nearly a quarter of global GDP and it spent $596B on defence on 2015, which was more than double China’s total of $215B. Nevertheless, even if it is a global superpower, the US is certainly not a global hegemon. To be a hegemon, as Mearsheimer and R. Harrison Wagner acknowledge, America would need to be able to defeat any combination of other powers in a given confrontation. But as its military spending in 2014 amounted to only 43% of the global total, this would not be enough to defeat all other states combined. Perhaps, therefore, the US remains a modest superpower, without the capacity to exercise total global dominance. Far from it: the US was unable to stop Russia’s annexation of Crimea in March 2014, as well as the seizure of Mosul by so-called Islamic State three months later.

Nevertheless, the global reach of the US, although not characteristic of a global hegemon, is certainly characteristic of a superpower. Whilst America’s soft power has declined – particularly following President Trump’s withdrawal from the Paris climate accord on 1 June 2017 – its hard power remains intact. After the Syrian government launched the Khan Shaykhun chemical attack on 4 April 2017, the US launched a Thomahawk cruise missile strike three days later. Syria has not dared to launch any chemical attacks since. Thus the US is capable of using targeted unilateral force with devastating effect.

On balance, therefore, the US remains a superpower to this day. But it nevertheless ‘has been in a state of almost continuous relative decline since the end of World War II’ (Antony Best et al, 2015). It is certainly a great power, and retains its superpower status thanks to its being ahead of any other state in crude economic, political and military terms. Nevertheless, in none of these areas does it exercise hegemony. Even America’s status as the world’s only regional hegemon may soon be cast away by a resurgent Middle Kingdom. America may be a superpower, but it is far from an unlimited one.

Word of the Week: Norms




Alexander VI: the normative Pope? (Creative Commons)

Traps and Norms

In previous posts, I have tried to investigate the underlying explanation for the Thucydides trap, or the hypothesis that, when a rising power confronts a ruling power, war becomes quite probable. This week, I ask which theory of international relations offers the best explanation for one particular case of peace enduring, despite the fact that a rising power (Spain) confronted a ruling power (Portugal) in the late 15th century AD. This is one of the Belter Center's 16 examples of the Thucydides trap. I find that social norms appear to play a part in explaining why the Thucydides trap did not spring in this particular case. 

Norms are patterns of behaviour. They have a descriptive element and a prescriptive element. One the one hand, norms describe patterns of behaviour, e.g. the respect that European states demonstrated for the Pope. On the other hand, norms prescribe patterns of behaviour, e.g. the respect that European states thought they ought to demonstrate for the Pope. Employing the constructivist theory of international relations, I argue that norms restrained the warlike tendencies of the Thucydides trap in the late 15th century. Nevertheless, there are many other possible explanations for why peace endured between Spain and Portugal against the odds, which I shall look at shortly. But first, before I look at the explanation for peace between Portugal and Spain, it's worth briefly describing how close Portugal and Spain got to war, and what they did to avoid it.

Portugal and Spain

In the late 15th century, Portugal, the 'ruling power' (according to Professor Allison's schema) began to feel threatened by the 'rising power' that was the Catholic monarchy of Ferdinand and Isabella, who had in 1469 united the crowns of Castile and Aragon. Since the days of Henry the Navigator, the mid-15th century Portuguese prince with exploratory ambitions, Portugal had aspired to be the world's dominant sea power. But now its neighbour was on the rise. In 1492, Ferdinand and Isabella completed their reconquest of the final emirate of the Iberian peninsula - Granada. That same year, Christopher Columbus reached the New World in 1492 with Spanish backing, after being spurned by Portuguese King John II. The balance of power 'changed almost overnight', and Portugal and Spain were set to fall into Thucydides's timeless trap.

However, the trap did not spring, and peace endured between Portugal and Spain. In 1494, Pope Alexander VI helped negotiate the Treaty of Tordesillas between Spain and Portugal, giving Portugal access to India and Africa whilst giving Spain access to most of the Americas, as the dividing line between the Spanish and Portuguese spheres of influence was the 46th meridian. Generally, lands discovered west of the meridian were to be Spanish, whereas lands discovered east of the meridian were to be Portuguese. It turned out that Portugal had pulled the short straw and got the worse end of the deal. Nevertheless, it provided the Portuguese with some comfort that the Spanish would recognise their territorial rights, allowing both powers to remain comfortable with their respective spheres of conquest.


Explaining the peace

How was the Thucydides trap avoided? This is difficult for the offensive realist, since one would presume that, if Spain really wanted to become hegemonic at sea, it would want to displace Portuguese power totally rather than reaching a compromise. Perhaps for the defensive realist the Treaty of Tordesillas makes sense, as it clearly delineated the limits of Spanish power, thus making Portugal comfortable that the balance of power remained in equilibrium. Thus the theory of  defensive (structural) realism offers the more convincing realist case on why the Thucydides trap did not have the effects that it had on Athens and Sparta in the 5th century BC. Perhaps what makes realism most convincing is that King John II was unprepared for another war with Spain.

But structural realism only looks at part of the picture.  Indeed, Portugal and Spain were trying to enhance their security in the Treaty of Tordesillas. But they were also trying to exploit the new world for prestige, and Tordesillas was not just a security treaty but also a 'charter of empire', allowing each power to go about conquering foreign lands unhindered (Disney, 2009). This supports classical realism, since prestige was more important than power in preventing the descent to war. 

Social constructivism has some backing. Spain was exhausted from its conquest of Granada and fearful of a repetition of the Wars of the Castilian Succession of 1475 - 79. This made Spain wary of war, due to its internal inhibitions and sentimental constraints on going to war again. Perhaps Isabella and Ferdinand thought that they could not justify yet another protracted war to their people. Nevertheless, this idea is, in the end, a flawed one: from 1500 to 1504, Spain was fighting with France for control of Naples, which was a fight that Spain won. Spain was, therefore, not prepared to stay at peace for the sake of peace. More realistic is the final possibility, from societal constructivism.

According to societal constructivists, like classical realists, inter-state sociology is important. Ideas such as prestige, respect, extravaganza and magisterial authority are far more significant than money, resources and power according to societal constructivists. The fact is that the international society of the time was oriented around papal authority. As such, when Alexander VI demanded a settlement, Spain and Portugal felt compelled, as fellow Catholic powers, to conform to the demands of their international society. They wanted to be good Christian members of Christendom, so did what they could to avoid war. The norm of Christian behaviour, in the end, may have been what drove Portugal and Spain to unite. 

As economic interdependence was not at its heights in the 15th century, and as neither Portugal nor Spain was democratic, the liberal theories of economic interdependence and democratic peace can be brushed aside. Therefore, on balance, the strongest three theories emerging from this case are:

Defensive (structural) realism;
Classical realism;
Societal constructivism.

Indeed, these three theories can be complementary, and no one of them need take precedence over the others. The material factors underpinning defensive realism, for instance, may be compatible with the sociological take of constructivists. Both survival and ideas were on the agenda of Spain and Portugal, suggesting that defensive realism, classical realism and societal constructivism can all offer lessons for explaining Thucydides's trap. 

Restraints on the trap

The norms of prestige and good behaviour, seen through these last two theories, were perhaps, therefore, the most important causes of peace in the Iberian peninsula. This means that the Thucydides trap, at least in this case, can be restrained by circumstances in which 

(a) war is unfeasible or unwinnable in the perception of both sides, 
(b) war risks decreasing, rather than increasing, the prestige of either state, or
(c) international society is structured in such a way that war would violate fundamental norms of good behaviour.

Which one of these circumstances has the greatest effect on restraining the Thucydides trap is a subject for future posts, when I hope to look at some of the Belfer Center's other case studies. As Professor Allison's team does not put immediate emphasis on theories such as constructivism, I hope that my analysis provides a broader look at what the mechanisms are for restraining the Thucydides trap. Providing an explanation for the age-old trap perhaps might prove useful for policymakers, who must avoid it ensnaring the United States and China. The alternative, if the historical record is reliable, is war.

References

Disney, A.R. (2009), A history of Portugal and the Portuguese empire, Cambridge: Cambridge University Press.
Harvard Thucydides’s Trap Project (2015 onwards), ‘Thucydides’s Trap Case File’. Available online at: http://www.belfercenter.org/thucydides-trap/resources/case-file-graphic

Wednesday, 31 May 2017

Word of the Week: Bipolarity

Perhaps bipolarity is more complex than cartoons such as this one, depicting Khrushchev and Kennedy at the height of Cold War tensions, tend to suggest. 

Bipolarity is generally used as a shorthand for a bipolar balance of power: the relative stability achieved when two states, or coalitions of states, have equivalent economic or military might. A system might be bipolar if states are polarised into two coalitions of states. And a system might possess a balance of power if the two coalitions have equivalent powers. Thus a bipolar balance of power promotes stability and, to some degree, peace. If two states, or coalitions of states, have equivalent powers, then they are less likely to challenge one another militarily, since this would risk large-scale retaliation. In this way, a bipolar balance, as structural realists such as Kenneth Waltz claim, makes pre-emptive strikes unlikely. This, in turn, makes war unlikely.

However, there is more than one type of balance of power. In his Theory of International Politics, Waltz claimed that the history of international relations could be divided into two halves. In the first half, up to 1945, the international system was multipolar. There were many centres of power, many opposing coalitions of states, and many individual states which had the capability of forming coalitions. From the Peace of Westphalia (1648), the defining value of international politics was state sovereignty. This was demonstrated by the maxim ‘cuius regio eius religio’, first coined at the Peace of Ausburg (1555) and approximately meaning ‘to each prince, his own religion. Therefore, no state ought to become so powerful that it could influence the ‘religio’ or sovereignty of other states. This means that a multipolar balance of power defined the world up to 1945, since states made sure that any potential hegemon – or dominant power – was counterbalanced by a coalition of other states committed to the balance of power.

In contrast, during the Cold War, according to Waltz, there was a bipolar balance of power. This is what made the Cold War unique. Whereas after the First World War no two states were vastly more powerful than all the others, after the second there was an approximate balance of power forming between the Soviet Union and the USA. So the USA and the USSR formed opposing coalitions which divided the world in two. These coalitions, NATO and the Warsaw Pact, succeeded in giving each power the military capability to deter the other from launching a first strike. This system of deterrence prevented any one centre of power from challenging the other, thus ensuring a continuation of the balance of power. This is exemplified by Article 5 of the North Atlantic Treaty, which says that ‘an armed attack against one or more of [the parties] shall be considered an attack against them all’. The bipolar balance between the USA and the USSR thus contributed to the polarisation which characterised the Cold War. But it also led to stability by reducing the incentive for launching a first, or pre-emptive, strike.

However, on balance it does not seem that the Cold War provided a perfect balance of power. Firstly, the balance of power was assymetrical, and thus was not a true ‘balance’ at all. The economic might of the USA outweighed that of the USSR, but, conversely, the USSR was on the verge of establishing hegemony over Europe. The USSR, in other words, was geopolitically secure, and was able to defend its 13 international borders partly by the threat of nuclear deterrence. The USA, on the other hand, was economically and militarily secure. As R. Harrison Wagner notes in ‘What Was Bipolarity?’, the balance of power during the Cold War is difficult to define, as the Cold War was different in practice from what it was meant to be in structural realist theory. A particular reason for this ambiguity was the asymmetrical distribution of power in practice.

Secondly, bipolarity suggests that only two states mattered during the Cold War. But it is clear that even the weaker states had a degree of freedom, making the system perhaps vaguely resemble a multipolar one. Even if the Cold War was bipolar, perhaps it was only weakly bipolar. However, Arthur Lee Burns suggests that the Cold War was characterised by bipolar tensions at the nuclear level, since only the USA and the USSR had enough nuclear weapons to each act as a ‘global-deterrent power’. By being able to deter military action against themselves through the threat of nuclear weapons, Burns suggests that the system was strongly bipolar, as only the USA and the USSR mattered in a nuclearized world. Nevertheless, Burns and Waltz, as well as other supporters of the bipolar view, fail to notice that even weak states without nuclear weapons mattered. After all, the Korean and Vietnam wars were fought for the sake of allies – namely, Korea and Vietnam – by both the USSR and the USA, who supported opposite proxies in the wars. Thus Stephen Walt says that the Cold War was a ‘competition for allies’. The Cold War, in other words, was not perfectly bipolar – even the weak states mattered. Rather, the Cold War was weakly bipolar, as peace, although made vastly more probable by the new balance of power, was never guaranteed.


The Cold War provided a bipolar balance of power in more ways than one. Nevertheless, the system in reality was more nuanced and multifaceted that structural realists make it out to be. Perhaps even the Cold War, one of the few examples in world history of a (weakly) bipolar system, may not be completely unique.