🎵 Music · Undergraduate · MUS 320

World Music & Ethnomusicology

A survey of musical traditions across Africa, Asia, the Middle East, the Americas, Europe, and the Pacific, taught alongside the field that studies them. You begin with what ethnomusicology is, how it differs from musicology, and how it grew out of comparative musicology, wax cylinder collecting, and a colonial archive whose terms of collection are still being renegotiated. You then build a…

Start the interactive course (quizzes, progress, videos) →

Free forever. No sign-up, no ads. 15 lessons. The full lesson text is below so you can read it right here.

Module 1: What Ethnomusicology Is and How It Works

The discipline that studies music as human activity: how it differs from musicology, how its archive was built out of colonial encounters and then reckoned with, how fieldwork is actually done and what it owes the people recorded, and a vocabulary of timbre, texture, meter, tuning, form, and improvisation that works on any music on earth.

What Ethnomusicology Is

  • Define ethnomusicology and distinguish it from historical musicology.
  • Apply Merriam's three-part model of concept, behavior, and sound to a single performance.
  • Explain why definitions of music vary across societies and what follows from that.
  • State honestly what a text-only course can and cannot give you, and plan around it.

The big picture

In 1959 Alan Merriam settled into the village of Lupupa Ngye, in Basongye country in what was then the Belgian Congo, with a tape recorder, a notebook, and a question he expected to be easy: what is music? The answers came back in a shape he had not planned for. Music requires human beings, his hosts told him. A bird singing in a tree is not making music, whatever it sounds like. And music is distinguished from noise not by any acoustic property you could measure but by the state of the person producing it. When you are content, you sing; when you are angry, you make noise.

Sit with that for a second. It is not a quaint saying. It is a working definition that sorts sounds by their human source rather than by their waveform, and it does a job that the definition you probably carry around (music is organized sound, music is an art form) does not do. Two men could produce identical sound and only one of them would be making music. If you want to know whether a sound is music in Lupupa Ngye, listening to it is insufficient. You have to know something about the person.

Ethnomusicology is the discipline built around taking that discovery seriously and then repeating it everywhere. It studies music as a human activity: what people make with sound, what they believe about it, what they do while it happens, who is allowed to make it, what it accomplishes, and how it is passed to the next person. Sound is part of the object of study but not the whole of it. This lesson defines the field, distinguishes it from its cousin disciplines, shows the model that has organized it for sixty years, and tells you plainly what this course, being made of text, cannot do for you.

The prefix problem

The name is awkward and worth confronting before it does damage. The Dutch scholar Jaap Kunst coined the hyphenated form ethno-musicology in 1950 to replace the older German term vergleichende Musikwissenschaft, comparative musicology, which he thought overpromised. The new name stuck. But the ethno- prefix carries an implication that has taken decades to shake off: that there is musicology, which studies music, and then ethnomusicology, which studies the music of other people, presumably ones with ethnicity, which the observer does not have.

That implication is dead in the discipline and alive in general usage, so mark it now. Ethnomusicologists study country music in Nashville, karaoke bars in Osaka, church choirs in Ohio, Norwegian black metal, and orchestral conservatories. Henry Kingsbury spent a year inside an American conservatory taking field notes on how talent was talked about, and published the results in 1988 as a study of a culture, which is exactly what it was. Bruno Nettl wrote a book about a midwestern American music school as if it were a distant society with its own rituals and hierarchies, and the joke landed because the description fit. There is no repertoire that belongs to ethnomusicology and no repertoire exempt from it.

Key idea: Ethnomusicology is defined by its questions and methods, not by the music it studies; any music can be its subject, including the music the researcher grew up in.

Merriam's model: concept, behavior, sound

In 1964 Merriam published The Anthropology of Music and proposed a model that the field has been arguing with productively ever since. Any musical practice, he said, has three levels, and they cycle into each other.

  • Concept: what people believe about music. Where it comes from, what it can do, who owns it, whether it is dangerous, whether it should be sold, what makes a performance good.
  • Behavior: what people physically and socially do. How the body moves, where people sit, who speaks during a performance, how a musician is paid, what a student does for a teacher for eight years before being taught anything.
  • Sound: the acoustic result. Pitches, timbres, rhythms, textures, forms, the thing a microphone can capture.

The cycle matters more than the list. Concepts shape behavior; behavior produces sound; the sound is then judged against the concepts, which adjust. Change any one level and the others move. When a Javanese gamelan is moved from a palace courtyard to a university room in California, the sound can be reproduced almost exactly while the concept and behavior change completely, and something real has happened that a recording will not show you.

John Blacking pushed in a related direction from a different site. He spent 1956 to 1958 among Venda communities in the northern Transvaal of South Africa and came away with a definition that fits on a business card: music is humanly organized sound. His larger claim in How Musical Is Man? (1973) was sharper and more uncomfortable for readers in Europe and North America. Venda children, he observed, acquired rhythmic and polyphonic skills that European musicians found difficult, and they acquired them the way children acquire language, without selection or special training. If musical capacity is that widely distributed, then the European habit of treating musicality as a rare gift possessed by a talented few is not a fact about human beings. It is a fact about how one society organizes access to music.

Key idea: Sound is only the visible third of a musical practice; concepts and behaviors are equally part of what a music is, and only sound survives on a recording.

How this differs from historical musicology

Atlas offers History of Western Music (MUS 210) and Music Theory I (MUS 101), and this course deliberately does not repeat them. If you want notation, functional harmony, and the line from plainchant to Stravinsky, those are the places. The contrast between those courses and this one is a useful way to see what changes when the discipline changes.

Historical musicologyEthnomusicology
Typical questionWhat did this composer write, when, and why does the piece work as it does?What do these people do with sound, what do they believe about it, and how did you find out?
Primary evidenceScores, manuscripts, letters, treatises, publishers' recordsFieldwork: participation, observation, interviews, recordings, lessons
Where the work happensArchives and librariesA house, a courtyard, a shrine, a rehearsal room, a bar, then an archive
Unit of studyThe work, the composer, the style periodThe performance, the practice, the community, the process of transmission
Attitude to notationThe score is the primary documentNotation is one tool, often absent, and always a translation that loses information

Two things about the table. First, the columns leak: historical musicologists do ethnographic work, ethnomusicologists work in archives, and plenty of scholars refuse the distinction outright. Second, the sharpest difference is the middle row. In this field, the researcher is a participant in the thing being studied, which means the researcher is part of the data and has to be accounted for. That has consequences the next two lessons are entirely about.

The field is not standing outside the room

Here is the uncomfortable structural fact, stated concretely rather than solemnly. A researcher arrives with a grant, a recorder, and a return ticket. A musician gives months of teaching, sings songs that took a lifetime to learn, and is recorded. The researcher goes home, writes a book, gets tenure, and the recordings enter a university archive under the researcher's name. The musician may receive a cassette copy, a payment for a few sessions, and a thank-you in the acknowledgments. This was the standard arrangement for most of the twentieth century, and it produced the recorded archive that this course depends on.

Nothing about that arrangement was inevitable, and a good deal of the field's recent energy has gone into changing it. Paul Berliner studied mbira with Shona musicians in Zimbabwe from the 1970s, and when the fullest version of that work appeared in 2020 it carried two names on the cover: Berliner and the mbira player Cosmas Magaya, whose knowledge it mostly contains. Recordings held in Washington and Berlin have been copied and returned to the communities they came from. Consent forms now specify what may be published, sold, streamed, and withheld. None of this makes the original asymmetry vanish, and you should read every recording you encounter in this course with a question attached: who made it, under what agreement, and who benefits from it now.

Key idea: Every recording is the residue of a relationship, and the terms of that relationship are part of what the recording documents.

What this course cannot do

Be clear about the shape of the hole. This course is made of text. Music is made of sound. No sentence about the buzz of an mbira is the buzz of an mbira, and no paragraph about Umm Kulthum holding a single line for eleven minutes while an audience shouts at her will do what eleven minutes of Umm Kulthum does. The descriptions here are aimed at one specific job: giving you enough structure that when you do listen, you can hear parts instead of a wash.

So use the course this way. Read a passage. Then go find the music and listen, ideally twice. Then come back and reread the passage, which will now be about something. Two archives make this free and legal:

  • Smithsonian Folkways at folkways.si.edu. When the Smithsonian acquired Moses Asch's Folkways Records in 1987, it committed to keeping every one of the roughly 2,000 titles permanently available. The catalog now runs to tens of thousands of tracks with liner notes, most of them streamable, and it covers a large share of the traditions in this course.
  • The Library of Congress at loc.gov, especially the American Folklife Center, which holds the cylinder recordings, field collections, and the archive of folk song built from the 1920s onward.

Where a lesson names a specific piece, that name is a search term, not decoration. Nhemamusasa, Etenraku, Erquan Yingyue, Rokudan no Shirabe: type them in.

One minute of mbira, described three ways

To see the model do work, take a single performance: the Shona piece Nhemamusasa played at a bira, an all-night ceremony in Zimbabwe held to call ancestral spirits into the gathering. You will meet this music properly in Module 2. Right now it is a demonstration.

At the level of sound, you hear a metal-keyed instrument: twenty-two to twenty-four forged keys in three ranks on a hardwood soundboard, plucked with two thumbs and one index finger, set inside a large gourd resonator studded with bottle caps that rattle so that every note arrives with a halo of buzz. Two players interlock. The first, kushaura, lays out the cycle; the second, kutsinhira, plays a part that falls between the first player's notes, so the two together produce a stream of pitches faster than either is playing. The cycle is forty-eight pulses long and repeats, with variation, for twenty minutes or two hours. Over it a hosho, a pair of gourd rattles filled with seeds, plays a fast triple pattern that never varies. Voices come and go: a low wordless line, a high line, a half-spoken recitation of family names.

At the level of behavior, it is one o'clock in the morning in a room packed with relatives. The players sit on the floor. People clap on the offbeat, sing when they know the part, get up to dance, sit down, drink millet beer, hold conversations. Nobody applauds. A woman begins to shake; two people move toward her and support her; she is helped to a mat. The music does not stop for any of this. It must not stop, and players spell each other so it does not.

At the level of concept, the music is doing a job. The vadzimu, the family's ancestral spirits, are understood to be drawn by mbira, and the shaking woman is a medium through whom a specific dead relative may now speak to the assembled family about a specific problem: an illness, a dispute over land, a debt. The performance is judged good if the spirits come. Not if it is in tune, not if the audience is moved, not if the recording sells.

Notice what a sound-only account would have said: a cyclic piece for two lamellophones and rattle, roughly 48 pulses, hypnotic. All true, all useless. It would have missed that the music is a telephone.

Three words to handle carefully

Traditional is the most abused word in this subject. Japanese kumi-daiko, the drum ensemble now performed on stages worldwide and universally received as ancient, was invented in 1951 by a jazz drummer named Daihachi Oguchi in Nagano. Korean samul nori was founded in February 1978 by four musicians in a small Seoul theater. Balinese kecak took its current form in the 1930s in collaboration with the German painter Walter Spies, partly for visitors. All three draw on genuinely old materials. All three are twentieth-century creations. Traditions are not fossils; they are practices people keep making, and the ones that stop changing are usually the ones in a museum.

Folk arrived from nineteenth-century European nationalism, where it named the pure rural voice of a nation as opposed to the cosmopolitan city. It carries that political freight wherever it goes.

Authentic is almost always a claim being made by someone about someone else, usually in a marketplace. Ask who is making the claim, to whom, and what they gain. When a record label calls an album authentic, it means it will sell to people who want that; it rarely means anything a musician in the tradition would recognize.

Key idea: Traditional, folk, and authentic are claims made by particular people for particular purposes, not neutral descriptions of a music's age or purity.

Common misconceptions

  • "Ethnomusicology is the study of non-Western music." It is the study of music as human activity, using ethnographic method. Its subject can be a Balinese village, a Berlin techno club, or a high school marching band.
  • "Music is a universal language." Music is universal in the sense that every known society has it. It is not a language and it does not translate: the same sounds carry different meanings, and a piece can be a lullaby in one place and a summoning in another.
  • "If a music has no notation it must be simple." Notation is a storage technology, not a measure of complexity. The Shona mbira repertoire and the Persian radif were transmitted without notation and are more intricate than plenty of notated music.
  • "A recording captures the music." It captures the sound. Merriam's other two levels, and the entire social event, fall outside the microphone.
  • "Traditional means unchanged." Most living traditions change constantly; several famous ones are less than a century old.

Recap

  • Ethnomusicology studies music as human activity in context, defined by method and question rather than by repertoire.
  • Merriam's model separates concept, behavior, and sound, and treats them as a cycle in which each shapes the others.
  • Blacking's humanly organized sound implies that musical capacity is ordinary rather than rare, and that scarcity of musicians is a social arrangement.
  • The researcher is inside the situation being studied, which is why fieldwork ethics is a technical subject in this field and not an afterthought.
  • A text course can give you structure to listen with; it cannot give you the sound. Smithsonian Folkways and the Library of Congress can, for free.

Sources

  1. "Ethnomusicology." Encyclopaedia Britannica. Retrieved from britannica.com.
  2. Society for Ethnomusicology. About the society and the field. ethnomusicology.org.
  3. "John Blacking." Wikipedia. Summary of How Musical Is Man? (1973) and the Venda fieldwork. en.wikipedia.org.
  4. Smithsonian Folkways Recordings. Streaming catalog and liner notes. Smithsonian Institution. folkways.si.edu.
  5. American Folklife Center. Collections and field recordings. Library of Congress. loc.gov.
Key terms
Ethnomusicology
The study of music as human activity in its social and cultural context, based on fieldwork.
Merriam's model
The three-level analysis of any music as concept, behavior, and sound, each shaping the others.
Humanly organized sound
John Blacking's definition of music, which locates the definition in human intention rather than acoustics.
Fieldwork
Extended firsthand research in which the researcher participates in the musical life being studied.
Bira
An all-night Shona ceremony in Zimbabwe at which mbira music is played to call ancestral spirits.
Kushaura and kutsinhira
The lead and interlocking second mbira parts, whose notes fall between each other's.
Comparative musicology
The earlier German-language name for the field, replaced by ethnomusicology after 1950.
Reflexivity
Accounting for the researcher's own position, funding, and effect on what is observed.

Fieldwork: Methods, Transcription, and Ethics

  • Describe the core methods of musical fieldwork, including participant observation and apprenticeship.
  • Explain what transcription can and cannot represent, and choose an appropriate notation for a given music.
  • Apply a practical ethics checklist covering consent, restriction, credit, and payment.
  • Explain why the ownership of a field recording is rarely settled by copyright alone.

The big picture

Simha Arom had a problem that no amount of careful listening was going to fix. In the forests of the Central African Republic in the early 1970s, he was recording Aka horn ensembles in which each player sounds one or two pitches, entering at a different point in a fast cycle, so that the melody you hear is not played by anybody. It emerges from the gaps. Arom could hear the result. He could not extract the parts, and neither could a spectrogram, because the parts overlap in pitch and timbre.

His solution used the recorder as an instrument of analysis. He recorded one player alone. Then he played that recording back through headphones to a second player, who performed their part against it while a second track captured only the new part. Then a third, and a fourth, each hearing the accumulating whole but recorded in isolation. At the end he had every part separately and the assembled piece, and he could finally see how the interlock worked. Then he did the step that matters most: he played his reconstruction back to the musicians and asked whether it was right. They corrected him.

That last move is the whole method in miniature. Fieldwork is not extraction followed by interpretation somewhere else. It is a long argument with people who know more than you do about the thing you came to study. This lesson covers how that argument is conducted: what fieldworkers actually do, what happens when sound is turned into notation, and what is owed to the people on the other side of the microphone.

Participant observation, and why it takes so long

The default expectation in this field is a year or more in one place, and the reason is not thoroughness for its own sake. A year covers a full ritual and agricultural calendar, so you see the music that happens once. It is roughly the time it takes to become unremarkable, which is when people stop performing for you. And it is about how long it takes to be wrong in public often enough to learn the categories that matter locally.

Consider the difference between two events that produce identical audio. In the first, a singer performs at a wedding while you happen to be present with a recorder. In the second, you ask a singer to perform the same song at ten in the morning in her kitchen so you can record it cleanly. The second is called an elicited performance, and it is not a lesser or fake version, but it is a different event: different audience, different stakes, often different length and different tempo, and frequently a different choice of what to sing, because people select for the recorder. Write down which one you have. A great deal of confusion in the literature comes from elicited recordings described as if they were ceremonies.

Key idea: The presence of a researcher and a recorder is part of the event, and the field notes have to say what kind of event it was.

Learning to play

Mantle Hood's bi-musicality, from Lesson 2, is a research method and not a hobby. Hands teach things ears cannot. You will not understand why the Shona mbira parts kushaura and kutsinhira are described as one thing rather than two until your thumbs have tried to play the second against the first and your ear has flipped, involuntarily, so that the composite pattern becomes the melody and your own part disappears. You will not understand a tabla theka until you have counted sixteen while your teacher claps on beats one, five, and thirteen and waves an open hand on nine.

Apprenticeship also teaches the social structure directly, because you have to enter it. In North Indian classical music you become a shishya, which historically meant living with a teacher and serving the household for years before receiving the good material. In a Mande jeli family, knowledge follows hereditary lines and an outsider's access has limits that are explained politely once and thereafter assumed. In a Ghanaian drumming ensemble you begin on the bell, which is where the learning is anyway.

Recording, and the metadata that makes it worth keeping

Field recordists argue about microphones. They agree about metadata. A beautiful recording with no documentation is close to worthless, and archives are full of them: a cylinder labelled "song, Africa, 1907" is a sound with no address. Whatever you record, capture at minimum:

  • Names of every performer, spelled as they want them spelled, plus roles and instruments.
  • Date, place, and the occasion, including whether the performance was elicited.
  • What the item is called, by whom, and what genre category it belongs to locally.
  • What permission was given, by whom, and for what uses. Specifically: archive, publication, streaming, commercial release, teaching.
  • Anything the performers say should not be circulated.

Record ambient sound too. The clapping, the talking, the dogs, the length of the pauses between pieces: these are data about behavior, and a recording engineered to isolate the music alone has thrown away half the evidence.

Transcription is analysis, not evidence

In 1958 Charles Seeger drew a distinction that clears up years of confusion. Prescriptive notation is a set of instructions to a performer: a score. Descriptive notation is a report of what was heard: a transcription. They look similar and do opposite jobs. A score can leave out everything a competent performer already knows. A transcription cannot leave out anything you want to claim.

Western staff notation is a superb prescriptive tool that carries assumptions everywhere it goes. It assumes twelve equal semitones, so a neutral third in maqam Rast becomes either E natural or E flat, both wrong. It assumes bars with a strong first beat, so a West African bell pattern acquires a downbeat it does not have. It assumes fixed durations, so the flexible pulse of an alap or a pansori line looks like error. Every one of these is a distortion introduced by the paper, not observed in the music.

Alternatives exist and you should pick by the job:

NotationGood forWhat it assumes
Staff notationPitch and rhythm in twelve-tone, metered musicEqual temperament, bars, fixed durations
Time Unit Box System (boxes on a grid)Cyclic percussion, interlocking parts, timelinesAn even underlying pulse and a repeating cycle
Cipher notation (numbers)Gamelan, Chinese jianpu, any fixed scale-degree systemA scale defined by degree rather than absolute pitch
TablatureInstruments where the action matters more than the pitch, such as the qinA specific instrument and technique
Cents graphs and spectrogramsMicrotonal shading, timbre, vocal ornamentNothing about scale, but hard to read musically

Here is a bell pattern in a box grid, twelve pulses across, x for a stroke. You will meet it properly in Lesson 5.

123456789101112
xxxxxxx

That grid is honest about what it claims: twelve pulses, strokes on seven of them, no assertion about which one is the downbeat. Staff notation cannot make that last refusal, because writing it down at all requires choosing a time signature.

Key idea: Any transcription is already an interpretation; choose a notation whose assumptions match the music rather than one whose assumptions you are used to.

Ethics, in specifics

Consent is the beginning and it is not a signature. Real informed consent means the person understands what will happen to the recording, which is hard to convey honestly when the answer includes a streaming platform that did not exist when you asked. Practical standards now in use:

  • Layered permission. Ask separately about archiving, teaching use, publication with the recording attached, and commercial release. People say yes to some and no to others.
  • Restricted repertoire is real. Many traditions have material that may be performed only by certain people, at certain times, or only in ceremony. Navajo, Hopi, and many Australian Aboriginal communities have categories of song whose circulation is a harm regardless of the singer's individual agreement. Ask what should not be recorded before you ask what can be.
  • Who can consent for a group. An individual can consent to their own performance. A community's repertoire is a different question, and the answer usually requires talking to people who are not in the room.
  • Pay people. Session fees are the minimum. If a recording earns money, the performers should share in it by written agreement, not by good intentions.
  • Credit by name. The convention of anonymous informants is dead. People who make music are musicians and belong on the label.

The Society for Ethnomusicology maintains a position statement on ethics and has taken public positions on specific abuses. In 2007 it adopted a statement condemning the use of music as an instrument of torture, after reports that recordings were used at high volume and duration against detainees in United States custody. The statement is short and worth reading as an example of a scholarly society deciding that its subject matter obliged it to say something.

Hugh Tracey and the long tail of a collection

The case that keeps this from being abstract: Hugh Tracey recorded across southern, central, and eastern Africa from 1929 into the 1970s, founded the International Library of African Music at Rhodes University in South Africa in 1954, and left roughly 35,000 recordings. For a great many practices his tapes are the only documentation that exists, and musicians and scholars across the continent use them constantly.

He also recorded as a white settler working in colonial and then apartheid southern Africa, with the access that gave him, and he licensed material commercially through arrangements the performers had little part in negotiating. Some of his released series named performers carefully; some did not. ILAM has spent decades on the consequences: identifying performers, returning recordings to communities and national archives, and building databases that let people find their grandparents' voices. Both halves of that sentence are the record. The collection is indispensable and the terms it was made on would not be acceptable now, and the work of the last thirty years has been to make the second fact do something rather than just be noted.

Key idea: Who owns a field recording is rarely settled by copyright, which usually vests in the recorder; deposit agreements, community protocols, and Traditional Knowledge labels do the work copyright will not.

A checklist you can actually use

If you ever record anyone, from a family member singing at a holiday to a band in a bar, run this:

  1. Ask before recording, and say specifically where the recording will go.
  2. Ask what should not be recorded or shared, and write down the answer.
  3. Get names and spellings. Get instruments and roles.
  4. Note date, place, occasion, and whether you asked for the performance.
  5. Offer a copy, and then actually deliver it.
  6. If you publish, name the musicians in the first line, not the acknowledgments.
  7. If money appears, go back and negotiate before it is spent.

Common misconceptions

  • "A signed consent form settles the ethics." It settles the paperwork. Consent has to be specific about uses and revisited when the uses change.
  • "Transcription is objective and prose description is subjective." Every transcription encodes decisions about what counts. Staff notation encodes a great many of them silently.
  • "You can analyze interlocking parts by listening harder." Sometimes you cannot, which is why Arom built a method to separate them, and why he checked it with the players.
  • "Field recordings belong to the people recorded." Legally they usually belong to whoever pressed record. That gap between law and justice is the reason for deposit agreements and TK labels.
  • "Old collections made under colonial conditions should be sealed." Communities themselves overwhelmingly want access to them. The argument is about who controls the terms, not about erasure.

Recap

  • Fieldwork is long because a year covers the calendar, and because it takes that long for people to stop performing for you.
  • Elicited and observed performances are different events; the documentation must say which.
  • Learning to play reveals structure that listening does not, and inserts you into the social system of transmission.
  • Transcription is analysis; choose the notation whose assumptions fit the music, and be explicit about what it refuses to claim.
  • Ethics is specific work: layered consent, restricted repertoire, group authority, payment, credit, and honest handling of legacy collections such as Tracey's.

Sources

  1. Society for Ethnomusicology. Position statements on ethics and on the use of music as torture. ethnomusicology.org.
  2. International Library of African Music. The Hugh Tracey collection and its repatriation programs. Rhodes University. ru.ac.za.
  3. Local Contexts. Traditional Knowledge Labels and community protocols. localcontexts.org.
  4. "Simha Arom." Wikipedia. The re-recording method for analyzing Central African polyphony. en.wikipedia.org.
  5. American Folklife Center. Fieldwork guidance and collection documentation standards. Library of Congress. loc.gov.
Key terms
Participant observation
Extended research in which the fieldworker takes part in the activity being studied rather than only watching it.
Elicited performance
A performance produced at the researcher's request, which is a different social event from a performance in its usual setting.
Prescriptive notation
Notation that instructs a performer, such as a score; it may omit what performers already know.
Descriptive notation
Notation that reports what was heard; it must include anything the analyst wants to claim.
Time Unit Box System
A grid notation marking strokes on an even pulse, useful for cyclic percussion because it need not assign a downbeat.
Layered permission
Consent obtained separately for archiving, teaching, publication, and commercial release.
Restricted repertoire
Material a tradition holds may be performed or heard only by certain people or at certain times.
ILAM
The International Library of African Music, founded by Hugh Tracey in 1954, holding roughly 35,000 recordings.

A Vocabulary for Any Music

  • Describe timbre, texture, meter, tuning, form, and improvisation in terms that do not assume any one tradition.
  • Measure and compare tunings in cents, treating equal temperament as one solution among several.
  • Distinguish divisive, additive, cyclic, and free approaches to time.
  • Apply a six-question listening protocol to an unfamiliar recording.

The big picture

In the summer of 1889 Claude Debussy, then twenty-six, spent time at the Paris Exposition Universelle in front of the Javanese pavilion, where a gamelan from the court of Solo played daily for visitors. He returned repeatedly. Years later he wrote that Javanese music followed a counterpoint next to which Palestrina looked like a child's game, and that the percussion of Java made European percussion sound like the noise of a travelling circus. He had heard something whose organization he could hear was rigorous and whose rules he had no words for.

Here are some of the words. A Javanese slendro tuning divides the octave into five steps of roughly 240 cents each. A piano divides it into twelve of exactly 100. Neither is a version of the other with errors. This lesson builds a vocabulary that lets you describe any music without smuggling in one tradition as the standard, organized under six headings: timbre, texture, time, pitch, form, and improvisation. It is the toolkit for the eleven lessons that follow, so it is worth the half hour.

Timbre: the thing you recognize first and discuss last

You can identify a voice on the phone in two syllables. That is timbre, and it is the fastest-acting dimension of music and the one most often left out of analysis because notation has no place to write it. Describe it along a few axes: attack (how the sound starts: struck, plucked, bowed, blown, breathy, sudden, gradual), decay (how it dies: immediate, ringing, sustained by breath or bow), brightness (how much energy is in the high partials: a nasal double reed is bright, a stopped flute is dark), and noise content (how much of the sound is not pitched at all).

That last one deserves emphasis, because it is where European conservatory habits mislead hardest. In many traditions deliberate noise is not a defect but the point:

  • The Shona mbira is played inside a gourd hung with bottle caps or shells so that every note carries a rattling halo. Remove the buzz and Shona musicians will tell you the instrument sounds thin and the ancestors will not attend.
  • The Japanese shamisen is built with a deliberate contact point called sawari that makes the lowest string buzz, and the sitar has a broad curved bridge, the jawari, that produces a shimmering edge on every note.
  • A shakuhachi player can use muraiki, a violent breathy attack, in which the audible rush of air is the expressive event.
  • Across West and Central Africa, rattles, metal jingles, and buzzing membranes are attached to xylophones, harps, and drums on purpose.

Key idea: A great deal of the world's music treats noise as a component of beauty; recording, mixing, or instrument design that cleans it up is destroying content, not improving quality.

Texture: how many things are happening and how they relate

TextureDefinitionExample in this course
MonophonyOne melodic line, whatever the number of performersPowwow singing in octaves; unaccompanied cantillation
HeterophonyOne melody, several simultaneous variants of itChinese jiangnan sizhu; Japanese gagaku; Cajun fiddle and accordion
PolyphonyIndependent lines of comparable weightBaAka forest polyphony; Georgian three-voice song; Sardinian cantu a tenore
HomophonyA melody with subordinate accompanimentZulu isicathamiya; most global popular music
Drone-basedMelody over a continuous fixed pitch or pitchesIndian tanpura; Highland pipes; Persian and Turkish practice
Interlocking or hocketA single line distributed between players, none of whom plays itBalinese kotekan; Andean sikuri panpipes; Aka horns

Heterophony is the one Western-trained listeners consistently mishear as sloppy unison. It is not. In a Shanghai teahouse ensemble each instrument plays the same tune in the idiom natural to it: the flute adds ornaments, the pipa breaks the line into plucked figures, the erhu slides between the notes. Nobody is out of time. The tune is being viewed from several angles simultaneously, and the slight blur is the aesthetic.

Time: four different ways to organize it

Divisive meter takes a span and divides it: four beats each split into two or three. This is the default in European and most Anglophone popular music, and the notation was built for it.

Additive meter builds a cycle by adding unequal groups. Bulgarian dance meters are the classic case, and they are counted as short and long rather than as fractions:

DanceTotal unitsGroupingFeel
Rachenitsa72 + 2 + 3quick quick sloooow
Daichovo92 + 2 + 2 + 3three quicks then a long
Kopanitsa112 + 2 + 3 + 2 + 2the long sits in the middle

Dancers do not count to eleven. They learn a foot pattern in which one step takes longer, exactly as an English speaker does not count syllables to say the word "elephant" correctly. Turkish usul cycles and many Middle Eastern rhythmic patterns work the same way.

Cyclic time organizes music into a repeating span marked by fixed events, where position within the cycle matters more than a bar line. A Javanese gongan ends when the great gong sounds; everything is heard as approaching that stroke. An Indian tala is a cycle of counted beats whose first beat, sam, is the point of arrival that improvisations aim for. A West African bell pattern is a fixed timeline that all other parts orient to. In each case the cycle is the frame, and the interesting question is where in the cycle an event falls.

Free or unmeasured time has pulse but no fixed meter: the alap that opens a raga performance, an Arab taqsim, plainchant, much pansori. It is not the absence of rhythm. It is rhythm governed by phrase and breath rather than by a grid.

Layered on top of these is cross-rhythm: two conflicting groupings sounding at once, most often three against two. Say "nice cup of tea" evenly against "pass the jam" evenly, both filling the same six-unit span, and you have it in your mouth. Lesson 5 builds an entire ensemble practice out of that relationship.

Key idea: Meter is not one thing. Divisive, additive, cyclic, and free are four distinct organizing principles, and reading all music as divisive is the most common analytic error a Western-trained listener makes.

Pitch: cents, and equal temperament as a local solution

Use Ellis's cents from Lesson 2: 1200 to the octave, 100 to an equal-tempered semitone. Now compare some intervals.

IntervalPure (small whole-number ratio)12-tone equal temperamentDifference
Octave (2:1)1200 cents1200 centsnone
Fifth (3:2)702 cents700 cents2 cents flat
Major third (5:4)386 cents400 cents14 cents sharp

Fourteen cents is clearly audible. Every major third on a modern piano is noticeably wide, and listeners raised on pianos have simply learned to accept it. Why accept it? Because equal temperament solves one specific problem extremely well: on a fixed-pitch keyboard with twelve notes per octave, it lets you play in all keys and modulate freely between them, at the cost of making almost every interval slightly wrong. That was a pressing European problem from roughly the seventeenth century onward. It is not a universal problem. A tradition with no keyboard, no modulation, and no need to transpose has no reason to pay that price.

It is also worth knowing who got there first. Zhu Zaiyu, a prince of the Ming dynasty, published the correct mathematics of equal temperament in 1584, computing the twelfth root of two to many decimal places, decades before Mersenne in France. The system that Europeans often present as the endpoint of musical progress was worked out in China and never widely adopted there, because Chinese music had no use for it.

Now some real tunings, in cents from the lowest tone:

SystemApproximate stepsNote
12-tone equal temperament0, 100, 200, 300, ... 1200Every semitone identical by construction
Javanese slendroFive steps near 240 eachNear-equal, but every gamelan is tuned individually and none matches another
Javanese pelogSeven markedly unequal stepsSome intervals near 120, others near 250 or more
Arab maqam RastIncludes a neutral third near 350Between a minor third (300) and a major third (400)
Thai classicalSeven near-equal steps near 171Fits neither Western semitones nor whole tones

The Javanese detail is the one people find hardest to accept: two gamelans in neighboring villages are tuned differently on purpose, each set has its own character called embat, and you cannot borrow one instrument for another ensemble. The set is the instrument.

Key idea: Equal temperament is a compromise engineered for keyboard modulation, not a natural baseline; measured in cents it is simply one tuning among many, and it is out of tune by the standards of most of them.

Melody: mode is more than a scale

In many traditions the melodic framework carries far more information than a list of pitches. A raga specifies which notes ascend and which descend, which notes are emphasized and which are touched in passing, characteristic short phrases, required ornaments, and often a time of day. A maqam specifies its building blocks and a typical path through them. A Javanese pathet governs which pitches can end a phrase and which register the melody occupies. Call all of these modal systems and remember that translating one as a scale throws away most of it. The scale is the alphabet; the mode is the handwriting, the vocabulary, and the accent.

Form and improvisation

Common formal principles across the course: strophic (repeat the tune with new words), cyclic (repeat a fixed span with variation, as in gongan, tala, or an mbira cycle), call and response (leader and group alternate), expanding or accelerating suite (slow unmeasured opening leading to progressively faster metered sections, as in Hindustani alap-jor-jhala, Korean sanjo, Turkish fasil, Persian dastgah), and theme and variation.

Improvisation is best understood as a spectrum with a model at one end and a realization at the other. Nothing is invented from nothing, and almost nothing is fully fixed. A taqsim, an alap, and a jazz solo are all governed by rules a listener inside the tradition can hear being followed and broken. The useful question is not whether something is improvised but what is fixed and what varies: which layer holds still (the bell, the drone, the gong cycle, the chord changes) and which layer moves.

A six-question listening protocol

  1. What is making the sound, and what does its timbre emphasize, including noise?
  2. How many independent things are happening, and how do they relate?
  3. Is there a pulse? Is it divisive, additive, cyclic, or free? Can you find the cycle length?
  4. What does the pitch material do: how many notes, is there a drone, do intervals sound like anything on a piano?
  5. What is fixed and what varies as the piece proceeds?
  6. What appears to be the occasion, and who is the music for?

Common misconceptions

  • "Music without harmony is simpler." Complexity is distributed differently. Traditions without functional harmony often carry extraordinary complexity in melody, ornament, timbre, and rhythm.
  • "Out-of-tune singing." Before saying that, measure. A neutral third at 350 cents is precisely placed, and calling it out of tune only says your reference is 12-tone equal temperament.
  • "Heterophony is unison played badly." It is a deliberate texture in which each instrument varies a shared melody in its own idiom.
  • "Additive meters are irregular." They are highly regular; the units are simply of two lengths rather than one.
  • "Improvised means made up on the spot with no rules." Improvised traditions are typically the most rule-governed, which is what makes the variation legible.

Recap

  • Timbre is describable by attack, decay, brightness, and noise content, and deliberate buzz is an aesthetic goal in many traditions.
  • Six textures cover most cases; heterophony and interlocking are the two most often misheard.
  • Time comes in divisive, additive, cyclic, and free varieties, with cross-rhythm layered over any of them.
  • Cents make tunings comparable; equal temperament is a keyboard compromise, mathematically solved in Ming China in 1584 and not adopted there.
  • Modal systems carry hierarchy, motion, and ornament, not just a set of pitches, and improvisation is a spectrum between model and realization.

Sources

  1. "Tuning and temperament." Encyclopaedia Britannica. britannica.com.
  2. "Cent (music)." Wikipedia. The logarithmic unit and its use in comparing scales. en.wikipedia.org.
  3. "Slendro." Wikipedia. Near-equidistant five-tone Javanese tuning and its variability between ensembles. en.wikipedia.org.
  4. "Heterophony." Wikipedia. Simultaneous variation of a single melody. en.wikipedia.org.
  5. Smithsonian Folkways Recordings. Streaming archive for the listening protocol in this lesson. folkways.si.edu.
Key terms
Timbre
Sound quality, describable by attack, decay, brightness, and noise content; often the first thing a listener recognizes.
Heterophony
A texture in which several performers play simultaneous variants of one melody.
Hocket
A single melodic line distributed between performers so that no one player produces it alone.
Additive meter
A cycle built from unequal groups, such as 2+2+3, counted as short and long rather than as equal beats.
Cyclic time
Organization by a repeating span marked by fixed events, such as a gong stroke or the sam of a tala.
Cross-rhythm
Two conflicting groupings of the same span sounding at once, most commonly three against two.
Equal temperament
A tuning dividing the octave into equal steps, adopted in Europe to permit keyboard modulation at the cost of pure intervals.
Modal system
A melodic framework specifying hierarchy, characteristic motion, and ornament, not merely a set of pitches.
Embat
The individual tuning character of a particular gamelan set, which makes its instruments non-interchangeable.

Module 2: Sub-Saharan Africa

West African ensemble drumming taught from the bell outward: cross-rhythm, entry points, and how parts lock, plus talking drums and the argument about what African rhythm actually is. Then the Shona mbira and the ceremony it serves, the Mande jeli and the kora, southern African vocal traditions, and the musical logic that crossed the Atlantic.

The Bell Does Not Move: Polyrhythm in West Africa

  • Notate and count the seven-stroke bell pattern and explain why a twelve-pulse cycle enables cross-rhythm.
  • Explain how supporting parts lock to a timeline through fixed entry points.
  • Describe how tonal-language drums reproduce speech and what listeners need to decode them.
  • Evaluate the scholarly disagreement over polymeter and the regulative beat.

The big picture

In an Ewe drum ensemble in southeastern Ghana, the part that looks least impressive is the one everything depends on. The gankogui is a double bell: two forged iron cones joined on a curved handle, struck with a short wooden stick, low cone and high cone. Its player sounds seven strokes spread across a twelve-pulse cycle, then does it again, and again, for twenty minutes or two hours, without variation and without acceleration. Beginners are usually put on the bell, which is not a demotion. C.K. Ladzekpo, who grew up in a drumming family in the coastal village of Anyako and has taught this repertoire at Berkeley for decades, teaches the bell first because every other part in the ensemble is defined by where it falls against it.

Learn this one pattern and a great deal of music opens up: Ewe and Akan drumming, Yoruba bata, Cuban rumba, Haitian vodou drumming, Brazilian candomble, and a fair slice of what happened next in the Americas. This lesson teaches it concretely, then builds an ensemble on top of it, then examines a live argument about what we are actually hearing.

The seven-stroke bell, written out

Count twelve even pulses: 1 2 3 4 5 6 7 8 9 10 11 12, all the same length, none louder than another. The bell strikes on 1, 3, 5, 6, 8, 10, and 12.

123456789101112
xxxxxxx

The gaps between strokes run 2, 2, 1, 2, 2, 2, 1. Notice that the pattern is asymmetrical: it cannot be cut into two identical halves, which means that at any moment the pattern itself tells an experienced listener where in the cycle they are. That is its job. Ethnomusicologists usually call it the standard pattern or the timeline; musicians call it the bell.

Now the crucial structural fact. Twelve divides by 2, 3, 4, and 6. So a twelve-pulse cycle can be heard, simultaneously and without contradiction, as four groups of three and as three groups of four.

Pulse123456789101112
Four groups of threeXXXX
Three groups of fourXXX
Bellxxxxxxx

The two grouping rows agree only on pulse 1 and diverge everywhere else. That is cross-rhythm, three against four in this case, and it is the engine of the entire style. The bell pattern is built to sit across both: its strokes include 1, 3, 5, 6, 8, 10, 12, which touch members of both groupings, so it can support either hearing and commit to neither.

Try three against two in your body before continuing. Tap four even beats with your left hand while your right hand fills the same span with three. Speak "nice cup of tea" against "pass the jam", both stretched evenly over the same six units. When it locks you will feel a particular gait, and that gait is what dancers in this tradition are stepping.

Key idea: A twelve-pulse cycle can carry four groups of three and three groups of four at once, and the asymmetrical seven-stroke bell is designed to sit across both without resolving the ambiguity.

Building an ensemble on top of it

An Ewe ensemble for a dance such as Agbekor or Gahu layers a fixed set of parts, each with its own timbre, register, and entry point:

PartInstrumentRole
GankoguiIron double bellThe timeline. Never varies, never accelerates.
AxatseGourd rattle in a bead netFills the pulse; struck down on the thigh and up against the hand, so it produces two timbres.
KaganSmall high drum, sticksA short repeating figure that deliberately avoids the main beats.
KidiMedium drum, sticksA fixed phrase that answers the master drum and changes when signalled.
Sogo / kloboto / totodziLarger supporting drumsAdditional fixed phrases at their own entry points.
AtsimevuTall master drum, played standingSignals: starts and stops sections, cues dancers, converses with the kidi.

Two features make this hold together. First, the supporting parts are fixed, not improvised, and each has a specific entry point relative to the bell. A kagan pattern that starts one pulse late is not a variation; it is wrong, and everyone hears it. Second, parts are learned as spoken syllables before they are played, so a teacher can hand you a phrase by saying it. The mouth learns first, the hands follow.

The master drummer is not soloing in the jazz sense. He is directing. A specific phrase means the dancers should change figure; another means the supporting drums move to their second pattern; another ends the section. In performance the dance, the drums, and the singing are one activity, and the ensemble exists to move people who are moving.

What the dancers are doing, and why it settles an argument

Here is the disagreement that has occupied this subject for seventy years, with each side's actual evidence.

The older position, associated with A. M. Jones, who transcribed Ewe drumming in detail in the 1950s, and with Richard Waterman's idea of a metronome sense, is that African ensemble music is polymetric: different parts are genuinely in different meters, with different downbeats, running simultaneously. Jones's evidence was his own transcriptions, made by having musicians play into a recording apparatus that marked strokes on a moving strip, which showed him parts whose accents never aligned into a shared bar. If you notate each part separately from its own starting point, that is what the page shows.

Kofi Agawu, a Ghanaian music theorist who grew up with this music and now writes about it from a position inside it, argued in 2003 that the polymetric reading is largely an artifact of outsiders transcribing sound and ignoring bodies. His evidence is the dance. The dancers step a single regulative beat, usually four to the twelve-pulse cycle, and everyone in the ensemble knows where that beat is. There is one meter. The parts are placed in complex relationships to it, including relationships that displace their accents, but a shared beat exists and is visible in the room. Agawu's further charge is that insisting on impenetrable polymeter has served to make African music seem exotic and unanalyzable, which is a political effect and not a musical finding.

The practical upshot for you as a listener: find the dance step. If you can locate the beat that feet are landing on, the rest of the texture organizes itself around it and stops sounding like a wall.

Key idea: The cleanest way into this music is not to count the drums but to find the beat the dancers are stepping, and hear every part as placed against that.

Drums that speak

Yoruba and Akan are tonal languages: the same syllable carries different meanings at different relative pitches. That makes a certain kind of drum into a speech instrument.

The Yoruba dundun is an hourglass drum with two heads joined by tension cords running along the body. Squeeze the cords under your arm and the head tightens and the pitch rises; release and it falls. A skilled player can trace the tone contour and rhythm of a spoken phrase. The Akan atumpan are a pair of tall drums, one higher and one lower, that similarly render the high and low tones of Twi.

How does anyone understand it? Not by decoding syllable by syllable, which would be hopeless, since many words share a tone pattern. Listeners recover meaning because the drums mostly deliver known material: proverbs, praise names, the appellations of a chief, standard calls to assemble. The drum supplies the tonal and rhythmic skeleton and the listener's stock of formulae supplies the rest. It is closer to recognizing a familiar melody from its rhythm alone than to reading a telegram.

Mande drumming and the djembe's global career

Further north and west, among Malinke, Bamana, and Susu communities in Guinea, Mali, and Senegal, the central ensemble is the djembe, a goblet drum of carved hardwood with a goatskin head, played with bare hands, together with dunun, cylindrical bass drums played with a stick while a bell is struck with the other hand. The djembe produces three basic tones, bass, tone, and slap, and much of the technique lies in the difference between them.

Its global spread is recent and traceable. The Ballets Africains, Guinea's national ensemble under Sekou Toure's post-independence cultural policy from 1958, toured internationally and put the djembe on stages worldwide. Master drummers who came out of that system, notably Mamady Keita and Famoudou Konate, spent decades teaching in Europe, Japan, and North America. The result is that the instrument most Westerners now picture when they think African drum is a specific instrument from a specific region, taught through a modern pedagogy built for outsiders. That is not fake. It is a real tradition that acquired a second life as an export, with all the questions that raises about which repertoire gets taught and who is paid.

Careful with the word African

This lesson has described practices from a band of West Africa perhaps 1,500 kilometers wide. The continent holds 54 countries and well over a thousand languages, and its musical practices include Ethiopian liturgical chant with its own notation, Central African forest polyphony that has nothing to do with drum ensembles, Sahelian string traditions, and the Arabic-influenced music of the north. Generalizing from Ewe drumming to Africa is like generalizing from Neapolitan song to Eurasia. When you read the phrase African rhythm, ask which Africa.

Key idea: There is no African music; there are specific traditions in specific places, and the drum-ensemble style taught here belongs to one region among many.

Go and listen

Search Smithsonian Folkways for Ewe drumming, Agbekor, and Ghanaian percussion, and the Library of Congress for West African field recordings. Listen once for the bell alone, ignoring everything else. Listen again for the highest drum. Listen a third time and try to hear the four-beat step. The parts should begin to separate; if they do not, you have not listened enough times, which is the honest answer.

Common misconceptions

  • "African drumming is improvised." The supporting parts are fixed and precisely placed. The master drummer varies, and does so to direct the dance, not to display.
  • "It is too complex to count." The cycle is usually twelve or sixteen pulses. The complexity is in placement, not in size.
  • "Polyrhythm means everyone plays in a different meter." That reading is contested. Agawu's case for a single regulative beat, visible in the dancers' feet, is the stronger one for listening purposes.
  • "Talking drums spell out words." They render tone and rhythm; comprehension depends on shared proverbs and formulae, not on decoding syllables.
  • "The djembe is the traditional African drum." It is a Mande instrument whose worldwide presence dates from state ensembles and teaching tours after 1958.

Recap

  • The seven-stroke bell across twelve pulses (1, 3, 5, 6, 8, 10, 12) is the reference everything else is measured against.
  • Twelve pulses carry four groups of three and three groups of four at once, which is what makes cross-rhythm possible.
  • Supporting drum parts are fixed phrases with specific entry points, learned as spoken syllables; the master drum directs rather than solos.
  • Jones and Waterman read the texture as polymetric; Agawu argues for a single regulative beat found in the dance, and warns that the mystifying reading has political effects.
  • Tonal-language drums reproduce speech contour, and listeners decode them through stock phrases.

Sources

  1. "African music." Encyclopaedia Britannica. Regional traditions, ensembles, and rhythmic organization. britannica.com.
  2. "Polyrhythm." Wikipedia. Cross-rhythm, the standard bell pattern, and its diaspora forms. en.wikipedia.org.
  3. "Talking drum." Wikipedia. Tension-cord construction and speech surrogacy in tonal languages. en.wikipedia.org.
  4. "Djembe." Wikipedia. Mande origins, construction, and twentieth-century international spread. en.wikipedia.org.
  5. Smithsonian Folkways Recordings. Ghanaian and West African percussion collections with liner notes. folkways.si.edu.
Key terms
Gankogui
The Ewe iron double bell that plays the unvarying timeline of a drum ensemble.
Standard pattern
The seven-stroke bell figure across twelve pulses, striking on 1, 3, 5, 6, 8, 10, and 12.
Timeline
A fixed, repeating rhythmic reference against which all other parts are positioned.
Cross-rhythm
Simultaneous conflicting groupings of one span, such as three groups of four against four groups of three.
Entry point
The fixed position in the cycle at which a given supporting part begins; changing it makes the part wrong.
Atsimevu
The tall Ewe master drum whose signalled phrases direct dancers and supporting drums.
Dundun
The Yoruba hourglass pressure drum whose tension cords let the player bend pitch to imitate speech tone.
Regulative beat
The single shared beat, usually shown by the dancers' steps, that Agawu argues organizes the whole texture.

Mbira, Kora, and Voice: From Zimbabwe to the Atlantic

  • Explain how mbira parts interlock and how the resulting inherent melodies arise.
  • Describe the social function of Shona mbira in the bira and its treatment under colonial rule.
  • Describe the kora's construction and the four musical layers of jeli performance.
  • Trace specific musical features from West and Central Africa into the Americas and back again.

The big picture

In 1974, in a country then called Rhodesia, a woman named Stella Rambisai Chiweshe recorded a piece called Kasahwa on an instrument she was not supposed to be playing. Mbira was men's work by convention, and it was disapproved of by mission churches and the colonial administration alike because of what it was for: calling ancestral spirits into a room full of their descendants. Chiweshe had learned it anyway, from teachers who taught her quietly, and she went on to a fifty-year career, mostly outside her own country, until her death in 2023.

Two threads run through this lesson. The first is technical: how three African instruments actually work, in enough detail that you can follow the music rather than just enjoy the atmosphere. The second is that in each case the instrument sits inside a social arrangement, and in each case that arrangement collided with colonial power in a specific, dateable way.

Mbira dzavadzimu: what is in your hands

The name means mbira of the ancestors. Twenty-two to twenty-four metal keys, historically hand-forged from scrap iron and now often from car spring steel, are laid across a hardwood soundboard in three ranks: a left-hand group of lower pitches, a left-hand upper row, and a right-hand row. The left thumb plucks downward, the right thumb plucks downward, and the right index finger plucks upward from beneath. Three fingers, three ranks.

The soundboard sits inside a deze, a large hemispherical gourd, and the rim of the gourd and the edge of the board carry machachara: bottle caps, shells, or metal rings that rattle. Every note therefore arrives inside a wash of buzz. To a listener raised on clean recorded sound this reads as distortion. In Shona terms it is the sound: it thickens the tone, it carries in a crowded room, and it is understood to be part of what makes spirits attend. Take the caps off and you have a quieter, thinner instrument that is not doing its job.

How two players make a third part

A typical mbira piece cycles through 48 pulses, heard as four phrases of twelve. Within the cycle a fixed progression of paired low notes functions something like a harmonic scheme, and every variation must fit it. The first player plays kushaura, from the verb to lead: the version of the cycle that establishes it. The second plays kutsinhira, to interlace: a part designed to fall in the gaps of the first, often entering a pulse or a half-cycle later.

What you hear is neither part. The interlocked pitches produce lines that ascend and descend across both players' notes, and the ear assembles them into melodies that no individual is playing. Ethnomusicologists call these inherent patterns, following Gerhard Kubik, who documented the effect across Central and southern Africa. They are not an illusion to be corrected; players aim for them, and they shift as variations accumulate, so the same cycle yields a different apparent tune ten minutes later. Meanwhile the hosho, a pair of seed-filled gourd rattles, plays a fast triple pattern that never changes and functions exactly like the bell in Lesson 5.

Core repertoire pieces you can search for: Nhemamusasa, Nyamaropa, Kariga Mombe, Taireva, Bukatiende. These are not compositions in the sense of fixed works with authors. They are cycles with names, centuries old, that every player learns and then inhabits differently.

Key idea: Mbira music is built so that the audible melody belongs to the combination rather than to either player, and the composite shifts as variations move.

The bira, and what the music is for

A bira is an all-night gathering, usually convened because something is wrong: an illness, a dispute, a death that needs explaining. The vadzimu, the family's ancestral spirits, are understood to be attracted by mbira and by the sung parts that go with it. Two vocal layers accompany the instruments: mahonyera, low wordless humming that shadows the cycle, and kudeketera, half-sung poetry that names ancestors, recites family history, complains, praises, and pleads.

At some point a svikiro, a spirit medium, may be possessed, and a specific dead relative speaks through them to the assembled family. The music continues throughout; players relieve each other so that it does not stop. Paul Berliner documented this world in The Soul of Mbira (1978) after years of study with Shona musicians, and the fullest version of that work, a vast archive of improvisations published in 2020, carries the name of the mbira player Cosmas Magaya alongside his own.

Mbira, banned and amplified

Mission churches in colonial Zimbabwe classified mbira as heathen practice, and its association with ancestor veneration made it politically suspect. What happened next is a case study in how music moves under pressure. In the 1970s, during the liberation war, Thomas Mapfumo began transposing mbira cycles onto electric guitars, keeping the interlocking parts and the cyclic harmonic scheme but playing them on a rock band's instruments, and singing in Shona rather than English. The lyrics used the indirection of traditional poetry, so that songs about a hare, or about people going to the fields, carried meanings the censors could not pin down but audiences understood.

He called the result chimurenga music, after the Shona word for the liberation struggle. Rhodesian radio banned records; he was detained in 1979 and released after protest. After independence in 1980 he became a national figure and then, when he turned the same technique against corruption under the new government, an exile. The instrument the missions called heathen became the sound of the nation and then a problem for the nation, in twenty years.

Key idea: Musical form carried political content here precisely because it was indirect: a cyclic ancestral idiom, sung in Shona with traditional metaphor, was legible to its audience and hard for censors to prosecute.

The jeli and the kora

In Mande societies across Mali, Guinea, Senegal, and the Gambia, certain families are hereditary specialists in speech and music. The Mande word is jeli; the French term griot is the one that traveled. A jeli's job is not entertainment. It is genealogy, historical narration, praise, mediation between parties who cannot address each other directly, and the maintenance of the epic of Sunjata Keita, founder of the Mali Empire in the thirteenth century. Knowledge passes within the family, and non-jeli families do not do this work.

The kora is a twenty-one-string harp-lute. A large calabash cut in half is covered with cowhide; a long hardwood neck passes through it; the strings, formerly of twisted antelope hide and now of fishing line, run over a tall notched bridge in two parallel ranks, eleven on one side and ten on the other. The player grips two upright handposts with the last three fingers of each hand and plays with both thumbs and both index fingers. Because the ranks alternate between the hands, the instrument is built for interlocking figures in the same way an mbira is.

Jeli performance layers four things, and once you can name them you can follow any kora recording:

LayerWhat it is
KumbengoThe repeating instrumental ostinato that identifies the piece and holds the cycle
BirimintingoFast virtuosic runs that break out of the ostinato and return to it
DonkiloThe fixed sung refrain, usually short and repeated
SataroRecited or half-spoken text: praise, genealogy, proverb, commentary on the people present

Kora tunings have names: Silaba, Tomora ba, Sauta, Hardino, each shifting one or two strings and producing a different set of available pieces. Toumani Diabate, who died in 2024 and whose family traces its line of players back many generations, recorded one of the first solo kora albums in 1988 and then spent thirty years collaborating outward, with Malian singers, a blues guitarist, a flamenco group, and a symphony orchestra. Sona Jobarteh, from a Gambian jeli family, became the first woman to build a career as a kora virtuoso in a tradition that reserved the instrument for men, and founded a school in the Gambia to teach it.

One honest note about praise singing. A jeli who names a wealthy man's ancestors at a wedding and is given money for it looks, to outsiders, like flattery for hire. Inside the system it is a professional service with a specific social function: the naming ties a living person to a documented lineage in public, and the payment is the fee for the archive being maintained. Whether the archive tilts toward whoever pays is a live question that Mande audiences themselves debate.

Voices in southern and central Africa

Isicathamiya is Zulu male a cappella singing that grew in the mine and factory hostels of Johannesburg and Durban among migrant workers separated from their families. Competitions run all Saturday night, with a judge sometimes recruited off the street to ensure impartiality, and the name refers to stepping lightly, because the tightly choreographed movement had to be quiet enough not to wake the hostel. Solomon Linda's group recorded Mbube in 1939, giving one branch its name and, as Lesson 15 will show, becoming the most consequential copyright case in the history of this subject. Ladysmith Black Mambazo, formed by Joseph Shabalala in 1960, is the ensemble most listeners outside South Africa have heard.

Umngqokolo is a Xhosa overtone singing technique in which a woman produces a low fundamental with a rough, constricted vocal quality and simultaneously selects audible harmonics above it, so that one singer sounds like two. It is related to the uhadi, a braced musical bow with a gourd resonator whose sound works the same way: one string gives two fundamentals and the mouth or gourd selects harmonics from them. The vocal technique imitates the bow, not the other way around, which tells you something about how instruments and voices co-evolve.

Among the BaAka and other forest peoples of Central Africa, polyphonic singing is contrapuntal, yodeled, hocketed, and leaderless: several interlocking vocal roles, each with its own melodic territory, combining into a dense texture that changes continuously as singers enter, drop out, and vary. This is the music Simha Arom took apart with playback in Lesson 3, and it is nothing like drum-ensemble music, which is the point about not saying African music.

Across the Atlantic, and back

What crossed the ocean was not repertoire but principles, and they are traceable:

  • Timelines: the asymmetrical bell figure survives as Cuban clave, Brazilian and Haitian bell patterns, and as the rhythmic bones of a great deal of American popular music.
  • Call and response: leader and group alternation persists in field hollers, the ring shout, gospel, and blues form.
  • Lamellophones: the mbira principle reappears in the Caribbean marimbula, a large box with metal tongues played as a bass instrument.
  • Plucked skin-headed lutes: West African instruments including the Jola akonting and the Mande ngoni are ancestors of the banjo, an instrument invented by enslaved Africans in the Caribbean and North America.
  • Interlocking parts and the buzz aesthetic: both are audible in Cuban rumba, candomble, and vodou drumming.

Then the loop closes. From the 1930s the British label His Master's Voice pressed Cuban son and other Latin records in its GV series and sold them in West and Central Africa, where they arrived as glamorous modern music. Congolese musicians in Kinshasa and Brazzaville built a new style on those records, added guitars, and produced Congolese rumba and later soukous, which became the dominant popular music of much of the continent. Le Grand Kalle's Independance Cha Cha, recorded in 1960 as the Belgian Congo became independent, is a Cuban dance form sung in Lingala celebrating African independence. Music that left Africa in the holds of slave ships came back seventy years later on shellac discs and was received as foreign and exciting.

Key idea: Atlantic musical exchange is not a one-way transmission from Africa to the Americas; it is a circuit, and Congolese rumba is the proof.

Common misconceptions

  • "The mbira buzz is a defect of old recordings." It is engineered into the instrument with bottle caps and is required for the sound to work as intended.
  • "One mbira player plays the melody and the other accompanies." The melody you hear belongs to the combination and is played by neither.
  • "Griot means African folk singer." A jeli belongs to a hereditary professional class with duties in genealogy, mediation, and history; the role is inherited, not chosen.
  • "Isicathamiya is ancient village music." It took shape in twentieth-century migrant labor hostels, and the Saturday-night competition is central to it.
  • "African influence flowed one way, to the Americas." Cuban records sold in Africa from the 1930s produced Congolese rumba, which then reshaped popular music across the continent.

Recap

  • Mbira dzavadzimu has 22 to 24 keys in three ranks, a buzzing gourd resonator, and a 48-pulse cycle carried by kushaura and kutsinhira parts that generate inherent melodies.
  • Its purpose in the bira is to bring ancestral spirits into a gathering, and mission and colonial authorities opposed it for exactly that reason.
  • Mapfumo moved mbira cycles onto electric guitars and sang in Shona, creating chimurenga music, banned before independence and awkward for the government after it.
  • The kora is a 21-string harp-lute played by hereditary jeli families, and jeli performance layers kumbengo, birimintingo, donkilo, and sataro.
  • Southern and central African vocal traditions range from hostel-bred isicathamiya to Xhosa overtone singing derived from the musical bow to leaderless BaAka polyphony.
  • African musical principles crossed the Atlantic and returned as Cuban records, producing Congolese rumba.

Sources

  1. "Mbira." Wikipedia. Construction, tuning, repertoire, and ceremonial use. en.wikipedia.org.
  2. "Kora (instrument)." Wikipedia. Construction, tunings, and jeli practice. en.wikipedia.org.
  3. "African music." Encyclopaedia Britannica. Regional traditions and instrument families. britannica.com.
  4. "Isicathamiya." Wikipedia. Migrant labor origins and competition format. en.wikipedia.org.
  5. Smithsonian Folkways Recordings. Zimbabwean, Malian, and southern African collections with liner notes. folkways.si.edu.
Key terms
Mbira dzavadzimu
The Shona lamellophone of the ancestors: 22 to 24 metal keys in three ranks, played inside a buzzing gourd.
Kushaura and kutsinhira
The leading and interlacing mbira parts, whose combination produces the audible melody.
Inherent patterns
Melodies a listener hears in an interlocked texture that no single player performs.
Bira
An all-night Shona ceremony at which mbira and singing call ancestral spirits, often through a medium.
Chimurenga music
Thomas Mapfumo's transposition of mbira cycles onto electric guitars, sung in Shona during the liberation war.
Jeli
A hereditary Mande specialist in genealogy, praise, mediation, and epic narration; the French term is griot.
Kumbengo
The repeating kora ostinato that identifies a piece and holds its cycle.
Umngqokolo
Xhosa overtone singing in which one voice produces a fundamental and audible harmonics, modeled on the uhadi bow.
GV series
The record line that sold Cuban and Latin music in Africa from the 1930s, seeding Congolese rumba.

Module 3: South and Southeast Asia

The two Indian classical systems taught properly: the drone, the anatomy of a raga, tala cycles you can count, the shape of a performance, and how the music passes from teacher to student. Then the bronze ensembles of Java and Bali, with their own tunings and cyclic forms, and the traditions of mainland and island Southeast Asia.

Raga and Tala: Indian Classical Music, North and South

  • Explain the function of the drone and why Indian classical melody is heard against a fixed tonic.
  • Describe the components of a raga beyond its pitch set, and count teental and adi tala.
  • Compare Hindustani and Carnatic practice on form, improvisation, and repertoire.
  • Explain guru-shishya transmission, the gharana system, and how institutions changed both.

The big picture

In 1938 a young man named Ravi Shankar, who had spent his teens touring Europe and America as a dancer in his brother's company, wearing tailored clothes and living in Paris, gave it all up and moved to Maihar, a small town in central India. He went to live in the household of Allauddin Khan and become his student. That meant sweeping, fetching, waiting, and practicing for many hours a day for years before being considered ready to perform in public. He stayed roughly seven years.

That arrangement is not a colorful detail; it is the transmission system, and it explains features of the music that would otherwise look arbitrary. This lesson covers what he was learning: a melodic system called raga, a rhythmic system called tala, and two related but distinct classical traditions, Hindustani in the north and Carnatic in the south.

First, the drone

Before a note of melody, a Hindustani or Carnatic performance begins with a continuous drone, usually from a tanpura: a long-necked lute with four or five strings that are not fretted, not melodic, and simply plucked in slow rotation forever. Its distinctive shimmer comes from a thread laid between each string and the flat bridge, which makes the string buzz against the bridge and throws off a cloud of overtones.

The drone sounds the tonic, called Sa, and usually the fifth, Pa. Sa is movable: a singer picks whatever pitch suits their voice and everything is built relative to it, so absolute pitch is irrelevant and the concept of a key signature has no work to do. More importantly, the drone never changes. There is no modulation, no chord progression, no harmonic departure and return. Every note you hear is heard permanently against Sa and Pa.

That constraint is what makes this music possible. Because the reference never moves, extremely small differences in pitch and approach become meaningful. A slightly low third heard against a constant tonic is unmistakable, and a slow oscillation around a note becomes an event. In a music that modulates, those details would be washed away by the changing harmony.

Key idea: The unchanging drone converts every melodic detail into a measurable relationship with a fixed tonic, which is why microtonal shading and ornament carry so much weight in this tradition.

Swara: the notes, and what is between them

The seven degrees are Sa, Re, Ga, Ma, Pa, Dha, Ni, abbreviated S R G M P D N and functioning like movable do. Five of them have lowered (komal) forms, Ma has a raised (tivra) form, and Sa and Pa are fixed, which yields twelve positions per octave, superficially like a chromatic scale.

Superficially. Indian theory also speaks of shruti, traditionally twenty-two audible pitch positions in the octave. These are not a twenty-two-note scale anyone plays. They describe the fact that the same named note is intoned differently in different ragas: the Ga of one raga sits perceptibly lower than the Ga of another, and getting that placement right is part of playing the raga correctly rather than a matter of expressive taste.

What a raga actually is

A raga is not a scale. Here is the standard demonstration. Raga Bhupali and Raga Deshkar use exactly the same five notes: S R G P D. They are different ragas, immediately distinguishable to a listener who knows them, because everything except the pitch set differs. Bhupali lives in the lower and middle register and emphasizes G and D; Deshkar lives higher and emphasizes D and P. Their characteristic phrases differ, their resting notes differ, their moods differ.

So a raga specifies, at minimum:

  • Aroha and avaroha: the ascending and descending forms, which are often different. A note may be omitted going up and present coming down, or approached only obliquely.
  • Vadi and samvadi: the most emphasized note and the second most emphasized.
  • Pakad: a short catch phrase that identifies the raga in a few seconds.
  • Nyasa: the notes where a phrase may come to rest.
  • Ornament requirements: which notes must be approached by a glide (meend), which must oscillate slowly (andolan), which are touched only in passing.
  • Time and season: Bhairav belongs to dawn, Yaman to early evening, Malkauns to deep night, and the Malhar family to the monsoon.

The time associations are real conventions of practice, not mysticism, and musicians observe them more strictly in some settings than others. In a broadcast studio at eleven in the morning, exceptions get made.

Tala: cycles you can count

Indian rhythm is cyclic. A tala is a fixed number of beats divided into sections, with a system of hand gestures that lets everyone track position: a clap for a stressed section, a wave of the open hand for the section called khali, or empty. Beat one is sam, the point of arrival that all improvisation aims at.

The most common Hindustani tala is teental, sixteen beats in four groups of four. The drum plays a standard pattern of stroke syllables called the theka:

Beats 1-4Beats 5-8Beats 9-12Beats 13-16
Dha Dhin Dhin DhaDha Dhin Dhin DhaDha Tin Tin TaTa Dhin Dhin Dha
clap (sam)clapwave (khali)clap

Listen for the khali section: the syllables change from Dha to Tin and Ta, which means the resonant left-hand bass drum stops sounding. That audible thinning at beat nine is how you find your place in a sixteen-beat cycle without counting. Other common talas include jhaptal (10, as 2+3+2+3), rupak (7, as 3+2+2), and ektal (12).

The characteristic cadence is the tihai: a phrase played three times in a row, timed so that the final note of the third repetition lands exactly on sam. Musicians calculate these on the fly. If you are eleven beats from sam and your phrase is three beats long, you play it three times with one beat of gap between repetitions and arrive precisely. When an audience shouts at the moment a long tihai lands, they are applauding arithmetic executed at speed.

Key idea: Tala is a cycle, not a meter of bars, and the interest lies in departure from sam and calculated return to it.

The shape of a Hindustani performance

A full performance moves from unmeasured to measured and from slow to fast:

  1. Alap: no drum, no pulse. The soloist introduces the raga from the low register upward, phrase by phrase, over the drone. In the dhrupad style this can run forty minutes.
  2. Jor: a pulse emerges, still without drum.
  3. Jhala: fast, driving, with rapid rhythmic strumming on the instrument's drone strings.
  4. Gat or bandish: a fixed composition enters with the tabla and the tala begins. Usually a slow (vilambit) piece first, then a fast (drut) one.
  5. Improvisation expands within the cycle: gradual melodic development, then fast runs (tans), then exchanges with the drummer in which soloist and tabla trade phrases and both land on sam.

Instruments to know: the sitar, with curved movable frets, sympathetic strings that ring untouched, and a broad curved bridge that gives its shimmering buzz; the sarod, fretless with a polished metal fingerboard, so pitch is continuous and glides are unbroken; the bowed sarangi, closest to the human voice; the bamboo bansuri; and the tabla pair, whose larger left drum is pitch-bent by pressing the heel of the hand into the head while striking, which is the sound you hear swooping under the beat. Vocal genres run from the austere dhrupad to the dominant khayal to the lighter, romantic thumri.

Carnatic music in the south

South Indian classical music shares raga and tala with the north and differs in emphasis. It is more composition-centered: the core repertoire consists of songs, mostly devotional, by named composers, above all the three known as the Trinity, who all worked in the Kaveri delta around Thanjavur in the late eighteenth and early nineteenth centuries. Tyagaraja (1767 to 1847) alone is credited with hundreds of songs still sung daily.

HindustaniCarnatic
Core unitRaga elaboration; compositions are often short pegsThe kriti, a composed song in three sections: pallavi, anupallavi, charanam
OpeningLong unmetered alap before the drum entersShorter alapana; the composition comes sooner
Scale theoryTen thats classify ragas looselySeventy-two melakarta parent scales generate the system exactly
OrnamentPresent and importantGamaka is obligatory: a note is usually a shape, not a steady pitch
PercussionTablaMridangam, with ghatam, kanjira, and morsing; tani avartanam is a full percussion feature
Melody instrumentsSitar, sarod, sarangi, bansuriVeena, violin, flute, nadaswaram

Two Carnatic specifics repay attention. First, the melakarta system, formalized in the seventeenth century, generates seventy-two parent scales by systematic combination and derives thousands of ragas from them. It is one of the most complete pieces of scale theory produced anywhere. Second, konnakol: rhythm is spoken aloud using syllable groups (ta ka di mi for four, ta ki ta for three) and performers recite complex rhythmic structures before or instead of playing them. Learning rhythm as speech means a musician can compose a rhythmic cadence in their mouth while walking.

The tala system counts with the hand: a clap, then finger counts, then claps and waves, in units called laghu, drutam, and anudrutam. Adi tala, the commonest, is eight beats: a clap plus three finger counts, then clap-wave, then clap-wave.

The violin arrived in South India around 1800 through European contact and is now completely Carnatic: held with the scroll braced against the ankle of a seated player, tuned differently, and played with the continuous slides that gamaka requires. Nobody in the tradition regards it as a foreign instrument. Keep that in mind when Module 6 takes up questions of borrowing and authenticity.

Key idea: Hindustani practice foregrounds the elaboration of the raga, Carnatic practice foregrounds a composed repertoire treated as a vehicle for improvisation, and both use the same raga and tala foundations.

How it is transmitted, and how that changed

The traditional route is guru-shishya parampara: a student attaches to one teacher, often lives with them, serves the household, and receives the material over many years in the order the teacher chooses. In some lineages the relationship is formalized by tying a sacred thread on the student's wrist. Instrumental and vocal lineages developed into gharanas, named schools with distinct repertoires and stylistic preferences, such as Gwalior, Kirana, Jaipur-Atrauli, and Agra for vocal music, or the Maihar lineage of Allauddin Khan for instruments.

The twentieth century added institutions. Vishnu Narayan Bhatkhande devised a written notation and a classification of ragas, and music colleges began teaching in classes with syllabi and examinations. All India Radio hired musicians on salary and broadcast recitals cut to a fixed length. Recording imposed the three-minute side and then the LP's twenty-two minutes on a music that had assumed unlimited time.

There is a loss and a gain here, and both are specific. Institutional teaching made the music available to students without family connections and reduced the total dependence of a student on one teacher's goodwill. It also standardized the material, weakened the distinctiveness of gharanas, and shifted authority from lineages toward examiners. Meanwhile a whole class of professional women performers, the courtesan singers who had maintained much of the thumri and khayal repertoire, were pushed out of public life by anti-nautch respectability campaigns from the late nineteenth century onward. Their repertoire survived; their claim to it largely did not.

Go and listen

Listen to a full alap in Raga Yaman without doing anything else, and notice how long it takes for the seventh note of the raga to appear. Then find a teental recording and try to clap the sixteen-beat cycle, listening for the thinning at beat nine. Then find any Tyagaraja kriti and follow the return of the pallavi. Smithsonian Folkways holds substantial Indian collections, and full-length concert recordings are widely available.

Common misconceptions

  • "A raga is an Indian scale." Two ragas can share an identical pitch set and remain wholly distinct through emphasis, phrase, ornament, and register.
  • "There is no rhythm in the alap, so it is free." The alap has no meter but is highly disciplined; the raga's rules apply strictly throughout.
  • "Improvisation means the performer plays whatever they want." The raga constrains pitch and motion, the tala constrains time, and the tihai has to arrive on the beat.
  • "Hindustani and Carnatic are the same music with different names." They share foundations and differ in form, repertoire, theory, ornament, and instruments.
  • "Shruti means Indian music uses twenty-two notes." Shruti describes how a named note is intoned differently in different ragas, not a scale of twenty-two playable pitches.

Recap

  • The tanpura drone fixes Sa and never changes, which makes small melodic differences meaningful.
  • A raga specifies ascent, descent, emphasized notes, catch phrases, resting points, obligatory ornaments, and often a time of day.
  • Tala is a counted cycle with claps and a waved empty section; teental is 16 in four groups of four with khali at beat nine; tihais land on sam.
  • Hindustani performance runs alap, jor, jhala, then composition with tabla; Carnatic performance centers on composed kritis, the seventy-two melakarta system, obligatory gamaka, and spoken konnakol rhythm.
  • Guru-shishya transmission and gharanas were partly displaced by colleges, notation, radio, and recording, with gains in access and losses in lineage authority, and courtesan performers were pushed out of the tradition they had sustained.

Sources

  1. "South Asian arts: Music." Encyclopaedia Britannica. Raga, tala, and the two classical systems. britannica.com.
  2. "Raga." Wikipedia. Structure, classification, and performance practice. en.wikipedia.org.
  3. "Tala (music)." Wikipedia. Cycles, clap patterns, khali, and common talas. en.wikipedia.org.
  4. "Carnatic music." Wikipedia. Melakarta, kriti form, gamaka, and the concert format. en.wikipedia.org.
  5. Smithsonian Folkways Recordings. Indian classical collections with liner notes. folkways.si.edu.
Key terms
Tanpura
The four or five string drone lute whose threaded bridge produces a shimmering overtone cloud.
Sa
The movable tonic against which all pitches are heard; chosen to suit the performer's voice or instrument.
Raga
A melodic framework specifying ascent and descent, emphasized notes, phrases, resting points, and required ornaments.
Shruti
The traditional twenty-two pitch positions describing how a named note is intoned differently across ragas.
Tala
A cyclic rhythmic framework of counted beats divided into sections and tracked by claps and waves.
Sam
Beat one of the tala cycle, the point of arrival that improvisation aims at.
Khali
The waved, unstressed section of a tala, audible in teental as a thinning of the drum sound at beat nine.
Tihai
A phrase played three times so that its final note lands exactly on sam.
Melakarta
The Carnatic system of seventy-two parent scales from which other ragas are derived.
Guru-shishya parampara
Transmission through long personal apprenticeship to one teacher, historically including living in the household.

Gamelan and the Gong-Chime Belt of Southeast Asia

  • Describe how a gamelan is tuned and why its instruments cannot be exchanged between sets.
  • Read a colotomic structure and explain how gong strokes define form.
  • Explain irama levels and the relationship between tempo and surface density.
  • Compare Javanese, Balinese, Thai, and Philippine gong-chime practices.

The big picture

In 1929 the Canadian composer Colin McPhee, then living in New York, heard a small set of newly issued 78 rpm discs of gamelan music recorded in Bali. He played them repeatedly. Two years later he sailed for the island, built a house at Sayan, and stayed for most of the decade, transcribing music that no European notation had been designed to hold. He was not the first outsider to be stopped in his tracks by this sound. In 1889 Claude Debussy had spent the summer in front of the Javanese pavilion at the Paris Exposition.

What both men encountered is often described as an orchestra, which is the first thing to unlearn. A gamelan is a single instrument that requires twenty or more people to play. The bronze is tuned as a set, to itself, and no piece of it belongs anywhere else.

What is in the room

A Central Javanese gamelan is a collection of bronze keys, kettles, and hanging gongs, plus drums, a bowed fiddle, a flute, a zither, and singers. The instruments divide by function, not by family, and this division is the key to hearing the music:

FunctionInstrumentsWhat they do
Core melody (balungan)Saron, demung, pekingPlay the skeletal melody in even notes, struck with a mallet and damped with the other hand
ElaborationBonang, gender, gambang, celempung, suling, rebabWeave faster figuration around and between the core notes, each in its own idiom
Punctuation (colotomic)Gong ageng, kempul, kenong, kethukMark fixed positions in the cycle; they define the form
DirectionKendhang (drums)Set tempo, signal transitions, cue endings; the drummer is the conductor
VoiceSindhen (solo female), gerong (male chorus)Sing lines related to but independent of the core melody

There is no conductor waving at anyone. The rebab or the singer leads melodically, the drummer leads structurally, and everyone else knows the piece.

Tuning: why a gamelan cannot borrow a saron

Javanese gamelan uses two tuning systems. Slendro has five tones per octave, spaced nearly but not exactly equally at roughly 240 cents. Pelog has seven, spaced very unequally, with some intervals near 120 cents and others near 250 or more. A large gamelan is really two gamelans side by side, one in each tuning, sharing drums and some gongs, and the players turn to face whichever set the next piece requires.

Now the part that surprises people. There is no standard slendro and no standard pelog. Each set is tuned to itself by its maker, and the particular character of a set's tuning, its embat, is prized and named. Two gamelans in neighboring villages will not match, and you cannot carry a saron from one to the other. The set is the instrument, which is also why sets receive proper names, are treated with respect, sometimes receive offerings, and are not stepped over.

Key idea: A gamelan is tuned as a unit with no external pitch standard, so its instruments are not interchangeable and the ensemble is best thought of as one large instrument.

Colotomic form: the gong tells you where you are

Javanese pieces are cycles, and the cycle is defined by which punctuating gong sounds at which position. The largest gong, the gong ageng, sounds once at the end of the cycle, and its arrival is the structural downbeat of the whole music: everything is heard as approaching it. Here is the ladrang form, a common thirty-two-beat cycle. Read the columns as beat positions.

Beat48121620242832
Kenong (medium kettle)NNNN
Kempul (small hanging gong)PPP
Gong agengG

The kethuk, a small damped kettle, ticks on the even-numbered beats between these. Notice that there is no kempul at beat 4: the space right after the gong is deliberately left open. Once you can hear kenong every eight and kempul in between, you can locate yourself in a piece you have never heard, and you will know a cycle is ending because the drum changes and the great gong arrives.

Smaller and larger forms work the same way with different densities: lancaran is a short bright cycle, ketawang is sixteen beats, and the largest ceremonial forms stretch a single gongan over many minutes.

Irama: the piece breathes

Here is the concept with no close Western parallel. Irama is the ratio between the core melody and the elaborating parts. As the drummer slows the core melody down, the elaborating instruments do not slow with it. They subdivide further, doubling their density, so the surface stays busy or gets busier while the skeleton stretches.

Irama levelElaborating notes per core noteEffect
Lancar1Fast and plain
Tanggung2Moderate
Dadi4The common expanded level
Wiled8Slow core melody, dense filigree
Rangkep16Extremely stretched; singers have room for long lines

A single piece can move up and down these levels at the drummer's signal. If you listen and think the music has slowed down and sped up at once, you have heard irama correctly: the gongs are further apart and the small instruments are playing twice as many notes.

Key idea: Slowing a Javanese piece does not thin it out; the elaborating instruments double their density, so tempo change is a change of texture as much as of speed.

Wayang kulit: nine hours of it

The most demanding context for gamelan is the shadow play. A dalang sits behind a lit cloth screen manipulating leather puppets, speaks every character's voice, narrates, sings mood-setting songs, and cues the ensemble by knocking a wooden mallet against the puppet chest and rattling hanging metal plates with his foot. A performance runs most of the night. The stories come from the Mahabharata and Ramayana, thoroughly localized, with Javanese clown-servant characters who comment on current events and are frequently the reason the audience came. The gamelan follows the dalang, changes pieces on cue, and provides the emotional register for scenes it cannot see.

Bali: the same bronze, a different aesthetic

Balinese gamelan sounds nothing like Javanese, and two features explain most of the difference.

First, paired tuning. Balinese instruments come in pairs, one tuned slightly higher than the other, typically by five to eight beats per second at a given pitch. Strike both and the two frequencies interfere, producing a pulsing shimmer called ombak, waves. This is deliberate acoustic beating, engineered into the metal, and it gives Balinese bronze its characteristic glitter.

Second, kotekan. Fast figuration is split between two parts, polos and sangsih, each playing an incomplete pattern with rests where the other plays. Neither player performs the line you hear. Rehearsed to the precision this requires, a Balinese ensemble produces streams of notes far faster than any single player could manage. It is the same interlocking principle as the mbira duet and the Aka horns, executed on bronze at high speed.

The dominant twentieth-century style, gong kebyar, emerged in north Bali around 1915. Kebyar means something like bursting into flame, and the style features sudden ensemble attacks, extreme dynamic contrast, abrupt tempo changes, and virtuosity as an explicit value.

Kecak deserves a candid paragraph. The famous chorus of dozens of men chanting interlocking cak syllables around a story from the Ramayana is presented to visitors as ancient. Its current form was assembled in the 1930s from the vocal chorus of a trance ritual, developed with the involvement of the German painter Walter Spies, partly to be filmed and shown to outsiders. It is now performed nightly at clifftop temples at sunset for paying audiences. It is also, by now, genuinely Balinese, performed by Balinese people, meaningful to them, and continuously developed for ninety years. Both statements are true, and the case is a good rehearsal for Module 6.

The gong-chime belt

Bronze gong ensembles run across mainland and island Southeast Asia, and they share a structural principle: several instruments play the same underlying melody at different densities, with punctuation from gongs and direction from drums.

  • Thailand. The piphat ensemble pairs the pi nai, a piercing quadruple-reed oboe, with the ranat ek xylophone and khong wong yai, a circle of tuned gongs around a seated player. Thai classical tuning uses seven nearly equal steps of about 171 cents, matching neither Western semitones nor whole tones.
  • Cambodia. The pinpeat ensemble accompanies court dance and shadow theater. Between 1975 and 1979 the Khmer Rouge killed an estimated eighty to ninety percent of the country's artists and musicians. The tradition was rebuilt after 1979 from the memories of the surviving teachers, at the Royal University of Fine Arts in Phnom Penh, which is one of the most consequential acts of musical reconstruction of the twentieth century.
  • Vietnam. The dan bau is a single-string zither played entirely in harmonics, with a flexible rod that bends the pitch continuously, so the instrument sings. Ca tru chamber song and the nha nhac court repertoire are both on UNESCO lists.
  • Philippines. Maguindanao and Maranao kulintang is a row of eight small kettle gongs on a frame, played melodically by one person, supported by larger hanging gongs and a drum. It belongs to the same bronze family as gamelan and piphat while sitting outside the Indic court traditions.
  • Myanmar. The hsaing waing ensemble centers on the pat waing, a circle of twenty-one tuned drums around a seated player who plays melody on them, and the saung gauk, an arched harp of a type that vanished elsewhere in Asia centuries ago.

Common misconceptions

  • "Gamelan is an orchestra." It is one tuned set, treated and often named as a single instrument, whose parts cannot be mixed with another set.
  • "Slendro is a pentatonic scale like any other." It is a near-equidistant five-tone division that matches no Western scale, and every set tunes it differently.
  • "The gong marks the beginning of a phrase." The gong marks arrival at the end of the cycle; the music is heard as approaching it.
  • "Slower means simpler." In irama, slowing the core melody doubles the density of the elaborating parts.
  • "Kecak is an ancient ritual." Its present form dates from the 1930s and involved a European collaborator, which does not make it any less Balinese now.

Recap

  • A gamelan is tuned as a unit in slendro or pelog with no external standard; embat is the individual character of a set.
  • Colotomic gongs define form: in ladrang, kenong every eight beats, kempul between, and the great gong at thirty-two.
  • Irama levels change the ratio of elaboration to core melody, so tempo change alters texture.
  • Balinese practice adds paired tuning for ombak shimmer and kotekan interlocking at high speed, with gong kebyar as the dominant modern style.
  • Related gong-chime traditions run from Thai piphat and Cambodian pinpeat to Philippine kulintang and Burmese hsaing waing.

Sources

  1. "Gamelan." Encyclopaedia Britannica. Instruments, tunings, and ensemble organization. britannica.com.
  2. "Gamelan." Wikipedia. Slendro and pelog, colotomic structure, irama, and Balinese practice. en.wikipedia.org.
  3. "Southeast Asian arts." Encyclopaedia Britannica. Regional ensembles and court traditions. britannica.com.
  4. "Kulintang." Wikipedia. Philippine gong-chime ensembles and repertoire. en.wikipedia.org.
  5. UNESCO. Lists of Intangible Cultural Heritage, including Vietnamese ca tru and nha nhac. ich.unesco.org.
Key terms
Slendro
The five-tone, near-equidistant Javanese tuning, with steps of roughly 240 cents.
Pelog
The seven-tone Javanese tuning with markedly unequal steps.
Embat
The individual tuning character of a particular gamelan set.
Balungan
The skeletal core melody played in even notes by the saron family.
Colotomic structure
Form defined by which punctuating gong sounds at which position in the cycle.
Gongan
One complete cycle, ending with the stroke of the great gong.
Irama
The ratio of elaborating notes to core melody notes, which doubles as the core melody slows.
Kotekan
Balinese interlocking figuration split between polos and sangsih parts, neither of which plays the audible line.
Ombak
The shimmering pulsation produced by deliberately detuned paired Balinese instruments.
Dalang
The shadow-play puppeteer who voices all characters, narrates, and cues the gamelan.

Module 4: East Asia and the Middle East

The Chinese literati zither and the tablature that records action but not time, the court and chamber traditions of Japan and Korea and the aesthetics behind them, and the maqam world from Cairo to Istanbul to Isfahan, including Sufi devotional practice and Jewish liturgical music.

China: The Qin, the Tablature, and the National Orchestra

  • Describe the qin's three timbral categories and explain what jianzipu tablature does and does not specify.
  • Explain dapu and why two players can produce different realizations of one notated piece.
  • Connect Confucian music ideology and imperial pitch standards to state power.
  • Explain how twentieth-century reform reshaped Chinese instruments and ensembles.

The big picture

In September 1950 two researchers from Beijing arrived in the city of Wuxi carrying a wire recorder and looked for a blind street musician named Hua Yanjun, known to everyone as Abing. He had been a Daoist temple musician, lost his sight in his thirties, and spent years playing on the streets for coins. They recorded six pieces, three on erhu and three on pipa, and planned to come back for more. Abing died three months later, in December 1950.

One of those six, Erquan Yingyue, usually translated as Moon Reflected on the Second Spring, became one of the most played instrumental pieces in China. It exists because two people with a recorder walked down the right street in the right year. Nearly everything else Abing knew is gone.

Hold that as you read the rest of this lesson, which is about a musical culture with the longest continuous written record of any on earth, and about what that written record does and does not preserve.

The qin: an instrument built for one room

The guqin is a long, narrow zither about 1.2 meters in length with seven strings, historically twisted silk and now often steel wound with nylon. It has no frets and no bridges under the strings. Instead, thirteen inlaid dots called hui run along one edge, marking the positions of the natural harmonics. It is extremely quiet. Played at full strength it is roughly the volume of conversation, which tells you what it is for: a room, one or two listeners, or nobody at all.

Its sounds divide into three named categories, and hearing this division is most of what a beginner needs:

  • San yin, open strings: full and resonant.
  • Fan yin, harmonics: the left hand touches a string lightly at a hui and the note rings pure and bell-like, with a distinctive floating quality.
  • An yin, stopped notes: the left hand presses the string to the board and then slides, vibrates, or presses further, producing a note that continues changing after it is struck.

The third category carries the tradition's aesthetic weight. A qin note is often not a fixed pitch at all but a gesture: attack, then a slide upward, a slow oscillation, a fade toward silence with the finger still moving on a string that has almost stopped sounding. The friction of skin on silk is audible and is not considered noise. The tradition classifies dozens of named left-hand techniques, and a player is judged on the quality of sounds that a spectrum analyzer would barely register.

The qin belonged to the scholar-official class, the literati, for whom it was one of the four accomplishments alongside the board game go, calligraphy, and painting. It was played for self-cultivation, not for audiences, and the literature around it is a literature of ethics as much as of technique.

Key idea: On the qin the sound after the attack is the music, which is why a tradition of extremely quiet, timbrally detailed playing developed alongside an ideology of private cultivation.

Tablature that specifies action but not time

Qin music is written in jianzipu, reduced-character notation. Each symbol is a composite glyph assembled from fragments of Chinese characters, and it tells the player: which string, which finger of the right hand and what stroke, where the left hand stops the string in relation to which hui, and which technique to apply afterward. It is a complete instruction for what the hands do.

It does not notate rhythm. There is no duration, no meter, no indication of how long to hold anything.

The consequence is a practice with no European equivalent. Dapu is the work of reconstructing a playable piece from a tablature, and it can take a player months for a single composition: deciding the phrasing, the pacing, the shape. Two accomplished players working from the same handbook page will produce recognizably different pieces, both legitimate. Handbooks such as the Shenqi Mipu of 1425 preserve hundreds of pieces this way, so the repertoire is simultaneously very old and, in performance, continuously re-created.

Two pieces to know. Liu Shui, Flowing Water, exists in a recording by Guan Pinghu that NASA placed on the Voyager Golden Record in 1977; at seven and a half minutes it is the longest music on the disc, and two spacecraft are now carrying it out of the solar system. Guangling San is attached to the scholar Ji Kang, executed in 262, who is said to have played it before his death and remarked that it would now be lost. It was not.

Music as government business

Confucian thought treated music as a moral and political instrument. The Record of Music, part of the Book of Rites, argues that music expresses the state of a society and shapes it in return: orderly music produces orderly people, and licentious music indicates a state in trouble. Court music, yayue, was distinguished from popular music, suyue, and maintained by an official bureau.

This produced one of the strangest facts in the history of tuning. The fundamental pitch of the empire, called huangzhong, yellow bell, was fixed by the length of a standard pitch pipe, and successive dynasties recalculated it. Setting the pitch standard was an act of legitimacy, like issuing a calendar or a currency: a dynasty that had received the mandate should be able to determine the correct pitch from first principles. Court scholars spent centuries on the mathematics.

One of them solved a problem Europe was still working on. In 1584 Zhu Zaiyu, a prince of the Ming, published the correct computation of equal temperament, deriving the ratio of the twelfth root of two and carrying it to many decimal places, decades before Mersenne. Chinese music had no use for it, since it had no keyboards needing to modulate, so it stayed a mathematical achievement rather than a musical revolution. That is a useful corrective to any story in which equal temperament is the natural endpoint of musical progress.

Key idea: Pitch standards in imperial China were instruments of statecraft, and the mathematics of equal temperament was solved there in 1584 without being adopted, because the music had no problem it solved.

Strings, winds, and the sound of a teahouse

Beyond the qin, the instruments most listeners meet first:

InstrumentWhat it isSound to listen for
PipaFour-string pear-shaped lute, played with taped-on false nailsRapid tremolo across all fingers; percussive effects in battle pieces such as Ambush from Ten Sides
ErhuTwo-string fiddle with a snakeskin-covered resonator; the bow hair passes between the stringsContinuous portamento, a vocal quality, no open-air brightness
GuzhengLong zither with movable bridges, plucked with picksLeft-hand pressure behind the bridge bends notes after they sound
DiziTransverse bamboo flute with an extra hole covered by a thin membraneA distinctive buzzing edge on every note, produced by the membrane
ShengMouth organ of vertical pipes with free reeds, sounding on both breath directionsChords and clusters; the free reed principle reached Europe in the eighteenth century and produced the accordion family

In Shanghai teahouses the jiangnan sizhu ensemble, silk and bamboo, plays a shared melody in which each instrument ornaments in its own idiom. This is textbook heterophony: the flute adds turns, the pipa breaks the line into plucked figures, the erhu slides between notes. The slight blur between versions is the aesthetic, not a coordination failure.

Peking opera

In 1790 opera troupes from Anhui traveled to Beijing to perform for the Qianlong emperor's birthday and stayed. What developed over the following century became jingju, Peking opera: a total theater of stylized speech, song, acrobatics, mime, and combat, with role types (sheng for male roles, dan for female, jing for painted-face characters, chou for comic) that determine vocal production, movement, makeup, and costume.

Two melodic systems, xipi and erhuang, supply the tunes; a small high-pitched fiddle, the jinghu, doubles and pushes the singer; and a percussion group led by clappers and a small gong controls the pacing of the stage action, so that entrances, fights, and emotional turns are cued by drum patterns. The singing uses a tight, high, deliberately nasal placement that Western listeners often find startling and that is trained for years.

Mei Lanfang, the most celebrated performer of female roles in the twentieth century, toured the United States in 1930 and the Soviet Union in 1935. Bertolt Brecht saw him in Moscow and drew from the experience part of his theory of theatrical estrangement, which is a small reminder that influence has never run in only one direction.

Reform, rupture, and revival

The twentieth century rebuilt Chinese music institutionally. Conservatories were founded from the 1920s. Instruments were redesigned: the erhu acquired steel strings and a louder projection, the pipa was refretted to equal temperament so it could play with Western instruments and in fixed keys, and a modern Chinese orchestra was assembled on the model of a symphony, with sections, a conductor, arranged scores, and newly invented bass instruments to fill a register that traditional ensembles never had.

During the Cultural Revolution, from 1966, most traditional repertoire was suppressed in favor of a small set of approved model works. Qin playing nearly stopped. Since the 1980s the tradition has been rebuilt, and UNESCO listed guqin music in 2003, which brought funding, students, and a market in expensive instruments.

Finally, scope. The People's Republic recognizes fifty-six ethnic groups, and this lesson has described Han traditions almost exclusively. Uyghur muqam in Xinjiang, a suite tradition related to the Central Asian and Middle Eastern maqam family you will meet in Lesson 11, is Chinese by citizenship and belongs to an entirely different musical world. Tibetan, Mongolian, Naxi, and Dong traditions are equally distinct; Dong grand song is a polyphonic choral practice with nothing Han about it. When a phrase like Chinese music appears, ask which China.

Key idea: Modern Chinese instruments and ensembles were substantially re-engineered in the twentieth century to play in equal temperament and in orchestral formats, so a concert Chinese orchestra is a modern institution rather than an ancient one.

Common misconceptions

  • "Qin tablature is a score." It specifies exactly what the hands do and says nothing about rhythm, which is why dapu reconstruction is a discipline.
  • "The Chinese orchestra is the traditional ensemble." It was assembled in the twentieth century on a symphonic model, with redesigned and newly invented instruments.
  • "Chinese music is pentatonic." Pentatonic frameworks are common, but the systems, tunings, and regional practices vary enormously, and the imperial theory of pitch was anything but simple.
  • "Equal temperament came from Europe to China." Zhu Zaiyu computed it in 1584; it was not adopted there because it solved no local problem.
  • "Peking opera is very old." It dates from after 1790 and took its recognizable form during the nineteenth century.

Recap

  • The qin is a quiet seven-string zither whose three sound categories are open strings, harmonics at the hui, and stopped notes that keep changing after the attack.
  • Jianzipu notates hand action without rhythm, so performance requires dapu, and multiple valid realizations coexist.
  • Confucian ideology made music state business, and imperial pitch standards were claims to legitimacy; Zhu Zaiyu solved equal temperament in 1584.
  • Pipa, erhu, guzheng, dizi, and sheng each have identifiable timbral signatures, and jiangnan sizhu is a clear case of heterophony.
  • Peking opera dates from after 1790; twentieth-century reform produced conservatories, refretted instruments, and the modern Chinese orchestra, and the Cultural Revolution interrupted traditional practice before its revival.

Sources

  1. "Chinese music." Encyclopaedia Britannica. Instruments, court theory, and historical periods. britannica.com.
  2. "Guqin." Wikipedia. Construction, hui, timbral categories, jianzipu, and dapu. en.wikipedia.org.
  3. "Peking opera." Wikipedia. Origins after 1790, role types, and musical systems. en.wikipedia.org.
  4. NASA Jet Propulsion Laboratory. Voyager Golden Record contents. voyager.jpl.nasa.gov.
  5. UNESCO. Lists of Intangible Cultural Heritage, including guqin music and Uyghur muqam. ich.unesco.org.
Key terms
Guqin
A quiet seven-string Chinese zither with thirteen hui marking harmonic positions, associated with the literati.
Hui
The thirteen inlaid markers along the qin that indicate the natural harmonic nodes.
Jianzipu
Qin tablature of composite characters specifying string, finger, position, and technique, but not rhythm.
Dapu
The reconstruction of a playable performance from qin tablature, which yields different valid realizations.
Yayue
Elegant court music maintained by the imperial bureaucracy, distinguished from popular suyue.
Huangzhong
The imperial fundamental pitch, recalculated by successive dynasties as an act of political legitimacy.
Jiangnan sizhu
Silk and bamboo chamber ensembles of the Shanghai region, a clear example of heterophony.
Jingju
Peking opera, developed after Anhui troupes came to Beijing in 1790, organized by role types.
Chinese orchestra
The twentieth-century concert ensemble built on a symphonic model with redesigned and newly invented instruments.

Japan and Korea: Court Music, Breath, and the Drum

  • Describe the gagaku ensemble and how its parts combine without a conductor.
  • Explain ma and jo-ha-kyu and identify them in specific Japanese practices.
  • Describe pansori performance, including the roles of the drummer and the audience.
  • Distinguish long-established practices from twentieth-century inventions in both countries.

The big picture

Before a gagaku performance, one of the musicians holds a mouth organ over a small electric heater. The sho has seventeen slender bamboo pipes rising from a wind chamber, sounded by free metal reeds, and the player's breath condenses inside it. A wet reed will not speak. So the instrument is warmed before playing and returned to the heater whenever it is idle, and if you attend a performance you will see it happen at the edge of the stage.

The ensemble this instrument belongs to has been maintained by the Japanese court since the ninth century, when the imperial bureau reorganized a body of music that had arrived over the preceding centuries from Tang China, the Korean peninsula, and further west. Today about two dozen musicians employed by the Imperial Household Agency perform it, many from families that have held the positions for generations. It is plausibly the oldest continuously performed ensemble music in the world, and the sho player still has to deal with condensation.

Gagaku: how it is built

The repertoire divides into togaku, music of Chinese derivation, and komagaku, of Korean derivation, historically staged as the music of the left and of the right. The ensemble has three layers:

LayerInstrumentsFunction
MelodyHichiriki (short double reed, loud and piercing), ryuteki (transverse flute)Play heterophonic versions of the same melody, slightly apart from each other
Sustained harmonyShoHolds five- and six-note clusters called aitake that change slowly, floating under the melody
PunctuationKakko, taiko, shoko; biwa and kotoMark positions in the cycle; the plucked strings arpeggiate rather than play tunes

Two things surprise new listeners. First, the tempo. Etenraku, the best-known piece, unfolds at roughly the pace of slow breathing, and the melody instruments phrase by breath rather than by beat. Second, the sho clusters. They are not chords in any functional sense and they do not resolve; they are a sustained sonic environment inside which the melody moves. There is no conductor. The ensemble coordinates by watching the percussion and by breathing together, and the small variations in exactly when the hichiriki and ryuteki arrive at each note are part of the sound rather than a flaw in it.

Key idea: Gagaku is layered like the gong-chime ensembles of Southeast Asia, with a slow melody, a sustained cluster layer, and punctuating percussion that marks the cycle, and no one directs it.

Two aesthetic principles worth carrying everywhere

Ma is the interval between things: the space between two sounds, treated as charged rather than empty. A shakuhachi phrase followed by four seconds of nothing has not stopped; the silence is part of the phrase and is shaped by what preceded it. Performers speak of the quality of a particular silence as they would of the quality of a note. If you listen to Japanese traditional music expecting continuous sound, you will hear it as full of holes; the holes are load-bearing.

Jo-ha-kyu is a principle of pacing articulated in noh theory by Zeami in the fifteenth century and applied at every scale: an introduction that begins slowly (jo), a development that breaks and expands (ha), and a rapid conclusion (kyu). A single phrase can be shaped jo-ha-kyu; so can a piece; so can a whole day's program. It is a rule about acceleration, and once you know it you will hear it structuring performances that seem otherwise formless.

The shakuhachi: an instrument that makes a virtue of breath

The shakuhachi is an end-blown bamboo flute with four finger holes in front and one behind. The standard length, one shaku eight sun, about 54.5 centimeters, gives the instrument its name. The player blows across a sharp edge cut outward from the rim, and pitch is controlled not only by the holes but by the angle of the head: lowering the chin, called meri, flattens a note substantially and darkens its timbre; raising it, kari, does the reverse. The same fingering therefore produces several different pitches with several different colors, and the tradition exploits that.

Its solo repertoire, honkyoku, comes from the Fuke sect of the Edo period, whose komuso monks played as a form of meditation, blowing zen, wore basket hats that covered the face, and held travel privileges that the government eventually suspected of being cover for spies. The sect was abolished in 1871 and the music survived as an art practice. Listen to Kyorei or to Shika no Tone, a piece for two players representing distant deer calling to each other across a valley, in which the two instruments answer from opposite sides of the space.

The crucial point for a listener trained on Western flute tone: the audible rush of breath is not a defect. A player can attack a note with muraiki, a deliberately violent breathy blast, and the noise is the expressive content. Clean tone is available and is one option among many.

Koto, shamisen, and taiko

The koto is a long zither with thirteen strings over movable bridges, plucked with picks on three fingers of the right hand while the left presses behind the bridges to bend pitch after the attack. Yatsuhashi Kengyo, a blind seventeenth-century musician, is credited with the tunings and repertoire that founded the modern solo tradition; Rokudan no Shirabe, six variations over a fixed structure, is the piece most students learn. Michio Miyagi's Haru no Umi of 1929, for koto and shakuhachi, is a modern composition that has become standard new year listening in Japan.

The shamisen is a three-string lute with a skin-covered body played with a large wooden plectrum, the bachi, which strikes the skin as well as the string, so every note has a percussive slap in it. The lowest string is set to buzz against a deliberately shaped groove, the sawari: another engineered noise, like the mbira's bottle caps and the sitar's bridge. Its genres are theatrical: gidayu-bushi narrates bunraku puppet plays, nagauta accompanies kabuki, and Tsugaru-jamisen from the far north is a fast, percussive solo style now played by touring virtuosos.

Taiko needs a plain statement. Japanese drums are ancient and appear in festivals, temples, and theater. The ensemble of many drummers playing choreographed arrangements, which is what the word means to most people outside Japan, was invented in 1951 by Daihachi Oguchi, a jazz drummer in Nagano who arranged an old shrine score for multiple drums. Ondekoza formed in 1969 and Kodo in 1981 on Sado Island, adding rigorous physical training and long-distance running to the practice. The result is a thrilling and genuinely Japanese art form roughly seventy years old.

Key idea: Several Japanese and Korean practices presented internationally as ancient are twentieth-century creations built from older materials, and saying so does not diminish them.

Pansori: one singer, one drummer, eight hours

Korean pansori is sung epic narrative. The performer, the sorikkun, stands with a folding fan and performs an entire story alone. The only other person is the gosu, seated with a barrel drum. Of an original twelve stories, five survive complete, including Chunhyangga, the tale of a faithful lover, and Simcheongga, the story of a daughter who sacrifices herself to restore her blind father's sight. A full performance can run from three to eight hours.

The performance alternates three modes: sori, the sung passages; aniri, spoken narration and dialogue that also lets the singer rest; and ballim, gesture with the fan and body. The drummer is not accompaniment in the background sense. He interjects chuimsae, spoken calls of encouragement and approval, at structurally appropriate places, and the audience is expected to join in. A silent pansori audience is a failed one.

The vocal ideal is a rough, cracked, powerful sound. Training traditionally involved singing against a waterfall for hours a day until the voice broke down and re-formed with the desired grain. A conservatory-trained Western ear hears damage; the tradition hears the sound of a voice that has been through something, which is exactly the aesthetic point given what the stories are about.

Jangdan and sanjo

Korean rhythm is organized by jangdan, cycle patterns played on the hourglass drum, most of them subdividing in threes, which is why so much Korean music has a lilting rather than a square feel.

JangdanCharacterWhere you meet it
JinyangjoVery slow, long cycleThe opening of a sanjo; the most inward passages of pansori
JungmoriModerateNarrative sections
JajinmoriFastAction and excitement
HwimoriFastestClimaxes

A sanjo, meaning scattered melodies, is a solo instrumental form that moves through these cycles from slowest to fastest across thirty minutes or more, accompanied only by the janggu drum. It was developed in the late nineteenth century and is credited in its gayageum form to Kim Chang-jo. The gayageum has twelve silk strings over movable bridges, plucked with the fingers, with a deep left-hand press behind the bridge that makes notes sag and swell after they sound; the geomungo has six strings struck with a short bamboo stick and a much darker, more percussive voice.

Samul nori is the other case worth naming precisely. In February 1978 four musicians in a small Seoul theater took the instruments of outdoor farmers' band music, the small gong kkwaenggwari, the large gong jing, the hourglass janggu, and the barrel drum buk, sat down, and played them as concert music. The group's name became the genre's name. Within a decade it was touring the world and being taught as Korean traditional percussion. The materials are old, the format is from 1978.

A word about han

You will read that Korean music expresses han, an untranslatable national sorrow. Treat that claim carefully. Korean scholars have argued at length that han as a defining national essence is a twentieth-century construction that took shape under and after Japanese colonial rule between 1910 and 1945, when a discourse of Korean sadness suited the colonizer, and that it was then absorbed into Korean self-description. Pansori certainly contains grief, and it also contains obscene jokes, satire of officials, and comic set pieces. If a program note tells you an entire nation's music expresses one emotion, ask who first said so and when.

Key idea: Claims that a national music expresses a single essential emotion usually have a datable political history, and han is a case where that history is documented.

Common misconceptions

  • "Gagaku is slow because it is ceremonial filler." The pacing is the form; melody instruments phrase by breath and the sho cluster layer is a sustained environment, not a chord progression.
  • "Silence in Japanese music is where nothing happens." Ma is treated as shaped and charged, and performers work on it directly.
  • "Taiko ensembles are ancient." The drums and festivals are old; kumi-daiko was invented in 1951.
  • "A pansori audience listens quietly." Calls of encouragement from drummer and audience are part of the performance.
  • "The rough pansori voice is untrained." It is the product of deliberate and punishing training toward a specific ideal.

Recap

  • Gagaku layers heterophonic melody, sustained sho clusters, and punctuating percussion, maintained by the Japanese court since the ninth century and coordinated without a conductor.
  • Ma treats silence as shaped material; jo-ha-kyu organizes pacing at every scale from a phrase to a program.
  • Shakuhachi pitch and color depend on head angle as much as fingering, and breath noise is expressive content.
  • Shamisen sawari and koto after-attack bending show the same interest in what happens to a sound once it exists.
  • Pansori is solo epic narration with a drummer whose calls, and the audience's, belong to the performance; jangdan cycles organize Korean rhythm and sanjo moves through them slow to fast.
  • Kumi-daiko (1951) and samul nori (1978) are modern formats built from old materials, and han as national essence has a datable colonial history.

Sources

  1. "Japanese music." Encyclopaedia Britannica. Court, theatrical, and instrumental traditions. britannica.com.
  2. "Gagaku." Wikipedia. Ensemble, repertoire divisions, and instruments. en.wikipedia.org.
  3. "Korean music." Encyclopaedia Britannica. Court and folk genres, instruments, and rhythmic cycles. britannica.com.
  4. "Pansori." Wikipedia. Performance practice, surviving stories, and vocal training. en.wikipedia.org.
  5. UNESCO. Lists of Intangible Cultural Heritage, including pansori epic chant and gagaku. ich.unesco.org.
Key terms
Sho
The Japanese seventeen-pipe free-reed mouth organ that sustains slow clusters called aitake in gagaku.
Gagaku
Japanese court music of Chinese and Korean derivation, maintained by the imperial household since the ninth century.
Ma
The charged interval between sounds, treated as shaped material rather than empty space.
Jo-ha-kyu
A pacing principle of slow introduction, expanding development, and rapid conclusion, applied at every structural scale.
Honkyoku
The solo shakuhachi repertoire of the Fuke sect, played as a meditative practice.
Sawari
The deliberate buzz built into the shamisen's lowest string.
Pansori
Korean sung epic narrative for one singer and one drummer, alternating sung, spoken, and gestural modes.
Chuimsae
Calls of encouragement from the pansori drummer and audience that form part of the performance.
Jangdan
Korean rhythmic cycle patterns, mostly subdividing in threes, played on the janggu drum.
Sanjo
A solo instrumental form that moves through jangdan cycles from slowest to fastest.

The Maqam World: Arab, Turkish, Persian, Sufi, and Jewish Traditions

  • Explain how a maqam is built from ajnas and describe neutral intervals in cents.
  • Describe iqa cycles using dum and tak, and the role of taqsim and tarab in performance.
  • Compare Arab maqam, Turkish makam, and the Persian radif and dastgah systems.
  • Describe qawwali and Mevlevi practice and the political conditions each has faced.

The big picture

On the first Thursday of the month, for decades from the 1930s, Egyptian radio broadcast a live concert by Umm Kulthum, and cafes from Casablanca to Baghdad filled with people who had arranged their evening around it. A single song might run an hour. She would sing one line of a poem, then sing it again shaped differently, then again, ten or fifteen times, while the audience shouted for more, and the orchestra waited for her decision about where to go next. Nothing was being padded. The repetition was the art: the same words rebuilt until the room reached the state the whole tradition is organized around.

That state is called tarab, and it is the key to the music covered here, a family of related systems running from Morocco to Central Asia. This lesson takes them in order: the Arab maqam system and its intervals, the rhythmic cycles, Turkish and Persian practice, Sufi devotional music, and Jewish traditions that grew inside the same soundworld.

Maqam: modes built from smaller pieces

A maqam is a melodic mode, and it is assembled rather than simply listed. The building blocks are ajnas (singular jins): short scale fragments of three, four, or five notes with a characteristic internal shape. A maqam is typically one jins sitting on the tonic and another jins starting on its fourth or fifth degree, plus a sayr, the customary path a melody takes through it: where to begin, which notes to dwell on, where to pause, how to descend to the end.

Take maqam Rast, the reference maqam of the system, built on C:

  • Lower jins Rast: C, D, E half-flat, F
  • Upper jins Rast on G: G, A, B half-flat, C

Maqam Bayati on D uses jins Bayati (D, E half-flat, F, G) below and typically jins Nahawand above. Maqam Hijaz on D uses D, E flat, F sharp, G, whose augmented second between the second and third degrees is the interval that Hollywood spent a century using as shorthand for the East.

The neutral intervals, in cents

That half-flat is the feature most listeners notice first. It is a real, stable, precisely intended pitch, not an approximation of anything.

Interval above the tonicCents (approximate)Comment
Minor second100Same as a piano semitone
Neutral second140 to 160Between a semitone and a whole tone; the sound people call a quarter tone
Major second200Same as a piano whole tone
Minor third300
Neutral third (Rast)340 to 355Sits deliberately between minor and major
Major third400

Two points follow. First, this is not a twenty-four-note equal-tempered scale. The 1932 Congress of Arab Music in Cairo adopted twenty-four quarter tones as a theoretical convenience, and practice never matched it: the neutral second in Aleppo is not the neutral second in Cairo, and a fine performer varies it within a phrase. Second, the instruments are built for that flexibility. The oud is fretless. The nay is a rim-blown reed whose pitch bends with the angle of the head. The qanun, a plucked zither with dozens of triple-course strings, carries rows of small metal levers called mandal that let the player raise or lower a pitch by a fraction of a semitone in the middle of a piece.

That Cairo congress is worth pausing on. Convened by the Egyptian king in 1932, it brought Arab musicians together with European scholars including Bela Bartok, Paul Hindemith, Curt Sachs, and Erich von Hornbostel, and it argued about whether Arab music should adopt Western harmony, notation, tempered pianos, and conservatory pedagogy. The congress declined wholesale adoption and made a large body of recordings that survive. It is a rare, dated moment where a tradition debated its own modernization in public with the comparative musicologists of Module 1 in the room.

Key idea: Neutral intervals are deliberate and variable, not quarter-tone approximations, and the core instruments are built to place them freely rather than to fix them.

Iqa: the rhythmic cycles

Rhythm runs on cycles called iqa, notated with two syllables: dum, a low resonant stroke in the center of the drumhead, and tak, a high crisp stroke at the rim. Learn a cycle by saying it.

IqaBeatsPattern
Maqsum8dum tak - tak / - dum - tak
Baladi8dum dum - tak / - dum - tak
Samai thaqil10dum - - tak - / dum dum tak - -

Samai thaqil, in ten, is the signature cycle of the Ottoman-Arab instrumental samai form, and once you can count it the form's shape becomes obvious.

Taqsim and tarab

A taqsim is a solo improvisation, usually unmetered, in which a player explores a maqam: establishing the lower jins, testing the upper, modulating to neighboring maqamat, and eventually descending to the tonic. It is the most demanding thing an instrumentalist does, because there is nowhere to hide and the audience knows the sayr as well as the player.

The traditional ensemble is the takht: oud, qanun, nay, violin, and riqq tambourine, four or five musicians in a room. In the twentieth century Umm Kulthum and Mohammed Abdel Wahab worked with a firqa, a large orchestra with massed violins, cello, and eventually electric guitar and organ, which is why mid-century Egyptian recordings sound orchestral while remaining entirely modal.

Tarab is the goal, and it is a joint achievement. Audiences respond audibly, calling out, asking for a line again, naming the maqam. Musicians speak of saltana, a state of complete immersion in a maqam that a performer needs to reach and that listeners can hear the presence or absence of. A recording that captures the audience is therefore more complete than a studio take, which is a nice practical example of Merriam's behavior level from Lesson 1.

Turkey: makam and usul

Ottoman court music developed a parallel and more theoretically elaborate system. Turkish theory divides the octave into fifty-three commas and uses a set of accidentals to raise and lower notes by these small increments, which lets it distinguish makams that Arab practice treats as one. Rhythmic cycles, usul, range from two beats to enormous spans: Devr-i Kebir runs twenty-eight beats and there are longer.

The instruments are close cousins of the Arab ones: the ney, the long-necked fretted tanbur whose many frets include microtonal positions, the kanun, and the small upright kemence. A concert suite, the fasil, alternates composed pieces with taksim improvisations.

The Mevlevi ayin is the ceremonial music of the Sufi order founded in the wake of Rumi, who died in Konya in 1273. The ceremony opens with a ney taksim, deliberately evoking the first lines of Rumi's Masnavi, in which the reed flute cries because it has been cut from the reed bed and separated from its source. The dervishes turn, one hand raised and one lowered, for extended periods. In 1925 the Turkish Republic closed the dervish lodges and banned the orders as part of secularization. The music survived privately and was permitted again from the 1950s as a cultural and touristic performance in Konya, which is how most people now encounter it, and it entered the UNESCO lists in 2008. A ceremony banned by a state, revived as heritage, and now performed for ticket holders is the whole argument of Module 6 in one institution.

Key idea: A performance context can be outlawed, survive privately, and return as officially sponsored heritage, and the music will not be the same practice on the far side of that journey.

Iran: the radif and the dastgah

Persian classical music is organized into twelve dastgah and avaz, each a collection of short melodic models called gushe arranged in a canonical sequence. The whole memorized body of these models is the radif, and there is no single radif: there are named versions transmitted through particular teachers, most influentially that of Mirza Abdollah, written down in the twentieth century by Nur-Ali Borumand.

A performance selects from the radif rather than reciting it. A player opens with the daramad, the introductory material that establishes the dastgah, and then moves through chosen gushe, generally rising in register and intensity before descending to close. Vocal performance is set to classical Persian poetry, above all Hafez, Rumi, and Saadi, and features tahrir, a rapid break in the voice between registers comparable in mechanism to yodeling and entirely different in effect: it ornaments the ends of lines like a catch in the throat. Instruments include the tar and setar (plucked lutes), the hammered santur, the spike fiddle kamancheh, the ney, and the goblet drum tombak. UNESCO listed the radif in 2009.

After 1979 music's public status in Iran became restricted and remains contested: solo singing by women before mixed public audiences is not permitted, and the boundaries around permissible performance have shifted repeatedly. Classical music continues under those conditions, with a large audience.

Sufi devotional music

Qawwali is the devotional music of the Chishti Sufi order in South Asia, developed at the Delhi shrine circle of Nizamuddin Auliya and traditionally credited to the poet and musician Amir Khusrau, who died in 1325. A qawwali party is eight to ten men seated on the floor: a lead singer, side singers who answer him, one or two harmoniums, tabla or dholak, and the entire group clapping.

The performance has a shape you can follow without knowing Urdu or Persian. An instrumental prelude, then an unmetered vocal introduction, then the composed poem enters over a cycle. As it builds, the lead singer inserts girah, couplets borrowed from other poems, to intensify a particular line, and repeats key phrases with increasing pitch and speed while the clapping drives underneath. The intended result is hal, an ecstatic state in listeners, who may rise, weep, or offer money. Nusrat Fateh Ali Khan, who died in 1997, made this music internationally famous while remaining a working shrine performer.

Two honest notes. Devotional music at shrines is contested within Islam: schools differ on the permissibility of instrumental music and of the practices around shrines, and some reformist movements reject them entirely. That disagreement has at times turned violent. The qawwal Amjad Sabri was shot dead in Karachi in 2016, and a 2017 bombing at the shrine of Lal Shahbaz Qalandar in Sehwan killed scores of people during the evening devotional dance.

In Morocco, Gnawa practice descends from sub-Saharan Africans brought north through the trans-Saharan slave trade. In an all-night lila ceremony, a maalem leads with the guembri, a three-string bass lute with a camel-skin face, accompanied by iron castanets called qraqeb, working through a sequence of spirits associated with colors, with participants entering trance. Since 1998 the annual Essaouira festival has made Gnawa a global export and a site of continuing argument about what happens when a healing ceremony becomes a stage act.

Jewish traditions in the same soundworld

Torah cantillation is a system of accent marks, the te'amim, fixed by the Masoretic scholars of Tiberias around the ninth and tenth centuries. The marks indicate phrase structure and melodic motifs, and they are identical in every Torah scroll and printed Bible worldwide. The melodies attached to them are not: Ashkenazi, Sephardi, Yemenite, and Iraqi communities sing the same signs to entirely different tunes. A notation system that preserves structure across a thousand years while allowing regional sound to diverge completely is an unusually clean example of what notation does and does not carry.

Abraham Zvi Idelsohn began recording the Jewish communities of Jerusalem in 1907 and published his Thesaurus of Hebrew Oriental Melodies over the following decades, which makes him both a founding figure in Jewish music scholarship and one of the comparative musicologists of Module 2, with the same mixed inheritance.

Among Syrian Jews, particularly the Aleppo tradition, sacred songs called pizmonim are set to Arab maqam, and a specific maqam is assigned to the Sabbath liturgy each week according to the Torah portion, chosen for the emotional character it carries. The melodies are frequently shared with the surrounding Arab repertoire. Here Jewish and Arab music are not two traditions in contact; they are one musical system used by two communities.

Klezmer is the instrumental wedding music of Ashkenazi Eastern Europe, built on modes including the one usually called freygish, with its lowered second and raised third, and forms including the unmetered doina and the dance-driving bulgar. Its performance tradition nearly disappeared through emigration, the Holocaust, and assimilation, and was rebuilt from the 1970s by musicians who learned repertoire from surviving elders and from old 78s: a revival that was explicitly archival in method and is now a living dance music again.

Common misconceptions

  • "Arab music uses quarter tones." It uses neutral intervals that vary by region, performer, and phrase; the 24-tone division was a 1932 theoretical convenience that practice never followed.
  • "A maqam is a scale." It is ajnas plus a customary path, and two maqamat can share pitches while behaving completely differently.
  • "Taqsim is free improvisation." It follows the maqam's sayr closely, and knowledgeable listeners hear every departure.
  • "Sufi music is universally accepted in the Muslim world." Shrine-based devotional practice is contested within Islam, sometimes violently.
  • "Jewish and Arab music are separate traditions." Aleppo pizmonim are sung in Arab maqam with shared melodies; they are one system used by two communities.

Recap

  • A maqam is built from ajnas plus a sayr; Rast, Bayati, and Hijaz are the entry points.
  • Neutral seconds sit around 140 to 160 cents and neutral thirds around 340 to 355, deliberately and variably placed; fretless oud, rim-blown nay, and the qanun's mandal levers make that possible.
  • Iqa cycles are learned as dum and tak; taqsim explores a maqam without meter; tarab is a shared state that the audience helps produce.
  • Turkish makam theory uses fifty-three commas and long usul cycles; the Mevlevi ayin was banned in 1925 and returned as sponsored heritage.
  • Persian practice organizes gushe into dastgah, memorized as a radif transmitted in named versions.
  • Qawwali builds toward ecstatic states through repetition, inserted couplets, and rising intensity; Gnawa lila ceremonies use guembri and qraqeb; both now also exist as festival music.
  • Torah cantillation preserves identical signs with divergent melodies; Aleppo pizmonim use Arab maqam; klezmer was rebuilt from elders and old recordings.

Sources

  1. "Islamic arts: Music." Encyclopaedia Britannica. Maqam, instruments, and regional traditions. britannica.com.
  2. "Maqam." Wikipedia. Ajnas, sayr, intervals, and regional practice. en.wikipedia.org.
  3. "Qawwali." Wikipedia. Ensemble, structure, and shrine context. en.wikipedia.org.
  4. "Radif (music)." Wikipedia. Dastgah, gushe, and transmitted versions. en.wikipedia.org.
  5. UNESCO. Lists of Intangible Cultural Heritage, including the Mevlevi sema ceremony and the Persian radif. ich.unesco.org.
Key terms
Maqam
An Arab melodic mode built from ajnas plus a customary path through them.
Jins
A three to five note scale fragment with a characteristic shape; the building block of a maqam.
Sayr
The customary route a melody takes through a maqam: where to start, dwell, pause, and close.
Neutral interval
A pitch deliberately placed between the standard minor and major sizes, around 150 or 350 cents.
Iqa
An Arab rhythmic cycle, learned and recited as dum (low) and tak (high) strokes.
Taqsim
An unmetered solo improvisation that explores a maqam and its modulations.
Tarab
The state of shared musical enchantment that Arab performance aims to produce with its audience.
Usul
A Turkish rhythmic cycle, ranging from two beats to spans of dozens of beats.
Radif
The memorized Persian repertoire of gushe organized into dastgah, transmitted in named teacher versions.
Pizmonim
Syrian Jewish sacred songs set in Arab maqam, with a maqam assigned to each week's liturgy.

Module 5: The Americas, Europe, and Oceania

Indigenous musics of North and South America, including how songs are owned, what was banned and when, and what respectful listening requires. Then the African and Iberian streams that became son, samba, tango, cumbia, and salsa, the Caribbean traditions that produced calypso, steelpan, and reggae, European collecting and revival, and Pacific song.

Indigenous Musics of the Americas

  • Describe the structure of a Northern Plains powwow song and the protocols around the drum.
  • Explain song ownership and restricted repertoire and what they imply for recording and circulation.
  • Summarize the legal suppression of Indigenous ceremony in the United States and Canada and its musical consequences.
  • Evaluate specific cases of appropriation and of genuine collaboration.

The big picture

On the Northern Plains, a drum is not only an object. A single large drum sits on a low stand with eight to twelve singers seated around it, each holding a padded stick, and the word names the group as well as the instrument: the Black Lodge Singers are a drum, Northern Cree is a drum. When a host says that four drums will be at the powwow, they mean four ensembles, each with its own singers, its own repertoire, and its own reputation.

That small linguistic fact opens onto most of what this lesson covers. In these traditions, songs are things people own, receive, care for, and pass on under rules, and the social unit and the musical unit are the same thing. This lesson looks at how the music is built, how it is governed, what happened when governments tried to end it, and what a listener outside these communities should and should not do.

How a powwow song is built

Listen to any Northern Plains intertribal song and you will hear a repeatable structure:

  1. The lead. One singer begins alone, very high, at the top of his range, setting the melody.
  2. The seconds. The group enters, overlapping the end of the lead's phrase, so the join is a slight collision rather than a clean handoff.
  3. The descent. The melody works downward in terraces: a phrase high, then the same shape lower, then lower again, arriving near the bottom of the range.
  4. Honor beats. Partway through, the drum plays a set of accented strokes, and dancers respond to them with a specific movement.
  5. Push-ups. The whole song is repeated, usually four times. A push-up is one complete rendition.

The text is often entirely vocables, syllables such as hey and ya that carry no lexical meaning. These are not improvised filler. A given song has its own fixed vocables, and singing them wrong is singing the song wrong. Many songs also contain words, and some are bilingual. Northern style sits higher and tenser than Southern style, which is lower and slower, and experienced listeners identify a drum's home region from the vocal placement alone.

Protocol surrounds the drum. It is kept and cared for, not left unattended, and there are rules about who sits at it, about alcohol, and about which songs may be sung when. A song may be given to a person or a family as a gift, which transfers a right to sing it. None of this is folklore around the music; it is part of the music's existence conditions.

Key idea: In many Indigenous American traditions a song is property with an owner and a set of permissions, so the question of who may sing it has a definite answer that is not the singer's preference.

Ownership, restriction, and the limits of recording

Songs may be owned by individuals, families, clans, or medicine societies. Some are received in dreams or visions and belong to the person who received them. Some may be sung only in a particular season, or only within a ceremony, or only by someone who has been given the right.

Navajo ceremonial practice shows the scale involved. Major ceremonial complexes such as Blessingway and Enemy Way involve long fixed sequences of songs performed in strict order across days, where the correctness of the sequence is the point: a mistake can require starting over, because the ceremony works through exact performance. David McAllester's study of Enemy Way music in 1954 was among the first to build an account of an Indigenous aesthetic from what Navajo people themselves said made music good, rather than from a visiting analyst's categories.

The practical consequence for anyone handling recordings: some material is not for general circulation, and the fact that a recording exists, is old, is out of copyright, and sits on a public server settles nothing. This is exactly the gap that the Traditional Knowledge labels from Lesson 3 were built for.

What was banned, and when

This is not distant history, and the dates matter.

  • In 1883 the United States established Courts of Indian Offenses with rules that made participation in ceremonies including the sun dance a punishable offense on reservations. Enforcement varied, but ceremonies moved out of sight, were compressed, or lapsed.
  • The Ghost Dance movement, whose songs were received in visions and spread across the Plains from 1889, ended for many communities at Wounded Knee in December 1890, where the United States Army killed some 250 to 300 Lakota, mostly noncombatants.
  • Boarding schools, beginning with Carlisle in 1879 in the United States and the residential school system in Canada, separated children from families and punished the use of Indigenous languages and songs, while teaching European band and choral music. Many schools had excellent brass bands, and that irony is part of the record.
  • Canada banned the potlatch from 1885 to 1951. Since songs, dances, names, and crests were transferred at potlatches, criminalizing the ceremony directly attacked the mechanism by which musical property changed hands. After a 1921 potlatch at Village Island, dozens of Kwakwaka'wakw people were prosecuted and regalia was confiscated; much of it was returned decades later to community-run cultural centers.
  • The American Indian Religious Freedom Act was passed in 1978. That is within living memory of most of the people singing today.

Set that beside Lesson 2. The same decades in which the Bureau of American Ethnology was recording thousands of songs onto wax cylinders were the decades in which federal policy was working to end the ceremonies those songs belonged to. Some communities today recover repertoire from those very cylinders. The archive is both evidence of harm and a resource against it.

Key idea: The suppression of Indigenous ceremony was legal policy within the last century, and the recordings that now support revival were made under it.

Katajjaq: a game with two voices

Inuit throat singing, katajjaq in Nunavik and Nunavut, is performed by two women standing face to face, often holding each other's arms, close enough that each uses the other's mouth as a resonating chamber. The sounds are short rhythmic patterns using both voiced tone and unvoiced breath, produced on the inhale as well as the exhale, which is what allows the pattern to continue without pause. One leads, the other follows a beat behind, so the two interlock.

It is a game. It continues until one of them runs out of breath, stumbles, or laughs, and laughing is the usual ending. Missionaries discouraged and in places banned it. It came back strongly from the 1980s through community teaching and festivals, and the Nunavut singer Tanya Tagaq built a solo improvised practice out of it that won the Polaris Prize in 2014 and has nothing traditional about its context while being unmistakably rooted in the technique.

The Andes, and an instrument that requires two people

In Aymara and Quechua communities of Peru and Bolivia, the siku panpipe is built in two complementary halves. One, the ira, has certain pitches; the other, the arka, has the pitches in between. Neither can play a scale. A melody exists only when two players alternate notes fast enough to make one line, which means the instrument physically cannot be played alone.

Large sikuri ensembles multiply this across dozens of players in parallel octaves, producing a massive interlocked sound at village festivals. It is the same hocket principle as Balinese kotekan and Aka horns, but here it is built into the object: reciprocity is not a metaphor about the community, it is a fact about whether a note happens. Alongside the sikuri you will meet the kena notched flute, the charango, a small high lute, and song genres such as the huayno.

Amazonia: songs that come from elsewhere

Among the Kisedje, formerly written Suya, of central Brazil, Anthony Seeger asked in the early 1970s who had composed a song and got an answer that reorganized his questions. Songs are not composed by people. They are heard from animals, or from spirits, or brought back by someone who has traveled outside ordinary human society, and they are then taught to the village.

His book Why Suya Sing (1987) argues that singing is not an expression of Kisedje society but a means of producing it: the ceremonial cycle of shout songs sung simultaneously by individuals and unison songs sung collectively creates the categories of person, name group, and time that structure the village. That is a genuinely different account of what music is for, and it does not translate into European aesthetics without loss.

Appropriation, with cases

Do not treat this as a general topic. Look at what actually happened.

The Indianist movement. Around 1900, American composers built concert pieces on Native melodies collected on cylinders. Charles Wakefield Cadman's From the Land of the Sky-Blue Water, published in 1909, was based on an Omaha melody and became a parlor hit. The melody had been collected by Alice Fletcher working with Francis La Flesche, who was Omaha, a scholar, and a co-author rather than an informant. So the same body of recordings supported both a genuinely collaborative scholarly project and a stream of concert arrangements in which the source communities were credited vaguely if at all and paid nothing.

Generic Indian music. The wooden flute and synthesizer genre marketed as Native American music, the tomahawk chop, and the powwow-flavored sports chant abstract a sound away from any specific nation. The harm is not that a flute was used; it is that hundreds of distinct musical traditions get replaced in public imagination by one invented sound belonging to nobody.

What good practice looks like now. Tribal cultural preservation offices set conditions for recording and publication. Archives apply Traditional Knowledge labels authored by the communities. Collections such as the Library of Congress holdings of Omaha music are online with documentation built with community involvement. And contemporary Indigenous musicians work in every genre: the electronic powwow of A Tribe Called Red, later The Halluci Nation, put powwow drum groups into club music on their own terms and with credited collaborators.

Key idea: The operative questions in an appropriation case are who consented, who is credited, who is paid, and who controls the result, and those questions have specific answers in specific cases.

Common misconceptions

  • "Native American music is one tradition." There are hundreds of distinct musical cultures across the hemisphere, as different from each other as Portuguese and Mongolian music.
  • "Vocables are meaningless improvisation." They carry no lexical meaning and are fixed for a given song; changing them is an error.
  • "These practices died out long ago." They were legally suppressed within living memory and are widely practiced today, including by young musicians in contemporary genres.
  • "If a recording is public domain, it can be used freely." Legal status and community permission are different questions, which is why TK labels exist.
  • "Powwow is ancient and unchanged." The intertribal powwow circuit is a twentieth-century pan-Indian institution built from older dance and song traditions.

Recap

  • A Plains song runs lead, seconds, terraced descent, honor beats, and four push-ups, with fixed vocables and protocol around the drum.
  • Songs are owned and may be restricted by season, ceremony, or right of use, which limits what may be recorded and circulated regardless of copyright status.
  • Ceremonies were criminalized in the United States from 1883 and in Canada from 1885, with the American Indian Religious Freedom Act arriving only in 1978.
  • Katajjaq is a two-woman breathing game, suppressed by missionaries and revived from the 1980s; the siku panpipe cannot be played by one person.
  • Kisedje song is received from animals and spirits and produces social categories rather than expressing them.
  • Appropriation cases turn on consent, credit, payment, and control, and collaboration such as Fletcher and La Flesche shows the alternative existed from the beginning.

Sources

  1. "Native American music." Encyclopaedia Britannica. Regional traditions, song structure, and ceremonial contexts. britannica.com.
  2. Library of Congress. Omaha Indian Music collection, recorded with Alice Fletcher and Francis La Flesche. loc.gov.
  3. "Inuit throat singing." Wikipedia. Katajjaq technique, suppression, and revival. en.wikipedia.org.
  4. "Pow wow." Wikipedia. Drum groups, song structure, and the intertribal circuit. en.wikipedia.org.
  5. Local Contexts. Traditional Knowledge Labels for Indigenous collections. localcontexts.org.
Key terms
Drum
On the Plains, both the large drum itself and the group of singers seated around it.
Push-up
One complete rendition of a powwow song; songs are usually sung four times through.
Vocables
Non-lexical sung syllables that are fixed for a given song rather than improvised.
Honor beats
A set of accented drum strokes partway through a song, to which dancers respond.
Song ownership
The principle that songs belong to individuals, families, or societies and may be transferred as gifts.
Katajjaq
Inuit throat singing performed as a breathing game between two women facing each other.
Siku
An Andean panpipe split into ira and arka halves so that two players are required to produce one melody.
Potlatch ban
The Canadian prohibition from 1885 to 1951 of the ceremony at which songs, dances, and names were transferred.

Atlantic Rhythms, European Revivals, and Pacific Song

  • Notate son and rumba clave and explain how a timeline organizes an Afro-Caribbean ensemble.
  • Trace specific Latin American and Caribbean genres to their African, Indigenous, and Iberian components.
  • Explain how the steelpan was invented and what that says about prohibition and creativity.
  • Distinguish collected village practice from staged revival in European and Pacific traditions.

The big picture

In 1884 the colonial government of Trinidad passed an ordinance restricting drumming and the torchlit Canboulay processions of carnival, following riots three years earlier. Skin drums went off the streets. What came back was bamboo: tamboo bamboo bands, playing tuned lengths of cut bamboo stamped against the ground and struck, in the pitch relationships the drums had used. Bamboo was then also restricted, being useful in a fight. So bands moved to metal: biscuit tins, paint cans, brake drums, dustbin lids.

Sometime around 1940, players in Port of Spain noticed that a dented metal container did not just clang, it produced pitches, and different dents produced different ones. Within a decade that observation had been developed into the steelpan: the head of a fifty-five gallon oil drum sunk into a bowl, the surface grooved into separate note areas, each hammered and heat-treated until it rings at a specific pitch. In 1951 the Trinidad All Steel Percussion Orchestra traveled to Britain and played it in front of European audiences.

Two prohibitions produced the only wholly new acoustic instrument family of the twentieth century. Keep that in view through this lesson, which follows what happened when African, Indigenous, and European musical systems were forced into contact across the Atlantic world, and then looks at Europe's own collected traditions and the Pacific.

Clave: the Cuban timeline

Lesson 5 gave you the West African bell. Here is its Caribbean descendant, written across sixteen pulses. Two versions, distinguished by one stroke:

Pulse12345678910111213141516
Son clave (3-2)xxxxx
Rumba clave (3-2)xxxxx

The pattern is asymmetrical: a side with three strokes and a side with two. Which side comes first matters, and musicians speak of a piece being in 3-2 or 2-3. Every other part in the ensemble is positioned relative to it, and a horn line written against the wrong direction of clave is described as being crossed, which is a mistake, not a style. As with the Ewe bell, the timeline does not move.

Son is the Cuban form where the Iberian and African streams meet visibly. It opens with a composed verse section derived from Spanish song, then shifts to the montuno, a call-and-response section over a repeating cycle where the improvisation lives. Its instruments come from both sides: the tres, a Cuban guitar with three doubled courses; bongo and later congas; a bass; and the claves and maracas holding the timeline. The son sextets and septets of 1920s Havana are the model for most of what follows.

Alongside it, rumba is percussion and voice only, originally on packing crates and later on congas, in styles including the slow yambu, the couple-dance guaguanco, and the acrobatic solo columbia. And in the Santeria religious tradition, three double-headed hourglass bata drums, iya, itotele, and okonkolo, play specific rhythms for specific Yoruba deities, which is Lesson 6's West Africa arriving intact rather than in translation.

Key idea: Clave is a timeline in the West African sense, and everything in an Afro-Cuban ensemble is placed relative to it, including its direction.

Salsa, samba, tango, cumbia

Salsa is what happened when this music was played by Puerto Rican and Cuban musicians in New York in the 1960s and 1970s. Fania Records, founded in 1964 by the flautist Johnny Pacheco and the lawyer Jerry Masucci, built the label, the stars, and the name. Celia Cruz, Willie Colon, and Hector Lavoe made records that were harder, faster, and more urban than their Havana models. The word salsa is a marketing term, applied to Cuban-derived forms by a New York industry, which is worth remembering when Lesson 14 takes up genre labels.

Samba came out of Afro-Brazilian communities in Rio, many with roots in Bahia. Pelo Telefone, registered in 1916, is generally cited as the first recorded samba. Within twenty years, a music that police had harassed as vagrancy was national symbolism under the Vargas government. The samba school parade is now an enormous competitive institution, and its bateria is a percussion ensemble with a division of labor as strict as the Ewe one: surdo bass drums in three interlocking parts, tamborim, caixa snare, agogo bells, the friction drum cuica, and the pandeiro tambourine. Bossa nova came later, from Joao Gilberto's guitar pattern on Chega de Saudade in 1958: the whole bateria compressed into one hand while the other played jazz-influenced harmony. Then Tropicalia, whose leading figures Caetano Veloso and Gilberto Gil were jailed and exiled by the military dictatorship in 1969.

Tango grew from the 1880s in the port neighborhoods of Buenos Aires and Montevideo among immigrants and Afro-Argentine communities. Its defining sound belongs to an instrument nobody designed for it: the bandoneon, a square free-reed instrument developed in Germany in the 1840s for religious and popular music, which arrived with migrants and turned out to do exactly what tango needed. Astor Piazzolla's nuevo tango, from the late 1950s, added dissonance and counterpoint and drew genuine hostility from traditionalists who felt he had made the music undanceable, which he largely had.

Cumbia comes from Colombia's Caribbean coast and wears its three sources openly: Indigenous gaita flutes, African drums including the tambor alegre and the timekeeping llamador, and Spanish song. Recorded in orchestrated form from the 1950s, it spread across Latin America and mutated everywhere it landed, into Peruvian chicha with electric guitars, Mexican cumbia sonidera, and Argentine cumbia villera. It is now probably the most widely played popular music in Latin America.

The Anglophone Caribbean

Calypso, older name kaiso, is Trinidadian topical song: sharp, funny, and political, sung in calypso tents in the carnival season, with a tradition of extempore competition in which singers improvise verses attacking each other. When the Empire Windrush docked at Tilbury in June 1948, a newsreel camera caught the calypsonian Lord Kitchener singing London Is the Place for Me on the dockside. He had composed it on the voyage.

In Jamaica, the chain runs mento, then ska in the early 1960s, then the slower rocksteady, then reggae, with the one drop pattern that leaves beat one empty and puts the weight on beat three, and the bass functioning as the lead melodic voice. Two Jamaican inventions changed global music. The sound system, a mobile disco with enormous speakers competing on exclusive one-off records, made the selector and the record more important than the band. And dub, developed in the early 1970s by King Tubby, an electronics repairman with a homemade mixing desk, and Lee Perry, took a finished mix apart, dropped instruments in and out, and drenched fragments in reverb and delay. The remix, as a form, starts here. Underneath it all sits Rastafari nyabinghi drumming, with bass, funde, and repeater drums, and its own religious world.

Europe: collected, revived, and staged

European traditional music reaches us mostly through collectors, which means it reaches us shaped by their purposes. Cecil Sharp collected in Somerset from 1903 and in the southern Appalachians from 1916, looking for survivals of an English past and quietly filtering out material that did not fit. Bela Bartok and Zoltan Kodaly went into Hungarian, Slovak, and Romanian villages with Edison cylinder machines from 1906 and produced transcriptions of extraordinary precision, notating ornaments and pitch deviations that a nationalist collector would have smoothed away. Bartok is a founding figure of comparative musicology as well as a major composer, and his fieldwork and his string quartets are the same project.

Traditions worth hearing, each doing something specific:

  • Sardinian cantu a tenore: four men in a circle, one singing the melody and three producing a chord beneath, two of them with a deliberately guttural, buzzing vocal production that sounds mechanical and is entirely human.
  • Georgian polyphony: three independent voices in non-tempered tuning, with harmonies that sound simultaneously ancient and unstable to ears expecting European chords. The song Chakrulo is on the Voyager Golden Record alongside the Chinese qin piece from Lesson 9.
  • Sami joik: a joik is not a song about a person, an animal, or a place; it is understood as a way of evoking or presenting that subject. Lutheran missionaries condemned it as sinful and it was suppressed for centuries, and it returned in the twentieth century as both cultural assertion and popular music.
  • Irish sean-nos: solo unaccompanied singing with dense ornamentation and flexible rhythm, judged on how a singer bends a line rather than on volume or beauty of tone.

And a case in the honest column. The recordings marketed internationally from 1986 as Le Mystere des Voix Bulgares are performances by a Bulgarian state radio choir singing arrangements written by trained composers, developed within an ensemble founded in 1951 to present village music in a concert form. The vocal technique is genuinely traditional; the arrangements, the choir, and the concept are the products of a state cultural policy. It won a Grammy in 1990 and was widely received in the West as the sound of untouched village Europe.

Key idea: Much of what circulates as European folk music passed through a collector or a state ensemble, and knowing whose purposes shaped it is part of hearing it accurately.

Oceania

In Hawaii, mele is chanted poetry and hula is its danced realization, in an older form accompanied by the ipu gourd and the pahu drum, and a modern form with guitar and ukulele. Missionary pressure from the 1820s drove hula out of public performance; King Kalakaua deliberately restored it at his coronation and jubilee in the 1880s, which is why he is called the Merrie Monarch and why the major hula competition carries that name.

Two Hawaiian instrumental inventions traveled further than almost any other music in this course. Slack-key guitar, ki hoalu, retunes the instrument so open strings sound a chord. And around 1889 a teenager named Joseph Kekuku began sliding a piece of metal along the strings of a guitar held flat, producing a continuous glide. Hawaiian steel guitar became a craze on the American mainland in the 1910s and 1920s and from there entered country music as the pedal steel, entered gospel as the sacred steel tradition, and entered Indian film and later Hindustani classical music, where it was rebuilt into new instruments. One Honolulu schoolboy's discovery is audible in Nashville, Chennai, and Chicago.

In Papua New Guinea, a country with over eight hundred languages, the Kaluli of the Bosavi region gave ethnomusicology one of its most useful borrowed concepts. Steven Feld's Sound and Sentiment (1982) describes dulugu ganalan, usually translated as lift-up-over-sounding: a texture in which voices, instruments, and the forest itself overlap in staggered, echoing entries so that nothing lands together and the sound is continuously layered. It is an indigenous theory of texture, and it describes what is happening better than any European term.

In Aotearoa New Zealand, Maori waiata and haka carry lineage, history, and challenge. Ka Mate, the haka known worldwide from rugby, was composed by the Ngati Toa chief Te Rauparaha around 1820. After long disputes about its commercial use by companies with no connection to the tribe, the New Zealand government reached a settlement with Ngati Toa recognized in legislation in 2014 that gives the tribe a right of attribution when Ka Mate is used commercially. Not a royalty; a right to be named. That is the shape of many outcomes in this area, and it leads directly into the last module.

Common misconceptions

  • "Clave is just a rhythm you can play anywhere in the bar." Its direction matters, and parts written against the wrong side are crossed.
  • "Steel pans were made because Trinidadians had no instruments." They were made because drums and then bamboo were restricted by law, and metal was what remained.
  • "Samba was always Brazil's national music." It was policed as vagrancy before being adopted as national symbolism in the 1930s.
  • "The bandoneon is an Argentine instrument." It was built in Germany in the 1840s and became central to tango after migration carried it across.
  • "Le Mystere des Voix Bulgares is untouched village singing." The technique is traditional; the choir and arrangements are products of state cultural policy from 1951.

Recap

  • Trinidad's drum prohibitions of the 1880s led through tamboo bamboo to the steelpan, developed from oil drums around 1940 and exported by 1951.
  • Son and rumba clave are asymmetrical timelines with a direction, and Afro-Cuban ensembles are organized around them.
  • Salsa is a New York marketing name for Cuban-derived forms; samba, tango, and cumbia each combine African, European, and in cumbia's case Indigenous components in traceable ways.
  • Jamaican sound systems and dub made the record and the mix into instruments and invented remix practice.
  • European traditional music reaches us through collectors such as Sharp and Bartok and sometimes through state ensembles, and knowing which matters.
  • Hawaiian steel guitar spread worldwide from around 1889; Kaluli lift-up-over-sounding is an indigenous theory of texture; Ka Mate produced an attribution right rather than a royalty.

Sources

  1. "Latin American music." Encyclopaedia Britannica. Regional genres and their African, Indigenous, and Iberian components. britannica.com.
  2. "Steelpan." Wikipedia. Drum prohibitions, tamboo bamboo, and the development of tuned pans. en.wikipedia.org.
  3. "Clave (rhythm)." Wikipedia. Son and rumba patterns, direction, and West African antecedents. en.wikipedia.org.
  4. "Ka Mate." Wikipedia. Composition by Te Rauparaha and the Ngati Toa attribution settlement. en.wikipedia.org.
  5. Smithsonian Folkways Recordings. Caribbean, Latin American, European, and Pacific collections. folkways.si.edu.
Key terms
Clave
The asymmetrical five-stroke Afro-Cuban timeline, in son and rumba versions, with a direction of 3-2 or 2-3.
Montuno
The call-and-response section of a son over a repeating cycle, where improvisation happens.
Bata
Three double-headed hourglass drums playing specific rhythms for specific Yoruba deities in Cuban practice.
Tamboo bamboo
Trinidadian bands using tuned bamboo lengths after skin drums were restricted by ordinance.
Steelpan
An instrument made from an oil drum head sunk and grooved into separately tuned note areas.
Bateria
The percussion section of a samba school, with interlocking surdo, tamborim, caixa, agogo, cuica, and pandeiro parts.
Dub
Jamaican studio practice of deconstructing a mix with drops, reverb, and delay; the origin of remix culture.
Joik
A Sami vocal form understood as evoking or presenting its subject rather than describing it.
Dulugu ganalan
Kaluli term, lift-up-over-sounding, for a staggered overlapping texture in which nothing lands together.

Module 6: Music in a Connected World

How the world music category was invented in a London pub in 1987 and what it did to the music it named, the appropriation cases that defined the argument, and then ownership, technology, migration, censorship, heritage listing, and a practical answer to the question of what to do next.

Selling the World: A Category, Three Cases, and the Argument

  • Explain how the world music category was created and what defining a genre by exclusion does.
  • State the strongest case on each side of the Graceland dispute using its own evidence.
  • Compare the Deep Forest and Enigma sampling cases and explain why the outcomes differed.
  • Apply a five-question framework to a borrowing case and identify power asymmetry as the operative variable.

The big picture

On 29 June 1987, roughly twenty-five people met in a room above a pub called the Empress of Russia in Islington, north London. They ran small independent record labels, booked concerts, and shared a specific commercial problem: shops had nowhere to file their releases. A record of Bulgarian choral singing or Zairean guitar bands went into a bin marked international, or ethnic, or nowhere, and customers never found it.

They voted on a name. World music beat several alternatives, including worldbeat and tropical. That autumn they ran a campaign: a compilation cassette, press advertising, and racking cards sent to record shops. The category worked immediately and commercially, and within five years it was a section in every large record store on earth, a Grammy award, a festival circuit, and a marketing description applied to musicians who had never heard of the meeting.

Notice what kind of category it is. World music is not defined by a sound, an instrument, a region, or a period. It is defined by what it is not: not Anglo-American pop, not the Western classical canon. A Malian kora player, a Bulgarian choir, an Indian classical vocalist, and a Colombian accordion band have almost nothing musically in common, and the one thing they share is where they are not from. This lesson works through what that framing did, and then through three cases in which the arguments about borrowing became concrete.

What a marketing bin does to the music inside it

The upside is not hypothetical. Musicians who had no route to European and North American audiences got distribution, festival fees, touring income, and press. Several careers in this course, including some named in Module 2, were built on it.

The costs were structural. A category defined by not being Western tends to reward whatever sounds most unlike Western pop, so an artist recording with drum machines and synthesizers in Abidjan or Lagos could find that Western buyers preferred an acoustic record made with older instruments, because that read as authentic. The reward went to the past-facing option. Steven Feld, who has written about this at length, points out that the discourse around the category oscillates between celebration and anxiety while rarely addressing the ordinary commercial question of who is being paid what.

The category also flattens. A listener who buys three world music albums has heard three unrelated musical systems and may reasonably think they have sampled a coherent thing. That is the exotic sampler problem this course has been built to avoid, which is why every earlier lesson gave you tunings, cycles, and structures rather than atmosphere.

Key idea: World music is a distribution category defined by exclusion, which brought real income and audiences to musicians while rewarding whatever sounded least modern and grouping unrelated systems as one shelf.

Case one: Graceland

In February 1985 Paul Simon recorded in Johannesburg with South African musicians, and the resulting album sold around sixteen million copies. The dispute that followed is the standard reference case, and it is worth setting out each position with its own evidence rather than as two symmetrical opinions.

The case against. The African National Congress had called for a cultural boycott of South Africa, and the United Nations maintained a register of entertainers who performed there. The point of a boycott is that it binds even sympathetic people, because selective compliance destroys it. Simon did not seek clearance from the ANC or the anti-apartheid movement before going, and the movement said so loudly at the time. Beyond the boycott, the money: session musicians were paid well, at roughly triple the American union rate, but session fees end. Songwriting credit and copyright, which generate income for decades, went overwhelmingly to Simon. Los Lobos, an American band, made a related complaint about one track, saying the music came out of their own jam session and they received no writing credit. The pattern is the same in both cases and it is about who owns the asset afterward.

The case for. The South African musicians involved were not consulted objects; they were professionals who chose to record and then defended the decision publicly and repeatedly for decades. Ladysmith Black Mambazo and members of the backing band said the album gave their music a global audience at a moment when the apartheid state was working to erase Black South African cultural achievement, and that the boycott, applied to them, would have meant enforced silence. Musicians were credited on the record, brought on the world tour, and paid at rates far above local norms. Exiled anti-apartheid musicians including Miriam Makeba and Hugh Masekela joined that tour. The United Nations reviewed Simon's case and removed him from its register in 1987.

Both accounts are accurate. The disagreement is not about facts but about which harm counts more: weakening a collective instrument aimed at a state, or silencing individual musicians in the name of it. Nothing in this lesson resolves that, and a reader who finishes it with a settled opinion has probably not held both sets of evidence at once.

Case two and three: two samples, two outcomes

In 1992 the French duo Deep Forest released an album built on samples of African and other field recordings, marketed with rainforest imagery and vague claims of association with international cultural bodies. Its best-known track uses a lullaby called Rorogwela, sung by a woman named Afunakwa of Malaita in the Solomon Islands, recorded in 1970 by the ethnomusicologist Hugo Zemp. Zemp published a detailed account in 1996 describing how the permissions had been sought from him and, in his telling, how the eventual use exceeded what was agreed. The album sold in the millions. Afunakwa, who died in 1998, received nothing. The framing was also false in a plain factual way: material presented as the music of Central African forest peoples included a Solomon Islands lullaby from the other side of the planet.

In 1993 the German project Enigma released Return to Innocence, whose central vocal hook is a sample of a 1988 recording of Difang Duana, also known as Kuo Ying-nan, and his wife Kuo Hsiu-chu, singing an Amis song in Taiwan. The recording had been made during a European cultural tour. The track was a worldwide hit and was used in promotional material for the 1996 Atlanta Olympics, which is how Difang, then in his seventies, came to hear his own voice on television. The Duanas sued in 1998 and the case was settled in 1999 with acknowledgment and compensation.

Compare them. The practice was the same: an archival recording of a named person, sampled into a commercial track without their knowledge. The outcomes diverged because the Duanas were alive, identifiable, physically able to encounter the product, and had access to a legal system that would hear them. That is not a moral difference between the two cases. It is a difference in whether the person had any leverage, which is precisely the variable that determines these outcomes.

Key idea: In sampling disputes the decisive factor is usually not the ethics of the borrower but whether the source is alive, named, and able to reach a court.

The strongest case for borrowing, stated properly

An argument that treats all cross-cultural borrowing as theft cannot survive contact with the material in this course, because nearly every tradition here is already a hybrid:

  • The violin entered South India around 1800 and is now a core Carnatic instrument, played seated with the scroll on the ankle.
  • The German bandoneon defines the sound of Argentine tango.
  • The European accordion is central to Colombian vallenato, Mexican norteno, and Louisiana zydeco.
  • The harmonium, a European reed organ, is indispensable to qawwali. All India Radio banned it from broadcasts from 1940 until 1971 on the grounds that it was foreign and could not produce the correct intonation. The ban did not remove the instrument from Indian music; it removed Indian music from the radio.
  • Cuban records sold in West and Central Africa produced Congolese rumba, as Lesson 6 described.

There is also the matter of musicians' agency. Toumani Diabate spent his career collaborating outward, including the album with Ali Farka Toure that won a Grammy in 1994. Ravi Shankar recorded with Yehudi Menuhin. Nusrat Fateh Ali Khan worked with European producers throughout his last decade. Treating such musicians as needing protection from collaboration is its own condescension, and freezing a tradition in the form an outsider finds most authentic is a way of denying living artists both income and the right to change.

But note that willing participants can still object to the terms. Shankar performed at Monterey in 1967 and Woodstock in 1969 to enormous audiences, and then spent years unhappy about what he had joined: audiences treating his music as an accessory to drug taking, applauding tuning as though it were a performance, and consuming him as exotic. He kept collaborating and kept objecting. Both things at once.

Five questions that do actual work

Rather than asking whether something is appropriation, ask these and see what the answers are:

  1. Consent. Did the source musicians know and agree, specifically, to this use?
  2. Credit. Are they named, correctly, where a listener will see it?
  3. Compensation. Were they paid, and do they share in continuing income or only a one-time fee?
  4. Control. Can they say no, later, to a use they object to?
  5. Context. Was the material restricted, ceremonial, or presented as something it is not?

Run Graceland, Deep Forest, and Enigma through those five and the differences become concrete rather than atmospheric. The operative variable across all of them is power asymmetry: who could have refused, and what it would have cost them. Most disputes filed under appropriation are, when examined, disputes about credit and money, which are tractable problems with known remedies.

Key idea: Consent, credit, compensation, control, and context turn an unfalsifiable argument about authenticity into specific questions with checkable answers.

Common misconceptions

  • "World music is a genre." It is a retail category invented in 1987 and defined by not being Anglo-American pop or Western classical.
  • "Musical borrowing is inherently exploitative." Nearly every tradition in this course incorporates borrowed instruments, and musicians actively seek collaboration.
  • "If the musicians were paid, the ethics are settled." A session fee and a share of copyright are different things, and the second is where the long-term money lives.
  • "The Enigma case shows the system works." It shows that a living, identifiable source with access to a court can win; Afunakwa had none of those.
  • "Willing collaborators cannot complain afterward." Ravi Shankar collaborated enthusiastically and objected publicly to how he was received, and both positions were coherent.

Recap

  • The world music category was created at a London meeting on 29 June 1987 to solve a record-shop filing problem, and it defines music by exclusion.
  • It delivered real income and audiences while rewarding whatever sounded least modern and presenting unrelated systems as one shelf.
  • Graceland is genuinely contested: a boycott weakened and credit concentrated on one side, musicians' own agency and global exposure on the other.
  • Deep Forest and Enigma involved the same practice with different outcomes, decided by whether the source could be identified and could reach a court.
  • Borrowing is the normal condition of music, from the Carnatic violin to the qawwali harmonium, and the useful analysis asks about consent, credit, compensation, control, and context.

Sources

  1. "World music." Wikipedia. The 1987 London meeting, the marketing campaign, and subsequent critiques. en.wikipedia.org.
  2. "Graceland (album)." Wikipedia. Recording in South Africa, the cultural boycott dispute, and the musicians' responses. en.wikipedia.org.
  3. "Return to Innocence." Wikipedia. The Amis sample, the 1996 Olympics use, and the 1999 settlement. en.wikipedia.org.
  4. "Deep Forest." Wikipedia. The Rorogwela sample and the permissions dispute. en.wikipedia.org.
  5. Society for Ethnomusicology. Position statements and professional guidance on ethics. ethnomusicology.org.
Key terms
World music
A retail and marketing category agreed at a London meeting in 1987, defined by not being Anglo-American pop or Western classical.
Authenticity premium
The market tendency to reward whatever sounds least modern, penalizing contemporary production by non-Western artists.
Cultural boycott
A collective refusal to perform in or with a state, whose logic requires that even sympathetic parties comply.
Session fee versus copyright
The distinction between one-time payment for performing and continuing ownership of the composition or recording.
Power asymmetry
The difference in who could have refused a use and at what cost; usually the decisive variable in borrowing disputes.
Hybridity
The normal condition of musical traditions, which routinely absorb foreign instruments and forms.
Five-question framework
Consent, credit, compensation, control, and context: checkable questions that replace arguments about authenticity.

Who Owns a Song? Rights, Technology, Power, and How to Keep Listening

  • Explain why copyright systematically mishandles collectively held and unfixed music.
  • Describe how recording formats and distribution technologies reshaped musical practice.
  • Give specific cases of music censored, weaponized, or used in political resistance.
  • Evaluate UNESCO intangible heritage listing on the evidence for and against, and build a personal listening practice.

The big picture

In 1939, at the Gallo studio in Johannesburg, a Zulu migrant worker named Solomon Linda recorded a song with his group the Evening Birds. He had improvised the main vocal line over three takes. The record, called Mbube, sold something like a hundred thousand copies across southern Africa over the following decade, and Linda signed the copyright over to the record company for roughly ten shillings.

In 1951 an American folk group recorded a version, having misheard the Zulu word as Wimoweh. In 1961 an American vocal group turned that into The Lion Sleeps Tonight, which went to number one. The song was recorded by dozens of artists, used in a Disney animated film in 1994 and its stage adaptation, and has been estimated to have earned in the region of fifteen million dollars. Solomon Linda died in 1962. He had, by the account of the journalist who reconstructed the story in 2000, about twenty-five dollars to his name, and his family could not afford a headstone. His heirs sued in 2004, using an old provision of British imperial copyright law that returned rights to a composer's estate decades after death, and settled in 2006.

Everything in this lesson is downstream of that story: what copyright does to music it was not built for, what technology does to circulation, what states do to musicians, and what any of us can do about listening.

Why copyright fits this music badly

Copyright is a specific legal machine with specific assumptions, and traditional music fails almost all of them:

Copyright assumesTraditional practice often hasResult
An identifiable individual authorRepertoire held collectively, developed over generationsNo one qualifies as the author, so nothing is protected
Fixation in a tangible formTransmission by ear and memoryUnfixed music is unprotected until an outsider records it
A limited term after which work is public domainPractices centuries oldThe underlying music is already public domain everywhere
Protection for new arrangementsOutsiders arranging or recordingThe arranger or label owns a fresh copyright in the version that circulates

Read the last two rows together and you have the mechanism. Because the traditional material is public domain and the arrangement is not, the legal system converts communally held heritage into privately owned property, and the owner is whoever wrote it down or pressed record. That is not a loophole; it is the machine working as designed on inputs it was not designed for.

Attempts to fix it have been slow. The World Intellectual Property Organization has run an intergovernmental committee on traditional knowledge and folklore since 2001 without producing a binding treaty. Some countries claim state ownership of national folklore, which protects against foreign exploitation and raises the immediate question of why a government should own a community's songs. The most practical current tools are extralegal: Traditional Knowledge labels that state community expectations regardless of copyright status, deposit agreements that bind archives, and attribution rights of the kind Ngati Toa obtained over Ka Mate.

Key idea: Copyright treats communal, unfixed repertoire as public domain and grants fresh ownership to whoever records or arranges it, which is why the remedies that work are contractual and community-authored rather than statutory.

What the technology did

Every format has reshaped the music that passed through it.

The 78 rpm disc gave roughly three minutes a side. Traditions that assumed an hour had to choose what fit: Indian classical artists recorded compressed sketches of ragas, and Arab singers cut down songs built on repetition. Generations of listeners learned those short forms as the pieces themselves.

The cassette broke the record industry's gatekeeping. Peter Manuel's study of India in the 1980s documents what happened when duplication became cheap: the near-monopoly of one large label gave way to thousands of small operations recording regional languages, devotional genres, and communities the national industry had never served. The technology that made piracy easy also made a vastly wider range of music commercially viable. Cassettes carried political speech as well, moving sermons and songs through networks that states could not easily police.

Radio made states into patrons and censors simultaneously. National broadcasters commissioned ensembles, salaried musicians, standardized repertoire to fit programming slots, and decided what would not be aired.

Streaming has given listeners something unprecedented: nearly everything, everywhere, immediately, which is what makes this course's listening assignments possible at all. It has also concentrated income, pays fractions of a cent per play, recommends toward what is already popular, and uses metadata systems that struggle with collective authorship, non-Latin scripts, and performers who are not the composer. Meanwhile the largest informal archive of live traditional performance in existence is now video shot at weddings and festivals by participants and uploaded without any rights framework at all.

Music that moves with people

Diaspora musics are not thinner versions of what was left behind. They are new forms answering new conditions. British bhangra in the 1980s took Punjabi drumming and song, put it over drum machines, and was consumed by South Asian teenagers in Birmingham and west London at daytime events held because evening clubs were not an option for many of them. Algerian rai moved with migration to Marseille and Paris and became a mainstream French pop form, while remaining dangerous at home: the singer Cheb Hasni was shot dead in Oran in 1994 during Algeria's civil conflict. In each case the music indexes a specific situation of arrival, and studying it as a degraded copy of a homeland tradition misses everything interesting about it.

Power, in cases

  • Fela Kuti, Nigeria. His compound outside Lagos, declared an independent republic, was raided by roughly a thousand soldiers in February 1977 after his song Zombie mocked the army. Buildings were burned, Fela was beaten, and his mother, the veteran activist Funmilayo Ransome-Kuti, was thrown from a window and died of her injuries in 1978.
  • Victor Jara, Chile. Detained after the coup of 11 September 1973, held in a Santiago stadium, tortured and shot. In 2003 the stadium was renamed after him.
  • Miriam Makeba, South Africa. After she testified about apartheid at the United Nations in 1963, her citizenship and right of return were withdrawn. She lived in exile for about three decades.
  • Afghanistan. The Taliban banned most music from 1996 to 2001, destroying instruments and recordings, and imposed restrictions again after 2021, driving musicians and the national music institute into exile.
  • Northern Mali, 2012. Armed groups controlling the north banned music outright in a region whose musicians are internationally known. The Festival in the Desert went into exile.
  • Music as an instrument of harm. After reports that recordings were played at extreme volume and duration against detainees in United States custody, the Society for Ethnomusicology adopted a formal position statement in 2007 condemning the use of music as torture.

Key idea: States and armed movements attack music because it organizes people, and the same properties that make a song useful to a community make it worth banning.

Heritage listing: the case for and against

UNESCO adopted the Convention for the Safeguarding of the Intangible Cultural Heritage in 2003, and it came into force in 2006. Practices are nominated by states and inscribed on a Representative List or, where a tradition is genuinely endangered, an Urgent Safeguarding List. Many traditions in this course are inscribed: pansori, the guqin, the Mevlevi sema, the Persian radif, Georgian polyphony, Sardinian cantu a tenore.

What listing does well. It brings money, and money buys teaching. Safeguarding plans have funded apprenticeship programs, instrument-making workshops, and school curricula. It gives communities a recognized status they can use in arguments with their own governments about land, funding, and language policy. And it has demonstrably raised the number of young people entering some traditions.

What critics say, at strength. Only states can nominate, so listings reflect state narratives, and a minority practice a government finds inconvenient will not be nominated by that government. Inscription tends to fix a canonical version of a living practice, because a nomination file has to describe what is being protected, and what is described becomes what is funded. It creates a competitive hierarchy between nations, and neighboring states have made rival claims over shared traditions. Listing accelerates tourism, which converts a ceremony into a scheduled performance for people who did not come for the ceremony. And the requirement to identify bearers imposes a structure of designated masters on traditions that had no such category.

The Mevlevi ayin from Lesson 11 shows both edges at once: a ceremony banned by a state in 1925, revived as a tourist performance, and inscribed as heritage in 2008. That trajectory saved the music and changed what it is.

Applied ethnomusicology is the field's name for practical work in this space: repatriating recordings, helping communities build their own archives, supporting teaching programs, advising on nominations, testifying in land and cultural rights cases, and working with refugee and displaced communities. It is where the ethics of Lesson 3 become a job description.

How to keep listening

The last thing this course owes you is a way to continue without it. Here is a plan you can actually run.

  1. Use the free archives properly. Smithsonian Folkways has kept its catalog permanently available since the Smithsonian acquired it in 1987, and its liner notes are scholarly documents. The Library of Congress American Folklife Center holds the cylinder collections and field recordings from a century of work. Read the notes before and after listening.
  2. Listen three times. Once for the whole, once for one layer, once for what changed. Use the six-question protocol from Lesson 4.
  3. Learn one thing with your body. The twelve-pulse bell. Son clave. The sixteen-beat teental cycle with the wave at nine. Physical knowledge of one timeline changes how you hear a hundred recordings.
  4. Follow musicians, not genres. Genres are retail categories, as Lesson 14 showed. Individual performers have teachers, lineages, and collaborators you can follow outward.
  5. Go to where music is happening within an hour of you. A gurdwara on a Sunday, a Greek or Ethiopian church festival, a powwow open to the public, a Chinese cultural center concert, a Balkan dance night, a mosque during Ramadan. Almost every city has more than you think. Attend, pay, sit through it, and do not record without asking.
  6. Ask the five questions from Lesson 14 about anything you buy or stream.
  7. Take a lesson. One hour with a tabla, kora, or gamelan teacher will teach you more about the relevant lesson in this course than rereading it.

One last honest word. Every description in this course, including the careful ones about mbira buzz and neutral thirds and the pause between gong strokes, is a set of instructions for hearing something you have not heard. The course is a map. It has been accurate about scale and legend and it is not the ground. Go find the ground.

Common misconceptions

  • "Traditional music is protected as cultural property." In most jurisdictions it is public domain, and the recording or arrangement made by an outsider is what carries protection.
  • "Streaming solved access, so the problems are over." Access improved enormously while payment, metadata, and recommendation systems remain badly matched to this music.
  • "UNESCO listing is straightforwardly good." It funds transmission and it also freezes practice, follows state priorities, and accelerates tourism.
  • "Censorship of music is historical." Bans and killings named here run from 1960 to the present decade.
  • "Diaspora music is a diluted version of the original." It answers different conditions and is its own thing, which is why bhangra sounds like Birmingham in the 1980s.

Recap

  • Solomon Linda's Mbube became one of the most lucrative songs of the century while its composer died with nothing; his heirs settled only in 2006.
  • Copyright's assumptions of individual authorship, fixation, and limited term convert communal repertoire into private property owned by the recorder or arranger.
  • Formats shape music: three-minute sides compressed long forms, cassettes broke industry gatekeeping, radio made states patrons and censors, and streaming gave access while concentrating income.
  • Musicians have been exiled, banned, and killed for their work in living memory, and music has been used as an instrument of abuse, which the field formally condemned in 2007.
  • UNESCO listing funds transmission and simultaneously fixes practice, reflects state priorities, and drives tourism.
  • A workable listening practice: free archives, three listens, one timeline learned physically, follow musicians, go to live music near you, and ask the five questions.

Sources

  1. "The Lion Sleeps Tonight." Wikipedia. Solomon Linda's Mbube, the subsequent versions, and the 2006 settlement. en.wikipedia.org.
  2. World Intellectual Property Organization. Traditional knowledge, traditional cultural expressions, and the intergovernmental committee. wipo.int.
  3. UNESCO. Convention for the Safeguarding of the Intangible Cultural Heritage (2003). ich.unesco.org.
  4. Smithsonian Folkways Recordings. Permanent catalog, liner notes, and educational resources. folkways.si.edu.
  5. American Folklife Center. Field collections and archival access. Library of Congress. loc.gov.
Key terms
Fixation
The copyright requirement that a work be recorded in tangible form, which unfixed oral traditions do not meet.
Arranger's copyright
The fresh protection granted to a new arrangement of public-domain material, owned by the arranger.
Traditional cultural expressions
The term used at WIPO for communally held cultural material that standard copyright does not fit.
Cassette culture
The decentralization of music production and distribution that cheap duplication produced, especially in South Asia.
Intangible cultural heritage
Practices, knowledge, and performance protected under the UNESCO convention adopted in 2003.
Heritagization
The conversion of a living practice into a presented, scheduled version of itself under heritage status.
Applied ethnomusicology
Practical work including repatriation, community archiving, teaching support, and expert testimony.
Safeguarding plan
The transmission and training program a state commits to when a practice is inscribed on a UNESCO list.

Open the interactive version with quizzes and progress →