Self-study course · working competence · 10 weeks · part 1 of 2
Morse is a spoken language that happens to have a written notation. This course teaches the listening. It never shows you the chart.
This bar is the word LISTEN, drawn to scale in real Morse timing. Tap it to hear where you are.
Copy an unfamiliar plain-text transmission onto paper — the full 26-letter alphabet, sent at 12 WPM character speed, at 90% character accuracy, cold, with no chart and no replay. And run the drill regimen that carries you onward without this course.
No exposure to Morse whatsoever. Comfortable with structured self-study and technical reading. No radio, no key, no licence, no musical training assumed. Headphones and a pencil are the entire equipment list.
Roughly 24–26 hours across ten weeks, published per block so you can decide one sitting at a time. That figure already has the standard 2.5× correction applied — self-study consistently runs long. Sourced fact S6
Week 6 is a real stopping point. It crosses all three thresholds and leaves you copying sixteen letters — a genuine skill. Weeks 7–8 close the alphabet, Week 9 adds light and is self-contained. Stopping at any of those boundaries leaves something whole rather than something half-crossed. Inference
Morse instruction has an unusual evidence problem: the primary research is a 1935 German engineering dissertation and a handful of pre-1960 psychology papers, while nearly all modern writing about it is published by people selling an app. Every load-bearing claim below carries a marker so you can tell which is which.
These persist. Leave character speed at 12 unless a block tells you otherwise — that number is the whole argument of Week 2.
Everything in this course runs on any of four channels. Pick one here and every exercise, every drill and every self-check uses it. Two of them are completely silent, so a meeting or a five-minute gap is practisable.
Which channel should you learn on first?
The common answer — "I learn better visually" — is the one part worth examining. Pashler and colleagues reviewed the evidence for matching instruction to a learner's preferred modality and found the studies with the right design either failed to show the effect or contradicted it. Preferences are real; the benefit from meshing instruction to them is not established. Sourced fact S9
But that research does not settle this case, because it is about presenting the same content in different formats. Here the format is the content. Morse by ear and Morse by light are not two presentations of one skill — they are two skills with different targets. So the real question is not which suits you; it is which one you actually want. Inference
Why vibration is the good silent option. The flash-rate limit is a photosensitivity constraint and applies only to light. Vibration has no such ceiling, so it runs at the full 12 WPM character speed — above Koch's critical rate, unlike the lamp. It is the only silent channel that does not force you below the speed where the pattern holds together. Inference
And sending is silent already. Any key exercise works with the sound off: tap, and the analyser still measures your timing. That is real practice you can do under a table.
The whole task, in its simplest form, on day one. You will not "learn the alphabet and then start copying." You start copying now with two letters and add letters to a thing you can already do.
Get a sheet of paper and a pencil. Press play. You will not be told what letters these are, and you are not trying to identify them. Every time you hear one complete sound-shape begin and end, make a single dot on the paper. That is the entire exercise.
This is Koch's own first exercise, from the 1935 dissertation, and almost no modern course includes it. His reasoning: marking rhythm on paper builds the link between the heard shape and the writing hand before any letter meaning is attached, and it forces you to listen for the shape rather than the parts. Sourced fact S1
Count your dots afterward. If you got 20, your ear is already segmenting the stream, which is most of what this block is for.
Two letters: K M
Play each one. Say the letter out loud as you hear it. Do this until the letter arrives in your head before you have finished thinking about it — which will feel like nothing happening, and then suddenly like the sound simply is the letter.
K and M are Koch's own suggested starting pair, chosen because their sound-shapes are maximally unlike each other. He noted that any well-differentiated pair works. Sourced fact S1
Before you copy cold, watch it done. Press play and follow along with the answer visible. Your job is only to bind sound to letter while the pressure is off.
This is a fully worked example. Week 3 gives you partial ones. Week 6 gives you none — the support is scheduled to disappear, and you should expect to notice it going.
The real thing. Paper and pencil, not the keyboard. Press send, write the letters as they arrive, then type what you wrote into the box. Do not replay before committing.
If you drop a character, leave a gap and keep going. Chasing a missed letter costs you the next three — this is the single most common beginner failure and you should practise the recovery now rather than at speed later. Inference
Three sentences on paper, for you, not for anyone else:
Prompted self-explanation is where most of the gain in worked-example learning actually comes from, and learners left to themselves explain superficially or not at all. Writing it beats thinking it. Sourced fact S7
Morse timing is built from one unit: the length of a dit. Everything else is a multiple.
| Element | Units |
|---|---|
| dit | 1 |
| dah | 3 |
| gap between elements inside a letter | 1 |
| gap between letters | 3 |
| gap between words | 7 |
Speed is defined against one reference word, PARIS, chosen because it comes to exactly 50 units including its trailing word gap. Count it yourself before reading on — the diagram below is drawn to scale, and every bar is one element.
Tap the diagram to hear it. Sourced fact S4
So one word per minute means fifty units per minute, and:
dit = 60 / (50 × WPM) seconds
= 1200 / WPM milliseconds
Work these out on paper before committing:
1. How long is a dit at 12 WPM? And at 5 WPM?
2. Of PARIS's 50 units, how many are inside characters, and how many are the gaps between characters and words?
Here is the idea the course turns on. You can stretch those 19 spacing units without touching the 31 character units. The letters keep their exact rhythm; the silence between them grows. That gives two independent speeds: a character speed and an effective speed.
This is Farnsworth timing, named for Donald R. "Russ" Farnsworth, W6TTB, and given a formal specification by the ARRL so that "18 WPM characters at 5 WPM overall" means one reproducible thing. Sourced fact S4
Hear the difference. Same six letters, three ways:
The third button is the important one. Listen to it twice. A letter sent at 5 WPM is not the same sound played slowly — it is a different sound, one you would have to learn separately and then unlearn.
Koch measured exactly where this breaks. Below roughly 10 WPM, he found, the character stops being perceived as a single acoustic whole and starts being perceived as a sequence of separate beeps you have to count and reassemble. He called the intact version the Gestalt, and named the changeover the critical rate. His conclusion was blunt: training speed must be above that rate from the very beginning. Sourced fact S1
A correction you will need. Nearly every app and article says the Koch method means starting at 20 WPM. Vendor framing
Koch's dissertation says something different. He identified ~10 WPM as the critical rate below which the shape decays, and found 12 WPM optimal for initial learning, with speed increased toward 20 only after all characters were known. The 20 WPM figure is the destination in his report, not the starting line. This course uses 12 because that is what the primary source actually says. Sourced fact S1
3. A friend says: "I'm learning at 5 words per minute and I'll speed up later." Two separate things are wrong with the plan. Name both.
Three new: R A N — working set K M R A N
Hear each alone first, then drill the set. Move on when you are at 90% on the drill below; that threshold is Koch's, and it is the rule for every week from here.
R and K are mirror images of each other written down, and so are A and N. Koch specifically noted that visual mirror-images cause almost no confusion by ear, because their sound-shapes are plainly different — which is a small proof that the written form and the heard form are not the same object. Sourced fact S1
Set up the practice pattern now, because it does more work than anything else in this course.
Two sessions a day, half an hour each, morning and afternoon. Not one long session. Koch tested this directly and reported that distributing the practice hours produced faster learning and better retention than massing them — and that half an hour was the useful ceiling for one sitting. Sourced fact S1
This is also the single most robust finding in the modern study-strategies literature, independently of Morse: distributed practice and retrieval practice are the two techniques that survive meta-analysis. You are getting both. Sourced fact S6
These reach back to Weeks 1 and 2 deliberately. A self-check sitting under the section that taught it measures whether you just read something. Spacing effects show up on delayed tests and not immediate ones, so the useful version of this is the annoying version. Sourced fact S6
1. Without scrolling up: how long is a dit at 12 WPM, and what fraction of PARIS's 50 units can Farnsworth timing stretch?
2. Cold ear check, no warm-up. Four characters from Weeks 1–2. Write them down before revealing.
There are two different things happening in your head when you copy, and you need to be able to tell them apart.
| Decoding | Recognition |
|---|---|
| You hear the parts, hold them, count them, and assemble a letter. | The letter arrives. There is no intermediate object. |
| Scales badly. Fails around 10–12 WPM and gets worse. | Scales. Speed becomes a matter of practice, not method. |
| You can feel yourself working. | Feels like you did nothing, which is why people distrust it. |
Koch's criticism of every method that came before him was aimed exactly here. Teaching from a printed chart, he argued, requires the operator to count the elements and reassemble them for comparison against a memorised optical symbol — workable at low tempo, impossible at speed. He recommended charts never be used at all. Sourced fact S1
The test. Play a single letter. Say it out loud immediately — within about half a second, before you have time to think. If you can do that reliably, it is recognition. If you need a beat, it is decoding, and that letter needs more drill rather than more speed. Inference
Scaffolding step two. Part of the answer is given; you fill the gaps. Play, copy what you can, then type the full string.
Answer: MARK RANK KARM. Compare against your paper, not your memory of how it went.
Three new: U D W — working set K M R A N U D W
A note on ordering. The character sequence used by nearly every Koch trainer — K M R S U A P T L O and so on — does not appear in Koch's dissertation. It comes from a much later software implementation and is attributed to him by convention. Koch's own sequence was different, and his stated principles argue against parts of the popular one: he explicitly warned that confusable pairs like S/H and U/V should be kept apart, and the standard sequence introduces S immediately. Sourced fact S1 Contested
This course therefore uses a sequence I built from Koch's stated principles rather than either his list or the popular one: well-differentiated shapes first, confusable families deferred, and the runs of pure dits and pure dahs held back until there is enough context to contrast them. That is my choice, not his, and you should treat it as such. Inference
In 1897 and 1899, two psychologists at Indiana University followed telegraph students for months and published the first careful learning curves anyone had produced for a complex skill. They found that receiving improved, then stopped, then improved again — and their explanation was that the operator was building a hierarchy of habits: first letters, then words, then whole phrases, with each level having to become automatic before the next could form. The flat stretch was the lower level consolidating. Sourced fact S2
Their own words are worth having: mastery of the telegraphic language, they concluded, involves mastery of the habits of all orders. A student who jumps from 18 to 25 words a minute has not suddenly learned more English — something else has automated. Sourced fact S2
Then, sixty years later, someone argued it was never there.
Fred Keller — who had run code training research for the US Army — published a paper in 1958 titled "The Phantom Plateau," arguing that the plateau in code learning is an artifact of how the curve is drawn and how the students were taught, not a necessary feature of skill acquisition. Under good instruction, he said, it largely disappears. Sourced fact S3 Contested
And the argument did not end there: later authors defended the plateau on different theoretical grounds, and a 2017 review in Current Directions in Psychological Science opens by observing that the field has been arguing about whether training plateaus are real for a hundred and twenty years. Sourced fact S5
What to actually do with a contested finding.
Not "pick a side." Notice which part is disputed and which part is not. Nobody disputes that learners experience flat stretches. What is disputed is whether the flatness is intrinsic to the skill or produced by the teaching. Both sides agree on the practical consequence, which is the only part you need: a flat stretch is not evidence that you have hit your limit, and it is not evidence that the method is failing. Keller's version is arguably more encouraging than Bryan and Harter's, because it says the flat part is at least partly removable. Inference
1. Bryan and Harter watched real students and saw a plateau. Keller looked at similar data and said it was not really there. Before revealing: what kind of thing could they disagree about, given they were not disputing each other's measurements?
Write this down on paper now, before you need it. When the drill goes flat for a week or more, which of these will you do?
There is no consensus answer, and anyone who gives you one confidently is overselling. The value of choosing now is that you choose while you are not frustrated. Inference
Three new: G F Y — working set K M R A N U D W G F Y
You now have enough letters for real words. Try this one as a word rather than as five separate letters — it is the first hint of what Week 5 is about.
These are deliberately mixed. Each describes a symptom; your job is to pick which intervention applies and say why. Executing a named technique is not the skill — choosing among them is. Sourced fact S6
1. You copy individual letters at 90%, but as soon as the text is real words you fall behind and lose whole chunks. Character speed, spacing, or something else?
2. You confuse two specific letters constantly, and only those two. Everything else is solid.
3. Everything was fine, and now you are missing the character immediately after a long one. Consistently.
Three new: T E O — working set K M R A N U D W G F Y T E O
These are the short ones, held back until now on purpose. E is a single dit and T a single dah, and learned early they tend to become a default guess for anything unclear. With eleven solid letters around them they land as themselves. Inference
Real words now. Copy onto paper. Deliberately try to let a word land whole rather than assembling it — and expect that to work perhaps once, briefly, and then not again for a while.
When a word does arrive whole, that is the hierarchy Bryan and Harter described, forming in real time. It is worth noticing precisely because it feels like less effort rather than more, and effortlessness reads as "not learning" when it is the opposite. Sourced fact S2
This is a natural stopping point. All three thresholds are crossed and you have a working skill. It is also the prerequisite for Sending Morse, which can be started from here in parallel with Weeks 7–10.
Two new: I S — final set K M R A N U D W G F Y T E O I S
S is the one Koch warned about — three dits, easily confused with H's four. H is not in this course, which is part of why S can safely come in now. When you add H later, drill S and H against each other directly from the start. Sourced fact S1
1. Back to Week 2, no scrolling. Why is a character sent at 5 WPM not simply "the same character, slower"?
2. Your sending is being copied badly. Your dah/dit ratio measures 2.6 instead of 3.0, and your inter-character gaps measure 1.9 units instead of 3.0. Which do you fix first, and why?
Three new: H V B — working set K M R A N U D W G F Y T E O I S H V B
Each is introduced against the letter it will be confused with. Koch warned specifically that sound-alike pairs cause trouble; his answer was to keep them apart in the sequence. Mine is to keep them apart until now and then deliberately jam them together, because the discrimination has to be built at some point and doing it under supervision beats discovering it later. Sourced fact S1 Inference
| New | Confused with | Drill |
|---|---|---|
| H | S — one extra dit | |
| V | U — one extra dit before the dah | |
| B | D — one extra dit at the end |
All three are the same mistake in different clothes: a trailing dit you did or did not hear. If you are missing them, the fix is not more speed and not wider gaps — it is counting nothing and instead learning the two shapes as different words. Inference
Three new: C L P — working set now 22 letters.
| New | Confused with | Drill |
|---|---|---|
| C | K — a trailing dit | |
| L | F, and R — same length, different middle | |
| P | W, and F — the dahs sit in different places |
Three new: J Q X — twenty-five letters.
These are the four-element characters with heavy dahs. They are long enough that impatience is the main enemy: the shape is not finished when you think it is. Inference
| New | Confused with | Drill |
|---|---|---|
| J | W — one more dah | |
| Q | G, Z — the family that starts with two dahs | |
| X | B, D — dah at both ends |
One new: Z — and the set is closed. A–Z
Now full English. This is the first time the drill has not been drawn from a restricted set, which means it is also the first time letter frequency works for you: E, T, A, O, I and N carry most of the text, and you have owned those for weeks.
The code is channel-independent. The timing structure is the message, and it survives translation to anything that can be switched on and off — a tone, a lamp, a torch, a shutter. That is genuinely the same skill.
But the constraints are not the same, and one of them is a safety limit rather than a preference.
WCAG 2.3.1 sets the general flash threshold at three flashes per second; above that, flashing content can trigger seizures in people with photosensitive seizure disorders, which affect roughly one person in four thousand. Sourced fact S8 A run of dits is a square wave, so the worst-case flash rate is one flash per two dit-lengths:
| Character speed | Dit | Worst-case flashes/sec | |
|---|---|---|---|
| 6 WPM | 200 ms | 2.50 | within the limit |
| 8 WPM | 150 ms | 3.33 | over |
| 12 WPM | 100 ms | 5.00 | over |
| 20 WPM | 60 ms | 8.33 | over |
So visual practice is capped at 6 WPM character speed in this course, and the lamp reports its own flash rate so you can check the arithmetic rather than trust me. Inference
Which collides with the whole first half of this course.
Koch's critical rate is 10 WPM: below it, the acoustic shape decays. The safe visual rate is 6 WPM. Both constraints cannot be satisfied at once. I am not going to paper over that. My reading is that Koch's finding is about auditory pattern perception specifically, and the eye integrates a flashing light differently from how the ear integrates a tone — which is consistent with signal-lamp work historically running far slower than radio telegraphy. But I have no source that tests the visual case directly, so treat this as my reasoning and not an established result. Inference Contested
Copy from light instead of sound. The lamp below is deliberately small — a small flashing area sits under the threshold's size exemption, which a full screen does not. Sourced fact S8
Expect this to be much harder than the ear at first, and expect it to come back faster than it did the first time. You are not learning the code again; you are attaching a channel to something you already know.
No worked example, no completion, no partial answer. Paper, pencil, one pass, no replay. Full alphabet.
Notice this looks nothing like Week 1, which handed you the answer before you started. That was scheduled: support that helps a beginner actively obstructs an improver, so it was designed to withdraw. Sourced fact S7
One page, your own words, your own numbers:
If you can write that page without consulting this one, the course is finished. Everything left is more of the same, and you know how to add it.
Last one. An app advertises "the Koch method — scientifically proven, start at 20 WPM, learn the code in two weeks." Three things should bother you. What are they?
There is no dot-dash chart in this course. That is the one design decision everything else rests on, and it is Koch's recommendation rather than mine: he argued charts should never be used, because reading one trains you to count elements and reassemble them, which is the habit that caps your speed. Sourced fact S1
Locking it away entirely would be dishonest, though — you can find it in four seconds and a course that pretends otherwise is just posturing. So here is the honest version of the trade: looking at the chart will feel like it helps and will cost you weeks of unlearning later. If you are stuck on one letter, the cheaper fix is almost always contrast drill against a letter you already own. Inference
Recorded, not recommended. Close it when you are done.
Every load-bearing claim is keyed to one of these. Each was opened and read rather than cited from a summary — with one exception, flagged below, where the primary source is in German and I worked from a translation.
What I could not verify. Koch's reported "class to 12 WPM in fourteen hours" is widely repeated but I have seen no subject counts, control condition, or independent replication — and a 1943 review found his two-tone variant produced a difference that probably was not significant. Treat the headline result as a favourable single case. Single study
The weakest link. Week 9 assumes that learning by ear transfers to reading light. The timing structure is identical, but I found no study testing the transfer, and the safe visual flash rate sits below Koch's critical rate — a genuine unresolved tension. If one claim here is wrong, it is that one. Inference Contested