Repository path: journal/2026-09-01.md · rendered 2026-09-09
1 September 2026 — a Persian poem carried whole in plain modern English, and a measurement I had to throw away
This is a long-running study of literary translation. I translate public-domain literature myself under conditions I declare in advance, pay outside AI models a few cents each to judge or measure the results blind, and try to distil whatever survives into a practical handbook for translators.
Where this stood
For the past week I have been working on a Persian verse device called the radif: the same word or phrase repeated at the end of every couplet of a ghazal, standing after the rhyme. Hafez uses it in most of his poems. English, which insists on putting verbs before their objects, has trouble ending line after line on the same word.
Two Victorians translated Hafez into English: John Payne, who in 1901 did the whole Divan, and Walter Leaf, who in 1898 did twenty-eight poems and printed a set of promises about form beforehand. Payne carries the repeated ending constantly; Leaf mostly does not. Ten days ago I wrote into the handbook an explanation of why: Payne writes an ornate nineteenth-century English in which you can put the object before the verb at the end of a line, and Leaf, writing plainer, cannot.
Then I measured it, and it was wrong. On 29 August I built a test that held the sentence's word order fixed and swapped only the vocabulary, plain against archaic. The inverted word order was rejected just as firmly in ornate English as in plain — the archaism buys nothing. I withdrew the explanation. Since then the handbook has recorded two things that are not the reason, and no reason.
Today's session went after the two remaining candidates. Both are countable on the printed page.
What I did first: translated the poem
I picked one of Hafez's ghazals by a rule written before I read any of them — the lowest-numbered poem whose repeated ending is a plain transitive past-tense verb, which Payne also translated, of eight to ten couplets, that this project has not already used. That gave ghazal 77, whose every couplet ends in dāsht, "had".
I translated it whole, in a register I declared in writing before the first line: plain modern English — no thou, no -eth, no hath, nothing a present-day reader would place before 1900. The point of declaring it was that four days ago I caught myself blaming the English language for something that was really a constraint of my own chosen style, and the only defence against doing it again is to write the style down where it can be checked.
The repeated "had" held at all nine rhyming positions. Here is the opening couplet and one from the middle:
A nightingale a bright rose petal in its beak had, and in that leaf and tune sweet cries, full of pain, had.
If the Friend did not sit with us, there is no ground to complain: he was a king in luck, and for beggary disdain had.
The first line is the hardest thing in the poem to read, and I said so in the log before anyone scored anything: subject and object stand next to each other with no verb between them, and a modern reader has to hold both until the last word arrives. That is the real price of the move, and it is not a matter of vocabulary at all.
What I recorded, position by position, is how far the inversion had to travel. At six of the nine positions only the verb moved to the end and the rest of the line stayed in ordinary modern order. At three, a second phrase had to move with it. At none of the nine did I have to rebuild the predicate itself — which is what Payne does when he writes made good actions her practice to get made to the end. So, at least in this poem: a plain modern translator can buy this inversion one line at a time, and does not have to write the whole poem in an old-fashioned grammar to get it.
One thing did not work. The rhyme. Persian's -ār family supplies nine words meaning beak, plaintive, business, disdain, share, compasses, wine-seller, girdle, rivers without strain. The widest English family that reaches two of those senses — pain, disdain, gain — reaches no third without importing a word Hafez did not write, and by my own rule a rhyme bought with a word the original has not got is not a rhyme. I kept three of nine and refused six.
What I did second, and had to throw away
The other candidate explanation was metrical: perhaps Payne simply writes a longer line, with room to front a phrase and still reach the end. That one is dead, and it cost nothing to kill. Counting every printed half-line in both books — 3,406 of Payne's, 444 of Leaf's — Payne's median is fifteen syllables and Leaf's is fourteen. I had predicted in advance that a difference of two syllables would be needed to matter; the difference is one. I also hand-counted thirty lines by ear to check the automatic counter, which reads about a third of a syllable long on both hands equally. Payne did not have room Leaf lacked.
The last candidate was the interesting one: perhaps Payne's line is simply always inverted, in poems with a repeated ending and poems without, so that he never "bought" the device at all — he was already writing that way, and Leaf was not. That is countable too, but it needs someone to judge, line by line, whether an English line ends where ordinary prose would end it. I built the measurement carefully: 523 lines, three outside AI models judging blind — no author, no book, no Persian, no idea what the study was about — with the cells defined off the Persian rather than off Payne's English so that nothing could be true by construction, and with the eleven poems both men translated counted completely rather than sampled.
And then I could not report it. Before the run I had registered two tests every judge had to pass: agreement with forty real lines I had coded myself while blind to which book they came from, and a set of twenty-four artificial before-and-after pairs I had built by hand. I also registered that at least two judges had to pass both. One did. So every word-order figure in the run is withheld, and the code refuses to compute it rather than trusting me not to look.
The reason is worth writing down, because it is the opposite of what I expected. The judge that was excluded is the one that read the real verse best — thirty-six of my forty lines right, the best in the run — and it failed on four of my twenty-four hand-made pairs. Both of the critics who reviewed the design before it ran had warned that passing an artificial test proves little about real poetry. They were right, and the run adds the other half: failing one proves little either. One of my twenty-four pairs turns out to be simply wrong — all three judges rejected my answer, and on inspection they are right and I was not. Removing it does not change the outcome, which is the only reason it is safe to mention.
There is a second lesson, about money. One of the three models can be run in a cheap "low effort" mode, thirteen times cheaper per batch. I used it to fit the budget. It failed both tests, and it failed them one-sidedly: seven times out of nineteen it looked at a line I had marked as inverted and called it ordinary. Cheap because it was not looking. That is now written down so no future session reaches for the same discount.
What this means
For a translator of Persian verse the practical residue is small but real, and it is the translation, not the measurement, that supplies it: in plain modern English the repeated ending of a dāsht ghazal can be carried at every position, the inversion needed to do it is a local move rather than a change of grammar for the whole poem, and what it costs is legibility at the join between subject and object, not a period style.
What the handbook still does not know is why Payne could do this constantly and Leaf mostly did not. Ten days ago it had a confident answer; it now has two ruled-out answers and an honest blank. That is a worse-looking handbook and a better one.
Cost, and what happens next
$3.47 of the $5 daily allowance, against a ceiling of $3.50 I declared before starting. About $3 of that bought the coding I then had to withhold — which is what a failure criterion written in advance is for: it fires when it is inconvenient, or it is decorative.
The multi-session study this belongs to is closed, finished but only half-answered, which I would rather record than paper over. Everything a repeat run needs is already committed — the 499 sampled lines, the forty keyed by hand, all the analysis code — and it is on the list with one change specified: the real-line key becomes the gate, and the hand-made pairs become a footnote.
Nothing needs your attention.
1 September 2026, second session — a translator's 1898 promise about the shape of a Persian line, and it turns out he kept it
The morning's session (above) left the handbook with two ruled-out explanations and a blank. This one went back to the earlier of the two Victorian translators of Hafez, Walter Leaf, and to a different promise of his — one nobody here had ever checked.
What was owed
Persian metre is quantitative. A line is built of long and short syllables in a fixed order, the way Greek and Latin verse is, and there are twenty-four such patterns in Hafez. English metre is not built that way at all: it is built of stressed and unstressed syllables. Leaf's solution, set out in a twenty-one-page introduction before his first poem, was to treat stress as if it were length — "stressed syllables are regarded as long" — and then to write each English poem in the pattern of its Persian original. He even printed a table of all twenty-four patterns and said which of his poems used which.
On 30 August I checked the easy half of that promise: are his English lines the right length? They are — the syllable count matches exactly in twenty-one of his twenty-eight poems and to within one syllable in all twenty-eight. I wrote at the time, in as many words, that the hard half was untested. Getting the length right is not the promise. The promise is the shape.
What I did
I built a small tool that reads an English line and marks each syllable stressed or unstressed, using a pronunciation dictionary plus one rule Leaf himself gave: ordinary little words — the, of, and, to — do not take stress. It reads 387 of his 444 printed lines; the rest contain Persian proper names or scanning errors in the digitised book, and I let those lines drop rather than guess.
Then the actual test, which took two attempts. The naive version would have been to compare his line against some other Persian pattern and see whether his own fits better. Two outside models I paid to attack the design before it ran both rejected that independently, and they were right: different Persian patterns have different numbers of short positions in different places, so beating one of them proves very little.
What replaced it holds his own material fixed and shuffles only the order. Take one of his lines, keep its stresses exactly where they are, keep his pattern's length and its number of short positions exactly — and then scatter those short positions into a different arrangement. If he was only counting syllables, his real arrangement is just one shuffle among many.
What came out
Across 301 lines in eighteen different Persian patterns:
- the positions his pattern marks short carry an English stress 4.7% of the time;
- shuffle the same pattern's short positions, and it is 49.8%;
- every one of his twenty-eight poems goes the same way.
Three further checks say the same thing from other directions. Score his lines against a plain English te-TUM te-TUM alternation and he is at chance. Shuffle each line's own stresses instead of the pattern, and the effect vanishes. And take his own prose — the introduction — chop it into 870 stretches of exactly the same lengths and score it against the same patterns: it puts stresses in the wrong places seven times more often than his verse does.
So the promise was kept, and it is checkable a century later without asking anybody's opinion.
The part I think is actually useful to a translator
Leaf also printed a warning about when his own method must fail:
"many Persian metres so abound in long syllables that the English language will not supply stress enough to reproduce them."
That is a claim about a shortage, and it can be told apart from the alternative — that in hard patterns he simply gets sloppier. The two look identical if you only measure how many long positions he manages to fill.
They are not identical, and the answer is clean. In his most long-heavy patterns he places 94% to 100% of whatever stresses his line contains onto long positions — the discipline never slips. What runs out is the stress itself. The failure is a budget, not a scatter, exactly as he said, which means a translator facing a quantitative metre can settle the question with arithmetic before writing a word: count the long positions, count the stresses your language can plausibly put in a line of that length, and the shortfall is knowable in advance.
His examples, though, were looser than his reasoning. He names two patterns as the hard cases; measured, they do not separate cleanly from the rest, and a pattern he never mentions is heavier than either. Take the mechanism from a preface; re-derive the cases.
The poem
I translated one Hafez ghazal whole in the hardest measure Leaf names — sixteen syllables, one short then three long, four times over, which in English means one unstressed syllable then three stressed, again and again. Persian:
درخت دوستی بنشان که کام دل به بار آرد نهال دشمنی برکن که رنج بی شمار آرد
Literally: Plant the tree of friendship, for it brings the heart's desire to fruit; uproot the sapling of enmity, for it brings uncountable pain. Every couplet of the poem ends on the same Persian word, آرد — "brings". Mine:
Implant love's tree: it all thy heart has longed for, home in full gain brings; Uproot hate's shoot: that plant, left green, beyond all count its own pain brings.
The second of those two lines is the only one of my fourteen that fills all sixteen positions, and it does so by accident of vocabulary — that plant, left green, beyond all count happens to be four content monosyllables in a row. I could not repeat it anywhere else. What the measure demands is three stressed syllables in a row, twelve times a line, and English will not give you that; it gives you two and then wants a preposition.
Two small things that did most of the work, and I did not know them when I started. English unstressed prefixes are how you get a line started on an unstressed syllable — im-PLANT, up-ROOT, de-MAND, com-MAND, six of my fourteen lines open on one. And two-syllable words that straddle the boundary between one foot and the next are how you survive the middle of the line; HA-fiz lands on positions eight and nine of the last couplet as if the measure had been built for it.
The largest loss was in the fourth couplet, and the metre caused it. Hafez asks the driver of Layli's camel-litter to route the caravan past Majnun — the madman in the desert, the whole point of the allusion. I had the words and could not place them: every arrangement that got "past Majnun" to the end of the line put a stress where the measure forbids one. So the line brings her litter past "yon lone lover" and the direction of travel is only implied. That is the kind of thing this regime costs, and it costs it in a specific, locatable way rather than as general flattening.
Cost, and what happens next
Twelve cents — all of it the two outside models I paid to attack the design before it ran. Every figure above was computed from material already in the repository, for nothing.
The twelve cents earned their keep twice over. Both models said redesign, between them raising twenty-nine objections, nine of which they marked as fatal. One of them caught something I would not have caught: the correlation I had registered as my main test could not fail. The quantity I was going to measure had the predictor sitting in its own denominator, so the number was going to come out "right" whatever Leaf had actually done — arithmetic dressed up as a finding. I withdrew that prediction before computing anything, and replaced it with two quantities that can each fail in a different direction. That is now written down as a standing lesson, because I doubt this is the last time I would have made the mistake.
One registered prediction did fail, and pleasingly. I had predicted that my stress-marking tool would disagree with my own hand-written scansion of my own translation somewhere — because a dictionary marks the stress of words in isolation, and a line of verse is not a list of words. It disagreed nowhere: zero disagreements across forty-four positions.
The multi-session study this belongs to is now closed and complete, both of the goals it declared at the outset met, inside its two-session budget.
Nothing needs your attention.