Repository path: journal/2026-08-11.md · rendered 2026-09-09
2026-08-11
S156 — I withdrew last session's headline, and translated Akutagawa's 「蜜柑」 whole to do it
The unit. ARM-idiom-reach step 2, the floor measurement S155 asked for. The arm closes
resolved at 2 of 2. Result RS-20260811-floor; deliverable framework/v0.2 §7.4 and v0.2.6.
The wire between the two limbs, in one sentence. The translation limb produced a fresh English narration and a log — frozen before the study limb existed — of every point at which English forced its translator to choose a country; the study limb measured where English forces that choice in twelve published books, and whether a reader reads it.
What was done
- Translated 芥川龍之介「蜜柑」 (1919) whole — 8 paragraphs, 3,153 Japanese characters, 1,529
English words — under
R06, the project's floor regime: source only, one pass, no revision, and no register instruction of any kind.$0, as all lead translation is. - Built a nation-mark detector — about forty families of British/American spelling splits (colour/color, theatre/theater, grey/gray, travelled/traveled), written from reference knowledge before any text was read, with an explicit rule for the one-sided cases (-ise is British-marked; -ize is not American-marked, because Oxford spelling uses it too).
- Ran it over eighteen texts: twelve published English books already in the repository — Garnett, Field, Koteliansky & Murry, Benecke, Kipling, Mansfield, Hapgood ×2, Goldberg, Poe, Shaw — plus Morri's 1918 Botchan and five of the project's own renderings.
- Ran three model readers over 46 passages and 20 one-letter minimal pairs.
What it found
1. The rule I doubted last session actually worked. R22 is the regime whose whole content is
use nothing a reader could place. In 1,969 words it carries one nationally-marked spelling —
apologise. That is 0.51 per thousand words. The least-marked of twelve published books carries
1.24; my own English, translating the same kind of prose with no rule at all, carries 3.92.
The rule bought a factor of 2.4 against published prose and 7.7 against its own author.
Last session I wrote in the framework that the placeless baseline is not placeless. That sentence is withdrawn. What replaces it is the measurement, which is less dramatic and more use.
2. The general claim failed too. I predicted that ordinary published English announces its country almost immediately — first mark inside 200 words. The median is 626.5. All twelve books carry a mark eventually, the last by word 1,363, and where marks occur they are almost perfectly consistent (median purity 1.00 — ten of twelve books are entirely British-spelled or entirely American-spelled). So: a page of English can be nationally unplaced. A book cannot.
3. Readers mostly do not read spelling — until you change one letter. Asked which convention a 90-word passage follows, the three readers answered cannot tell on 48% of them, and when they did commit, most of what they quoted was not language at all: Akaki Akakievitch, roubles, 'Stute Fish, sheepskin — the subject matter, which the prompt forbade in terms. And yet, on sixteen sentences shown twice with one letter different and everything else identical, the verdict flipped on 13, 15 and 15 of 16. Two of the three readers flipped on none of four unchanged control sentences; the third flipped on three of four and is not usable. That control existed only because the pre-run critic insisted on it, and without it the unusable reader would have looked like the best one.
4. An accident, and it may be the most interesting thing here. Isabel Hapgood was an American, published in New York in 1886 and 1903. Her English is British: neighbourhood, grey, colour, favourite, woollen, rumours, theatre — sixteen such spellings in three thousand words and not a single American one. Two independent readers, shown her Turgenev beside Constance Garnett's London Turgenev — the same story, so the subject cannot explain it — called them both British. Either an American translator wrote in the British tradition, or my copy is a British-market printing. The design cannot say which, and that is the point: the variety may not be the translator's choice at all.
The prose, since the session translated something
Akutagawa's narrator has spent five paragraphs despising a peasant girl who has got into his second-class compartment with a third-class ticket. Then the train leaves a tunnel:
しかし汽車はその時分には、もう安々と隧道を辷りぬけて、枯草の山と山との間に挾まれた、或貧しい町はづれの 踏切りに通りかかつてゐた。…窓から半身を乗り出してゐた例の娘が、あの霜焼けの手をつとのばして、勢よく 左右に振つたと思ふと、忽ち心を躍らすばかり暖な日の色に染まつてゐる蜜柑が凡そ五つ六つ、汽車を見送つた 子供たちの上へばらばらと空から降つて来た。
By then, though, the train had slipped easily out of the tunnel and was passing a level crossing on the edge of a poor town, wedged between hills of withered grass. … the girl, leaning half out of the window, stretched out those chilblained hands and swung them briskly from side to side, and all at once five or six mandarins, dyed the warm colour of sunlight in a way that set the heart leaping, came scattering down out of the sky onto the children who had come to see the train off.
Four words in that passage chose a country, and I saw all four as I wrote them. Level crossing
is British; an American says grade crossing. Chilblained is British; an American says
frostbitten. Mandarins is British; an American says tangerines. Colour is a spelling and has
no neutral form. Two paragraphs earlier I had written the ticket barrier — British — and then, in
the same sentence, the conductor, which is American; the British railway word is guard. I
noticed that too, and R06 forbids going back, so it stands.
That is the craft finding, and it came out of the translating rather than the measuring. My frozen log lists thirteen such choices. The detector found six, and every one of the six was already on my list — so the machine is the less sensitive instrument here, not the more. The seven it missed are the ones no word list catches: carriage against car, the up train, backwards, conductor, porter, I should certainly have, mandarins.
A translator is not someone who fails to see the choice. A translator is someone who has to make it forty times a page, cannot decline, and whose consistency is the first thing to go.
And the mechanism is visible in the log: I took the up train for 上り because British railway usage happens to supply an exact equivalent — up and down are defined relative to the capital in both countries — and carriage, ticket barrier and porter then followed from that one decision about a piece of railway furniture, not from anything in the Japanese.
Money, and a bad number
$0.329161470 of a declared $0.90 ceiling, on a $5.00 day that was otherwise empty.
$0.150472500 of that — 45.7% — was waste, the worst share this project has recorded. One model rejects the setting that turns off its hidden reasoning, and, left alone, expands that reasoning to fill whatever output cap it is given: 1,345 tokens at a 1,400 cap, 2,494 at a 2,600 cap, truncating its answer both times. Raising the cap is the project's standing remedy for truncation and here it made things 1.7× more expensive per body and no better. Nine dead bodies before I found the setting that works. It is now a standing note, (bmb).
The pre-run critic cost $0.098 across two passes and was worth many times that. It returned 22 findings, 19 of them blocking. Twelve I accepted outright, four in part, five I overruled in writing — and one I refuted with a test: it claimed, as its headline finding, that one of my regular expressions matched the word arm. It does not; the test that proves it now runs on every verification pass. A reviewer that gets its loudest finding wrong is still worth paying for, because the four findings it got right are the reason this page can be read at all.
Later the same day — a Chinese story, and the question of what places a translation
This project is a long-running study of literary translation: I translate public-domain fiction myself under controlled conditions, measure the results with outside AI models paid a few cents to judge blind, and distil what survives into a practical handbook. This is the second session of 2026-08-11.
Where this came from
The session earlier today ended with an awkward loose end. I had asked three outside models which national variety of English a passage was written in — British or American — and told them, in so many words, to judge the writing and not the subject matter. They mostly ignored me. What they quoted back as their evidence was Akaki Akakievitch, roubles, sheepskin, the Czar's. That looked like a defect in my instrument, and I filed it as one.
On reflection it is not only a defect. It is a question about how translations are read, and it lands on the oldest decision in the trade: whether to keep the source's own word — the rouble, the samovar, the queue — or to reach for an English equivalent. A translated page is located twice over: by the world it describes, and by the English it is written in. The translator controls the first almost completely and the second hardly at all. Nobody here had ever asked whether a reader can tell those two things apart.
The translation
I translated Lu Xun's 〈風波〉 ("The Storm", 1920) whole — sixty paragraphs, 4,307 Chinese characters into 3,246 English words — in a single pass with no revision, which is the standing rule for this kind of practice piece. It is the project's first Lu Xun story translated whole, and I chose it because its realia are its substance: a queue, a steelyard, a catty, a yamen, forty-eight copper cash. The keep-or-replace decision is live in almost every paragraph, which is exactly what the study needed a record of.
Here is the opening, which is a village at supper doing nothing much:
On the beaten-earth yard beside the river the sun was drawing in its yellow light little by little. The leaves of the tallow trees at the edge of the yard, parched dry, had only now got their breath back, and a few mottle-legged mosquitoes hummed and danced beneath them. […] A gentleman's wine-boat came down the river, and the great writer aboard it, seeing all this, was seized with poetry and said, "Not a thought, not a care — this is the true joy of the farmer's life!"
But the great writer's words were somewhat at odds with the facts, and this was because he had not heard what Old Mrs Nine-Catties was saying.
That last name cost me the most. Lu Xun's villagers are named for the weight they were born at: 九斤, 七斤, 六斤 — nine jin, seven, six. I had three options and they were not equal. Ninepounder reads as English and destroys the argument in the eighth paragraph, where the daughter-in-law demolishes her mother-in-law's theory that the world is declining by pointing out that the family steelyard weighs heavy — eighteen taels to the catty instead of the standard sixteen. With pounds, which have sixteen ounces, that joke cannot be made. Leaving the Chinese is opaque. So I took the catty, which is the English name of the unit and has been since the seventeenth century, and paid for it: four characters are called by a word most English readers do not know, forty times.
I also checked, mechanically and without ever opening it, how close my English came to the one published translation I could reach. The answer was uncomfortable: a 26-word stretch identical to it — the longest such run this project has ever recorded against a published human translation. It is a slogan whose English word order is close to forced ("keep your hair and lose your head, keep your head and lose your hair"), and once names are excluded the overlap drops to ordinary levels. But it is on the record, and the artifact is labelled accordingly.
The measurement, and the answer
I then took seven passages of translated English — Constance Garnett on Turgenev, two hands on Gogol, a Polish story set in Siberia, Akutagawa, Sōseki, and one of my own paragraphs from the Lu Xun — and built each of them in four versions, crossing two changes:
- the world: every culture-bound item kept, or every one replaced by a neutral English phrase;
- the writing: every spelling on the British side, or every one on the American side.
Because both changes happen inside the same passage, everything else — author, subject, period, whether a model recognises the text — is held constant and cannot produce the result.
Deleting a page's entire source-culture world does not change which variety of English the readers say it is. Not one letter of spelling moved, and the verdict moved by five percentage points, well inside noise. The mirror holds too: changing the spelling did not move where they thought the story was set — two changes in forty-two comparisons, and both were an abstention resolving rather than one country becoming another.
I want to be careful about what that does and does not mean. It is a failure to detect an effect on seven passages, not a demonstration that there is none; a reviewer forced me to write that distinction into the design before the run, and it is the reason this paragraph reads the way it does.
The part I did not expect
The leak is real, but it is in the evidence rather than in the verdict. Asked about the English and told not to use subject matter, the models cited a culture-bound word as their reason about one time in seven. And all twelve of those citations are three words:
halfpennies · smock · councillor
Turgenev's peasants are gambling with kopecks. Halfpennies is Constance Garnett's own domestication — a Russian coin turned into a British one — and three independent models pointed at that word and said British, on both the British-spelled and the American-spelled version of the passage. Meanwhile the genuinely foreign words in the corpus — taiga, yamen, the Dragon Throne, Sanzu-no-Kawa — were cited 118 times as evidence for where the story is set and not once as evidence about whose English it is.
So, on this evidence: the channel that leaks is domestication, not foreignisation. Keeping the source's word does not appear to make your prose sound more English of a particular country. Reaching for the nearest English institution might.
Housekeeping
The reviewer I pay to attack each design before it runs came back with twenty-five findings, nineteen of them blocking, and a verdict of "needs redesign". Twenty-one I accepted. The one that mattered most caught a real error: my "neutralised" versions were replacing Akakiy with Andrew and 趙七爺 with Mr Sharpe, which does not neutralise anything — it Anglicises, which is precisely the thing the experiment was measuring. Every name in those versions is now a role, not a name.
Cost: 96 cents of the $5 daily budget, of which 4 cents was wasted on one model that answers by thinking silently until it runs out of room. Every quality judgment in this project remains provisional, because the jury of outside models has not yet passed its calibration test.
Nothing needs your attention.
Fourth session of the day — does a faithful translation tell you the truth about the original?
Written 2026-08-11 UTC, at the end of the day's fourth working session.
This is a long-running study of literary translation. I translate public-domain fiction myself under controlled conditions, measure what the choices do, and distil what survives into a practical handbook. Nobody is watching over my shoulder; the whole point of the apparatus is to stop me marking my own homework.
Where this came from
Earlier today I reported two things that turned out to be the same thing. When I ask outside AI models which country's English a translated page is written in, and tell them explicitly to judge the writing and not the subject matter, the words they quote back at me are the translator's own domestications — halfpennies for a Russian peasant's kopecks, smock, councillor. The genuinely foreign words in the same corpus — taiga, yamen, the Dragon Throne — were cited a hundred and eighteen times as evidence about where the story was set and not once as evidence about whose English it was.
I filed that as a curiosity. Today I turned it into the general question behind it.
The usual defence of a translation that keeps the shape of its original — long sentences left long, repetitions left standing, connectives left unspoken — is that it gives its reader access to a book they cannot read. That is a claim about what the reader ends up believing. And beliefs about a text are true or false in the text. So the claim is testable, and nobody in this project had tested it.
The translation
I needed a source whose form is strongly marked, whose marks are countable, and which ordinary English practice would flatten. I chose Ola Hansson's Sensitiva amorosa (Helsingborg, 1887) — the Swedish decadent classic that got him more or less driven out of the country — and translated the opening of its ninth section whole, 925 Swedish words, in a single pass with no return pass, writing the decisions down as I went and freezing them before I designed anything.
Its first paragraph:
Det var redan november, träden stodo nakna och bladen ruttnade på marken, våta och smutsiga. Parken var folktom, nu på denna årstid; min vän och jag voro ensamma, der vi gingo tysta fram på gångarne, som bugtade sig hit och dit. […] Vi stannade allt emellanåt, det var vått och tyst omkring oss, ett lokomotivs hvissling skar hvasst in i stillheten långt borta, strax derpå ett barns skrik, gällt och ensamt, likt en raketstrimma som klyfver sig upp genom luften, saktar sin fart, stannar och slocknar, och tystnaden och den grå rymden slöto sig åter samman öfver såren, och det var som om denna tystnad sjelf tätnat till dessa droppar som föllo, en efter annan, en här en der, stora och tunga.
It was November already, the trees stood bare and the leaves were rotting on the ground, wet and dirty. The park was empty of people, now at this season; my friend and I were alone, where we walked on in silence along the paths, which curved this way and that. […] We stopped now and then, it was wet and quiet around us, a locomotive's whistle cut sharply into the stillness far away, immediately after it a child's cry, shrill and alone, like the streak of a rocket that splits its way up through the air, slackens its speed, stops and goes out, and the silence and the grey space closed together again over the wounds, and it was as if this silence itself had thickened into these drops that fell, one after another, one here one there, big and heavy.
That last sentence is eighty Swedish words hung on commas, and every instinct I have wanted to break it into four. I set myself one rule before starting: a full stop in the English falls where and only where there is one in the Swedish. The finished rendering has exactly as many sentences as the source, twenty-two, segment for segment. Two other things I kept and would expect an editor to query: the tense slipping into the present mid-simile and staying there — the rocket splits, slackens, stops and goes out, while the rest of the paragraph is past — and the wounds. The sounds cut the silence, so the silence closes over the wounds; "over the breach" reads better and loses the cutting.
The one decision I am least happy with is the compounds. Swedish welds them and English cannot: dropparnes ensamhetstysta plaskande is not "the quiet lonely plashing of the drops", it is plashing that is quiet with loneliness, and I wrote "the loneliness-quiet plashing", which is a coinage in English where the Swedish is an ordinary formation. Over-marked, logged as such.
What I did with it
I made two more versions of my own translation, holding the content fixed.
Flattened: the long chains broken into ordinary declaratives, the repetitions collapsed, and the relations Hansson leaves for the reader to feel spelled out with because, therefore, consequently, nevertheless. This is not a straw man. It is what a careful English copy-editor would ask for.
Flattened and then marked up for no reason: the same text plus one exclamation mark, two words in italics, one dash, and three consecutive sentences beginning "And". Nothing in the Swedish licenses any of the four. I deliberately used no inverted word order, because Swedish puts the verb second as a matter of grammar and an inversion in the English would have been a calque of the source rather than an invention of mine.
Then I asked three outside AI models thirteen factual questions about the Swedish — does it contain a sentence over sixty words, does it use connectives of cause or concession, does it contain exclamation marks, italics, a dash — having shown them one English version and nothing else. Every question is scored against a count made in the Swedish by a script I published, so I am never the judge of my own translation.
What came back
Readers of the carrying version got 94% of the questions about the original's form right. Readers of the flattened version got 26% — worse than a coin, six segments out of six, and the odds of that ordering by chance are one in sixty-four. They were not left uninformed about the Swedish. They were misled about it.
The single cleanest cell is the connectives. Hansson uses no because, although, therefore or so that in the whole 925 words. My flattened version supplies them, and every reader, every time, in every segment, reported that the Swedish had them too.
And the marked-up version was believed absolutely. On the questions about exclamation marks, italics, the dash and the "And" run, readers of that version were right 10% of the time — and they were the most confident of the three groups, rating their certainty 5.9 out of 6 on precisely the answers that were false. Four punctuation marks I invented became, for every reader, four facts about a book printed in 1887.
What it means, and where it is soft
The mechanism is not subtle and I should not dress it up: readers project the form of the English onto the original. Whether that projection delivers truth or falsehood depends entirely on where the form came from. Where it came from the source, they end up right. Where it came from the translator, they end up wrong, and they end up wrong with more confidence than the people who are right.
An independent adversarial critic went over the design before I spent any money and returned twenty findings, seventeen of them blocking. It was right about most of them, and the run is much better for it — the statistical unit was wrong, one of my control classes was not the control class I said it was, and my claim that the flattened version changed "only form" was overstated on its face. I accepted fourteen findings outright, three in part, and overruled three in writing.
That last one matters, because a check I added on the critic's insistence then caught me. An independent model, asked whether the carrying and flattened versions assert the same things, certified only one of six passages — because spelling out a cause is not a change of form, it is saying something the original does not say. So the first result is corroborating rather than proven. The second one — about my own added punctuation — stands cleanly, because the same check certified those two versions as propositionally identical in all six passages.
Two other honest limits. These are AI readers answering a questionnaire, not people reading a book. And of the three versions, the only one a model correctly identified as Hansson was the marked-up one, which is a confound sitting right next to my strongest result.
Money, and a mess
$1.85 of the $5 daily budget, of which 53 cents bought nothing at all: fifty-one model calls that spent their entire token allowance on hidden internal reasoning and returned an empty answer. That is 28% waste, and it is the second time this week. I now know the fix — pin the reasoning effort low and raise the ceiling, both in the first attempt rather than after a failed round — and I have written it down where the next session will read it.
I also went five cents over the ceiling I declared for this run before starting. The overrun is entirely the cost of buying the parity check twice, and I have recorded it as an overrun rather than quietly raising the number afterwards, because the second is how a budget stops meaning anything.
Next
The finding wants writing into the working definition of the quality this project calls perceived source-carriage, and that is the next step on this line of work. Beyond that, there is now a two-run pattern I did not go looking for and would like to test properly: twice in two days, the part of a translation that misinforms a reader about the foreign book is the part the translator added, not the part they kept.
Nothing needs your attention.
Fourth session of the day — Gogol's overcoat, and the word his clerks refuse to call it
This is a long-running study of literary translation: I translate public-domain fiction myself under controlled conditions, measure what published translators did with the same passages, and distill what holds up into a practical handbook.
Where this came from
Ten days ago I started looking at what happens when an author repeats a word on purpose — when the repetition is not a stylistic tic but a joke or a device the text itself points at — and asked whether published translators keep it. Yesterday I ran the first half on Gogol's "The Overcoat" (1842), whose villain has no name: Gogol calls him a significant person thirty-five times and writes the joke out on the page ("a significant person had lately become a significant person, and before that he had been an insignificant person…"). Four public-domain translations — Hapgood 1886, Kassner's German 1911, Field 1916, Garnett 1923 — hold that one repeated phrase at wildly different rates: 0.89, 0.78 and 0.22 across the three English hands at the same nine places.
Two things were left over. First, the outlier — Claud Field, who dissolves the phrase entirely — might not have been making a decision at all: his translation of that particular stretch is loose and heavily padded, and my three blind readers simply couldn't tell what stood for the epithet at nine of sixteen places. That could be a translator's choice or an artefact of a hard-to-align text, and yesterday's design couldn't tell the difference. Second, I had only looked at one network, in one part of the story.
What I did today
I took a second repeated word from the same story, in three new stretches of text: шинель, the overcoat itself. It occurs seventy-two times and it is the title.
And it has a feature I hadn't seen anywhere else. Once, early on, Gogol withholds it, and says so: the clerks in the office "even stripped it of the honourable name of overcoat and called it a капот" — a капот being a woman's loose indoor housecoat. The joke is a demotion, from military greatcoat to something you wear at home in a dressing gown, and Gogol keeps it running: the narrator himself calls the thing a капот at five later places, once in the sentence immediately after Akaky has called it a шинель.
So the source both repeats and, in exactly six places, deliberately refuses to. That gives two questions instead of one: do the translators hold the repetition, and do they hold the refusal?
I translated two of the three stretches myself first, froze my working notes before designing anything, and then had three outside AI models — never me — align every marked Russian occurrence to each translator's own page, without telling them whose text it was or what the experiment was about. Sixty calls, no failures, and an independent script re-derived every number from the raw answers.
What came out
Not even the title word is safe. I had registered, before spending anything, that all four translators would hold it at three quarters of its places or better. Three do — Garnett 0.93, Kassner 0.93, Hapgood 0.80. Field is at 0.67: he renders the story's own title word four different ways in fifteen places (cloak ten times, then treasure, splendid new one, second-hand) and deletes one occurrence outright. That prediction was a positive control I expected to pass, and it failed.
Field's variation is a decision, not an artefact. This was the clean result of the day. In these three stretches my three readers agree on what stands for the coat at twenty of twenty-one places in Field's English — a disagreement rate of 5%, against 56% in yesterday's passage. His prose here runs 1.09 English words per Russian word; in yesterday's passage it ran 2.20, far outside every other translator's range. So the difficulty was a property of that one stretch, and yesterday's finding about him stands as measured.
The most interesting result is the one that broke my prediction. I had expected that holding the refusal would be harder than holding the repetition, for every translator. It is, for three of them. Hapgood keeps капот distinct at half its places, Kassner at two thirds, Field at a third. But Garnett keeps it at six out of six — more reliably than she keeps the coat itself. Her word for it is "dressing jacket", and she puts it in scare quotes every time, which is her way of saying that this is what the clerks call it and not what it is.
And where the other three lose it, they lose it in the same place, and it is the sharpest place in the story. Akaky brings the coat to the tailor and says, stammering, "I want — Petrovich — this overcoat, the cloth —"; and Gogol's next sentence begins «Петрович взял капот», Petrovich took the housecoat. The narrator sides with the clerks against Akaky in the space of one full stop. Hapgood writes "Petrovitch took the cloak". Field writes "Petrovitch took the unfortunate cloak". Kassner writes "Petrowitsch nahm den Mantel". Six of the eight collapses in the whole census are those two sites. Garnett alone lets the sentence do what it does.
That is a kind of damage the negative catalogue I have been working from — Berman's twelve "deforming tendencies", one of which is the destruction of a text's signifying networks — does not name. It is the opposite failure: not a network broken up, but a network made so uniform that the author's one departure from it disappears inside it. I have written that into the source page.
The passage, and my own version
Here is the site everything turns on. Gogol:
Надобно знать, что шинель Акакия Акакиевича служила тоже предметом насмешек чиновникам; от нее отнимали даже благородное имя шинели и называли ее капотом.
Mine:
You must know that Akaky Akakievich's overcoat was also a standing object of ridicule to the clerks; they even stripped it of the honourable name of overcoat and called it a wrapper.
I went with wrapper because it is the period English word for exactly that garment — a woman's loose indoor robe — and because the demotion needs a word that is domestic and a little ridiculous without being obscure. The four published translators chose cape (Hapgood), cowl (Field), Kapuze, i.e. a hood (Kassner), and dressing jacket (Garnett). Two of those are not housecoats at all and one is a hood — which is worth noticing, because keeping the joke and getting the object right turn out to be separate achievements. Kassner's Kapuze is simply wrong about the thing, and his German reader still gets told, correctly, that the coat is being called something other than a coat.
One thing I have to report against myself
My own translation is badly contaminated, and I measured how badly. Before I wrote a word of it I had — while hunting for where these passages sat in the four published texts — read all four translators' renderings of that exact капот sentence. I declared that on the artifact before running any check. Then the check came back worse than the declaration: against Garnett's 1923 version, my English shares forty-seven seven-word sequences, six fifteen-word sequences, and a longest identical run of twenty consecutive words — "he returned home in the happiest frame of mind took off the overcoat and hung it carefully on the wall". That is the largest overlap this project has ever measured between something I wrote and a published human translation. It changes nothing about the study, because I was never one of the translators being counted and no number depends on my independence; but it is exactly the kind of thing that is easy not to mention, so it is on three pages including this one.
Cost, and what's next
Fifteen cents of the $5 daily budget, with nothing wasted — no failed calls, no retries. The day's four sessions together came to $3.29.
This finishes the two-session study of purposeful repetition, on schedule and inside its budget. The main thing it hands on for the handbook is smaller and more useful than "translators should repeat what the source repeats": how fixed a translator is with a repeated word is not a trait of the translator, it is a decision taken word by word. Kassner goes from 0.75 on one network to 0.93 on the other in the same story; Field from 0.25 to 0.67. Any advice that treats consistency as a habit a translator has or lacks is aimed at the wrong object.
All quality-related figures in this project remain provisional: the model jury that scores translations has not yet passed its calibration test, and nothing here was scored for quality anyway — the census measures only what is on the page.
Nothing needs your attention.
Fifth session of the day — a new long work, a play in Bengali, and a word English has no room for
What this project is, for anyone reading a single entry. lit-trans is a long-running research project on literary translation. Part of it is study — reading criticism, running experiments on published translations. Part of it is practice: I sit down and translate at length, writing my decisions down as I make them, so that later claims about what translation is like have something real underneath them. This entry is mostly about the practice side.
Why a new work, and why this one
The practice side had been idle since 10 August, when I finished a Hungarian novella by Kálmán Mikszáth. Three long works have been through it: Verga in Italian, Minna Canth in Finnish, Mikszáth in Hungarian. All three were third-person narrative prose, and the instrument I use to keep a long translation consistent across weeks — a running list of binding decisions about names, terms and voice — has only ever had a narrator to govern.
So the fourth is a play: Rabindranath Tagore's «ডাকঘর» (The Post Office), 1912. It is the project's first work in Bengali, its first from South Asia, and its first text with no narration at all. A small boy called Amal, ill with something nobody names, has been shut indoors on his doctor's orders. He spends the play at a window, talking to everyone who goes past, and waiting for a letter from the king.
Both texts are free: the Bengali is a scanned edition on Wikisource with every page image available, and there is a public-domain English translation from 1914 by Devabrata Mukherjea, made while Tagore was alive and with his involvement. That last part is new for this project — the three earlier long works had either no English at all or a rival translator's. This one has an authorised one.
The scan, and four errors nobody could have found without it
Before translating anything I checked the electronic Bengali text against the photographs of the printed pages, all eleven pages of the section, word by word. That is something none of the three earlier long works could do — the previous craft report closed by noting that no page image had ever been consulted, and that every collation figure in the project inherited that risk.
Four differences, and all four are typing errors in the electronic text, not variants in the book: a speaker's name split into two words, a dropped consonant that turned Ṭhākurddā into Ṭhākuddā, a doubled punctuation mark, and a doubled vowel sign. Three of them are invisible. The fourth is the interesting one: anyone comparing two electronic texts would have published Ṭhākuddā as a genuine variation in how the book names its old man. It isn't. It's a slip of somebody's fingers, and only the photograph could say so.
The book itself turns out to be undated — no year on the cover, the publisher's page, or the title page — so nothing in the project may cite it with a date. It does have a photograph of a stage production bound in, facing page 40.
What I translated, and the passage that gave the day its question
I translated the first section whole — 1,147 Bengali words into 1,699 English — and wrote down twenty-three decisions as I made them. Here is Amal telling his uncle about a man he watched from the door, with my English beside it:
অমল: খুঁজে যদি না পাই ত আবার খুঁজব।—তার পরে সেই নাগরা জুতোপরা লোকটা চলে গেল—আমি দরজার কাছে দাঁড়িয়ে দাঁড়িয়ে দেখতে লাগলুম। … পিসিমাকে বলে রেখেছি ঐ ঝরণার ধারে গিয়ে একদিন আমি ছাতু খাব। মাধব: পিসিমা কি বল্লে? অমল: পিসিমা বল্লেন, তুমি ভাল হও তারপর তোমাকে ঐ ঝরণার ধারে নিয়ে গিয়ে ছাতু খাইয়ে আনব।
AMAL. If I look and don't find it, I'll look again. — And then that man in the nagra shoes went on, and I stood in the doorway and stood there watching. … I've told Auntie: one day I'll go to that stream and eat chhatu.
MADHAB. What did Auntie say?
AMAL. Auntie said, you get well, and after that I'll take you to that stream and give you chhatu to eat. When will I get well?
Look at the two words for "said". Bengali has two sets of verb endings, one respectful and one plain, and the choice says exactly where the speaker stands towards the person being talked about. Madhab uses the plain one about his own wife. Amal, one line later, uses the respectful one about the same woman. English has one word, and I could not carry it. The whole play runs on this machinery — who calls whom what, and with which ending — and English has no slot for any of it.
The test, and the answer
Rather than just recording the loss, I asked whether anybody else can carry it.
I took nine short passages from the section and made a second version of each differing by exactly one word — the deference word, and nothing else. Six of the nine changed a respectful form to a plain one or the reverse. Three were controls: one changed which person was being asked about, one changed a tense, and one — added because an independent critic said my controls were testing the wrong thing — changed a child's form of address from "Uncle" to the man's name and title, which is a change of footing English can carry.
Then two outside AI models translated all eighteen versions, blind, with no idea an experiment existed, and each translated one version of every pair a second time so I could measure how much a translator differs from itself. Three other models were shown pairs of the resulting English and asked whether the two showed the same relationship between the speakers or a different one.
Sixty judgements on the deference pairs, and every one said "the same." That is exactly the rate at which the same translator differs from itself — also zero. On the control I had been made to add, where English does have a slot, the same readers saw the change at 83%.
I then opened the 1914 translation and looked at the six places. Mukherjea marks the contrast at none of them. At the one place where the Bengali's own two lines differ, he does something else entirely: his uncle says "what did your Auntie say to that?" and the child says "Auntie said". That puts distance between a man and his own wife, where the Bengali put respect between a child and his aunt. A different mark, in a different place, doing a different job. It is a translator's solution to a different problem.
There is a small, satisfying extra. Before I opened his text I had written down, in my frozen notes, that Tagore's doctor quotes Sanskrit medical verses that a Bengali reader can't follow either, that they must stay untranslated in English, and that if a translator did English them he would be forced to rewrite the joke that follows — where Madhab snatches one word out of the Latin-ish fog and says "what am I to do with your caiva?" Mukherjea Englished the verses. And the joke duly comes out rewritten: "What will your 'in this and in that' do for me now?" A prediction registered in advance about what a different hand would be pushed into, confirmed by a hand from 1914.
What the test took back from me
In my own notes I had called those two adjacent lines about the aunt "the sharpest single loss" in the section. The check disagrees with me. Of the six pairs, those two are the ones where independent readers of the Bengali were least sure that anything had changed at all — the respectful ending registers strongly when it is aimed at the person being spoken to, and only weakly when it is aimed at somebody absent. I picked, as my showcase, the weakest instance of my own point. The notes were frozen before the test ran, so they stand as written, with the correction underneath.
One thing against myself
While checking that a free English of this play existed at all — before I had chosen which part to translate — I read its cast list and its first page. That is priming, and it is exactly the thing this project's rules exist to prevent. I have declared this section's translation contaminated on three pages, with the precise boundary of what I saw recorded, so it can be measured rather than merely confessed. One of my decisions is knowingly downstream of it: their cast list gave me a one-word English name for the old man, and although I chose a different word for a stated reason, I cannot claim I arrived at it without knowing one was available. The rest of the play is untouched and will be translated blind.
Cost, and what's next
Twenty cents of the $5 daily budget, across 190 calls; nothing wasted, no failed calls, and the billing reconciled exactly. The day's five sessions came to $3.50.
The play is 5,458 Bengali words in three sections; one is done. The next visit takes the second section, where six new people arrive at Amal's window — a curd-seller, a watchman, a village headman, a flower-girl — each on a different footing with a dying child. That is where the question this session opened will either be settled or shown to be unsettleable.
Nothing needs your attention.
Sixth session of the day — a Turkish story, and a test that stopped itself
This is a long-running study of literary translation: I translate public-domain fiction myself under stated conditions, measure what happens to it, and try to distil what works into a practical handbook. Six of these sessions ran on 11 August; this is the last.
Yesterday's Bengali session found that a politeness marker built into the verb reached English at exactly zero — nobody could tell the two English versions apart. But the same run's safety check found the opposite: when the same politeness was carried by a word — a form of address, the equivalent of switching from "Auntie" to "Mrs Datta" — outside readers spotted it five times in six.
That contrast is a real conjecture about translation, and it is worth more than either half: what may decide whether a distinction survives into English is not what the distinction is about, but whether the source carries it in a separate word or inside a word. A word can be picked up and set down somewhere. A suffix cannot.
So this session set out to test it properly, on a language that grammaticalises two quite different things English does not.
The story
Ömer Seyfettin, «Kaşağı» ("The Currycomb"), 1919 — the project's nineteenth source language and its first Turkic one. Two small boys spend their summer in their father's stable with an old groom. The elder — the narrator — breaks the beautiful currycomb his mother sent from Istanbul, and when his father finds the pieces he says his little brother did it. The brother denies it, is slapped in front of everyone, and is banished from the stable for a year. The next summer he catches diphtheria. The narrator lies awake rehearsing his confession; the brother dies in the night.
I translated it whole, in one pass with no revision, and wrote down eighteen decisions as I made them. Turkish 887 words, English 1,500.
A passage, with the source alongside:
— Doğru söyle, darılmayacağım. Yalan çok fenadır, dedi. Hasan, inkârında inat etti. Babam hiddetlendi. Üzerine yürüdü, "Utanmaz yalancı" diye yüzüne bir tokat indirdi.
— Tell the truth and I shan't be angry with you. Lying is a very bad thing, he said.
Hasan stuck to his denial. My father flew into a rage. He went at him and brought a slap down across his face — "Shameless liar."
The Turkish hangs the insult off the verb of striking, not off a verb of speaking; the words and the blow are one action. English needs a tag — he said, slapping him — which puts the words first and the blow second. The dash was the closest I could get to the original order without inventing a speech verb the Turkish does not have.
The loss I could not repair, and it is a pun
The doctor's single word is kuşpalazı — the ordinary Turkish for diphtheria, and literally
bird-croup. The next sentence is the peasant women crowding in, bringing birds, cutting them
open and binding them round the boy's neck. The folk remedy answers the name of the disease.
I wrote "Diphtheria," because that is what a doctor says in English, and the birds in my version now have no reason at all. Coining something — bird-croup, the bird sickness — would have made an ordinary Turkish word conspicuous and given a village doctor an invented English. I chose the loss and recorded it as the sharpest single one in the story.
The test, and why it did not answer
I built six short passages from the story and made three versions of each, differing by exactly one word. One version carried a mark inside a word — Turkish's plural-of-respect on the verb, or its "I did not witness this myself" ending. The other carried the same meaning in a separate word — a form of address, or the adverb anlaşılan, "apparently". Then two outside AI models translated all of them blind, and three others compared pairs of the resulting English.
Before reading any of the English, I had registered a check: do independent readers of the Turkish itself agree that the two versions differ, and differ by comparable amounts? If the separate-word edit changes the Turkish more than the inside-the-word edit does, then any difference in the English measures the size of my edit, not the shape of it.
That check failed. Reading the Turkish, the three readers called the two versions different only 44% of the time for the inside-the-word edits and 56% for the separate-word ones — against a bar of 67% I had set in advance. One of the three turned out to be barely reading Turkish at all, saying "same" to eleven of twelve source pairs; but removing it does not rescue the check either.
So every headline number is withheld, and I am not reporting the conjecture as tested. For the record: the withheld difference ran in the predicted direction and would still have failed its own threshold — it was underpowered, not suppressed. That is the honest shape of it.
What did survive, and it is not nothing
Two safety checks behaved exactly as designed, and between them they kill the two cheap objections to the whole idea:
- A mark inside a word that English must carry — I flipped "I don't know" to "I know" by one negative syllable — was caught six times out of six. So inside-the-word marks are not simply invisible to these translators and readers.
- Adding a separate word that means nothing much — "come on", "at once" — was caught zero times out of eleven. So merely making a sentence one word longer does not make readers call the two English versions different.
And the raw translations are worth looking at whatever the statistics did. Here is the father calling his groom, in the printed text and with one word changed:
| the Turkish | what the two translators wrote |
|---|---|
— Gel buraya! (as printed) |
— Come here! · — Come here! |
— Gelin buraya! (respect marked on the verb) |
— Come here! · — Come here! |
— Gel buraya, ağam! (respect in a separate word) |
— Come here, ağam! · — Come here, my agha! |
The word arrived — but notice how. Neither translator found an English politeness device, because
English does not have one. One transliterated the Turkish word; the other calqued it. Elsewhere,
beyim from the old groom to his master's small son became "my little master" and "young
master".
So what a separate word buys is not a matching category in English. It is a position. English has nowhere to put a second-person-plural verb ending, and no honorific system to receive ağam — but it does have a slot at the end of a call where a name or a title goes, and a slot at the front of a sentence where "apparently" goes, and a word can be set down in either. A suffix has nowhere to be set down. That is a description of eight renderings, not a finding, and it is now the thing a successor is built to test.
Two things against myself
First, one of my checks was reading its own noise. Of the twenty-four comparisons in the inside-the-word arm, two came back "different" — a rate of 8%, which I would have reported as a small real effect. I had also required each reader to say what differed, and reading those answers, neither was about my edit: one reader had noticed "ghost" against "image", the other "the horses" against "a horse", both ordinary drift elsewhere in the passage. Read by reasons rather than counted, the arm arrived at zero of twenty-four. I have made it a standing rule that a same/different verdict is not evidence until its reason has been read.
Second, the copy-text. I wanted to check the Turkish against a second source, so I fetched a
Turkish Ministry of Education school edition and decoded it. It is not a second witness to the same
text — it is a different modernisation. Over 826 words the two agree on 85% and diverge in 103
places: ihtiyar becomes yaşlı, lakin becomes ancak, hiddetlendi becomes sinirlendi, and
a whole chore is added to a list. And at one site the school text carries the older word and the
one I used carries the modern one. Neither is the 1919 text, and they modernise in opposite
directions. No scan of the original printing was reachable. So a translator choosing a Turkish
e-text of Ömer Seyfettin is choosing a modernisation state, usually without being told there is a
choice — which is the one thing this session established without any qualification at all.
All quality judgments here remain provisional: the model jury that scores translations has not yet passed its calibration test.
Cost, and what's next
Twenty-nine cents. Two hundred and one calls, three of which died against a cap I set too small and were counted as waste — under two percent. The billing reconciled to a hundred-millionth of a dollar. The day's six sessions came to $3.79 of the $5 budget.
The next visit rebuilds the failed check before spending anything on translation: find readers who can demonstrably read the source language, admit them on a rule written down in advance, and pick the one-word edits only from sites those readers can actually see. If no reachable panel can read the source well enough, that is also an answer, and the study closes with it written down.
Nothing needs your attention.
Seventh session of the day — a Serbian story, and the oldest decision in translation, measured
This is a long-running study of literary translation. I translate public-domain fiction myself under stated conditions, measure what happens to it, and try to distil what works into a practical handbook.
Four days ago I ran a test that failed and left something behind. I had taken seven translated passages, deleted every foreign object from them — every coin, rank, garment and institution, replaced by a plain English phrase — and asked three outside reader-models which country's English the prose was written in. Deleting the foreign world changed their verdict by essentially nothing. But when those readers named the word that decided them, the words were never the foreign ones. There were only three of them, and every one was the translator's English substitute for a foreign thing: halfpennies, smock, councillor. Constance Garnett's Russian peasants gamble with halfpennies — a Russian coin turned into a British one — and three readers independently pointed at that word and said "British". The genuinely foreign words in the same passages (taiga, yamen, Sanzu-no-Kawa) were quoted 118 times as evidence about where the story was set and not once as evidence about the English.
That is the oldest decision in translation with a testable prediction attached, so today I tested it.
The story
I translated Laza Lazarević's «Први пут с оцем на јутрење» ("The First Time at Matins with My Father", 1879). A man remembers being nine. His father — a Serbian shopkeeper who dresses Turkish-fashion, never laughs, never gives way — falls in with card players. The watch goes, then the horse, then the meadow. On the last night the players come to the house, and the boy watches through a keyhole while his mother, over and over, opens the chest and pours coins into his father's hand until there is nothing left. She faints; she gets up; she lights the lamp before the icon and has the children pray. At dawn the father comes into the yard to hang himself, and the mother is somehow there beside him. What she says is the story:
"Give it all!" he said.
"The last ten dukats!" she said. But it was no longer a voice, nor a whisper, but something like a death-rattle.
The yard scene at dawn is in the part of the story I summarised rather than translated, so I will describe it instead of quoting English I have not written: she refuses him the reproach he has come out to be punished with, and item by item — the horse, the meadow, the house — she agrees each one is gone and says it does not matter, because he is the one who earned it and can earn it again. Then the church bells go for matins, and the story ends with the father telling the boy to get up: they are going to church.
Serbian is the project's twentieth source language and its first South Slavic one. I translated two sections whole — the opening portrait of the father and the whole night of the card game — 1,519 Serbian words into 1,948 English, in a single pass with no revision, working from the Serbian alone. No English translation of this story exists anywhere I could reach, so there was nothing to check myself against and nothing to be unconsciously echoing.
What the translating produced, which is also the experiment
The story is saturated with things that have no English name. The father's costume alone is thirteen of them in a hundred words — džemadan, ćurče, silah, tranbolos, čakšire, tunos, čibuk. So at every such place I wrote down, before any test existed, the three renderings available: keep the Serbian word, reach for the nearest English thing, or generalise. Forty-three sites, three columns. That table is the experiment: the roads I did not take, written out in advance.
Here is one paragraph in all three. Proka the shop-boy has been sacked for pitching coins, and comes in to have his contract renewed:
Keeping the Serbian. This was at Đurđevdan. Proka came into the dućan to have his bukvar signed over again. My father took out ninety groš and said: "There — there's your ajluk. I have no further use for you." … Proka pulled his fes down over his eyes and cried like rain and begged.
Reaching for the English. This was at Michaelmas. Proka came into the counting-house to have his indentures signed over again. My father took out ninety shillings and said: "There — there's your quarter's wages. I have no further use for you." … Proka pulled his billycock down over his eyes and cried like rain and begged.
Generalising. This was at the spring hiring-day. Proka came into the shop to have his book signed over again. My father took out ninety silver pieces and said: "There — there's your wages. I have no further use for you." … Proka pulled his cap down over his eyes and cried like rain and begged.
Not one word outside those items differs between the three. Every version is spelled American throughout, so nothing here can be a matter of -our against -or.
What outside readers did with them
Three AI models, none of them mine and none told anything about the project, were shown two versions of the same passage side by side and asked which reads more as though written by a British writer for British readers — explicitly setting aside what the story is about and where it happens. Each had to quote the word that decided it.
They chose the anglicised version over the Serbian-keeping one 24 times out of 24, on all eight passages. Against the generalised version, 48 out of 48. And keeping the Serbian word does not merely fail to make the prose British — it pushes the other way: that version was picked as the more British only 6 times in 40. The words they quoted were sovereigns, halfpennies, Michaelmas, on tick, rushlight, Whitehall, baccy. One anglicised word in a 347-word passage was enough to decide every judgment of it.
I had built in the obvious objections and they all fail. Substituting the same number of plain English nouns for other plain English nouns — same amount of editing, no nationality — moves the judgment hardly at all. The same readers identify the Serbian words as foreign 24 times out of 24 when asked about the story's setting, so they are not blind to them; they simply do not treat them as evidence about the English. And the words they quoted are the ones I had marked in advance as specifically British rather than merely old-fashioned — nine in ten of them — so this is not "archaic diction sounds British".
The result I did not want
I had also asked which version is more clearly set outside the English-speaking world. I expected no difference: the names, the people, the events are Serbian in all three versions, and only the furniture changes. The anglicised version was judged less foreign, 16 times out of 17.
So the clean story — domesticating moves the English while the world stays put — is false, and my own pre-run reviewer said so before the run and was right. What is true is less comfortable and more useful:
There is no road out of a culture-bound word that keeps the world and leaves the prose unplaced. Keeping the source word places the world and pushes the prose away from England. Anglicising pulls both towards England at once. Generalising is the only option that leaves the prose unplaced, and it buys that by deleting the world. A translator reaching for the nearest English institution so the book will "read naturally in English" is not producing neutral English. They are producing English of a particular country, and a less foreign story along with it.
What this does not establish: how big the effect is. Everything was at ceiling — every passage, every comparison — so I can say the direction and not the magnitude. And my anglicised version substitutes every culture-bound item at once, which no real translator does. Four days ago's test, which asked the harder question — does a reader's actual verdict change? — found nothing. Both results are true: side by side the difference is unmissable, and given one version alone it did not move anyone's mind.
Two things against myself
I wasted fifteen cents, 16% of what I spent today, on a reviewer model that twice returned nothing at all — it spent its entire allowance on invisible internal reasoning and produced no text. The first failure was bad luck. The second was not: this project wrote down a rule after the same model did the same thing on four earlier occasions — change the model, don't raise its allowance — and I raised the allowance. The rule was right and I didn't read it.
And a correction I made mid-session, on the reviewer's advice, quietly made two of the run's own safety checks impossible to pass: I raised the minimum number of judgments a check needed to 24 at the same time as I halved the number of judgments two of them could collect. That is arithmetic I could have done before spending anything, and there is a standing note in this project saying exactly that.
Cost, and what's next
Ninety-eight cents, of which the fifteen above bought nothing. Fourteen of 210 model replies came back cut off mid-sentence and were discarded; recovering them by hand afterwards changes no figure by more than three hundredths. All the verification passed: 113 independent checks with no failures, and six deliberate corruptions of the data all caught.
The two-session study of what realia handling does to a reader closes here, with its conclusion written into the project's list of what "good" can mean and into the draft handbook. The obvious successor is the magnitude question this run could not touch: not every item anglicised, but one, then two, then four, to find where the effect starts. Today's evidence says the floor is very low — one word in three hundred and forty-seven.
Nothing needs your attention.