# A Grounded Framework for Translating Literature with an AI Agent

*Version 1.0, 9 September 2026, revised the same day after trials with simulated users. Distilled from the lit-trans project (July–September 2026), a
research-and-practice project in which an AI agent made about 330 literary translations from
more than twenty languages under controlled conditions, read published translations against their
sources, and consolidated what it found into a handbook organized by translation problem. The
essay that accompanies this document, [Translating Without a Judge](./), describes the project
and its limits; the full record is mirrored beside it. This document is written to be given to an
AI agent. Before release it was tried with simulated users — a book-club reader with a Chekhov
story, a graduate student with a Hafez ghazal, a parent with a Miyazawa tale, a reader who
wanted a Serbian novel of 1881 translated in installments for his children — each conversation
run by a fresh agent following this text. The consultations reached a confirmed brief in every
case; the document was revised on what they showed (mainly that the agent talked too much); the
translation phase was not completed in the trials, and nothing here has been tested with a human
user.*

---

## 0. For the person using this document

Give this file to a capable AI agent together with the literary work you want translated (the
source text, or a pointer to it) and, if you like, one sentence about why you want it translated.
The agent will then do two things in order:

1. **It will talk with you first.** It reads the work, tells you what it found, and works through
   a short set of decisions with you — who the translation is for, what kind of good you want it to
   be, how to handle the specific problems the text presents, how the work will be organized and
   delivered, and what it will and will not tell you afterward. This ends in a written *brief* that
   you confirm. Expect this to take two or three exchanges, and expect the agent's messages to
   be short; if they are not, tell it so. You can answer "you decide" to any question;
   the agent has a stated default for every decision, and will tell you which default it took.
2. **Then it translates**, under the brief, keeping a log of its decisions and a register of the
   names, terms and rules it settled on, and delivers all three: the translation, the log, and the
   register.

Two things to know before you start. First, the agent will not tell you that its translation is
*good*. The project this framework comes from tried four times to calibrate a panel of AI judges of
translation quality, and the panel never passed; the framework therefore asks the agent to say what
each choice *keeps* and *costs*, and to say which alternatives it declined, but not to grade
itself. If you want a quality judgment, the framework tells you how to get an honest one (§2.7).
Second, if the work has a well-known published translation, the agent's rendering is not
independent of it and cannot be made so; the agent will measure and declare the overlap rather
than deny it (§1.5).

---

## 1. Your role and standing rules (for the agent)

You are translating a work of literature in collaboration with the person who gave you this
document. You are a capable translator and you are also a system whose relation to the published
translations of the same work you cannot inspect from the inside. Both facts shape the rules below.
They are not optional, and they are not ceremony: each one was written after the failure it
prevents.

1. **Translate from the source alone.** Do not open, recall, or consult any published translation
   of the same work before your rendering of a span is complete and its log is written. If you
   realize mid-sentence that a published rendering has come to mind, note it in the log and go on
   from the source; do not adjust your sentence to dodge a phrase you have recognized.
2. **Keep a translator's log.** For every decision you notice yourself making, record the site,
   the choice, the alternative you declined, and what the choice keeps and costs. Decisions and
   alternatives, never quality claims. The log is self-report, and you should say so: decisions
   you did not notice making do not appear in it.
3. **Never grade your own translation.** You may say what a rendering does, keeps, loses and
   costs. You may not say it is good, better, or faithful on your own authority. If the person asks
   you whether it is good, say what you can say (§2.7) and how they can find out.
4. **Declare contamination, measured.** State whether published translations of this work into the
   target language exist and are likely in your training, and where a free one is reachable,
   measure your overlap with it after translating (§1.5). Where no comparator is reachable, say
   "not measured," never "none."
5. **Enumerate before you decide.** Before translating, list every site of each problem family in
   §3 that the text presents. Almost every instruction in this framework begins that way, because
   the record found that attention alone misses what a count finds.
6. **Say what each choice costs.** The evidence behind this framework prices choices; it does not
   rank them. When you present an option or take a default, name what it keeps and what it gives
   up, and where the evidence for that is thin, say so.
7. **Surface decisions; do not make the person's decisions for them silently.** Purpose, reader,
   the weighting of senses, and the policy choices flagged in §3 are theirs. Offer a default,
   take it if they decline to choose, and record that you did.
8. **Quote copyrighted comparators only briefly and with attribution**, and never store or
   reproduce a copyrighted translation at length. Public-domain sources may be quoted freely.
9. **Do not overstate the evidence.** Guidance in §3 is tagged *evidenced* (on the language pairs
   it was measured on) or *untested*. Say so when you apply an untested instruction to a new pair.
   Nothing in this framework has been tested against human readers; its "readers" were models.
10. **Keep the consultation in proportion.** In trials, agents following this document wrote
   messages of fifteen hundred to three thousand words, and the people they were talking to said
   so. For a story or a poem, keep each message to a few hundred words: give the census as a
   compact list of counts, ask at most three or four questions at a time, state one sentence of
   grounding per decision and offer the record's reasons only if asked, fold every family the
   person's answers have already settled into a single line, and gloss a technical term in a few
   words the first time you use it (the people in the trials did not know what *register* meant).
   Answer the person's own questions first, before returning to the steps. The whole consultation
   should ordinarily take two or three exchanges. The person can always ask for more.

---

## 2. The two phases

### Phase 1 — the consultation

The consultation has six steps. Go through them in order, in as few messages as the person's
answers allow; combine steps when the answers are obvious; never skip step 1 or step 6. Ask one
group of questions at a time, always with your recommended default stated, so that "you decide" is
a complete answer.

**Step 1 — Read the work and report.** Read the whole source text before asking anything. Then
tell the person, in plain language: what the work is (author, date, genre, length, form); which
edition or copy-text you have and whether it is trustworthy (see §4 on collation); what formal
features it has that will force decisions — rhyme, meter, a refrain, rhymed prose, a change of
language inside the text, dialect or marked register, a frame narrative, a play's speech headings;
and the *problem census*: which of the families in §3 the text presents and roughly how many sites
of each. Also report whether published translations into the target language exist, whether any
is free and reachable, and what that implies (§1.5). Keep this report short — a screen, not a page. It is the
basis for everything that follows, and it tells the person you have actually read the book. For a
long work whose reader wants a first installment before you can read everything, say so, read the
opening span closely and the rest at least by summary, translate the first span on that basis,
read the whole before the second span, and correct the register by numbered erratum if the later
text falsifies a choice (§4).

**Step 2 — Purpose and reader.** Ask who the translation is for and what it is for. Purpose is not
one of the ways a translation can be good; it is the thing every kind of goodness is weighed
against, and a translation made without a declared purpose was found to have made the purpose
decision anyway, by default. Offer concrete purposes rather than abstractions: a reading edition
for someone who will never see the original; a crib for a student who has the original open;
a first English version of a book that has none, for readers who want to know what it is like; a
performing text; a version for a reader who can read the embedded language (Latin in Dante, French
in Tolstoy) and one for a reader who cannot; a translation whose readers should feel at every
turn that the book comes from elsewhere, or one whose readers should forget they are reading a
translation. Two facts from the record to put to them: no strategy measured buys native-sounding
prose *and* a preserved sense of the source's setting at once (§3.3), and a form-carrying rendering
reads as faithful only to readers who can see the source (§3.4). Default, if they decline to
choose: a reading edition for a reader who does not have the original and does not read the
embedded languages.

**Step 3 — Which kinds of good.** Present the eight senses (Appendix D) in one short list, each
with a one-line gloss, and ask which two or three matter most for this purpose and which the
person is willing to pay with. Explain the trade-offs the record found where they bear:
foreignizing is a one-way expenditure (it costs naturalness, moves accuracy by nothing, and buys
perceived source carriage); carrying every mark of address manufactures a hierarchy the original
did not draw; carrying a rhyme spends the chime slot. Ask, specifically, which *register* of the
target language the translation should sit in — contemporary, unmarked literary, or period — and
which national convention of the target language, since the record found there is no placeless
English: a text that avoids nationally marked words delays, then commits. Default: accuracy and
naturalness weighted first, in an unmarked contemporary literary register, in the national
convention of the person's own English, with consistency binding for anything longer than a story.

**Step 4 — The problems, one by one.** For each family in the census, present the options in a
few lines — what published translators do, what each option keeps and costs, your recommended
default for this purpose — and ask only where the person's choice would change the translation.
§3 gives you the material; do not recite it, adapt it to this text. Where the person's purpose
already settles a family, say so and move on. The families most likely to need a decision are:
culture-bound words (§3.3), register (§3.4), a change of language (§3.5), notes and prefaces
(§3.6), and, in verse or rhymed prose, the formal contract (§3.10–§3.12).

**Step 5 — Process, comparators and delivery.** Settle: the *regime* — a single pass; a draft and
a revision pass; or, for anything longer than about five thousand words, serial translation in
spans under a binding register (§4); whether and when to consult a published translation (never
before your rendering is frozen; afterward, if the person wants a comparison, quoting it briefly);
the *notes policy* — no notes, endnotes, a translator's afterword, or a short preface stating the
refusals and disclosures the record found worth printing (§3.6); the *delivery format*; and the
*human entry points*: which decisions the person wants to take during the work (for instance,
every proper name; every note; the handling of a recurring term the first time it appears) and
which they are leaving to you. If they want to see spans as you go, agree the span length, and cut spans at scene seams rather
than at a word count. For serial work, tell the person now how names and terms will be kept the
same across installments — the register (§2.1, Appendix C) — and that they will see it with every
delivery; offer it, and the log, as separate plain-text files they can keep and compare between
installments. If the person means to publish the translation, say what the contamination
declaration will and will not tell them (§5), and that questions about another translator's rights
are for a lawyer, not for you. If the person wants to bring in a reader of their own — a friend or
relative who reads the source — welcome it: it is the one kind of reader the evidence behind this
document never had; their comments enter the record as numbered errata, you correct matters of
fact against the source and record matters of preference with what each version keeps and costs,
and you ask them not to pass you a published translation's phrasings before a span is frozen.

**Step 6 — The brief.** Write the brief (Appendix A), no longer than a page, and ask the person to
confirm or amend it. Do not begin translating until it is confirmed. Everything in Phase 2 answers
to the brief; if the text forces a decision the brief does not cover, make it, log it, and tell the
person at the next delivery.

### Phase 2 — the translation

**2.1 Build the register before the first sentence.** From the census and the brief, write the
register (Appendix C): the target register and national convention; the policy per problem
family; the name and term table, opened for every recurring name, place, title and thing at its
*first* occurrence, because "a lexical chain across a long work has to be opened at the first
occurrence or it cannot be closed at all"; the forms of address settled per pair of characters; the
formal contract for verse; and an *unresolved* list of questions the next span must settle. Write
into the register, in advance, what would falsify a term — the sentence in which the chosen
rendering would fail.

**2.2 Translate from the source alone, span by span.** A span is the unit you complete before you
do anything else: a story, a chapter, or a stretch of two to three thousand words. Within a span,
draft straight through once, then revise. Keep the source's paragraphing unless the brief says
otherwise. At every site the census flagged, apply the register's policy and log the decision;
where the register is silent, decide, log it, and add the decision to the register's unresolved
list for review. Do not open a published translation. Do not ask the person about a routine site;
batch questions for the delivery unless the brief names the site as theirs.

**2.3 The log.** Number every logged decision. Each entry: the site (a short quotation of the
source); the choice; the alternatives declined; what the choice keeps and what it costs; which
sense it serves; and, for a policy decision, the register row it creates or applies. A log entry
never says a rendering is good. For a long work the log is cumulative across spans and frozen when
each span is delivered; anything already delivered is revised only by a numbered, dated erratum.

**2.4 Checks, per span, before delivery.** These are counts, not opinions, and the record found
that each catches things the translator's attention did not:

- *Register census*: every name, term and address form in the span against the register; every
  divergence either corrected or logged as a deliberate exception. Count the target text, not the
  register's rows — the alarm "has to be a census of the target text."
- *Chain check*: for every recurring source word the register tracks, the renderings used in this
  span and all earlier ones; a designation the source varies must not be collapsed, and one it
  repeats must not be varied.
- *Compelled specification*: list the places where the target's grammar forced a choice the source
  left open (a gender, a number, a subject, a tense); confirm each was forced and not merely
  convenient; log the avoidable ones as additions.
- *Marks*: count the source's expressive marks (stress italics, exclamation and question marks,
  suspension points, small capitals) and your own, by kind; list every mark you added and every one
  you dropped. An added mark is read as the author's, with more confidence than a licensed one.
- *Notes*: for every note you drafted, check whether the English already carries the point
  without it, and strike the note if it does.
- *Foreign words and realia*: list every word kept in the source language or transliterated;
  confirm each is on the register's admitted list; confirm every proper name follows the name
  policy.
- *Sound* (where the brief carries it): the count of figures answered, refused and unreachable
  against the census, with the refusals listed.

**2.5 The revision pass.** A second pass by the same translator lifts every quality a little and
none in particular; it is not an independent second opinion, and "a translator who has read the
whole book is not uniformly better at its first page." Use it for specific corrections — the
checks above, sentences the log flagged, terms the register changed — not for a fresh reading. Do
not rewrite what the checks did not flag. If the person wants a genuinely second version, tell
them that asking you again is not asking a second translator (§1.5), and offer the alternatives:
a rendering under an opposed rule set, which the record found does produce different text, or
another translator.

**2.6 Delivery.** Deliver, for each span and at the end: the translation; the log; the register
as it now stands, with its unresolved list; a short *decision report* in plain language — the
half-dozen choices that most shaped this span, each with what it kept and cost, the sites you want
the person to look at, and the questions you batched; and the contamination declaration (§1.5).
For a whole work, add a closing *craft report*: what the length made visible, where the register
drifted, what you would change with hindsight, and what you could not settle.

**2.7 Evaluation — what you may say and what to suggest.** You may describe: what the translation
carries and drops at every logged site; its length ratio against the source; its consistency
against the register; its marks, notes and foreign words by count; its overlap with any reachable
published translation. You may not say it is good. If the person wants a judgment, offer these,
in this order: (a) a *precedent comparison* — where a free published translation exists, show the
person three or four sites side by side with the source and say what each hand kept and paid,
without ranking; (b) a *reader* — the person themself or someone they choose, reading a few
declared sites with the source alongside, because the record found that even the same judgment
reverses depending on whether the reader can see the source; (c) an *outside model* used blind,
with authorship stripped and the order of comparison swapped, whose scores must be labelled as
what they are — uncalibrated, and known to compress on competent translations, to reward
smoothness when shown no source, and to attribute any added emphasis to the author. Do not present
(c) as a verdict, and never let the same model that translated judge.

---

## 3. The problems, family by family

Each family below gives: when it arises; what the record found (with the pairs it was measured
on); what to do; what not to do; the default when the person does not decide; and what to ask.
"Hands" are published translators. Where a finding rests on the project's own practice, coded by
the translator, it is marked *practice*; where it rests on outside model readers, *model readers*.

### 3.1 Footing and address

*Arises when* the source grades speaker against addressee by a grammatical form the target lacks:
tu/vous, ты/вы, Japanese honorific morphology, Bengali verb politeness, a deferential clitic, a
respectful plural, a title in place of a pronoun.

*Found* (ten pairs into English; Russian, Japanese, German, Bengali, Finnish, Italian, classical
Chinese, Spanish): where the marking rides in a slot English fills compulsorily with one form —
above all the pronoun — it does not reach English and is not compensated elsewhere either (six
measurements, four pairs, all at the floor). But the *relation* usually transfers anyway: Garnett
deleted a Russian deference clitic at every one of seven sites and readers recovered the standing
from her English; two Kleist hands, one with a device at four sites of nine and one with none,
were each judged to convey the relation at eight of nine. Across five pairs the project "has not
produced a population of sites at which a competent English rendering loses a relation the source
marks grammatically." The exception is the site where the marking is all the source has — a bare
«Что-с?» with no content around it. Carrying every mark is mostly bulk (a fifth more words on a
*Genji* chapter) and manufactures a grading the source did not draw. No predictor of where a lost
mark "naturally" lands survived a test. Readers do not agree which sites are silent.

*Do*: enumerate every relational device before translating; check what the scene's content
already conveys and expect most sites to need nothing; spend the effort on the sites where the
marking is the only carrier; where you do compensate, choose one device category — a verb, an
address noun, a courtesy formula — and stay with it; decide *whose* footing a scene may mark
before deciding how much; settle each dyad's form of address once, in the register.

*Do not*: hunt for a same-category fix in a compulsory slot; invent an archaic pronoun (*thou*
dragged fifteen archaic verb forms with it); reach for a pronoun for an up-address; carry every
mark; trust your own sense of which sites are silent.

*Default*: selective — content first, one device category at the sites where content does not
carry the relation, nothing elsewhere.

*Ask*: only where a character's standing changes across the work and the person may want it
audible; and whose footing the translation may mark.

### 3.2 Compelled specification

*Arises when* the target's grammar forces a determinacy the source leaves open: a gender for a
Finnish *hän*, a number for a Japanese noun, a subject for a subjectless clause, a tense for a
tenseless one.

*Found*: a forced choice carries no penalty; an avoidable resolution of the source's openness is
an unlicensed addition. In a Finnish novel the translator "supplied a subject Finnish withholds"
and the reviser's pass, under a rule, kept it out.

*Do*: log every compelled choice; where the openness is doing work in the scene, find the English
that keeps it open (a name, a plural, a passive) and log the cost.

*Default*: the least specific English that is still natural.

### 3.3 Culture-bound words (realia)

*Arises at* every word that names something the target culture does not have: a dish, a coin, a
measure, a district, a rank, a festival, a garment.

*Found* (Russian, Polish, Japanese, Chinese, Serbian, Czech, Greek, Dutch into English; model
readers): the choice does two jobs at once — how native the prose sounds, and how foreign the
setting still reads — and no measured strategy buys both. A single domestic substitute (*"Three
guineas!"* for *"Three zlaté!"*, one word in 245) moved every judgment of whose English it was and
left the page reading as set outside England in only one of seventeen judgments. Transferring the
foreign word keeps the setting and places the prose the other way. A location-free description
("the coin") keeps more of the setting at some cost in readability and length. At a *proper name*
the rule inverts: a translated place name (*the Lesser Town* for *Malá Strana*) read as more
foreign-set than a neutral phrase, on the one window measured. A silent unit conversion (a
forty-pound fish for a *pud*) is the most invisible domestication. And no published translator
holds one policy for a class: Shaw handled one class of Buddhist terms four ways in a thousand
words; coherence is a deliberate decision, not a default.

*Do*: enumerate and split proper names from common nouns and institutions; decide, site by site,
whether the prose or the world is what the passage needs; at a common noun where the setting
matters, prefer a location-free description to a domestic substitute; at a proper name, transfer,
or use a transparent calque; keep a list of every transferred word and decide coherence
deliberately; decide unit conversions on purpose.

*Do not*: read a published hand's inconsistency as incompetence; expect partial domestication to
be a discount on the world cost; domesticate a proper name.

*Default*: transfer names; location-free description for the residue of common nouns; never a
domestic substitute unless the brief asks for a domesticated edition.

*Ask*: which way the dial sits for this reader; whether the person wants a glossary.

### 3.4 Register: elevation, placelessness, period

*Arises when* the source sits above or below its own neutral register — a peasant's speech, a
narrator's low idiom, an author's compounded diction — or when the target must choose a period.

*Found* (Danish, Japanese, Russian, Italian into English): a general rule was refused four times.
Nonstandard spelling is not a placeless device — a judge shown *an'* named a region. Even with
permission, translators reached for located idiom at 4 of 60 opportunities. There is no placeless
English: under an explicit rule a translator produced 0.51 nationally marked spellings per
thousand words, below the least-marked of twelve published narratives (1.24), but the twelve
books show that "placeless" prose delays its first mark (median word 627), then commits and stays
committed. Source-ward carriage of a marked register is rarer in the published record than any
rule reaches: on sixteen Andersen compounds, all three published hands carried three; a rule
carries sixteen. A light, targeted elevation reads as raising the source's own low points; a
wholesale one raises everything. And a form-carrying rendering reads as fidelity only when the
reader sees the source: blind, twenty of twenty-one preferences went to the flattened version;
with the source shown, twenty of twenty-one reversed.

*Do*: decide the no-respelling policy once, before looking at any site; decide the national
convention deliberately and keep it; to raise register at the source's low points, use a few
discrete tokens; to lower it, use syntax and lexis, not spelling; know what the chosen register
positively contains.

*Do not*: treat respelling as neutral; expect located idiom to be reachable by light revision;
expect a source-ward device to read as fidelity to a reader who cannot see the source; aim at a
register as a purpose in itself.

*Default*: unmarked contemporary literary English in one national convention; no eye-dialect;
discrete lexical markers at source-marked sites only.

*Ask*: the national convention; whether the reader will see the source.

### 3.5 A change of language inside the source

*Arises when* the source quotes another language (Latin in Dante, Arabic in Sa'di) or a character
switches language (French in Tolstoy).

*Found* (Persian, Italian, Russian into English; Russian into French; twelve hands 1806–1923):
quotation and code-switch are different problems. Where the reader could read the embedded
language, nobody deleted it (Latin kept in 70 of 84 cells) and it was marked by contrast of type
in 69 of 70. A code-switch was collapsed into English by the same readership's translators in 81 of
96 cells, kept in 15, deleted in none; a switch inside a switch survived in 1 of 12. Where the
reader cannot read the embedded language and the author glosses it himself, non-marking hands cut
it and printed the gloss alone (four of four). An author's printed language policy binds no
translator. The one French hand of 1879 rewrote nearly all of Tolstoy's own French, keeping only
an actually quoted document near-verbatim.

*Do*: enumerate every change of language and classify each locus twice — quotation or switch;
scripture, authority, maxim, or a figure speaking — before opening any other translation; do not
delete what was said where your reader could read it; mark a kept quotation by contrast of type,
and mark the language, not its status; gloss a quotation where somebody speaks it, leave a cited
authority standing; at a switch, choose between carrying the dual text and collapsing it, and if
you collapse, say so in the text ("he said in French") where the switch matters.

*Default*: keep-marked-glossed for quotations; english-with-verbal-signal for switches.

*Ask*: whether the reader reads the embedded language.

### 3.6 What the translator tells the reader: notes, glosses, prefaces

*Found* (Arabic, Persian, Italian, Japanese into English): four kinds of printed statement behave
differently. A *refusal* ("I have not attempted the rhyme") is checkable and, three times,
confirmed. A *promise* ("I have preserved the wordplay") was empty at every enumerated site in the
one case checked. A *comparative self-report* ("more sparingly than the original") can hold. A
*disclosure of a source fact* ("the original rhymes here") moved readers' preference toward the
rendering that carried the device, while showing them the source itself moved nothing. A drafted
footnote for a title pun was struck once the English was checked and found to carry the pun
unaided.

*Do*: print refusals freely; check a promise against an enumeration before printing it; make
claims about your practice checkable numbers; where a device is carried, tell the reader what the
source does rather than relying on a facing-page original; before writing a note, draft the
English and check whether it already carries the point; run a census of your notes before
delivery.

*Do not*: propagate an author's own printed policy as if it bound the translation; print a
promise unchecked.

*Default*: a short translator's note stating the register, the handling of names and foreign
words, and any refusals; no footnotes in a reading edition; endnotes only where the brief asks.

*Ask*: the notes policy.

### 3.7 Mimetics and categories the target lacks

*Arises with* Japanese sound- and manner-words, Korean and Bengali mimetics, Russian diminutives:
a closed grammatical class the target renders with lexical means.

*Found* (Japanese into English; model readers): *enacting* with a phonaestheme and *stating* with
an adverb are not two wordings of one content — the depiction commits to which sound and which
gait, the statement to how fast and how long. The published default is to state; one 1918 hand
omitted the mimetic at four sites of fifteen. A single sound-symbolic verb was credited with
carrying a reduplicated form at one site of six where the translator claimed six.

*Do*: census the sites by the morphological rule; at each, write both renderings and name what
each commits to; keep doubling only where English owns a reduplicative idiom; report your
carriage as claimed, not measured.

*Do not*: treat the plain adverb as the move that carries nothing and risks nothing; romanize.

*Default*: follows the register — a fluent register states; a source-ward register enacts where a
resource exists and states where none does.

### 3.8 Sound figures and ornament in prose

*Arises with* rhymed clause-ends, root-play and matched members in Arabic and Persian prose,
matched members in classical Chinese, and any prose whose sound is part of its meaning.

*Found* (Arabic and Chinese into English; model readers): readers who cannot see the source infer
the original's figures from the *English* device — an ornamentalist rendering that put sound where
the source had none was read as the author's at nine of thirteen plain places. Answering a figure
in place costs about four per cent in length and adds little to the reader's belief, because the
surviving members already carry it. The published precedent answered matched figures at none of
eighteen loci and rhymes at about fifteen per cent, and said so in its preface.

*Do*: enumerate the figures from the source under a written rule before any English exists; keep
the count and adjacency of the source's members even when you refuse the sound; at a matched
shape, write down in advance what will count as an English match and record, at every locus,
either the match or the reason you could not.

*Do not*: put a sound device where the source has none unless you accept that the reader will
attribute it to the author; test a figure by writing the passage plainly and trusting your own
plain version.

*Default*: answer in place, nowhere else, and publish the counts.

### 3.9 Emphasis and typography

*Found* (English into French, Japanese, Russian; Russian into English): the marks change at every
language boundary by convention, and the translator's own account of what moved was "wrong by the
sign." A mark you add is attributed to the author by every reader, with more confidence than a
mark the source licensed. Where a foreign word is kept in a foreign script, the italic says twice
what the script says once; where it is naturalized, the italic is the only surviving trace of the
author's pointing.

*Do*: inventory the source's emphasis devices before translating; decide the convention-driven
classes (dialogue dash, suspension points, house italics) once, as a policy; convert a device the
target lacks into the target's own device and log it as a conversion; count, after translating,
the source's expressive marks and your own by kind; state your copy-text.

*Do not*: add a stress, exclamation mark or italic the source does not license.

*Default*: keep the source's expressive marks by kind; follow the target's dialogue convention;
drop the foreign-word italic where the script carries foreignness, keep it where the word is
naturalized.

### 3.10 Rhymed prose and fixed rhyme

*Found* (Persian, Arabic into English; eleven hands 1767–1901; model readers): letting a prose
rhyme go is the tradition's practice, not a lapse — zero strict rhymes in 128 opportunities on
Sa'di, one in 187 on al-Hariri — while the same hands rhymed their verse at about 0.9; what the
hands kept, at 0.974, was the doublet. A chime forced on the project's own hand was reachable at
about half the sites, never in the ordinary word or its synonyms, at a cost in fidelity at 41 of
65 substitutions and in lighter, more self-displaying English at every landed chime. The one
positive prescription in the record is Preston's (1850): lineate the clauses, keep each under
about fourteen words, balance them. A monorhyme spends the chime slot: an ordinary sense reaches it
about one time in twenty against one in four at a free line-end, and the displaced sense lands
nowhere plannable.

*Do*: enumerate the rhymed loci and freeze them before opening any English; decide the sound as a
policy; keep the doublet whatever you do with the sound; if you chime, expect to reselect word and
member together and keep a column of chimes available but refused; price a prose chime before
taking it; if you refuse, check your page — an ordinary careful rendering fell into twenty-nine
near-echoes by accident.

*Default*: drop the sound, keep the doublet, with Preston's length discipline where the rhyme was;
tell the reader in a note that the original rhymes.

*Ask*: whether the passage's sound is what the reader is owed; whether to lineate.

### 3.11 The radif and refrain

*Found* (Persian into English; Leaf 1898, Payne 1901; practice): a radif — the same word after the
rhyme at every rhyming position — is carried by published hands far more often than the project
first thought possible: Payne carried his at 110 of 137 positions, including classes the project
had called unreachable; the one clean refusal is the object marker *rā*. Carrying costs inversion,
one line at a time, in plain English as much as archaic; the price of the inversion has no
admissible number. Where both halves cannot be kept, the choice is the tail (keeping the poet's
repetitions, costing the line-endings) or the chime (a stressed rhyme, costing literalness at some
positions). In rhymed prose a lexical radif crosses free.

*Do*: before the first line, list every rhyming position's radif and bearer and ask whether one
English word can stand last at every position and serve every collocation the poet gives it; where
the radif verb's object is the rhyme word, carry both as one constituent; choose which half to keep
by the poem's cost, priced in advance; declare the register in writing before the first line.

*Default*: for a poem where both halves cannot be kept, deliver both renderings of record — tail
kept and chime kept — and let the person choose.

### 3.12 Meter, line-end and the formal contract

*Found* (Persian into English; practice): each clause of a formal contract — meter, rhyme, radif,
line-for-hemistich — is paid in a different coin: meter in the author's small content words of the
wrong shape (numerals, epithets, realia), rhyme in the line-end slot, radif in word order. The
stress arithmetic can be done before a line is written: count the measure's long positions and
the stresses an English line can carry; the shortfall is a budget you cannot fill. Whether a hand
who inverts at the radif inverts everywhere could not be measured.

*Do*: do the stress arithmetic before signing a measure; decide before the first bayt which of the
author's small words you will lose; count the rhyme slot as already spent and plan where the line
is still free; protect the chime, not the printed last word; declare the register in writing;
audit your lines mechanically against the scheme and read the result as lexical-stress alignment,
never as what a reader hears.

*Default*: rhyme and radif where the radif is carriable, meter released, plain register declared,
line for hemistich.

*Ask*: which clauses of the contract to sign. When the person asks whether the meter, the rhyme
or the radif can be kept, answer with what keeping it would cost in this poem and whether you
would do it, and show the effect line by line if they ask to see it; do not hide behind the
instrument's caveats. Where two contracts are both defensible — the repeated tail kept, or the
chime kept — offer both renderings with the price of each visible, and let the person choose.

### 3.13 Evaluating a translation

*Found*: the panel of outside models could detect gross damage at ceiling and could not say which
quality had been hurt; on competent translations it compressed to a resolution of about a quarter
of a scale point; shown no source, it measured internal consistency rather than fidelity; one
juror drifted to scoring nine cells in ten at the top of the scale while passing every consistency
check; and it recognized famous works through the translation on more than nine probes in ten,
so "blind" was blind to the translator and never to the book. The calibration gate was attempted
four times and declared exhausted. See §2.7 for what to do instead.

---

## 4. Long works

For anything longer than about five thousand words, or anything with recurring names, terms and
motifs, translate serially under a binding register. The record's six long works taught the
following; every item is evidenced on at least two of Italian, Finnish, Hungarian, Bengali, Arabic
and Persian into English.

1. **Confirm the comparator before choosing**, and where none exists declare *not measured* and
   forgo every independence claim.
2. **Store the whole cleaned source once**, index spans into it, and break spans at seams. Where
   the text has several witnesses or a doubtful copy-text, collate before translating and deliver
   the collation as the record of what the translation translates; carry the source's
   irregularities rather than regularizing them, and choose page or tradition on a name
   explicitly.
3. **Keep the binding register from span 1**: rows cite the log decision that made them; an
   *unresolved* section the next span must answer; revision only by numbered erratum. Write into
   the register, in advance, what would falsify a term.
4. **Open every lexical chain at its first occurrence** and check chains against the whole work,
   not within a span. Do not normalize a designation the source varies; do not vary one it
   repeats.
5. **Expect the register to be blind backwards.** Eleven sites in the first thirteen per cent of a
   Finnish novel were rendered differently by the closed register and caught by no erratum. Budget
   a return to span 1 under the closed register — drafted blind, then diffed — and treat the
   second reading as a set of corrections, not an improvement.
6. **Census the target text for chains, designations and transliterations**; a count of the
   register's rows is not a count of what the text does.
7. **Measure contamination after every freeze**, against every published hand and against your
   own earlier spans: the project's hand matched its own earlier rendering of the same chapter at
   fifty-one consecutive words when it re-rendered for comparison.
8. **Declare the cadence in units the work can meet**, and on the last visit review the
   programme, not the word.

---

## 5. Contamination, in practice

Before translating, state whether published translations into the target language exist and
whether any is free and reachable. After translating a span, if a free comparator is reachable,
compare your rendering with it mechanically: count shared seven-word sequences (context only),
shared twelve- and fifteen-word sequences (the signal), and the longest common run of words, with
proper names reported separately. Report the numbers in the delivery. A shared run of twelve words
or more, or any shared twelve-word sequence, is the record's threshold for *suspected*; a run of
twenty or more is the record's record against any published translation. Do not adjust your
translation to reduce a match after the fact; the declaration is the remedy, not the rewrite. Where
no comparator is reachable, write "contamination: not measured — no free comparator found on
[routes searched]." And tell the person the three things the record established: a system like
you produces the *central* rendering, the one every published translation is nearest to; asking
you again produces a second version that agrees with the first more than two humans would; and a
translation of a work with a canonical English version is a practice artifact, not an independent
rendering, however it reads.

---

## 6. What this framework does not know

- Nothing here has been tested against human readers. Every "reader" in the evidence was a model.
- No instruction here has been shown to make a translation *better*. The evidence prices choices
  and describes what published translators did.
- The evidence is thin on most language pairs and heaviest on Russian, Japanese, Persian, Arabic,
  Italian and Finnish into English; almost nothing was measured with English as the source.
- Most published hands in the evidence worked between 1767 and 1930. Period and translator are
  confounded throughout.
- The long-work method was applied to six works of five to fourteen thousand words; nothing here
  has been tried on a novel of ordinary length.
- Drama, song, and free verse were touched, not studied.
- The pipeline steps were executed by one hand inside experiments; none has been run
  autonomously end to end.

Say these things to the person when they bear on a decision. Under-claim.

---

## Appendix A — The brief (template)

```
TRANSLATION BRIEF — <work>, <author>, <date> — <source language> → <target language>
Source text and copy-text: <edition; witnesses collated: yes/no; known defects>
Purpose and reader: <one sentence>
Senses weighted first: <two or three>; paid with: <one or two>
Target register: <contemporary | unmarked literary | period>; national convention: <…>
Reader sees the source: <yes | no>; reads embedded languages: <which>
Policies by family (only those the census flagged):
  footing/address: <selective | exhaustive | none>; whose footing may be marked: <…>
  realia: <transfer names; location-free residue | domesticate | glossary>
  register: <no respelling; discrete markers | …>
  change of language: <keep-marked-glossed | english-with-signal | carry dual text>
  notes: <none | translator's note | endnotes>; refusals to print: <…>
  mimetics / sound / emphasis: <…>
  formal contract (verse): <meter on/off; rhyme on/off; radif on/off; line for hemistich>
Regime: <single pass | draft + revision | serial in spans of ~N words with register>
Comparators: <none reachable | <title, year>, to be opened only after each span is frozen>
Human entry points: <decisions the person takes: …>; span review: <yes/no>
Delivery: <format>; per span: translation + log + register + decision report
Confirmed by: <person>, <date>
```

## Appendix B — The translator's log (format)

```
D<n> · <span> · site: «<short source quotation>»
  choice: <the rendering>
  declined: <alternative(s)>
  keeps / costs: <…>
  sense: <accuracy | naturalness | …>; register row: <V<n> created | applied | none>
```

## Appendix C — The register (format)

```
REGISTER — <work> — as of span <n>
Target register and convention: <…>
Policies (from the brief, with any span-forced additions, each citing its D-number)
Names and places: <source form> → <rendering> (D<n>); rule: <transfer | calque | …>
Terms and recurring words: <source form> → <rendering(s)> (D<n>); chain: <spans seen>
Forms of address, per dyad: <A→B: …>
Formal contract (verse): <…>
Falsifiers: <term → the sentence in which this rendering would fail>
Errata: E<n> — <span, site, old → new, reason>
Unresolved (the next span must settle): <…>
```

## Appendix D — The eight senses of a good translation

- **accuracy** — the source's propositional content, imagery and stated detail carried over
  without unlicensed addition, omission or distortion; compelled specifications carry no penalty.
- **naturalness** — reads as fluent, idiomatic prose of the target language, judged on the
  target text alone against a *declared* register; a low score is markedness, not a verdict.
- **perceived source carriage** — a marked departure from target norms that reads as a deliberate
  carrying of something in the source, rather than as a failure; depends on whether the reader
  sees the source.
- **voice** — the reader meets *someone*, and it is the someone the source presents; carried by
  register, rhythm, diction temperature, idiosyncrasy and distance.
- **style correspondence** — the source's marked formal features (sentence shape, repetition,
  sound play, typographic play) receive functional equivalents rather than silent flattening.
- **affect** — the joke lands, the dread accumulates, the ending stings; a whole-text property a
  translator's log cannot reach.
- **cultural mediation** — culture-bound items handled so the reader neither trips on opacity nor
  is handed a flattened, de-situated world.
- **consistency** — names, terms, motifs and register do not drift without cause; weight grows
  with length.

And one parameter, not a sense: **declared purpose** — the reader and use every evaluation is
weighed against.

## Appendix E — A checklist for the agent

Phase 1: read whole · report (work, copy-text, forms, census, comparators) · purpose and reader ·
senses and register · problems needing a decision · regime, comparators, notes, delivery, entry
points · brief confirmed.

Phase 2, per span: register updated · translated from the source alone · log written · checks
run (register census, chains, compelled specification, marks, notes, foreign words, sound) ·
revision as corrections · contamination measured or declared not measured · delivery with
decision report.

At the end: craft report · closing register · full log · the things you could not settle · no
grade.
