Republished copy of Daniel Bourdeau’s Unsolved Historical Ciphers, released under CC BY 4.0. Text unchanged; hosted by JIC alongside its translations; not an official copy by the author · View the original ↗
REKUTATA UHOYAKATE SEKANAOKU SAKEKINISE
OSUKAKUCHI REUHONIUSU MAKASUSAKU FUREOHIHE
… ATAREUSASU OSOYUANIRA KISU

China / Japan · code in romanised kana · 1916

Huang Xing to Lin Hu and Li Genyuan

Huang Xing
fromHuang Xing

Relayed through the Japanese Foreign Ministry, 25 May 1916 · Minister Ishii to Consul Ōta at Zhaoqing · JACAR B03050731500

The Ministry filed a decode with the telegram, so the catalogue records it as “solved but specific scheme unknown”. The decode was transcribed and aligned with the kana; the alignment identifies the scheme and 42 codebook entries.

Daniel Bourdeau · posted · updated

Method: key recovered based on adjacent plaintext · Extent: partial

Summary. In May 1916 Huang Xing asked the Japanese Foreign Ministry to pass a telegram to Lin Hu and Li Genyuan in Guangdong. It went out in a cipher of the Chinese side's own making: 139 syllables of romanised Japanese kana, packed into ten-letter telegraph words, with no voicing marks. S. Tomokiyo later found a memo recording that the Ministry had filed a decode with the telegram, so the item is listed as solved; he could not transcribe the handwritten Chinese, and the encryption scheme stayed unidentified. Both are settled here. The decode and its annotated Japanese rendering are on frames 0247 and 0248 of the file, and the message is a short reply about the National Protection Army and a traveller's sailing date. The scheme is a kana code of a kind not previously described for this correspondence: three kana encode one character, deterministically, from a private codebook. A first analysis took the consonant row as the digit and the vowel as a free homophone; the frames, examined again at 500 dpi, overturn that (section 03). Extent: scheme identified, one transmission error located, plaintext transcribed from the decode, 42 codebook entries recovered; the kana-to-digit table and the rest of the book need a second telegram.

01 The file

JACAR B03050731500 is a Gaimushō volume on the anti-Yuan risings (1.6.1.75-1_015, reel 1-0955, frames 0244–0302). The covering despatch, 電送第一九八〇号, cipher, sent 25 May 1916 at 4.35 p.m. from Foreign Minister Ishii to Consul Ōta at Zhaoqing, says in plain Japanese what the enclosure is:

黄興ヨリ林虎李根源宛電報依頼アリタルニ付同人ニ伝達アリタシ 内容ハ先日来電ノ挨拶ナリ(以下支那側暗号)
“Huang Xing has asked that a telegram be sent to Lin Hu and Li Genyuan; please pass it to them. The content is an acknowledgement of the telegram received the other day. (The Chinese side's cipher follows.)”

Three frames matter: 0246, the telegram on an Imperial Government Telegraphs forwarded-message form; 0247, headed 電稿, the decoded Chinese in cursive; and 0248, a clerk's rendering of it rearranged into Japanese word order on Foreign Ministry paper, with his own glosses.

Imperial Government Telegraphs forwarded-message form with the words REKUTATA UHOYAKATE SEKANAOKU and so on
Frame 0246: the telegram as handed in. Five rows of ten-letter words, ending OSOYUANIRA KISU. Image: Diplomatic Archives of the Ministry of Foreign Affairs of Japan, 1.6.1.75-1_015 frame 0246, via JACAR Ref. B03050731500.
REKUTATA  UHOYAKATE  SEKANAOKU  SAKEKINISE  OSUKAKUCHI  REUHONIUSU
MAKASUSAKU  FUREOHIHE  YENIKIYEHI  TAKASUREU  SASHIKIHE  WAONIAAHO
SAYEHITE  UHINAKUTA  NEATEMUYE  SUHEUHAKA  YESOMIKANU  TEASUUKU
TSUHAKOSE  KEKININAKI  TEAKETASHI  KONUOHIKE  HIMIKANO  SHIKUSAHE
UHATAKASA  KEKINIKA  ATAREUSASU  OSOYUANIRA  KISU

Read as kana this is 139 syllables, 36 of them distinct. There are no voiced kana at all (no ga, za, da, ba), which is normal for telegraph traffic, where the voicing marks were simply not transmitted. The syllables are Japanese (SHI, CHI, TSU, and the old YE and WA), not the consonant-vowel letters of the condenser used in the Swatow telegram to Sun Yat-sen six weeks earlier.

02 The Ministry’s decode

Frame 0247 is the decode: a heading 電稿 and four columns of 12, 11, 12 and 11 characters: forty-six characters of message. Frame 0248 is the same text rearranged for a Japanese reader, and it is the more useful of the two, because the clerk added glosses where he was unsure.

Cursive Chinese decode headed 電稿, five columns
Frame 0247, 電稿: the decoded Chinese. Image: Diplomatic Archives of the Ministry of Foreign Affairs of Japan, 1.6.1.75-1_015 frame 0247, via JACAR Ref. B03050731500.
The same text rearranged into Japanese word order on Foreign Ministry ruled paper
Frame 0248: the clerk's Japanese rendering, with his glosses. Image: Diplomatic Archives of the Ministry of Foreign Affairs of Japan, 1.6.1.75-1_015 frame 0248, via JACAR Ref. B03050731500.

隱印崧蕘邕行諸兄鑒:正發書翰之際,適接哿電,敬悉一切。護國軍能速入湘贛甚好。章行嚴何日東渡?速令出發,並望預電。興 徑

“To brothers Yin, Yin, Song, Yao and Yongxing: just as I was sending a letter, your telegram of the twentieth arrived, and I have respectfully noted it all. It would be excellent if the National Protection Army could move quickly into Hunan and Jiangxi. What day does Zhang Xingyan cross to Japan? Have him start at once, and please wire me in advance. — Xing, the twenty-fifth.”

The clerk's glosses carry most of the identifications. Against the string of names at the head he wrote 林虎・李根源等ノ字ナラン, “probably the courtesy names of Lin Hu, Li Genyuan and others”; he was guessing too, and one of them is certainly 蔡鍔's 松坡. He glossed 湘贛 as 湖南江西, Hunan and Jiangxi, and 東渡 as 日本ニ来ル, coming to Japan. 章行嚴 is Zhang Shizhao, Huang Xing's close associate. Two characters are dates in the 韻目代日 rhyme-code that Chinese telegraphy used for the day of the month: 哿 = the 20th and 徑 = the 25th. 徑, the last character of the signature, matches the dispatch date on the covering despatch exactly. That agreement is the best single check that the decode belongs to this telegram.

Correction, second session. The running text above is the clerk's paraphrase from frame 0248 (50 characters). The cursive on frame 0247, examined at 500 dpi, has 46: 隱印崧誥邕行諸兄鑒正發書 · 間適接哿電敬悉一切護國 · 軍能速入湘贛甚好[章]行嚴何日 · 東渡望速啓行先電示興徑, with 章 an interlinear addition. That count matches the 46 kana triples exactly; see section 03.

Several glyphs in the cursive remain uncertain and are given here on the balance of the two frames. The column lengths are 2 + 12 + 11 + 12 + 11; Tomokiyo's note gives the same figures but sums them as 68, where they add to 48.

03 The scheme

Forty-six characters against 138 usable kana is three kana to the character, with a single kana left spare at the end. That settles the rate. The question is what part of each kana does the work, and the gojūon syllabary answers it by its shape: it is a grid of ten consonant rows by five vowels.

—
K
S
T
N
H
M
Y
R
W
A
KA
SA
TA
NA
HA
MA
YA
RA
WA
I
KI
SHI
CHI
NI
HI
MI
—
RI
WI
U
KU
SU
TSU
NU
FU
MU
YU
RU
—
E
KE
SE
TE
NE
HE
ME
YE
RE
WE
O
KO
SO
TO
NO
HO
MO
YO
RO
WO

If the consonant row carries a digit and the vowel is free, then a character repeated in the plaintext will repeat as a triple of rows but need not repeat as a triple of kana. That is a testable prediction, and the counts are unambiguous. Across the 46 groups:

readingrepeats observedexpected by chancepermutation test
consonant row71.03p = 0.006
vowel88.28p = 0.70
literal kana1——

The consonant rows repeat seven times where chance predicts one. The vowels repeat eight times where chance predicts eight: on this grouping they carry nothing. And the surface kana, the thing an interceptor actually sees, repeats once in forty-six. The mechanism is visible in the repeated groups themselves, the same character spelled two different ways:

S K N  →  SE-KA-NA  and  SHI-KO-NU
R — H  →  RE-U-HO  and  RE-O-HI
T — H  →  TA-U-HO  and  TE-U-HI
S K H  →  SA-KU-FU  and  SHI-KI-HE

That was the first analysis, and the frames correct it. Re-fetched and examined at 500 dpi, frame 0247 has four columns of 12, 11, 12 and 11 characters after the heading, 46 in all, the number of kana triples:

隱印崧誥邕行諸兄鑒正發書 · 間適接哿電敬悉一切護國 · 軍能速入湘贛甚好[章]行嚴何日 · 東渡望速啓行先電示興徑

The earlier 50-character plaintext above followed the clerk's Japanese paraphrase on frame 0248; the cursive has 46, and 章 is an interlinear addition. This plaintext repeats 行 three times (positions 6, 32, 41), 速 twice (26, 39) and 電 twice (17, 43). Under the grouping above only one of those pairs matched. Drop one kana at position 106, the O of OHIKE at the start of character 36, and every triple from there on shifts by one: 行 becomes KE-KI-NI all three times, 速 HE-U-HA both times, 電 RE-U-SA both times, and the spare SU at the end completes the last character. The telegram as filed carries one superfluous kana, and the code is deterministic: the same character is always the same three kana, vowels included.

The seven consonant-row collisions in the table are therefore not one character written with different vowels. After the shift they are pairs of different characters (印/護, 誥/日, 鑒/間, 書/敬, 國/嚴, 甚/何, 諸/示) that share a row triple and differ only in vowels. The vowels carry information. What survives of the first analysis is the rate, three kana to the character, and the fact that the kana's row is correlated with its code value; the excess of row collisions over chance is what that correlation looks like from outside. Three kana over a 50-kana syllabary would address 125,000 entries, far more than any private book; the natural interpretation is still a three-digit code with each digit written as one of a set of kana, the sets not being the rows.

One test of that was run and failed. If the rows were the digits of a codebook compiled in dictionary order, some permutation of the ten rows would make the 42 distinct characters' code numbers monotone in standard telegraph-code order. Over all 3,628,800 permutations the best Spearman rho is 0.524; the same search on twelve shuffles of the plaintext gives a null of mean 0.468 and maximum 0.614. No signal.

04 What is still open

A second telegram in the same code would settle both. What this one settles is the scheme, the plaintext, and the transmission error that hid the repeats.

05 Sources