(9 hours ago)chenxiang Wrote: You are not allowed to view links. Register or Login to view.In science, this is called "constantly patching a theory in order to save it"
But that is in fact how most scientific theories are developed!
They are not conceived from the start in the clean form that you read in textbooks and scientific journals.
Consider someone trying to decipher a medieval encrypted document. He may first guess the letters that correspond to 20 Latin letters.. From a few phrases that he can read he guesses that "$" means "King" and from that he guesses that "%" means "Queen". But after working for a while he realizes that the symbol that he thought meant "s" actually means "z", and there are two different symbols for "s", and there seems to be no pattern for where each is used, and the symbols "#" and "&" are nulls. And then he discovers that the writer often misspelled "there" with "their" and "its" with "it's", and used two different symbols for the "th" or "this" and the "th" of "think", but on the other hand he wrote both the "ch" or "chair" and the "sh" of "show" with the same symbol. And then he finds also that "%" was not "Queen" but "Prince", and that the sign that he thought was "x" was actually an abbreviation for "-ing". And he never finds what the symbol "@" meant.
That is in fact how most medieval encrypted documents are normally deciphered: by constantly patching and re-patching an initial key guess. And that is how the solutions often look like: a very complicated, messy, and inconsistent process, with mappings that are not one-to-one and not even algorithmic. Check the Rohonc Codex, for example.
That is the case even when the author tried to be consistent and avoid errors. Check the recent decipherment of the Zodiac Killer letters.
Quote:When the spellings were inconsistent, you added "daiin/kaiin/laiin are scribal errors."
Or variant pronunciations. If you speak Mandarin, you must be familiar with tone sandhi, yes?
Quote:When the sentence structure was incomprehensible, you added "qo is 'and.'"
The idea that
qo was a non-phonetic symbol meaning "and" (like our &), and that it had been added by the Author because he found the Chinese way of listing things too confusing, occurred to me when I noticed this recipe
青石赤石黄石白石黑石脂等味甘平主黄疸泄利肠澼脓血阴蚀下血赤白邪气痈肿疽痔恶疮头疡
疥瘙久服补髓益气肥健不饥轻身延年五石脂各随五色补五脏生山谷
and a paragraph that seemed to match it seemed to have a title with five similar parts, with four
qo between them. That structure is exactly how the "and" syntax works in Arabic or Hebrew.
That theory explained many puzzling facts that had been known about
qo, including why it is so common. And, as a side effect, it seems to reduce some gap errors in my matches by a few EVA letters.
Quote:When the spellings of high-frequency words were chaotic, you added "it may be a dialect or Vietnamese."
I first concluded that the language was an East Asian monosyllabic one when trying to prove the opposite.
I had just plotted the number of different Voynichese word types with a given length k, and got a nice symmetric curve that coincided almost perfectly with the graph of comb(n,k), the number of ways to choose k elements from a set of n elements, for n = 9. At that point I thought that I had found a hard proof that Voynichese could not be a natural language in the plain -- because no
natural language could have such a perfectly symmetrical plot of anything.
To show this point, I created the same plot for several other languages, and indeed they all come out asymmetric, some even with multiple humps. Just for completeness, I did that also with a text in Vietnamese, and ...
... You are not allowed to view links.
Register or
Login to view. that coincided with the graph of comb(n,k).
And that turned out to be the case also of Mandarin in pinyin (after rewriting "wu" and "yi" as "u" and "i", to match their pronunciations).
And that is when I came up with the Chinese Theory and the Dictation Scenario. At that point I pretty much lost interest in the VMS -- because I thought that, in order to make any progress, one would have to guess the language and then know how the language was pronounced in the 1400s.
I came back after 25 years only because I saw on YouTube Koen claim that the language could not be Chinese, no way. (Thanks Koen

)
I don't recall when I first thought that the Starred Parags section could be the Shennong Bencao.
The idea came because I read that it had 365 recipes and the SPS seemed to have had about that number of parags initially. I was just a wild hunch then. But after I came back I got the CTP file of the Shennong, learned how to process Chinese text in Unicode, and plotted the number of SBJ entries with k hanzi and of and SPS parags with k words. The shapes of the two plots matched surprisingly well.
Then I compared the longest entries in each file, and noticed that the spacings of four of the seven 主 in it (the CTP file was missing the 8th) matched the spacings of four
daiin in the SPS recipe.
And the rest is history -- I mean, a huge pile of python code, still evolving, that tries to automate the matching of the other recipes to other parags.
Quote:A truly correct theory should be one that "maps the letters of the VMS to specific sounds, reads out complete sentences, and is immediately understood by Sinologists around the world," rather than relying on a programmer's "badness scoring."
Well, I
can read out the entire Starred Parags section (even the missing pages!) -- in Chinese characters.
I still don't know what are the sound values of the Voynichese glyphs, and therefore I still don't know which pronunciation of the SBJ it is recording. All 50+ East Asian monosyllabic languages (not just the Chinese "dialects") are still candidates. Since I don't speak any of those languages, much less how they were spoken 600 years ago, I will not even try to solve that part of the puzzle.
Unfortunately the recipes were all scrambled when the book was bound, and probably also when the Author tuned his dictation notes into a draft for the Scribe. There is no reason to think that they were in some sort of "alphabetic order". (Indeed, the three digital SBJ files I got list the recipes in three different orders.)
So I cannot yet tell which recipe matched which parag. That is where the program comes in.
Quote:This is my perhaps somewhat immature view; I am just guessing, because I am only 16 years old, and my perspective may not be as far-sighted as that of you scholars or professors. I have been trying hard to improve.
Well, congratulations -- you are already sounding like certain 70-year-olds who were in the Voynich mailing list when I first learned about the manuscript, in the 1990s, and still cannot give up their religious belief that the language is European.
All the best, --stolfi