ololololo > 24-09-2026, 12:27 PM
(24-09-2026, 04:24 AM)Jorge_Stolfi Wrote: You are not allowed to view links. Register or Login to view.the statistics of Voynichese are quite compatible with an East Asian monosyllabic language, and quite INcompatible with any European language, any plausible encryption scheme, and any plausible gibberish generation methodI agree with the similarity of the statistics, but I disagree with this list. If Voynich is not consistent with known encryption methods, this does not mean that Voynichese is definitively Chinese. This could be an unknown encryption method, couldn’t it?
Jorge_Stolfi > 24-09-2026, 03:07 PM
(24-09-2026, 12:27 PM)ololololo Wrote: You are not allowed to view links. Register or Login to view.(24-09-2026, 04:24 AM)Jorge_Stolfi Wrote: You are not allowed to view links. Register or Login to view.the statistics of Voynichese are ... quite INcompatible with any European language, any plausible encryption scheme, and any plausible gibberish generation methodIf Voynich is not consistent with known encryption methods, this does not mean that Voynichese is definitively Chinese. This could be an unknown encryption method, couldn’t it? Maybe it’s a special way to generate gibberish, isn’t it? you’re saying more “I don’t know of anything similar,” rather than “there’s nothing similar"
chenxiang > 24-09-2026, 04:28 PM
chenxiang > 24-09-2026, 04:34 PM
(20-09-2026, 09:20 AM)Jorge_Stolfi Wrote: You are not allowed to view links. Register or Login to view.(19-09-2026, 04:53 PM)chenxiang Wrote: You are not allowed to view links. Register or Login to view.what the first character of the next word is depends mainly on semantics ... "remote effect" in Chinese [comes from] is grammatical structure, with clear semantic boundaries.
As I pointed out in my previous message, the point is that there are marked short- and long-range non-trivial statistical dependencies between hanzi in Chinese texts, including in the Shennong Bencao.
Yes they are due to semantics: common multiple-hanzi phrases that may or may not be tied to the topic of the text, and common grammatical constructs. And that is why those "remote effects" are seen in the VMS, even in other sections besides the SPS.
Quote:Tone sandhi: Chinese has tone sandhi for "一" (yi) and "不" (bu), but this is strictly limited to the immediately following syllable (adjacent), and is absolutely never affected by prefixes several words prior (such as the hypothetical affixes "sh", "ch").
Wikipedia says that tone sandhi in Mandarin is much more pervasive. In addition to those two cases,Are these statemenrs wrong?
- When there are two 3rd tones in a row, the first one becomes 2nd tone. E.g. 你好 (nǐ + hǎo > ní hǎo)You are not allowed to view links. Register or Login to view.You are not allowed to view links. Register or Login to view.
- The You are not allowed to view links. Register or Login to view. is pronounced at different pitches depending on what tone it follows.
And I read somewhere that the first rule applies to multiple syllables. That is, nǚ zǐ rǔ would be pronounced nǘ zí rǔ. Is that true?
Even these tone sandhi effects between adjacent syllables could explain "remote effects" in a phonetic transcription, depending on how tones are encoded.
In the "numeric" version of pinyin (also used for Cantonese jyutping and other phonetic systems for other languages), the tone is indicated by a digit at the end of the letter. Thus the Wikipedia rules above would beThe position of the tone digit in this version of pinyin is arbitrary. It could be placed at the start of the syllable, or anywhere inside it, and it would still be understood. And the placement would not have to be consistent. Thus the "remote effect" could be between the syllable-initial symbols rather than syllable-final symbols;
- A syllable that would normally end with '3' changes its ending to '2' if the next syllable ends in '3'.
- For certain common syllables, their final digit will be the same as that of the previous syllable.
But there is another way to encode Mandarin tones, which is to assign digits 1-3 to pitch levels (low, medium, high) and use them to indicate the tone profile. For example, hǎo could be written 2ha1o3 (or 213hao, or ha213o, or any of many other ways).
In this notation, one can also omit the pitch digits when they don't change. For example fú zhī qīng shēn bù could be written 1fu3 2zhi qing shen2 3bu1 or even 1fu3 zhi qing shen 3bu1 I suspect that the Vounichese letters a, o, y may be encoding tones this way.
In this notation, tone sandhi can have effect over several syllables.
Quote:Counter-evidence: If a Chinese text describing different plants has character sequence statistics that do not vary with the type of plant (semantic content), but are strongly constrained by prefixes like "sh", "ch", "k", etc., spanning multiple characters to determine the probability of the subsequent "q(o)", then this book is absolutely not a natural language botanical monograph, but can only be an encrypted text or generated pseudo-text.
See again the first recommendation in my previous post...
Quote:does this "remote effect" mean the clue "daiin = 主" is completely dead too?
As explained above, definitely not.
But note that not every daiin means 主, just as not every "shi" means teacher (or stone, or lion, or whatever).
It is the other way around: as far as I have checked, a 主 in the SBJ corresponds in the SPS to a daiin or one of its few recognized variants.
All the best, --stolfi
(22-09-2026, 11:52 AM)Jorge_Stolfi Wrote: You are not allowed to view links. Register or Login to view.(22-09-2026, 09:59 AM)eggyk Wrote: You are not allowed to view links. Register or Login to view.As far as I can tell from a look online, the older system did not divide the ecliptic into 24 equal parts. It divided the year into 24 parts of just over 15 days each.
But this doesn't really change that much with regards to that fact that there was a 24 * ~15 system.
Exactly. I previously retracted the claim that the VMS "things" (label+star+nymph groups) are Babylonian degrees (which would have implied a division of the Ecliptic rather than the year).
I have also suspended the claim that the Chinese explicitly divided every solar term interval into 15 parts (which would each be a 365.25/369 = ~1.01458 days, and would be the meaning of the VMS "things"), because I could not find a detailed description of the system that could have said so. Maybe @chenxiang can help me find it?
All the best, --stolfi
ololololo > 24-09-2026, 05:59 PM
(24-09-2026, 03:07 PM)Jorge_Stolfi Wrote: You are not allowed to view links. Register or Login to view.The encryption method would have to be specifically designed to get that result. For instance, if every token contains some null letters -- even if just one, in a fixed position -- each word type of the original will become any of N different word types in the cyphertext, for each of the N possible choices of nulls. That multiplication of word types would be plainly visible in the Zipf plot -- unless the possible null choices are made with different, carefully tuned probabilities.I can't disagree with that. This is especially applicable to the Naibbe cipher, which uses full‑fledged tokens like aiin as glyphs. This is a very easy way to achieve the desired result... Nevertheless, this shows that, with certain assumptions, it is quite easy to reproduce the laws of VMS.
And the same "plausibility" objection applies to gibberish generators. In the copy-and-mutate method, for instance, the seed text must already be statistically similar to Voynichese (which already creates a chicken-and-egg problem) and the mutate procedure must be complex and fine-tuned to preserve the statistics.
Jorge_Stolfi > 2 hours ago
(24-09-2026, 04:34 PM)chenxiang Wrote: You are not allowed to view links. Register or Login to view.First, regarding your Pinyin data: it is true that adjacent syllables in Chinese show statistical tendencies (e.g., a syllable ending in 'g' is often followed by one starting with 's'). However, this is a phenomenon of adjacent phonetic assimilation and high-frequency word collocations (词汇搭配), not a grammatical or structural dependency.
In Chinese, the influence of a syllable's final sound strictly applies to the immediately following syllable within the same phrase. It does not extend across word boundaries to mechanically govern the distant structure of an unrelated word ... not a "long-distance effect." In the VMS, as JoJo_Jost and others have shown, the sh/ch at the end of a word strongly affects the qo at the beginning of the next word, ... This kind of mechanical, cross-boundary, long-range statistical dependency is not a feature of natural Chinese syntax or phonology.
Jorge_Stolfi > 44 minutes ago
(24-09-2026, 04:34 PM)chenxiang Wrote: You are not allowed to view links. Register or Login to view.I saw your request regarding whether there is a detailed description in Chinese texts of dividing every solar term interval into 15 parts. I would be happy to help clarify this.
To be direct: No, such a system does not exist in traditional Chinese astronomy or calendar literature.
While the number 15 is extremely significant in Chinese culture (e.g., 15 days per solar term, 15 minutes per quarter-hour), ancient Chinese calendars never divided the ~15 days of a solar term into 15 equal "things" or units.
Quote:Here is how Chinese calendar systems actually divided time:
Quote:Solar terms (节气): There are 24 solar terms in a tropical year. They are determined by the sun's ecliptic longitude (定气) or by simply dividing the tropical year into 24 equal parts (平气).
Quote:There is no classical text that supports the idea of mapping the VMS "things" to ...