Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

What surprises me (only a little) is that Japan continues to use Chinese characters when they've got a perfectly good native syllabary.
 help



I don't know if this is from the perspective of having learned Mandarin in elementary school first or if Japanese folks feel similarly, but reading words like, e.g., だいがくいんせい versus 大学院生 is actually a little more annoying because I have to scan more of the line to read the word in question. This is notwithstanding the effect that hiragana has in highlighting grammatical structures; it's a lot like setting off Finnish case endings and verb conjugations in a different color.

At least, that's the reason why I prefer still having kanji around.


Japanese as it currently exists is just really hard to read when written without kanji. A couple reasons: it has TONS of homophones, and it doesn't use spaces, and also it has lots of common words that are only 1 or 2 characters long. So if you're parsing a stream of kana, in the general case there will be several different ways to interpret the next couple characters and you'll constantly be backtracking.

So at a minimum, dropping kanji would mean the language needed to develop a system for where spaces go (which is less trivial than it sounds IMO). But also it would need to abandon a lot of homophones or find some way to disambiguate them. It could happen I guess, but I'm not sure how it could happen gradually.


Japanese could be written with the Latin alphabet very well, if a set of kanji radicals were retained and inserted before a part of the words, to disambiguate their meaning.

This is a not a new method of writing, but this is how writing was done in the past, for many millennia.

The oldest writing systems, like the cuneiform and the Egyptian, also used, besides signs that indicated the pronunciation of a word, determinative signs, which indicated the semantic domain of the word, e.g. it is a name of a human profession, it is a name of an animal, it is a name of a tool made of metal, it is a geographic name, and so on.

The determinative signs could be Chinese radicals to continue the tradition, or they could be a special set of emojis.


Sure - it's probably safe to say that any language can he written in any writing system, if one is willing to change both the language and the writing system :)

Korean could serve as a point of comparison, as modern Korean is written mostly without Chinese characters (you could mix them in if you really want to, but approximately nobody bothers), but it does have spaces between words. And yes, when you have a highly inflecting language, "where do the spaces go" is a very nontrivial problem.

It's a bit easier for Korean as it has far more available syllables (so fewer homonyms), but you could easily imagine these languages making different tradeoffs and ending up in different places.


A lot of material for younger readers uses similar spacing to Korean, like the Pokemon games. It works well enough when you know the vocabulary being used is limited, but even there it gets frustrating.

When you're reading, you're trying to identify a substring and map it back to the unit of meaning, which is effectively the kanji. That works seamlessly with Hangul, because it mostly blocks (as syllables) the same way as kanji, and, as you mentioned, the extra vowels and silent consonants provide more disambiguation. With kana, the length you have to scan to read, and then guess at (using only context to disambiguate), a root word, is a lot more arbitrary.

Mechanically kana is just less compressed, less aligned, and makes it harder to visually identify/scan for units of meaning. Obviously Korean shows that you can make different tradeoffs, but for Japanese losing the kanji would probably necessitate reworking Kana as well - and unless the pronunciation changed, it would probably need to be in a way that represents the pitch accent, which is a modifier on the whole syllable, rather than just a few extra consonants or vowels.


Agreed, I think that's a close analogy and it's wild to me how recently Korean dropped kanji. But I imagine having a lot more vowel sounds made it a lot easier. I've wondered whether the two interacted - e.g. if there were word pairs that had previously used the same vowel sounds, but changed after the switch to Hangul to be distinguishable. Do you know by any chance?

Sino-Korean words are different from Sino-Japanese words in that the sounds are very "rigid": most Chinese characters have one canonical sound, and it is pronounced that way regardless of where it appears (modulo some mostly predictable sound change rules). So, you can't just take a Sino-Korean word and "tweak" its sound to disambiguate - when such a word changes its sound it will no longer be perceived as Sino-Korean at all, which actually happened for some examples such as 산행 (san-haeng, 山行, "mountain + go" = "hike on a mountain") which gave birth to 사냥 (sanyang, "hunting"), a separate word.

What did happen was that we lost obscure words that could only be understood by looking at Chinese characters (because they're used so rarely, or because they sound the same as another more common word) - you could still use them if you wanted to, but unless you also write them in Hanja, it would be confusing, and if you mix gratuitous Hanja into your writing just because you like those words, most young readers will find it pretentious, and I fear, mostly unreadable.

I can't say we lost much by losing these words. Rather, I think it actually forced writers to select less ambiguous words, which is usually a good thing.

(Also, these days most Koreans know much more English than Classical Chinese, so English loanwords can sometimes fill the gap. It helps that most English (or other Western) loanwords sound different from native or Sino-Korean words, so there's less chance of collision.)


> it has TONS of homophones

I never get this line of reasoning - so are people listening to spoken japanese just in a perpetual state of confusion because they can't see the associated intended Kanji?

Also this coming from a person on a platform who's primary language is commonly pronounced with the last syllable/letter of each word used to start the next one.

The engli shlanguage ha sit sown nonsense.


Spoken Japanese uses pitch accent to distinguish homophones

In some cases, but note that (a) this is only true for a tiny handful of homophones, and (b) the distinguishing accent varies by region.

E.g. "hashi" famously can mean bridge/edge/chopsticks depending on accent, but which accent has which meaning is different in Tokyo vs Osaka.


So shouldn't the better solution be a few more pronounciation markers beyond two strokes, a circle and tiny letters instead of thousands of chinese characters?

But who is this solution for? Unless you change the pronunciation of the language, it would be adding markers around whole syllables to disambiguate - syllables that represent the kanji they came from (many of which already have phoentic components) - but those syllables would still be less dense and visually identifiable than kanji is for someone who already speaks the language.

It's similar to initialisms or emoji in English; it's of course frustrating for people who don't know them, and there are reasonable complaints about some of the extremes, but speakers/writers use them because they find them useful when communicating with an audience that's also assumed to know them.


I mean, do you want to know or are you venting? If the former: No, spoken Japanese is not constantly confusing, because (a) some words are distinguished by intonation, (b) Japanese speakers know what the homophones are, and either avoid them while speaking, establish the context before saying them, say extra words to clarify which word they mean, etc.

Kind of humorously, groups sometimes develop lingo to avoid all this. E.g. the JP words for science and chemistry both have the standard pronunciation "kagaku", so people in industries where both terms are used often use nonstandard pronunciations (bakegaku for chemistry).

Also the homophones thing is probably lesser of the things I mentioned. A typical sentence might only have 1 or 2 overloaded words, but it will have lots of substrings that can be parsed multiple ways.


Firstly, the native syllabary uses conventions that give rise to ambiguities. For instance えい (ei) can encode a long e sound like in "sensei" (teacher) or a separate e and i, like in "deiriguchi" (combined exit and entrance).

Secondly, there are many homonyms: words written with completely different kanji that sound exactly the same, very similar (e.g. same modulo pitch accent) or use the same spelling in kana.

Here are two words that don't sound the same at all. Let's use Hepburn, because it distinguishes them: kõri and kouri. One has a long o, the other has separate o and u. In hiragana, both are written こうり: exactly the same. kōri might be 公理 (axiom, self-evident truth) or 高利 (high interest rate). kouri is 小売 (retail, lit. "little selling").

If you're having a conversation with someone and don't know which "kõri" they are talking about, high interest rate or axiom, you can ask. But you cannot just ask a written text. You have to work it out from context.

Kanji eliminate the ambiguity, vastly simplifying reading. Reading Japanese that has been normalized to kana is murder. Especially if spaces are not introduced to mark word divisions. That is only done in hiragana books for small children!

Sometimes this happens in writing. A sentence happens to start with a couple of words that happen to be usually written in hiragana, with some particles after them, and you're left with a puzzle just working out where the word boundaries are. This requires backtracking in the general case; there is no principle like longest match.


I don't deny that reading Japanese written in kana alone is currently more painful than with kanji. However, it seems to me that the current writing system is a local optimum and not a global one. If people started writing Japanese using kana alone, some adjustments would eventually be made to make it more readable, and the end result would in my opinion be superior to the current kana/kanji combination. The same thing happened with the latin alphabet. We don't use it the same way the romans used it 2000 years ago: we introduced minuscule letters, spacing, punctuation, accents... At the end of the day, Japanese is a spoken language like all other languages, and there are no reasons why a purely phonetic system should not work for writing it.

Well the "native syllabary" is derived from Chinese characters, it's not something that was there and then Chinese characters were shoehorned in later.

I think at some early point most Japanese learners (that don't already know a Chinese language) think kanji should be dropped. And at a later point I think most would change their mind as the advantages outweigh the disadvantages.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: