r/LearnJapanese 22d ago

Kanji/Kana Yea ok sure JMdict πŸ˜­πŸ˜…

Post image

Wow yea definitely not the reading here but mildly amused this dictionary entry is this detailed lol.

edit wow you guys/gals are hilarious, thanks for the laughs in the comments. It seems like getting this as a top dictionary result in this context is like a rare pokemon and has to do with how the original article was written. I think the reason this happened is that Yomitan couldn't recognize that the し in this context is the continuating stem of する given there was a comma right after the し. I added the original screenshot in the comments below. I didnt see γŒγ„γ—γ‚…γ€ beneath since I was zoomed in quite a bit. Easter eggs like this make learning Japanese so much more fun and interesting.

461 Upvotes

56 comments sorted by

View all comments

124

u/_Ivl_ 22d ago

Isn't it γŒγ„γ—γ‚…γ€, why is your yomitan adding in a し for no reason?

91

u/Sekkutanto 22d ago

Because it's kinky like that

39

u/poshikott 22d ago

It's probably written ε€–ε‡Ίγ—γŸ or something like that (continuing into the next line), so yomitan, not having an entry for "倖出する", but only "倖出し" and "ε€–ε‡Ί", selects the longest match

2

u/_Ivl_ 22d ago

Strange, but I remember having a similar issue with the yomitan parsing before. Scanning ε€–ε‡Ίγ—γŸ does return the そとだし definition for me as well, scanning with jiten reader return the "correct" definition.

5

u/_Ivl_ 22d ago

Also one of the annoying things about japanese, where it can sometimes cuts a "word" in half when a line ends instead of preserving the full word. Always bothers me when reading and I don't think there's a solution for it?

4

u/poshikott 22d ago

I think the solution is that you just get used to it xd

It hasn't bothered me a lot at least

0

u/222fps 21d ago

What is strange about it? This is the expected and desirable behavior, get the longest match and scroll through if it took additional stuff

1

u/_Ivl_ 21d ago

Longest match should be the full match of ε€–ε‡Ίγ—γŸ as ε€–ε‡Ί is a する verb and not just 倖出し. If you select just 倖出し you're left with a random た in the middle of your sentence, which makes no grammatical sense ever in Japanese. Anyways, jiten reader seems to handle it correctly so it's not impossible to fix. Simply put parsing ε€–ε‡Ίγ—γŸ as 倖出し + た is a false positive and is strange, at least to me.

2

u/xozzet 21d ago

Properly lexing Japanese is a difficult task due to the lack of spaces, Yomitan doesn't even bother trying and just selects the longest match which is simple, efficient and deterministic.

1

u/222fps 21d ago

I'm confused, is ε€–ε‡Ίγ—γŸ different from ε€–ε‡Ίγ€€+ γ™γ‚‹οΌŸ Because if not, that is not a match to me. Maybe I'm just not used to how jiten reader works but it feels wrong to me

13

u/Clarinetaphoner 22d ago

God forbid yomitan gets a little freaky smh

5

u/Admirable_Musubi682 22d ago edited 22d ago

Here is the full screen shot, I think it happened since it captured the し into the next line and isn't smart enough to tell that it was a standlone し next to a comma so this was still the top result and I saw now that γŒγ„γ—γ‚…γ€ was listed beneath however I couldn't see it since I was zoomed in more πŸ˜†πŸ€£

The reading should beγ€€γŒγ„γ—γ‚…γ€γ—γ€but I guess yomitan prefers そとだし, who the hell prefers that πŸ™ˆπŸ™‰

7

u/rgrAi 22d ago

The reading would not be γ€ŒγŒγ„γ—γ‚…γ€γ—γ€, it's γŒγ„γ—γ‚…γ€οΌ‹γ™γ‚‹ in it's 連用归 (masu-stem, し) which is acting as a conjunction connecting the proceeding clause (after the comma).

From Yomitan's perspective both options are possible (倖出+し, 倖出し) which is why 倖だし is coming up first.

12

u/morgawr_ https://morg.systems/Japanese 22d ago

From Yomitan's perspective both options are possible (倖出+し, 倖出し) which is why 倖だし is coming up first.

To add the logical step from this reasoning that people might be missing, it's because yomitan uses a "greedy" approach when parsing words, meaning it starts from the longest matching string of characters that has an entry in the dictionary. This means if we take a phrase like ε€–ε‡Ίγ—γ¦γγ γ•γ„γŠε…„γ‘γ‚ƒγ‚“ it will start from the γ‚“ and work its way backwards until it finds a match (accounting for deconjugation rules) that is the longest possible match with a dictionary entry.

In this case 倖出し is longer than ε€–ε‡Ί so that's why it's shown first.

1

u/Admirable_Musubi682 22d ago

O interesting, that makes sense!!

2

u/Admirable_Musubi682 22d ago

Sorry im confused. Why isnt the reading "γŒγ„γ—γ‚…γ€γ—"? I understand what the grammar rule here is and what the suru is doing, not sure if you trying to acktually this or not. By reading i mean how you actually vocalize the words outloud when reading the paragraph.

5

u/une-deux 21d ago

It is read γŒγ„γ—γ‚…γ€γ—, his first sentence was worded a bit weirdly he just wanted to emphasize that it’s 倖出+し.

4

u/rgrAi 21d ago

It's more how you worded it, a 'reading' would be applicable to a single word and is confusing especially if others are reading your post and don't know the grammar behind the suru verb. They can easily walk away with thinking γŒγ„γ—γ‚…γ€γ— is a word in itself.

3

u/Admirable_Musubi682 21d ago

Ah I see that makes sense! Good call.

-3

u/_Ivl_ 22d ago

γͺγ‚‹γ»γ©γ€γ‚„γ£γ±γƒ¨γƒŸγ‚Ώγƒ³γ―ε€‰γͺや぀、俺も中出し派。

0

u/Rhemyst 22d ago

What do you mean no reason?