Search Ace Linguist

September 4, 2026

Link Haul - September 2026

I come bearing more articles on language-related posts and news!

We'll start by traveling back in time to medieval Britain. I enjoyed this curious video from Gavin the Medievalist about a difficult passage in Beowulf, featuring his attempted explanation.

Gavin also has a video on the earliest known mention of King Arthur, and discusses the unusual way in which he is brought up. I recommend following up with this video on King Arthur by Cambrian Chronicles, which discusses the complicated history of mentions of Arthur. Fun Old Welsh (and I'm wondering if the mention of Arthur is to rhyme with Gwawrddur.

Then looking at Indian music: I looked up the lyrics to one of my favorite Hindi songs recently ("Marjaani"). I found this blog post that provided not just a translation, but some cultural commentary as well. I thought this was pretty interesting:

1. So far as I've understood, the word 'Khasmanukhaye' originally meant a woman who'd eat (bring misfortune to) her husband (khasam). But now it simply mean someone wretched, or someone who brings misfortune.
2. I came across this really cute piece on the internet where a girl explains why Punjabis use so many swear words. In her words, "In punjabi culture (at least what I saw of it) complimenting or gushing over someone was supposed to jinx (nazar lag jaati hai). So they go out of their way to be abusive. I know my friends mom always calls her grandaughter Chudail (witch) and jhalli and she loves her to bits. My grandma would say "Kinni soni lagdi hai marjaani" to me all the time if I dressed up and was looking pretty. It is harmless namecalling and very cute too."
If you are not familiar with the song, I recommend listening to it - it's an all-time banger.

And finally, some more on AI stylistics. We discussed language mixing in LLMs last month. This post from LessWrong talks about concerns that LLMs will become unintelligible to humans, whether by developing a new language mix or by becoming so telegraphic and shortened that humans will not be able to understand. The author makes the argument that this is unlikely to happen, and that examples of language mixing or shortening may come from patterns in training data rather than being actually efficient ways to reason. There are some interesting counter examples in the comments if you'd like to delve into that.

And speaking of 'delve' (a common stylistic tic from ChatGPT,), this old tweet from 2024 came to my attention. While it is definitely a very straightforward way of trying to determine authorship, it's not rare. I've talked to a lot of people who are looking for ways to spot that something was AI-written.

 

However, the opposite phenomenon may well be arising. Using AI-influenced vocabulary may actually serve as a signal that you're clued in to the AI world and not merely posing as a vibe coder. This highly entertaining article from User Mag discusses "Claudlish," the style particular to Anthropic's Claude. "Claudlish" is seen as being closer to how agents actually write, so familiarity with it is useful to understand what Claude is saying, to avoid having to spend unnecessary tokens asking for clarifications, and as a proxy that you're genuinely plugged in to using agentic AI tools:

Being able to speak Claudish has become a meme on X, but fluency has also become a proxy for how heavily engaged with AI software development a person actually is. One startup founder I spoke to said that giving a potential job applicant a paragraph written in Claudish and asking them to translate would be a useful gauge for how heavily they actually use AI coding tools. He noted that a lot of people like to “LARP” as vibe coders, but if you can’t translate Claudish, you haven’t spent enough time in Claude Code.

Being able to understand and speak Claudish directly is also potentially important in a business context, Deng said. “If you ask [Claude Code] to explain or to translate the outputs back into human language, you’re wasting tokens,” he said, which can be costly.

 Of course, it's not all fun and games, and there's plenty of ribbing about Claudlish in the article as well.

July 1, 2026

Code-Switching for Robots - Language Mixing

I promise I'll write about things besides Japanese and LLMs. But first, a brief word on "language mixing."

I have occasionally encountered situations where LLMs suddenly and without warning begin mixing English and non-English words. Usually, the non-English language also has a non-English script. I've encountered Arabic, Korean, Japanese, and Chinese. The Japanese may be understandable in contexts where I've previously included Japanese text in the conversation for whatever reason, but the Arabic, Korean, and Chinese are quite head-scratching.

Chinese inserted. Image mine. 

Korean inserted. From source.

I've jokingly referred to this as "code-switching," the term for swapping between languages or dialects within a single discourse context. I say "jokingly" because "code-switching" is used by humans with each particular language selected for a reason. The term for it in academia is apparently "language mixing." Unlike human code-switching, language mixing is often "unintentional" in that it is an unwanted artifact. LLMs can produce text that swaps between two languages in a naturalistic way, which we could perhaps refer to as "code-switching," but "language mixing" is broader.

 


Apparently language mixing can actually help the models reason better. Suppressing bilingual chain-of-thought can degrade performance, which makes me think of how people who code-switch are actually people who are strong in both languages. (Of course, that is not to suggest we can generalize about LLM performance from human performance.)

At the same time, because the models are trained so aggressively on English, the chain-of-thought itself may be in English, and trying to force it to reason in a language other than English may also degrade performance, even if the output is meant to be in another language.

Users tend to feel unsettled when they see language mixing. Language mixing breaks the illusion that the LLM you are speaking to is a human-like companion because it makes a speech decision that no human ever would. Because it is so unexpected, it gives the impression that something has gone wrong with the LLM. Most language mixing today is minor slippage of a single word. However, there have been catastrophic instances of language mixing in the past.



June 1, 2026

"The moon is beautiful, isn't it?" - An odd translation myth

While watching a video, I encountered this anecdote:

When the novelist Soseki Natsume (1867–1916) was an English teacher, one of his students translated the English phrase “I love you” as 我君を愛す / ware kimi o aisu. Soseki pointed out that Japanese people don’t say 愛す / aisu (to love), and that the best translation would actually be 月が綺麗ですね / tsuki ga kirei desu ne (the moon is beautiful, isn’t it?).

This anecdote has become widespread as an example of how translation challenges are not just about words, but about culture. In its most reductive form, people have taken away that "In Japanese, 'the moon is beautiful, isn't it?' means 'I love you'."

Moon after rain, Kiyomizudera Temple, Kyoto

I first encountered this story many years ago, where I sort of accepted it as a curiosity and moved on. Upon encountering it again, I felt much more skeptical. For one, plenty of Japanese media use rather direct translations of "I love you", ranging from the more low-stakes "suki desu" to the serious declaration of "aishiteru." 

Secondly, while I understand that high-context cultures exist, "the moon is beautiful, isn't it?" is quite vague. No example is given of the text that the student was trying to translate. If he's a teacher of English, then I presume that his student is Japanese and not a native English speaker, yet in this example it sounds like he's telling someone who is not a native Japanese speaker how to translate this phrase.

The phrase itself also confuses me. Am I to believe that any shared appreciation of the moon is tantamount to a declaration of love? Or are there special conditions that turn lunar love into shared love? If it's about the context, why is the context absent from this anecdote? Are there actual examples of Japanese translations of English text where "I love you" is rendered as something non-literal, and if so, what are they and what are the circumstances? In short, the whole thing sounded oddly pat and exoticizing while also lacking crucial detail on what would make such a translation work.

After googling the phrase and following Reddit links, I found that the original anecdote is a misattribution. Soseki Natsume never said such a thing.

The earliest citation found for this formulation is from a 1961 book:

さらにいえば、日本の社交の基本は「見る」ことで成立する。
若い男女の恋人同士が愛の告白をするとき、西洋人のように、
「私はあなたを愛しています(I love you)」
などとはけっしていわない。
そんなことばを口に出さなくとも、満月を仰ぎ見て、
「いいお月さんですね」
そして、二人でじっと空を見上げるだけで、意思は十分通じるのだ。 

Furthermore, the foundation of Japanese social interaction is established through “looking.” When young men and women confess their love to one another, they do not, like Westerners, say things like: “I love you.” Even without putting such words into speech, they simply gaze up at the full moon and say, “What a beautiful moon.” And by quietly looking up at the sky together, their feelings are fully understood. 

There is also a book from 1922 that claims that "I love you" cannot be translated into Japanese:

日本語には英語の『ラヴ』に相當する言葉が全く無い。『戀』とか『愛』とか云ふ字では感じがひどくちがう。" I love you "や" Je t'aime "に至つては、何としても之を日本語に譯すことが出來ない。 

Japanese has no word that truly corresponds to the English “love.” Words like koi or ai feel quite different in nuance. As for “I love you” or “Je t’aime,” there is simply no way to translate them into Japanese. 

Neither of these are by Soseki Natsume, but he does have the following:

漱石文庫」に残された漱石メモ書きの中に、ジョージメレディスというイギリス小説家作品を取り上げて、

"I love you,Signora Laura."―Vittoria p.113.

I love you日本ニナキformulaナリ

と記した一節がある

In Sōseki’s notes preserved in the “Sōseki Bunko,” there is a passage where he takes up a work by the English novelist George Meredith and writes:

    “I love you, Signora Laura.” — Vittoria, p.113.

    “This ‘I love you’ is a formula that does not exist in Japanese.” 

A hatelabo user surmises that the story that Soseki believed "I love you" could not be translated to Japanese was conflated with the story that someone translated "I love you" as "What a beautiful moon," to create the hybrid story "Soseki said that 'I love you' should be translated as 'the moon is beautiful, isn't it'." 

The comments on the hatelabo post are quite revealing. Many of the users seem to think that it's quite normal to say "I love you", with one even talking about Japan's "confession culture," and some speculate that this may be a relic of a different Japanese culture. In short, the idea that 'ai' cannot be used to express "love" may have been true at one point in Japanese history and culture, but continued exposure to Western cultures has moved the meanings of those words closer, and Japanese culture itself changed to make direct declarations of love more acceptable. 

これは評価すべき増田だな。 しかしこれ読んで思ったが、『"当時"の愛という単語』は love に直訳できなかったが、 現代の日本語においては愛にloveの意味がおおむね正確に取り込まれているので、 "I love you"=私はあなたを愛しています、で問題なさそうだ。

“This is a commendable piece by Masuda.  But reading this made me think: the word <ai> ‘back then’ couldn’t be translated directly as love, but in modern Japanese, <ai> has more or less accurately absorbed the meaning of love. So ‘I love you’ = ‘watashi wa anata o aishiteimasu’ seems basically fine.”

ホンマや。隔世の感ありやね。むしろ現代の日本人には当時の日本人の感性がよくわからんてこっちゃなぁ。

True. It really feels like a world away. If anything, it means modern Japanese people don’t really understand the sensibility of Japanese people back then.

As for the significance of the moon, one person I discussed this with suggested that the moon may be significant because it was historically an image associated with lovers - two young people would meet clandestinely under the moon. This would make referencing the moon a way to make salient the fact that they're doing something a little dangerous for love. An interesting bit of speculation. 

In any case, this zombie story has taken on a life of its own such that saying "the moon is beautiful, isn't it" has become reinterpreted as a covert love declaration. It seems that this story is so sticky that it memed its way into reality!