Every language app in your pocket inherited a teaching method built for Latin. Understanding why that happened is a more useful design lesson than anything the apps themselves will teach you.你口袋里的每一个语言学习应用,都继承了一种为拉丁语设计的教学法。理解这件事为何发生,比应用本身能教给你的任何设计课都更有用。
In 1788, Prussia introduced the Abitur, a standardized national examination required for entry into universities and the civil service. To pass it, students needed to demonstrate measurable, gradable knowledge. The system needed to teach language to large classrooms, produce consistent outcomes, and do it with one teacher and thirty students. The educators responsible for designing this system reached for the only teaching template they had, one that had been used in European schools for two centuries: the method developed to teach Latin.1788年,普鲁士引入了Abitur(高中毕业会考),这是一种标准化的全国性考试,用于大学和公务员的入学资格。要考过,学生需要展示可衡量、可评分的知识。这个系统需要在大班级里教语言,产生一致的结果,且只有一个老师面对三十个学生。负责设计这个系统的教育者找到了他们唯一熟悉的教学模板,一个欧洲学校已用了两个世纪的方法:为教授拉丁语而开发的方法。
Latin, by 1788, was a dead language. Nobody needed to speak it. The scholars who studied it were reading Cicero and Virgil, not conducting conversations. The method built around it, memorizing grammar rules, constructing translations, analyzing written texts, reflected that reality exactly. Oral skills were irrelevant. Comprehension of written form was everything. The method was not designed to produce speakers. It was designed to produce readers of texts in a language nobody spoke.到1788年,拉丁语已经是一种死语言,没人需要说它。研究它的学者们阅读西塞罗和维吉尔,而不是进行对话。围绕它建立的方法——记忆语法规则、构造翻译、分析书面文本——恰好反映了这个现实:口语技能无关紧要,理解书面形式才是一切。这种方法不是为了培养说话者,而是为了培养一种没人说的语言的文本阅读者。
When Prussia applied this template to French and German, living languages spoken by living people, the premise did not change. Johann Valentin Meidinger’s textbook Praktische Französische Grammatik, published in 1804, ran to 37 editions across Europe by 1857. Karl Plotz formalized the approach into what became the dominant model for teaching modern languages across Europe and eventually the United States, where it became known simply as the Prussian Method [1]. Each institution that adopted it trained teachers in it, who trained students who became teachers. The constraint that created the method, how do you grade language at scale with limited resources, became invisible inside the method itself. What remained was the assumption: language is a body of rules to be learned consciously and measured. It was a design decision dressed up, over time, as a pedagogical truth.当普鲁士把这个模板应用于法语和德语——活人说的活语言——前提并没有改变。约翰·瓦伦丁·迈丁格的《实用法语语法》于1804年出版,到1857年已在欧洲出了37版。卡尔·普洛茨将其正式化为主导欧洲乃至美国的现代语言教学模式,后来被简单称为“普鲁士方法”。每个采用它的机构培训了教师,这些教师又培训了学生,学生再成为教师。创造这个方法的约束——如何用有限资源大规模评分语言——在方法内部变得不可见。剩下的只是一个假设:语言是一套需要有意识学习和衡量的规则。这是一个设计决策,经过时间积累,被伪装成了教学真理。
The observation that should have ended it#section2本应终结它的观察#section2
There are people in the world who cannot read or write a language and speak it fluently. There are children who hold full conversations years before they can read a single word. There are immigrants who arrive in a country knowing nothing of its language and come out, years later, speaking it naturally, not because they studied it, but because they lived inside it. Literacy and fluency are separate things produced by entirely separate mechanisms. The Grammar-Translation method, as it became known, assumed they were the same thing. That assumption was inherited from a method designed for a language nobody needed to speak, and it was wrong the moment it was applied to a language people actually used.世界上有些人不会读写一种语言却能流利地说它。有一些孩子能进行完整对话多年后才认识一个单词。有一些移民到达一个国家对语言一无所知,几年后却自然地说出来,不是因为他们学了,而是因为他们生活其中。识字和流利是两码事,由完全不同的机制产生。语法-翻译法(后来被如此称呼)假设它们是同一回事。这个假设继承自一种为没人需要说的语言设计的方法,一旦应用于人们实际使用的语言,就是错误的。
The evidence against it accumulated slowly. In the mid to late nineteenth century, reformers including François Gouin in France and Maximilian Berlitz in the United States argued independently that language should be taught the way it is actually acquired, through immersive exposure to real communication in the target language, not through analysis of its rules. Berlitz built an entire school network around this principle. The reformers were correct. They were also largely ignored by mainstream education systems, because the Grammar-Translation method had one decisive advantage that direct immersion did not: it could be graded.反对它的证据慢慢积累。19世纪中后期,法国的弗朗索瓦·古恩和美国的马克西米利安·伯利兹等改革者独立论证,语言应该按照实际习得的方式来教——通过沉浸式的真实交流,而不是通过分析规则。伯利兹围绕这个原则建立了一整套学校网络。改革者是对的。但他们基本上被主流教育系统忽略了,因为语法-翻译法有一个直接沉浸法没有的决定性优势:它可以被评分。
In 1982, the linguist Stephen Krashen gave the argument its most formal articulation in what he called the Monitor Model of second language acquisition. His distinction was precise: language acquisition, the unconscious process through which children absorb their native language and through which adults succeed in immersive environments, is categorically different from language learning, the conscious study of grammar rules and vocabulary that classrooms deliver [2]. Acquisition produces fluency. Learning, at best, produces the ability to pass a test. The evidence supporting this distinction, and the observation that immersive exposure to real native-speaker communication is the mechanism that produces genuine fluency, has only grown since.1982年,语言学家斯蒂芬·克拉申在他所谓的监控模型中给出了最正式的表达。他的区分很精确:语言习得——儿童吸收母语和成年人在沉浸环境中成功使用的无意识过程——与语言学习——课堂教学中语法规则和词汇的有意识学习——是截然不同的。习得产生流利。学习充其量产生通过考试的能力。支持这一区分的证据,以及沉浸式接触真实母语交流是产生真正流利的机制的观点,此后只增不减。
I went to Brazil without a word of Portuguese and came out speaking it. I studied French in a classroom for years and cannot hold a conversation in French today. This is not an unusual experience. It is the expected outcome, and it has been the expected outcome for as long as we have had formal language education.我去巴西时一个字不懂葡萄牙语,出来时能说它。我在教室里学了几年法语,今天无法用法语进行对话。这不是不寻常的经历。它是预期的结果,而且自我们拥有正规语言教育以来一直是预期的结果。
The same decision, made again in a different medium#section3同样的决定,在不同的媒介中再次做出#section3
Prussian educators faced the question: How do you deliver language learning at scale, measure progress, and retain users over time? The answer it arrived at was structurally identical to the one arrived at in 1788. Duolingo gamified the grammar drill into a streak. Anki formalized the translation exercise into a spaced-repetition flashcard. Babbel organized grammar lessons into structured modules. The interfaces were new. The underlying assumption, that language is a thing you study rather than an environment you inhabit, was not.普鲁士教育者面临的问题:如何大规模提供语言学习、衡量进度并长期留住用户?它得到的答案在结构上与1788年得到的一致。Duolingo把语法练习游戏化成连续打卡。Anki把翻译练习形式化为间隔重复的闪卡。Babbel把语法课程组织成结构化模块。界面是新的。但底层假设——语言是你学习的东西,而不是你居住的环境——没有变。
This was not a failure of design skill. The products that emerged from these decisions are, in many respects, genuinely well-crafted. Duolingo’s retention mechanics are sophisticated. Anki’s spaced repetition is grounded in real cognitive science. They are excellent at what they actually do. The problem is what they actually do: produce measurable engagement with a proxy for language rather than the conditions that produce language itself. A streak is measurable. A vocabulary score is measurable. The moment a user walks out of an app and holds a real conversation in another language, that happens in the world, outside the product, and cannot be instrumented.这不是设计技能的失败。这些决策产生的产品在许多方面确实是精心制作的。Duolingo的留存机制很复杂。Anki的间隔重复基于真正的认知科学。它们擅长自己实际做的事情。问题在于它们实际做的事情:产生与语言替代物的可衡量互动,而不是产生语言本身的条件。连续打卡是可衡量的。词汇分数是可衡量的。但用户走出应用、用另一种语言进行真正对话的那一刻,发生在世界中,在产品之外,无法被量化记录。
When the outcome a user needs is difficult to measure directly, the design process tends to reach for something that can be measured. The proxy becomes the goal. The interface optimizes for it. The gap between what the product delivers and what the user actually needed grows. This is not a pattern unique to language learning. It is a pattern that repeats across product categories whenever a design constraint—the need to measure, the need to scale, the need to produce a grade—gets built into a system so deeply that it stops being visible as a constraint and starts being mistaken for a truth about the problem itself.当用户需要的结果难以直接衡量时,设计过程往往倾向于寻找可以衡量的东西。替代物变成了目标。界面为其优化。产品交付的东西与用户实际需要之间的差距在扩大。这不是语言学习独有的模式。只要一个设计约束——需要衡量、需要扩展、需要产生分数——被深深嵌入系统,以至于不再可见,并开始被误认为是问题本身的真理,这个模式就会在多个产品类别中重复出现。
What happens when the constraint changes#section4当约束发生变化时#section4
The constraint that made the Grammar-Translation method necessary in 1788 was real and rational. One teacher. Thirty students. A standardized exam. You cannot grade a conversation at scale. You can grade a translation exercise. The method was not chosen because it produced fluency. It was chosen because it produced a score.1788年使语法-翻译法成为必要的约束是真实且合理的。一个老师。三十个学生。标准化考试。你无法大规模地给对话打分。但你可以给翻译练习打分。选择这个方法不是因为它能产生流利,而是因为它能产生分数。
That constraint no longer exists in the same form. Technology has made it possible to deliver immersive, real-time conversation practice to anyone with a smartphone, at a cost that continues to fall. The design problem is no longer how to make language learning gradable at scale. It is how to make the conditions of genuine language acquisition accessible to people who cannot move to another country or afford a native-speaker tutor.那个约束不再以同样的形式存在。技术使得向任何有智能手机的人提供沉浸式、实时的对话练习成为可能,且成本持续下降。设计问题不再是如何使语言学习在大规模下可评分。而是如何让真正的语言习得条件对那些无法搬到另一个国家或负担得起母语教师的人变得可及。
The products that are now closest to solving the actual problem are not the ones that invented a new pedagogy. They are the ones that removed the access barrier to an old one. Praktika builds AI conversation partners with distinct personalities, regional dialects, and cultural context, replicating the specificity of a real native speaker rather than a generic language-learning voice. Langua clones native speaker voices so that the interaction feels like a real conversation rather than a lesson. Rosetta Stone’s foundational methodology, image association in the target language with no translation, was built on the same insight Berlitz arrived at in the nineteenth century: language is acquired through immersive exposure, not through analysis of its rules [3]. A 2025 meta-analysis of 31 studies found that AI conversation tools produced a statistically significant improvement in language learning outcomes, a result that no amount of flashcard optimization has consistently matched [4].目前最接近解决实际问题的产品,不是那些发明了新教学法的产品,而是那些去除了旧教学法可及性障碍的产品。Praktika构建了具有不同个性、地区方言和文化背景的AI对话伙伴,复制真实母语者的具体性而非通用语言学习语音。Langua克隆母语者声音,使互动感觉像真正的对话而不是一课。Rosetta Stone的基础方法论——无翻译的目标语言图像关联——基于与伯利兹在19世纪得出的相同洞见:语言通过沉浸式暴露获得,而不是通过分析其规则。2025年一项包含31项研究的荟萃分析发现,AI对话工具在语言学习效果上产生了统计显著的改进,这是任何数量的闪卡优化都未能持续匹配的结果。
None of these products invented a new theory of language acquisition. They translated an existing one into something more people could reach.这些产品没有一个发明了新的语言习得理论。它们只是将一个现有理论翻译成更多人能够触及的东西。
The design question this leaves#section5这留下的设计问题#section5
The Grammar-Translation method persisted not because educators were wrong about design, but because a design decision made under a specific constraint became, over two centuries, indistinguishable from the thing itself. The constraint, how do you grade language at scale, was forgotten. The method it produced was inherited as if it were a description of how language works, passed from Prussia to Europe to America to the App Store, from the grammar drill to the streak.语法-翻译法之所以持续存在,不是因为教育者对设计有误解,而是因为一个在特定约束下做出的设计决定,在两个世纪后变得与事物本身难以区分。那个约束——如何大规模评分语言——被遗忘了。它产生的方法被继承下来,仿佛它是对语言运作方式的描述,从普鲁士传到欧洲,再到美国,再到应用商店,从语法练习到连续打卡。
Every time a design team optimizes for a metric because the actual outcome is hard to measure, they are making a version of the same decision. It is often the right decision given real constraints. The question worth asking is whether the constraint that made it necessary still exists, or whether it has simply become invisible inside the system it originally produced.每当一个设计团队因为实际结果难以衡量而优化一个指标时,他们就在做同一类决定。考虑到实际约束,这往往是正确的决定。值得问的问题是,当初使这个决定成为必要的约束是否仍然存在,还是它已经在最初产生的系统中变得不可见了。
Before reaching for what can be measured, it is worth asking what the user actually needs to do, and what stopped them from doing it before. Sometimes the answer is a new solution. More often it is an old one that was always out of reach.在追求可以衡量的事物之前,值得问问用户实际上需要做什么,以及什么以前阻止了他们这样做。有时答案是一个新的解决方案。更常见的是,它是一个总是遥不可及的旧方案。

This seems like a thinly veiled AI language tech advertisement.
Citations #1 and #5 seem to be AI hallucinations. Citation#1 is an article allegedly published in the Journal of Second Language Studies in 2013 however that journal only started operation in 2018. Citation #5 cannot be located anywhere on the web and is not listed in the Journal of Language and Linguistic Studies issues from 2024-2026
Great read! The depth here is refreshing. You connected the concepts in a way that’s actually actionable, which is rare. Thanks for sharing. Relevant to what I’m working on right now.
Really helpful piece. The way you walked through the logic step by step is exactly what I needed. Will definitely revisit this. Exactly the perspective I needed.