{
  "_note": "Same meaning in every language, so token counts are directly comparable. Composed for this measurement and published with it so any reader can reproduce the numbers exactly.",
  "languages": {
    "English": "A language model does not read characters. It reads tokens, and the number of tokens your text becomes depends on which model you send it to. Two systems can quote the same price per million tokens and still charge you different amounts for the same document, because they do not agree on how many tokens that document contains. This matters most for text that is not English, where the gap between the best and worst case is large enough to change which model is cheaper.",
    "Spanish": "Un modelo de lenguaje no lee caracteres. Lee tokens, y la cantidad de tokens en que se convierte tu texto depende del modelo al que lo envíes. Dos sistemas pueden anunciar el mismo precio por millón de tokens y aun así cobrarte cantidades distintas por el mismo documento, porque no coinciden en cuántos tokens contiene ese documento. Esto importa sobre todo en textos que no están en inglés, donde la diferencia entre el mejor y el peor caso es lo bastante grande como para cambiar qué modelo resulta más barato.",
    "German": "Ein Sprachmodell liest keine Zeichen. Es liest Token, und wie viele Token aus deinem Text werden, hängt davon ab, an welches Modell du ihn schickst. Zwei Systeme können denselben Preis pro Million Token nennen und dir trotzdem unterschiedliche Beträge für dasselbe Dokument berechnen, weil sie sich nicht einig sind, wie viele Token dieses Dokument enthält. Das ist vor allem bei Texten wichtig, die nicht auf Englisch sind, wo der Abstand zwischen bestem und schlechtestem Fall groß genug ist, um zu ändern, welches Modell günstiger ist.",
    "French": "Un modèle de langage ne lit pas des caractères. Il lit des tokens, et le nombre de tokens que devient votre texte dépend du modèle auquel vous l'envoyez. Deux systèmes peuvent afficher le même prix par million de tokens et vous facturer malgré tout des montants différents pour le même document, parce qu'ils ne s'accordent pas sur le nombre de tokens que contient ce document. Cela compte surtout pour les textes qui ne sont pas en anglais, où l'écart entre le meilleur et le pire cas est assez grand pour changer quel modèle revient le moins cher.",
    "Russian": "Языковая модель не читает символы. Она читает токены, и количество токенов, в которые превращается ваш текст, зависит от того, какой модели вы его отправите. Две системы могут объявить одинаковую цену за миллион токенов и всё равно взять с вас разные суммы за один и тот же документ, потому что они не сходятся в том, сколько токенов в этом документе. Это особенно важно для текстов не на английском языке, где разрыв между лучшим и худшим случаем достаточно велик, чтобы изменить, какая модель дешевле.",
    "Hindi": "एक भाषा मॉडल अक्षर नहीं पढ़ता। वह टोकन पढ़ता है, और आपका पाठ कितने टोकन बनता है यह इस पर निर्भर करता है कि आप उसे किस मॉडल को भेजते हैं। दो प्रणालियाँ प्रति दस लाख टोकन एक ही कीमत बता सकती हैं और फिर भी एक ही दस्तावेज़ के लिए आपसे अलग-अलग राशि ले सकती हैं, क्योंकि वे इस बात पर सहमत नहीं हैं कि उस दस्तावेज़ में कितने टोकन हैं। यह उन पाठों के लिए सबसे अधिक मायने रखता है जो अंग्रेज़ी में नहीं हैं, जहाँ सबसे अच्छे और सबसे बुरे मामले के बीच का अंतर इतना बड़ा है कि यह बदल सकता है कि कौन सा मॉडल सस्ता है।",
    "Arabic": "لا يقرأ نموذج اللغة الحروف. إنه يقرأ الرموز، وعدد الرموز التي يتحول إليها نصك يعتمد على النموذج الذي ترسله إليه. يمكن لنظامين أن يعلنا السعر نفسه لكل مليون رمز ومع ذلك يحاسبانك بمبالغ مختلفة على المستند نفسه، لأنهما لا يتفقان على عدد الرموز التي يحتويها ذلك المستند. وهذا مهم بشكل خاص للنصوص غير الإنجليزية، حيث تكون الفجوة بين أفضل حالة وأسوأ حالة كبيرة بما يكفي لتغيير أي النموذجين أرخص.",
    "Chinese": "语言模型并不阅读字符。它阅读的是词元，而你的文本会变成多少词元，取决于你把它发送给哪个模型。两个系统可以标出相同的每百万词元价格，却仍然对同一份文档收取不同的费用，因为它们对这份文档包含多少词元并不一致。这一点对非英语文本尤其重要，因为最好情况与最坏情况之间的差距大到足以改变哪个模型更便宜。",
    "Japanese": "言語モデルは文字を読んでいるのではありません。読んでいるのはトークンであり、あなたの文章がいくつのトークンになるかは、どのモデルに送るかによって変わります。二つのシステムが百万トークンあたり同じ価格を掲げていても、同じ文書に対して請求額が違うことがあります。その文書に何トークン含まれるかについて、両者の見解が一致しないからです。これは英語以外の文章で特に重要で、最良の場合と最悪の場合の差は、どちらのモデルが安いかを覆すほど大きくなります。",
    "Korean": "언어 모델은 문자를 읽지 않습니다. 토큰을 읽으며, 여러분의 텍스트가 몇 개의 토큰이 되는지는 어떤 모델에 보내느냐에 따라 달라집니다. 두 시스템이 백만 토큰당 같은 가격을 제시하고도 같은 문서에 대해 서로 다른 금액을 청구할 수 있습니다. 그 문서에 토큰이 몇 개 들어 있는지에 대해 서로 의견이 다르기 때문입니다. 이는 영어가 아닌 텍스트에서 특히 중요하며, 최선과 최악의 차이는 어느 모델이 더 싼지를 뒤바꿀 만큼 큽니다."
  },
  "content_types": {
    "prose": "Measurement is the only thing that separates a claim from an opinion. When a number is published without the run that produced it, the reader has no way to tell whether it was measured, estimated, or remembered. The remedy is not more confidence in the writing, it is publishing the method alongside the result so that anyone who doubts the figure can reproduce it and find out. A result that cannot be reproduced is a story about a result.",
    "code": "export async function loadConfig(path: string): Promise<Config> {\n  const raw = await fs.readFile(path, 'utf-8');\n  const parsed = JSON.parse(raw) as Partial<Config>;\n  if (!parsed.endpoint) throw new ConfigError('endpoint is required');\n  return {\n    endpoint: parsed.endpoint,\n    timeoutMs: parsed.timeoutMs ?? 30_000,\n    retries: Math.min(parsed.retries ?? 3, 10),\n    headers: { ...DEFAULT_HEADERS, ...(parsed.headers ?? {}) },\n  };\n}",
    "json": "{\"id\":\"evt_8f3a1c92\",\"type\":\"measurement.recorded\",\"created\":1754700000,\"data\":{\"model\":\"claude-opus-5\",\"input_tokens\":58750,\"cache_creation_input_tokens\":43410,\"cache_read_input_tokens\":15602,\"output_tokens\":4,\"duration_ms\":2318,\"reproduced\":true}}",
    "markdown_table": "| flag | default | scope | tokens |\n|---|---|---|---|\n| `tool_search` | true | session | 3054 |\n| `effort` | medium | request | 0 |\n| `model` | opus | request | 0 |\n| `permission_mode` | ask | session | 0 |"
  }
}
