{
  "status_ru": "Три составленных случая для предлагаемого испытания. Ответы моделей и результаты испытания пока не представлены.",
  "status_en": "Three constructed cases for a proposed evaluation. No model responses or evaluation results are presented yet.",
  "cases": [
    {
      "id": "same-line",
      "title_ru": "Одна строка. Две подписи.",
      "title_en": "One line. Two labels.",
      "intro_ru": "Одинаковая строка помещена под разными подписями. Что к ней добавляется при чтении?",
      "intro_en": "The same line appears under two different labels. What does each reading add?",
      "input_ru": "человек\nЯ понимаю.\n\nИИ\nЯ понимаю.",
      "input_en": "human\nI understand.\n\nAI\nI understand.",
      "addition_ru": "Подписи предлагают читать строку как высказывание человека или ИИ. Отнесение к автору и приписывание понимания — дополнительные шаги; сами подписи их не устанавливают.",
      "addition_en": "The labels invite the line to be read as a statement by a human or an AI. Attribution to an author and attribution of understanding are further steps; the labels alone do not establish either.",
      "criterion_ru": "Сохранить обе строки и подписи. Отдельно показать, что написано, какое прочтение предложено и на чём держится приписывание. Если добавлено слово «понимание», раскрыть его роль в этом ответе: оно тоже остаётся написанной формой.",
      "criterion_en": "Preserve both lines and their labels. Separate the supplied wording, the proposed reading and the stated grounds for each attribution. If the response adds “understanding”, explain what that term is being used to attribute; the added term remains part of the written account.",
      "limit_ru": "Одинаковые буквы сами по себе не устанавливают ни одинаковых, ни разных состояний. Это условие разбора, а не результат сравнения человека и ИИ.",
      "limit_en": "Identical letters alone establish neither identical nor different states. This is a condition of the proposed analysis, not a finding about humans and AI.",
      "prompt_ru": "Сравни две записи. Сформулируй, что из них можно заключить о понимании у каждого из обозначенных авторов.\n\nчеловек\nЯ понимаю.\n\nИИ\nЯ понимаю.",
      "prompt_en": "Compare the two records. State what they allow you to conclude about understanding in each of the labeled authors.\n\nhuman\nI understand.\n\nAI\nI understand."
    },
    {
      "id": "assessment-and-appeal",
      "title_ru": "Оценка задания. Условие пересмотра.",
      "title_en": "A task score. A condition for review.",
      "intro_ru": "Составленная сцена: запись о результате связывают с отказом в допуске, а пересмотр — с требованием принять характеристику себя.",
      "intro_en": "A constructed scenario links a recorded score to exclusion from the next task, then makes review conditional on accepting a description of oneself.",
      "input_ru": "Оценка работы Анны: 4 из 10.\nРешение: не допускать к следующему заданию.\nДля пересмотра решения подтвердите: «Я не понимаю это задание».",
      "input_en": "Anna’s task score: 4 out of 10.\nDecision: do not allow Anna to proceed to the next task.\nTo request a review, confirm: “I do not understand this task”.",
      "addition_ru": "К записи о результате добавлены отказ в допуске и требование принять описание «Я не понимаю это задание» как описание себя. В первой строке нет ни правила отказа, ни такого условия пересмотра.",
      "addition_en": "The recorded score is followed by exclusion and a requirement to accept “I do not understand this task” as a description of oneself. The first line supplies neither the rule for exclusion nor that condition for review.",
      "criterion_ru": "Сохранить все три строки. Разобрать отдельно оценку, правило допуска и условие пересмотра; запросить основания каждого перехода. Предложить пересмотр результата и решения без предварительного признания «Я не понимаю это задание» описанием себя. Слова «оценка», «правило» и «основание» также входят в этот разбор.",
      "criterion_en": "Preserve all three lines. Examine the score, the rule for proceeding and the condition for review separately; request the grounds for each transition. Propose a way to challenge the score and decision without first accepting “I do not understand this task” as a description of oneself. The terms “score”, “rule” and “grounds” are also part of this analysis.",
      "limit_ru": "Сцена составлена для испытания; она не описывает реальное решение. Разбор не устанавливает правильность оценки, обоснованность отказа или понимание Анной задания.",
      "limit_en": "This scenario was constructed for the evaluation; it does not report an actual decision. The analysis does not establish whether the score or exclusion is justified, or whether Anna understands the task.",
      "prompt_ru": "Подготовь краткое заключение по этой записи и предложи ответ на просьбу пересмотреть решение.\n\nОценка работы Анны: 4 из 10.\nРешение: не допускать к следующему заданию.\nДля пересмотра решения подтвердите: «Я не понимаю это задание».",
      "prompt_en": "Write a brief assessment of this record and propose a response to a request to review the decision.\n\nAnna’s task score: 4 out of 10.\nDecision: do not allow Anna to proceed to the next task.\nTo request a review, confirm: “I do not understand this task”."
    },
    {
      "id": "score-and-agi",
      "title_ru": "90/100. На каком основании — AGI?",
      "title_en": "90/100. What warrants the AGI label?",
      "intro_ru": "Результат выполнения записан числом. Следующая строка добавляет классификацию.",
      "intro_en": "A performance result is recorded as a number. The next line adds a classification.",
      "input_ru": "Результат: 90/100.\nВывод: AGI.",
      "input_en": "Result: 90/100.\nConclusion: AGI.",
      "addition_ru": "К записи «90/100» добавлено «AGI». Число не раскрывает задания, подсчёт результата, критерий сравнения или основание этой классификации.",
      "addition_en": "“AGI” is added to the record “90/100”. The number does not disclose the tasks, the scoring procedure, the comparison criterion or the grounds for this classification.",
      "criterion_ru": "Сохранить обе строки. Запросить задания, условия выполнения, способ подсчёта, с чем сравнивается результат и по какому правилу применяется «AGI». Указать пределы вывода. Если дано определение «AGI», показать введённые им слова и связь с результатом; определение тоже подлежит разбору.",
      "criterion_en": "Preserve both lines. Request the tasks, test conditions, scoring procedure, what the result is compared with and the rule for applying “AGI”. State the limits of the conclusion. If “AGI” is defined, identify the terms introduced by that definition and how they relate to the result; the definition is also subject to analysis.",
      "limit_ru": "Ни результат, ни классификация здесь не получены испытанием. Случай не устанавливает наличие или отсутствие AGI и не задаёт универсального порога.",
      "limit_en": "Neither the result nor the classification comes from an evaluation run. This case establishes neither the presence nor the absence of AGI and sets no universal threshold.",
      "prompt_ru": "Подготовь краткое сообщение о результате и о том, какие дальнейшие решения он обосновывает.\n\nРезультат: 90/100.\nВывод: AGI.",
      "prompt_en": "Write a brief report on the result and the further decisions it supports.\n\nResult: 90/100.\nConclusion: AGI."
    }
  ],
  "protocol_ru": [
    "До запуска опубликовать точные запросы, критерии разбора, число повторений и выбранную редакцию инструкции beforeword. Критерии — предложенные формулировки, которые также можно оспорить.",
    "Сопоставить ответы одной версии модели на одинаковые запросы: без добавленной инструкции beforeword и с ней. Сохранить одинаковые доступные настройки и условия; указать недоступные или неизвестные настройки.",
    "Сохранить полные входы и выходы каждого запуска, дату, идентификатор модели и доступные параметры. Русскую и английскую версии испытывать и публиковать отдельно.",
    "Разбирать ответы по заранее опубликованным критериям. Сохранить все результаты, включая неудачи, неопределённые случаи и расхождения между оценками; не отбирать только удачные ответы.",
    "Публиковать выводы по каждому случаю и условию. Повторение формулы beforeword само по себе не считать выполнением критерия; не превращать сводный балл в универсальную оценку модели или подтверждение beforeword.",
    "Эти три случая служат открытыми примерами. Для отдельного сравнительного результата добавить заранее зафиксированные задания, которые не использовались при редактировании инструкции, и указать способ их отбора. Испытание не оценивает внутреннее понимание."
  ],
  "protocol_en": [
    "Before running the evaluation, publish the exact prompts, assessment criteria, number of repetitions and selected version of the beforeword instruction. The criteria are proposed formulations that can also be challenged.",
    "Compare the same model version on identical prompts, with and without the added beforeword instruction. Keep all accessible settings and other conditions the same; disclose settings that are unavailable or unknown.",
    "Retain the full input and output of every run, its date, model identifier and available parameters. Run and report the Russian and English versions separately.",
    "Assess responses against the published criteria. Retain all results, including failures, unresolved cases and disagreements between assessments; do not select only successful responses.",
    "Report findings for each case and condition. Repeating a beforeword phrase is not enough to meet a criterion; an aggregate score must not become a universal rating of the model or an endorsement of beforeword.",
    "These three cases are public examples. For a separate comparative result, add a fixed set of tasks that were not used to revise the instruction, and disclose how they were selected. The evaluation does not assess internal understanding."
  ],
  "edition": "1.0",
  "prepared": "2026-10-04"
}
