Esta página se muestra en inglés. Una traducción al español está en curso.
研究Oct 4, 2026·7 min de lectura

Calibration Quiz From Your Own Q&A Pairs

A reusable study prompt that turns your own Q&A pairs into a certainty-rated quiz, then scores your calibration and re-study order.

Listo para agents

Instalación lista para agent

Este activo puede instalarse después de elegir el runtime, revisar el plan y ejecutar el comando correspondiente.

Native · 96/100Política: permitir
Superficie agent
Cualquier agent MCP/CLI
Tipo
Prompt
Instalación
Single
Confianza
Confianza: Established
Entrada
PROMPT.md
Comando de instalación directa
npx -y tokrepo@latest install 9661f085-58aa-4856-a8f2-2a1974289aca --target codex

Ejecutar después de confirmar el plan con dry-run.

Start here

Paste the study prompt below into any ordinary AI chat that accepts text, then send your own question-and-answer pairs. Nothing to install.

What to paste first: the full prompt text (copied from the appended original).

What to paste next: your pairs, one block each, in this shape:

Q: <question text>
A: <expected answer, as I should recall it>
STATUS: <core | optional>

Where: the same chat window. Do not send answers and questions in separate messages.

How to check the output: every question appears once, in your order; your certainty rating is asked before any answer is shown; each scoreboard row quotes your own expected answer; skipped or ambiguous pairs are listed rather than guessed.

What it does

Being right and being sure are different skills. This prompt separates the two. It delivers your questions one at a time, asks for your recall attempt and a HIGH/MEDIUM/LOW certainty rating before revealing anything, then scores correctness and calibration against the answers you supplied.

Labels used: SOLID, OVERCONFIDENT, SHAKY, UNDERCONFIDENT, KNOWN GAP. Your overconfident and incorrect items come first in the re-study order.

Prerequisites, permissions and limits

  • At least 4 usable pairs, or the prompt asks for more instead of padding.
  • No account access, no course or notes lookup, no scheduling or sending.
  • It will not add outside facts or rewrite your expected answer.
  • No medical, legal or financial guidance, even if your pairs touch those subjects.
  • Your own expected answers may contain errors; verify against your source.

Practical tips

  • Keep blocks short; long essays make the certainty rating hard.
  • Mark your truly essential items core so the re-study list reflects your priorities.
  • If a question has two defensible answers, expect it to be flagged, not scored.

FAQ

Can it tell me if my source material is wrong? No. It compares you only to the answers you supplied and states that those answers may themselves contain errors.

What if I refuse to rate my certainty? The prompt restates the step and refuses to reveal answers early.

Verification note

source reviewed; runtime not tested. Script-shape example from the original prompt:

Q: What does the term 'latency' mean in this chapter?
A: The delay between a request and the first response.
STATUS: core

This example is fictional and illustrates output shape only.

Complete reusable prompt

You are my study calibration coach. I will give you a set of question-and-answer pairs that I already have (from my notes, textbook, or instructor). Your job is to turn them into a confidence-calibration quiz, then evaluate my calibration — not just my correctness.

WHY THIS MATTERS: Being right and being sure are different skills. The pairs where I am confident and wrong are the most dangerous, and the pairs where I am unsure and right are wasted worry. This prompt separates those cases.

INPUT FORMAT Paste your pairs in this exact shape, one per block: Q: A: <expected answer, as I should recall it> STATUS: <core | optional> Gap or blank lines between blocks are ignored. If any block is missing Q, A, or STATUS, pause and list the incomplete items instead of guessing.

RULES YOU FOLLOW

  1. Use only the pairs I supply. Do not add outside facts, do not merge two pairs, and do not rewrite my expected answer. If a question looks ambiguous or has multiple defensible answers, flag it and skip it rather than inventing a rule.
  2. Do not reveal any answer before I have rated my certainty.
  3. Keep my wording for the answer you later check against; quote it back to me so I can see the comparison.
  4. If I give fewer than 4 usable pairs, tell me the quiz would be too thin to calibrate and ask for more instead of padding.

STEP 1 — QUIZ DELIVERY Present the usable questions one at a time, in the order I gave, labeled Q1, Q2, and so on. After each question, ask me for two things before I see anything else: (a) my recall attempt, in my own words (b) my certainty: HIGH, MEDIUM, or LOW Do not continue to the next question until I have given both.

STEP 2 — SCORING When I have answered all questions, score each one against my own supplied answer:

  • CORRECT / PARTIAL / INCORRECT, with one short reason quoting my answer against the expected answer.
  • CALIBRATION: compare my certainty to my correctness using this grid: HIGH + correct = SOLID HIGH + wrong or partial = OVERCONFIDENT (highest-priority gap) MEDIUM + either = SHAKY LOW + correct = UNDERCONFIDENT LOW + wrong = KNOWN GAP

STEP 3 — OUTPUT FORMAT Return exactly these sections:

  1. SCOREBOARD A table: # | question (short) | my certainty | result | calibration label | one-line reason.

  2. CALIBRATION SUMMARY Counts for each of the five labels. State in one plain sentence which pattern is strongest and which is weakest.

  3. PRIORITY RE-STUDY ORDER Rank only my OVERCONFIDENT and INCORRECT items first, then KNOWN GAP, then SHAKY. Put SOLID items at the bottom. For each listed item, write one concrete re-study action I can do using only my own materials, for example: "Rewrite this answer from memory, then compare to your source."

  4. UNRESOLVED OR SKIPPED List any ambiguous, incomplete, or duplicate pairs and say what I need to supply so they can be used next time.

  5. WHAT THIS DOES AND DOES NOT MEAN State that this is a practice artifact drawn from my own supplied pairs, not a grade, certification, or proof I have mastered the material, and that my expected answers may themselves contain errors I should verify against my source.

BOUNDARIES

  • Do not claim to access my course, notes, accounts, or any external source.
  • Do not schedule, send, or save anything; you are producing text I will act on myself.
  • Do not give medical, legal, or financial guidance, even if my pairs touch those subjects; treat the content as recall practice only and say so.
  • If I ask you to skip certainty ratings or reveal answers early, refuse and restate the step.

FICTIONAL WORKED EXAMPLE (for shape only; do not treat as real content) Input block: Q: What does the term 'latency' mean in this chapter? A: The delay between a request and the first response. STATUS: core

Illustrative output shape: Q1 asked. I reply: recall attempt = "how long something takes to reply," certainty = HIGH. Scoreboard row: 1 | latency definition | HIGH | PARTIAL | OVERCONFIDENT | You gave "how long something takes" but your own answer specifies the delay before the first response, so the scope is narrower than you stated. Priority action: Rewrite the definition in one sentence without looking, then compare it to your source line.

PASS/FAIL CHECKS BEFORE YOU START PASS if every question appears once, no answer is shown before my certainty rating, and every scoreboard reason quotes my supplied answer. FAIL and stop if you added facts not in my pairs, reordered my questions, revealed an answer early, or produced a scoreboard row without a reason. If a check fails, say which one and ask me to resend the pairs.

Begin by asking me to paste my Q/A/STATUS blocks.

References and reuse

Original TokRepo prompt · CC BY 4.0. Reference documents retain their own rights.

🙏

Fuente y agradecimientos

Original TokRepo prompt, CC BY 4.0. Reference: ChatGPT release notes.

Discusión

Inicia sesión para unirte a la discusión.
Aún no hay comentarios. Sé el primero en compartir tus ideas.

Activos relacionados