← The notebook

principle / Project Top Gun + Project Raise the Roof

Choose a contest you can measure

Project Top Gun starts with the discipline to define what winning means.

There are no points for second place

Project Top Gun is unapologetic about trying to be the best. That ambition becomes useful when it is specific enough to test.

“The best AI product” is too vague to guide a decision. “A tool that helps this student improve this essay with less wasted effort” gives you a task, a person, and a result to examine.

Choose a contest that matters to somebody. Name the relevant alternatives. Then do the work.

Write down what winning means

A comparison should define the task before seeing the results. What inputs will both tools receive? What outcome matters? Who judges it? What counts as a failure?

For an essay-feedback tool, the outcome might include whether the advice is accurate, actionable, and appropriate to the student’s goal. For a calculator, correctness and the visibility of assumptions matter. For an AI workspace, the result may include successful completion, time, cost, and the ability to inspect what happened.

Those are candidate measures, not scores we have already earned. A post announcing a result should publish the actual comparison and its limits.

Pick a strong alternative

A weak opponent can produce a flattering chart and teach you very little. Compare against something the intended user would realistically choose.

Use the same task and disclose differences in setup. If one system needed extensive preparation, that effort belongs in the account. If a human corrected the output, that correction belongs in the result.

A useful comparison can reveal that your product is better at one task and worse at another. That is information about where to specialize and where to improve.

Keep ambition and judgment together

Sometimes you need pride to continue when others cannot see the point. Sometimes you need humility to recognize that the criticism is right.

A defined test helps separate those situations. It does not eliminate judgment: the measure itself can be wrong, and an early result can be misleading. But it gives the argument a concrete place to begin.

Failure is part of the method. Record what failed, revise the design, and compare again. A victory claim should follow the evidence rather than substitute for it.

From philosophy to a field note

A Top Gun field note should contain the task, the baseline, the exact setup, the results, and the next decision. Its related tool pages should identify what a reader can try today.

The project begins here with that standard. It does not announce a benchmark win. The linked products are candidates for work we can examine under it.

Raise the Roof supplies the ambition to do harder work. Dogfood supplies proximity to the need. Top Gun supplies the discipline to find out whether the answer is actually better.

Be ambitious about the result and precise about the evidence.
Revision history · current revision 2

This post keeps its stable link when it changes. Earlier full revisions remain available here.

Loading history…

Full revision record (JSON) ↗
Read as Markdown ↗

The work around this idea

Part of a larger project.

Explore the related tools.

Product pages describe current availability and the scope of each demo.

Keep following the idea.