Skip to main content

Assess · architecture

How judging works behind the scenes

Two AI crews and a calculator, where plain arithmetic owns every published number.

Meet the judging agents

The Evidence team, five specialists in a relay

Writing a dossier is split into five small, single-purpose jobs, each sized so an affordable model does it well, instead of one giant “analyse this project” prompt only the most expensive models could attempt.

Surveyor

Quick model

“What is this codebase made of?” — languages, frameworks and logical groups, inferred from file names and dependencies alone. It never reads the code.

no tools

The judging

Three passes, then the maths takes over

Judging one project for one sponsor takes exactly three AI passes. The Panel and the Fit Specialist run in parallel, the Writer waits for their scores, and the Calculator is code.

Evidence dossier

Everything the judge is allowed to rely on. The judge never re-runs the investigation — it scores from this and only double-checks details.

The score

How the technical score is calculated

The four criterion scores go through a fixed weight table that sums to 100, and then a guardrail keyed to what the code actually proved, not to how the project presented itself.

The credibility cap

A polished pitch cannot buy an engineering score. Only verified code can.

Technical roll-up

The four criterion scores through the fixed weight table: one 0-100 engineering number, before any guardrail is applied.

Use the engine

Same judge, two ways in

Every score you have just read about is reachable either as a hosted event or as an API call, with the same citations behind it.

Hosted

Run it on your event

Bring the rubric. We read every repo, cite every score, and give your judges a board instead of a spreadsheet.

API

Call it from your stack

Point it at a repository, bring your own weights, and render the verdict in your own product.