Code craft
Would you be happy to inherit this?
Correctness, readability, tests, and whether the next person can change it safely.
Bonfire · the Record
Which is exactly why you can trust the rest of it.
Every builder leaves with a Record: six dimensions, rated and signed, each linked to real work.
The gap
It shows code someone wrote, at unknown difficulty, under unknown constraints, with unknown help, reviewed by nobody in particular.
What it cannot tell you: whether they understood the problem or just the ticket. Whether they argued for the simpler thing. How they explained a trade-off to someone non-technical. Whether they caught the agent's confident, plausible, wrong answer. Whether anyone competent ever looked at it.
What's measured
One of them is the software. The other five decide whether the software was worth writing, and they are the ones that got scarce.
Would you be happy to inherit this?
Correctness, readability, tests, and whether the next person can change it safely.
Do they build the right thing, or just the thing asked for?
Whether they speak up when the request and the real problem have come apart.
Does the team think more clearly because they are in it?
Explaining a trade-off to a non-engineer. Disagreeing well. Leaving a trail others can follow.
Do they see what this touches?
Failure modes, what else breaks, and the thing that goes wrong in production at 3am.
Will this still look like a good call in a year?
Choosing boundaries, resisting cleverness, knowing which decisions are hard to undo.
Do they catch the confident, plausible, wrong answer?
Directing AI, reviewing what comes back, and knowing when to throw it away.
The scale
Scores out of ten travel badly between companies. Every interview is really asking one thing: how much of my time will this person cost, and for how long? We answer it with evidence.
Does good work once the problem is framed and the approach is agreed. Learning quickly, and worth the time it takes.
Frames the smaller problems alone. Knows which decisions to bring to someone else, and brings them early rather than late.
Owns an outcome end to end. Escalates the right things at the right time, and very little else.
Raises the standard around them. Makes the calls the rest of the pod follows, and is usually right about them.
Why this scale
Nobody is uniformly senior.
Most people are two different levels on two different dimensions, and that shape matters more than any average. A builder who sets direction on communication and needs guidance on architecture is an excellent hire for one team and the wrong hire for another. A single seniority label hides that. The Record shows it.
A worked example
Built the way a real one is, including the parts that do not flatter.
Reviewer's recommendation
Ready for a mid-level product engineering role on a team that already has an established senior. Not yet ready to be the only engineer on a greenfield system, which is what the systems and architecture ratings above are describing. Three months on a service with real cross-boundary complexity would close most of that gap. Hire her for judgement and communication, and expect the architecture to catch up faster than you would assume.
Signed by the principal who reviewed the work, across all three engagements.
Illustrative sample. A real record names its reviewer and links every rating to the artefacts behind it.
Why it can be trusted
The pull request, the design doc, the demo. You can go and look.
Not the studio in the abstract. The principal who was hands-on with the work, with their name on the assessment.
A record that only says good things is marketing. Ours say what someone cannot do yet, in writing. That is what makes the rest of it worth reading.
If we vouched for everyone equally, we would be vouching for nobody. The reputation only works if it can be spent, which means sometimes saying the less comfortable thing about someone we like.
light the fire
If you are hiring, we will walk you through the Records of builders who fit. If you are building, we'll talk you through how it works.