Model finds
The judge extracts line-cited evidence. It never assigns a score.
The Bench · Code measurement
A throwaway script and a payment boundary can solve the same problem and score nothing alike, because they were built for different things. Facet refracts a file into 14 measured dimensions and reads it against the profile it was written for.
one number infourteen faces outnever a single grade
Profile a fileOpen a sample report
Free while in beta. No card. The standard web profile does not retain your plaintext source. A quick security check needs no account at all.
Building with an AI coding agent? Join the API, CLI, and MCP before-and-after beta.
Every codebase is optimising for something - speed, safety, throughput, changeability. If nobody chose the trade, the deadline chose it. The documented bill for leaving that unexamined:
The full case, with receipts - and what it means if you build fast with AI →
Every other tool hands you a single grade. That grade hides exactly the part you need: the tradeoffs. Facet puts a file on the bench and lets one beam of measurement refract into fourteen dimensions, each measured apart, each with its evidence cited to the line.
Hover or focus a dimension to see its cited evidence in the sample artifact. The probe pane lights the exact lines the judge read. Poles never net: capability and violation stay separate, always.
basis the account lookup is parameterised; no string-built SQL reaches the driver
The model never scores. It extracts line-cited evidence, and deterministic code sets every level. Same code, same profile: test-retest 0.94 on real code, frozen per content hash.
The weights are not our opinion. Where a dimension is weighted, the weights come from external standards: SonarQube rule severities, corroborated by NVD CVSS scores.
We attacked our own scoring. A 40-agent adversarial review tried to refute every security weight. It found real defects. Every one was fixed.
Security fails closed. One detected critical defect caps the security score no matter how much good practice surrounds it.
test-retest on real code · floor 0.90 enforced
two independent standards, one tier order - provisional
40-agent review · 0 defects survived
one critical defect sets the ceiling, whatever surrounds it
A short walkthrough: paste a file, watch the beam split into fourteen measured dimensions, and read each one against the profile the code was written for.
The beta question
Disagreement is useful when it is precise. Facet asks whether a dimension reads high, low, or about right, so beta feedback becomes calibration data. Free while in beta.
The judge extracts line-cited evidence. It never assigns a score.
Deterministic code sets each level. Same file, same profile, same result.
The same code can be right or wrong depending on what it was built for.
Managed durable scans temporarily hold an encrypted selected slice, then delete it after completion or expiry.