Syntaxe is a technology company: a systems lab building inspectable computational engines for security, law, science and critical infrastructure, plus an engineering practice for organisations solving difficult technical problems.
Reach the lab at hello@syntaxeltd.com · Syntaxe Ltd.
Every builder is structurally the worst possible judge of their own idea’s novelty. This is not a character flaw, it is a selection effect: an idea survives long enough to reach a founder’s attention specifically because that founder found it exciting, and excitement is not evidence of originality, only of unfamiliarity to the one person in the room least likely to have already seen it somewhere else. So we don’t ask the builder to grade the builder’s own work. We ask someone whose only incentive is to kill it.
The failure mode is not stupidity, it is structure. Building something requires sustained belief in it, and sustained belief is hard to maintain while simultaneously searching, in earnest, for the specific fact that would prove the belief wrong. Nobody sets out to skip the search. It happens because a team that has spent months building genuinely stops seeing the version of the search that would hurt, and starts, without noticing, running the version that confirms what they already believe. An external reviewer with zero investment in the outcome doesn’t have that problem. They have no months of work to protect.
We took twelve ideas, drawn from a dozen different domains, each one something a team inside the lab was genuinely excited to build, and put every one of them in front of an independent adversarial reviewer with a single instruction: find the reason this already exists, or the reason it won’t work, and do not stop looking until you either find it or run out of places to look.
The reviewer receives only the idea and its stated justification, deliberately withheld from the internal reasoning that produced it, so the review starts from the same position a competent outsider encountering the idea cold would start from. Their mandate has exactly one success condition: find the disqualifying fact, or exhaust the search. An idea that survives is marked SURVIVE. One with a real but non-fatal prior-art overlap is marked SKETCH, worth revisiting once the overlap is understood, not worth resourcing as-is. One that already exists, or fails for a concrete reason, is marked KILL.
Every one of the twelve had already cleared its own internal gate. The teams building them were confident, for good reason: the ideas were coherent, and nothing about the review process up to that point had surfaced a disqualifying problem.
The outside check killed or sketched all twelve. Not most. All. And it did so, almost without exception, on prior art the builder had no way of seeing from inside the process that produced the idea. In several cases the disqualifying prior art wasn’t obscure: it was published, sometimes years earlier, sometimes by a well known lab, using different vocabulary for the same underlying mechanism. In a smaller number of cases the idea hadn’t been published anywhere, but the reviewer identified a specific, near-term reason it would fail once it left the whiteboard, the kind of failure visible only to someone actively trying to find one.
Enthusiasm is not evidence, and internal confidence measures the wrong thing: it measures how thoroughly a team has convinced itself, not how thoroughly the world has already covered the same ground. A thing is novel only when someone whose entire job is to want it to fail, tries, and can’t. That is a much higher bar than “our team likes this,” and it is the only bar that actually correlates with the thing being worth building.
This is no longer a one-off exercise; it is a standing gate. Every candidate direction inside the lab now goes through the same adversarial check before it receives real resourcing, on the same terms: an outside reviewer, motivated only to kill it, given as long as it takes. Most of what gets proposed does not survive. What does survive has already been tested against exactly the kind of scrutiny it will eventually face from the outside world anyway, just earlier, and cheaper, than finding out after the fact.
This is the same claim the fragment on this site makes about the well: capable minds converge, so the idea was rarely going to be the moat. This report is that claim, tested on ourselves, with a number attached. Twelve for twelve.