Evidence ledger
TypeSafe reports Jev as materially faster than LLM workflows on its own evaluations.
The launch material reports large latency multiples; X discussion mostly repeats those figures.
Vendor ClaimClaim 2
Evidence
The launch material reports large latency multiples; X discussion mostly repeats those figures.
Counterarguments
No independent benchmark in the collected dataset reproduces the headline range on representative workloads.
Open questions
What are p50/p95 latency and accuracy under equal task definitions and concurrency?
This status describes how well the claim is supported by the collected evidence, not whether it is ultimately true. A Vendor Claim is a statement TypeSafe has made about its own product; repetition by others does not upgrade it.
Sources
Limitations
- No independent benchmark in the collected dataset reproduces the headline range on representative workloads.
- What are p50/p95 latency and accuracy under equal task definitions and concurrency?
Where this claim matters
- MAGI System on Jev ObservedAn open-source three-sage voting experiment inspired by Neon Genesis Evangelion.
- Probabilistic predicate + deterministic action PlausibleJev supplies fuzzy predicates while TypeScript, policies, and workflows execute constrained actions.
- Decision Regression Harness Authored HypothesisA test runner recording typed decisions and calibration metrics over versioned scenario suites.
- Research Claims Authored HypothesisImportant claims with status, evidence, counterarguments, sources, and open questions.