index / victorbvieira/system-one-lab
system-one-lab
Benchmarking System One models against LLMs for typed decisions in Python. Jev vs. LLM on routing, urgency triage and content selection, with reproducible evals.
Research and EvalsRoutingReview
Repository
Stars0
Language—
LicenseApache-2.0
Created2026-09-21
Last pushed2026-09-21
Officialno
Judgment
Genuine Jev project0.71
gate 0.5
Uses Jev at runtime0.29
Is a meta list0.09
Is a reimplementation0.20
Quality scales
Substance0.14 / 3
gate 0.5
Docs0.12 / 3
Novelty1.54 / 3
Composite score0.18
Category distribution
Research and Evals1.00
Learning0.00
Other0.00
Applications0.00
Agent and Dev Tooling0.00
Games and Simulation0.00
Integrations0.00
SDKs and Clients0.00
Assigned categoryResearch and Evals
Category probability1.00
Category confidence1.00
PatternRouting
Pattern confidence0.56
Curation
StatusReview
Reasonsubstance 0.14 below 0.5
README pickno
Overridenone
Sources
- search:"system one" jev
- search:jev in:name,description,readme created:2026-09-21
Provenance
Modeljev-1.13.0
Question setv2
Judged at2026-09-21 12:35 UTC