index / ngallodev-software/agent-workflow-benchmark
agent-workflow-benchmark
Reproducible comparative benchmark harness for Agent-Workflow with sealed evidence, finished software, scoring, timing/token capture, and Jev/TypeSafe studies.
Research and EvalsFeature ScoringReviewPython
Repository
Stars0
LanguagePython
LicenseApache-2.0
Created2026-09-19
Last pushed2026-10-07
Officialno
Judgment
Genuine Jev project0.48
gate 0.5
Uses Jev at runtime0.39
Is a meta list0.05
Is a reimplementation0.11
Quality scales
Substance2.89 / 3
gate 0.5
Docs2.31 / 3
Novelty2.12 / 3
Composite score0.84
Category distribution
Research and Evals0.99
Agent and Dev Tooling0.01
Integrations0.00
Applications0.00
Learning0.00
SDKs and Clients0.00
Games and Simulation0.00
Other0.00
Assigned categoryResearch and Evals
Category probability0.99
Category confidence0.99
PatternFeature Scoring
Pattern confidence0.55
Curation
StatusReview
Reasongenuine 0.48 below 0.5
README pickno
Overridenone
Sources
- search:topic:typesafe-ai
Provenance
Modeljev-1.13.0
Question setv2
Judged at2026-10-07 12:39 UTC