J
JevBench
Open SourceReproducible benchmark suite for evaluating typed decision models
๐ณ Self-Hostable๐ No Sign-upโก Traction Score: 81/100โ
143 Stars
pip install jevbenchJevBench provides a standardized, reproducible evaluation framework specifically designed for typed decision models. It bridges the validation gap in complex decision logic, enabling engineers and researchers to reliably measure model accuracy and consistency.
Evaluate models against strict type-safe decision schemas and constraints.
Standardized datasets and execution harnesses ensure consistent, deterministic benchmarking results.
Easily plug in custom decision models and domain-specific validation datasets.
Benchmarking newly trained typed decision models against established baselines
Validating decision logic consistency across multiple model iterations
Publishing reproducible evaluation results for academic or industry research
Machine learning researchers, data scientists, and systems engineers building automated decision-making pipelines.
Compare other trending developer tools and open-source projects in this space.