JevTracks

Submit your tool

← JevTracks

Python — Probability-aware evaluation for typed decision models: calibration, selective risk, latency, and reproducible benchmarks.

AbdelStarkevaluation-benchmarking12 stars · 0 forks

First discovered , last refreshed . Descriptions and stats are pulled from the project's own GitHub repo and refreshed automatically — they aren't independently verified by JevTracks beyond the initial eligibility check.