shikumi-eval: Typed evaluation framework for shikumi LM programs (EP-8)

[ ai, bsd3, library ] [ Propose Tags ] [ Report a vulnerability ]

The evaluation framework for shikumi: the owned data model (Example, Prediction, Dataset, Metric, Score, Report — MasterPlan integration point #5), built-in pure and LM-backed metrics, an evaluate runner that scores a Shikumi.Program.Program over a typed dataset with bounded parallelism and per-example error boundaries, and golden testing that pins a program's behaviour deterministically under a mock or replayed LM.

Downloads

Maintainer's Corner

Package maintainers

For package maintainers and hackage trustees

Candidates

  • No Candidates
Versions [RSS] 0.1.0.0, 0.1.0.1, 0.1.1.0, 0.2.0.0, 0.2.0.1, 0.2.0.2, 0.2.0.3
Change log CHANGELOG.md
Dependencies aeson (>=2.2 && <2.3), baikai (>=0.6 && <0.7), base (>=4.20 && <5), bytestring (>=0.11 && <0.13), containers (>=0.6 && <0.9), effectful (>=2.5 && <2.7), generic-lens (>=2.2 && <2.4), lens (>=5.3 && <5.4), shikumi (>=0.3.0.0 && <0.4), tasty (>=1.4 && <1.6), tasty-golden (>=2.3 && <2.4), text (>=2.1 && <2.2), vector (>=0.13 && <0.14) [details]
License BSD-3-Clause
Author Nadeem Bitar
Maintainer nadeem@gmail.com
Uploaded by shinzui at 2026-08-29T17:22:47Z
Category AI
Distributions
Reverse Dependencies 1 direct, 0 indirect [details]
Downloads 57 total (24 in the last 30 days)
Rating (no votes yet) [estimated by Bayesian average]
Your Rating
  • λ
  • λ
  • λ
Status Docs uploaded by user
Build status unknown [no reports yet]