Pinned Loading
-
metajudge
metajudge PublicA reliability and DIF report card for LLM-judge and human-rater scoring instruments.
Python
-
digital-health-survival-casebook
digital-health-survival-casebook PublicReproducible simulation study: survival analysis, psychometrics (CFA/IRT), behavioral segmentation, and churn prediction on a fully simulated smoking-cessation digital-health trial, scored as param…
Python
-
simply-put
simply-put PublicElixir/Phoenix rig for LLM plain-language rewriting: a deterministic Flesch-Kincaid gate for reading grade, plus a cross-vendor LLM judge calibrated against a human agreement ceiling for meaning-pr…
Elixir 1
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.