We have launched OpenEval (v1.0.0-rc.1), an open standard (Apache 2.0) for portable LLM evaluation datasets. Packages published on npm and PyPI. We have 17 issues filed across eval frameworks with active responses from Inspect AI, CrewAI, and Arize. Would love to collaborate on import/export support. Spec: https://github.com/adhabnr-ux/openeval/blob/main/spec/SPEC.md
We have launched OpenEval (v1.0.0-rc.1), an open standard (Apache 2.0) for portable LLM evaluation datasets. Packages published on npm and PyPI. We have 17 issues filed across eval frameworks with active responses from Inspect AI, CrewAI, and Arize. Would love to collaborate on import/export support. Spec: https://github.com/adhabnr-ux/openeval/blob/main/spec/SPEC.md