We have launched OpenEval (v1.0.0-rc.1), an open standard (Apache 2.0) for portable LLM evaluation datasets. Packages on npm and PyPI. 5 active conversations with contributors from Inspect AI, CrewAI, AutoGen, and Arize. Would love to collaborate. Spec: https://github.com/adhabnr-ux/openeval/blob/main/spec/SPEC.md
We have launched OpenEval (v1.0.0-rc.1), an open standard (Apache 2.0) for portable LLM evaluation datasets. Packages on npm and PyPI. 5 active conversations with contributors from Inspect AI, CrewAI, AutoGen, and Arize. Would love to collaborate. Spec: https://github.com/adhabnr-ux/openeval/blob/main/spec/SPEC.md