OpenEval v1.0.0 Proposes a Standard JSON Format to Unify LLM Evaluation Frameworks
A new open-source project called OpenEval aims to solve the lack of interoperability among popular LLM evaluation frameworks such as DeepEval, Promptfoo, Inspect AI, and lm-evaluation-harness. Currently, each framework uses its own incompatible format for test cases, grader definitions, and results, forcing teams to manually rewrite datasets when switching tools. OpenEval addresses this by defining a portable JSON Schema spec along with TypeScript and Python SDKs, a CLI, and converters for existing frameworks. The project has released version 1.0.0 on both npm and PyPI, making it available to developers immediately. Over 17 framework integration issues are open on the project's GitHub repository, inviting community contributions to expand converter support.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in