TrueIQ.si
Submit a tool

OpenAI Evals

github.com · LLM Evals

Open source

Framework and open registry of evaluations for testing language models and the systems built on them.

#framework#registry#openai#benchmarks

About the LLM Evals category

Frameworks and platforms for testing language models and LLM applications — from academic benchmark harnesses to CI-friendly unit tests and LLM-as-judge scoring.

OpenAI Evals is one of 15 llm evals tools indexed on TrueIQ. Facts on this page come from the tool's official site and public APIs; prices and features change, so confirm details on the official website before you commit.

// related

Related tools

All LLM Evals →