An open LLM benchmark with 249 tests that scores cost alongside accuracy, so teams can see what each correct answer actually costs across models.

Fund this project

Unverified URL

The funding manifest has not provided proof via wellKnown that this link is associated with it. Learn more.

Continue