Evaluation engine that measures the ROI of context for AI agent skills. Ruby · blind LLM judge
-
Updated
Aug 14, 2026 - Ruby
Evaluation engine that measures the ROI of context for AI agent skills. Ruby · blind LLM judge
Add a description, image, and links to the ruby-evals topic page so that developers can more easily learn about it.
To associate your repository with the ruby-evals topic, visit your repo's landing page and select "manage topics."