Skip to content
#

meta-evaluation

Here are 6 public repositories matching this topic...

Language: All
Filter by language

A local LLM benchmarking framework designed to evaluate model performance across multiple backends (LM Studio, Ollama, etc.), including metrics for speed, quality, and instruction adherence. Supports structured test runs, result analysis, and reproducible evaluations.

  • Updated May 3, 2026
  • HTML

Add this topic to your repo

To associate your repository with the meta-evaluation topic, visit your repo's landing page and select "manage topics."

Learn more