Benchmarking pipeline for image/video generation models — blind pairwise preference studies and Bradley–Terry ratings with bootstrap CIs
-
Updated
Aug 13, 2026 - Python
Benchmarking pipeline for image/video generation models — blind pairwise preference studies and Bradley–Terry ratings with bootstrap CIs
High-fidelity benchmarking & observability framework for 11-tier microservices across Public (Azure), Private (IITD Baadal), Multi-Cloud (Azure+GCP Mesh), and Edge (K3s) topologies using Prometheus, Grafana, and Locust. 🌩️📈
Not a product. Not a framework. Nothing here is packaged for you to deploy. This is my house, my desk, my power bill, and the machine that thinks in it. It exists so I can point at something on my wall and say that box runs my world.
To associate your repository with the benchmarkin topic, visit your repo's landing page and select "manage topics."