Benchmarking pipeline for image/video generation models — blind pairwise preference studies and Bradley–Terry ratings with bootstrap CIs
-
Updated
Aug 13, 2026 - Python
Benchmarking pipeline for image/video generation models — blind pairwise preference studies and Bradley–Terry ratings with bootstrap CIs
Not a product. Not a framework. Nothing here is packaged for you to deploy. This is my house, my desk, my power bill, and the machine that thinks in it. It exists so I can point at something on my wall and say that box runs my world.
High-fidelity benchmarking & observability framework for 11-tier microservices across Public (Azure), Private (IITD Baadal), Multi-Cloud (Azure+GCP Mesh), and Edge (K3s) topologies using Prometheus, Grafana, and Locust. 🌩️📈
To associate your repository with the benchmarkin topic, visit your repo's landing page and select "manage topics."