YPerf – Monitor LLM Inference API Performance
YPerf – Monitor LLM Inference API Performance
Our team operates several real-time AI applications, where both latency(TTFT) and throughput(TPS) are critical to most of our users. Unfortunately, nearly all of the major LLM APIs lack consistent stability. To address this, I developed YPerf—a simple webpage designed to monitor the performance of inference APIs. I hope it helps you select better models and discover new trending ones as well. The data is sourced from OpenRouter, an excellent provider that aggregates LLM API services.
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Incorrect prediction on native model
Similar products
LLM Inference Requirements Profiler
Healthyr - Rails app performance monitor
LLM Inference Performance Analytic Tool for Moe Models (DeepSeek/etc.)
Mashape Analytics - Visualize, Inspect and Monitor API Performance
Monitor your website downtime, ssl validity and performance
Speeding up LLM inference 2x times (possibly)
Open-source AMDGCN kernels for optimizing LLM inference
Onera – Private LLM Inference Inside AMD SEV-SNP Enclaves
Monitor website performance, downtime, SSL expiry for FREE
BonzAI – self-sovereign, local LLM inference in the browser