ComputeArena provides community-submitted throughput benchmarks for local AI models running on edge silicon. It allows users to decode and compare performance metrics like tokens per second, based on model, quantization, and chip. The platform uses BaseRT and llama.cpp runtimes for accurate, community-driven hardware performance data.
