Command-line benchmark for LLM APIs: send the same prompts to OpenAI and Anthropic models concurrently and compare p50/p99 latency, tokens per second, cost per request and success rate. Rust.
The sping command provides modern HTTP/TCP latency monitoring with real-time terminal visualization