inference-lab 0.5.0

High-performance LLM inference simulator for analyzing serving systems
Documentation