Authors

Yiren Zhao

Publications 3

Preprint Open access

AgentPerfBench: A Benchmarking and Evaluation Suite for Inference Performance of Agentic LLMs

The optimization of LLM serving engines, such as vLLM and SGLang, is largely benchmark-driven: optimizations, scheduling policies, hardware and system designs are all selected based on representative workloads. However, a significant mismatch has emerged in the agentic era. Existing benchmarks primarily focus on simple …

Co-authors