Blog posts

2026

2025

FlashInfer-Bench: Building the Virtuous Cycle for AI-driven LLM Systems Permalink

15 minute read

Published:

Introducing FlashInfer-Bench, a benchmark and infrastructure that lets AI systems optimize themselves: standardized GPU workload descriptions via FlashInfer Trace, realistic benchmarks from production LLM deployments, and a seamless path for deploying AI-generated kernels into live serving systems.