#
recompute
Here are 2 public repositories matching this topic...
Discrete-event simulation benchmark for decode preemption policies in LLM serving: when to interrupt a running decode to admit an urgent prefill, and whether checkpoint or recompute is cheaper.
performance-engineering benchmark simulation latency scheduling inference checkpoint systems decode slo fairness preemption discrete-event-simulation prefill kv-cache llm llm-serving machine-learning-infrastructure recompute serving-systems
-
Updated
Jul 26, 2026 - Python
Improve this page
Add a description, image, and links to the recompute topic page so that developers can more easily learn about it.
Add this topic to your repo
To associate your repository with the recompute topic, visit your repo's landing page and select "manage topics."