ExecuTorch 1.5 On-Device LLM Serving (2026): Batched Scheduling, Cancellation and Off-Graph KV Cache
ExecuTorch 1.5 shipped multi-method export, batched request scheduling, bounded cancellation and off-graph KV cache. What it changes for on-device LLM apps.
