flashinfer.cake_fmha.cake_batch_decode_with_kv_cache¶
- flashinfer.cake_fmha.cake_batch_decode_with_kv_cache(*args, **kwargs)¶
Run the FlashInfer TRTLLM paged-decode ABI through Cake FMHA.
Parameters and return values match
flashinfer.trtllm_batch_decode_with_kv_cache(). The selection is explicit and never replaces FlashInfer’s default backend selection.