AI Potluck
Infrastructure / Compilers & Model Optimization

FlashKDA

Moonshot AI

FlashKDA provides high-performance Kimi Delta Attention (KDA) kernels built on CUTLASS, powering Moonshot AI's Kimi model line. It is auto-dispatched as a backend from the flash-linear-attention library's chunk_kda operator.

Verified 2026-09-02 via the GitHub API and the LICENSE body.

Openness

5 high confidence
5.0
license
MIT(OSI)
source
public
core-gated
ungated

MIT license body confirmed. The repository is public and unarchived and builds the whole product, and the README describes no paid tier, enterprise edition or license-gated build beside it, so source is public and the core ungated.

Adoption

2 low confidence
2.0

1,243 GitHub stars, which lands in the 1K-10K stars band of the stars scale, level 2. FlashKDA has no registry package and installs from source, so no download count exists and the stars scale applies.

Capability

2 medium confidence
2.0

Banded on the category feature matrix as a narrow kernel set: one attention variant, built for one model family. Placed two bands below the tensorrt rung, the same level as thunderkittens and hummingbird.

  • https://github.com/MoonshotAI/FlashKDA/blob/master/README.md recorded 2026-09-02

    README describes FlashKDA as high-performance Kimi Delta Attention kernels built on CUTLASS, auto-dispatched from flash-linear-attention's chunk_kda, with no paid tier, enterprise edition or license-gated build beside the published source.

Verified 2026-09-02