AI Potluck
Model components / Inference code

RTP-LLM

Alibaba Cloud

RTP-LLM is an LLM inference acceleration engine developed by Alibaba's Foundation Model Inference Team, built on work from the FasterTransformer project.

A NOTICE file records that the engine is based on the FasterTransformer project. Verified 2026-08-31 via GitHub, the LICENSE body and the repository README.

Openness

5 medium confidence
5.0
license
Apache-2.0(OSI)
source
public(the published repository is the engine)
core-gated
ungated(no enterprise path in the repository root and no paid build of the engine in the README)

The LICENSE body is the Apache-2.0 text, read in full rather than taken from the API's label, and carries no appended condition. The repository is public and unarchived and builds the engine itself. Its root tree carries no enterprise, ee or commercial directory and its README describes no licence-gated build, so the core reads as ungated. Confidence is medium because that is a repository-and-README read rather than a pricing-page read.

Adoption

2 low confidence
2.0

1,321 GitHub stars, which lands in the 1K-10K stars band of the stars scale, where the scale caps at 3. No package artifact is declared for this record, so stars are the only instrument available and the band should be read as a floor.

Capability

4 medium confidence
4.0

One band below the vllm anchor. It accelerates inference as the anchor does, but over a narrower published surface: no multi-node serving record and no OpenAI-compatible API server is documented.

Verified 2026-08-31