DekaLLM: Models and Prices

Inference API hosted entirely in Indonesia, offered through Lintasarta Cloudeka.

DekaLLM served 164B tokens of other labs' open-weight models over Oct 3 to Oct 9, 2026 through a large public LLM router, #32 among providers.

Open-weight models

ModelFamilyShare of modelTokens, 7dInput $/MOutput $/MTokens/sUptime 3dContext
Qwen3.8 27BQwen8.1%47B$0.049$3.0014099.8%262K
Gemma 4 26B A4B Gemma10.7%28B$0.06$0.3362100.0%262K
gpt-oss-120bgpt-oss6.4%27B$0.03$0.184099.9%131K
DeepSeek V4.1 FlashDeepSeek<0.1%18B$0.12$1.207899.4%1.05M
GLM 5.3 FlashGLM0.1%14B$0.10$1.0010597.2%1.05M
Mistral NemoMistral5.2%13B$0.018$0.0319100.0%131K
gpt-oss-20bgpt-oss4.0%8B$0.029$0.142099.5%131K
Nemotron 3 SuperNemotron1.1%6B$0.08$0.452199.5%262K
Qwen3.6 35B A3BQwen5.4%2B$0.10$1.009599.2%262K
Qwen3 30B A3B Instruct 2507Qwen0.5%113M$0.09$0.308499.5%262K

List prices per million tokens as of October 9, 2026. Share: of all tokens served on that model; a dash where the volume is not reported.

Company

Headquarters
Indonesia

Sources: Cloudeka.

Finance your GPUs

Serving tokens on your own GPUs? Tell us the cluster, the models you serve and the contracts behind them, and we will share it with partner funders that fit.

Prefer email? hello@amcompute.com