Pricing

Pricing: free local AI, hosted GPUs from €19 a month

Free on your hardware. Effortless on ours. Self-host everything for €0, or let LU Labs run the GPUs so you don’t have to.

Flash chat: published daily allowance

Hosted includes 11 Flash models without credits; Pro and Max include 12. Unlimited prompt counts, up to 500,000 combined input and output tokens per account per UTC day in app sessions. All 47 chat models remain available on every plan. GLM 5.3 Flash uses credits on Hosted and is included in the daily allowance on Pro and Max. API keys and accounts without an active subscription always use credits.

App sessions share 500,000 input and output tokens per account per UTC day across the Flash models included on your plan without spending credits. The allowance resets at 00:00 UTC. It is a benefit of an active paid plan, not a free tier: an account without an active plan uses credits for these models, whatever it bought before. API key requests always use credits.

One free request can run at a time per account. A second concurrent request is refused, not silently billed. A request whose token budget exceeds the remaining allowance uses normal credit pricing, with a notice in the app.

We reserve the input estimate and output limit before generation. Without an explicit output limit, free requests allow up to 8,192 output tokens. Each free request lasts at most 240 seconds. Final provider usage settles the reservation. Failed or interrupted requests without final usage keep their reservation. This is a daily allowance, not unlimited usage.

Current model lists and exact monthly image and video budgets →

One app, your choice of where to run

  • Every model on every plan.
  • Chat, code, image, video and voice in one subscription, one app.
  • Runs the same open models on your own machine and in the cloud.
  • No five hour windows.

Hosted model access is shared by paid plans and credit packs. Local execution requires compatible hardware and downloaded models; it does not include hosted compute for free. Monthly credit grants, daily Flash allowance and the request limits below still apply. This is not an unlimited-use promise.

Refunds and withdrawal

Our current terms provide EU consumers with withdrawal within 14 days of purchase without giving reasons and a full refund. Contact hello@lu-labs.ai. Subscription cancellation and a refund request are separate actions.

How many images or clips does a budget buy?

Choose a credit pack or a plan, then an image or video model. Every result uses the published catalog and credit rules.

450,000 credits in one shared wallet.

1,500 images

Estimates spend the entire budget on this selection, not on all activities together. Chat draws on the same wallet at each model's published credit rate and is not part of this estimate. Media uses the standard image or 5-second clip rate, without extra operations. Longer clips and other operations can cost more. Plan video limits are included. Annual plans still grant credits monthly. Flash daily usage is separate and is not added to this estimate. Catalog prices can change. This is an estimate, not a guaranteed number of completed generations.

Request limits, not hidden windows

  • Chat: 60 requests per minute per account.
  • Media job submissions: 30 requests per minute per account.
  • Media uploads: 60 requests per minute per account.
  • Subscription checkout: 5 attempts per hour and 12 per day per account.
  • Credit-pack checkout: a separate 5 attempts per hour and 12 per day per account.

These are request counters, not guaranteed completed generations. They use windows starting with the first request and count requests that reach the limiter, including later validation failures. A rate-limit response includes a Retry-After header.

These burst counters currently live in each server process. Restarts reset them, and multiple server instances do not share them. They are not a global concurrency or availability guarantee. Provider capacity and wallet checks can also refuse a request. The separate Flash daily allowance and one-request limit above are enforced in the database across instances.

Chat model capabilities

Every model on every plan. All 47 chat models are available on paid plans and credit packs. Context and quantization below are from the public DeepInfra model list, checked 2026-09-11. Missing quantization is not a claim of full precision.

Tool transport, image input and reasoning controls describe the LU catalog, not a new live capability test. Provider-listed context is not a guarantee of a full-length prompt: input and output share context, and clients may trim history. No full-native-context promise is made here. Follow each model link for the provider's current information.

47 catalog chat models, provider facts checked 2026-09-11
Model and sourceProvider-listed contextProvider quantizationTool calling in LUImage input in LUReasoning controls in LU
Llama 3.1 8B Turbo131,072fp8Native tools parameterNoNo reasoning control
Ling 3.0 flash131,072Not stated by providerNative tools parameterNolow, medium, high; default high
Qwen3 30B A3B40,960fp8Native tools parameterNolow, medium, high; default high
Gemma 4 26B262,144fp8Native tools parameterYeslow, medium, high; default high
Qwen 3.6 35B A3B262,144fp8Native tools parameterYeslow, medium, high; default high
GLM 5.3 Flash1,048,576fp4Native tools parameterYeslow, medium, high, max; default high; no off switch
Lunaris 8B8,192fp8LU prompt translationNoNo reasoning control
MythoMax 13B4,096fp16LU prompt translationNoNo reasoning control
Hermes 3 70B131,072fp8LU prompt translationNoNo reasoning control
Euryale 70B131,072fp8LU prompt translationNoNo reasoning control
gpt-oss 120B131,072bfloat16Native tools parameterNolow, medium, high; default high
DeepSeek V3.2163,840fp4Native tools parameterNolow, medium, high; default high
Hermes 3 405B131,072fp8LU prompt translationNoNo reasoning control
Qwen3 Coder 480B262,144fp4Native tools parameterNoNo reasoning control
Kimi K31,048,576Not stated by providerNative tools parameterYeslow, medium, high; default high
DeepSeek V3.1163,840fp4Native tools parameterNolow, medium, high; default high
DeepSeek V4 Flash 07311,048,576fp8Native tools parameterNolow, medium, high; default high
DeepSeek V4.1 Flash1,048,576fp8Native tools parameterYeslow, medium, high; default high; no off switch
DeepSeek V4 Pro 08131,048,576fp8Native tools parameterNolow, medium, high; default high
DeepSeek R1163,840fp4Native tools parameterNolow, medium, high; default high
Qwen3 32B40,960fp8Native tools parameterNolow, medium, high; default high
Qwen3 235B A22B262,144fp8Native tools parameterNoNo reasoning control
Qwen 3.5 9B262,144bfloat16Native tools parameterYeslow, medium, high; default high
Qwen 3.5 35B A3B262,144fp8Native tools parameterYeslow, medium, high; default high
Qwen 3.5 397B A17B262,144fp8Native tools parameterYeslow, medium, high; default high
Qwen 3.6 27B262,144fp8Native tools parameterYeslow, medium, high; default high
Qwen3 VL 30B262,144fp8Native tools parameterYesNo reasoning control
Qwen3 VL 235B262,144fp8Native tools parameterYesNo reasoning control
Qwen 3.8 27B262,144Not stated by providerNative tools parameterYeslow, medium; default medium
Qwen 3.8 Max256,000Not stated by providerNative tools parameterNolow, medium, high; default high
Qwen 3.8 A95B262,144fp4Native tools parameterNolow, medium, high; default high
Llama 3.3 70B Turbo131,072fp8Native tools parameterNoNo reasoning control
Llama 4 Maverick1,048,576fp8LU prompt translationYesNo reasoning control
Llama 4 Scout327,680fp8Native tools parameterYesNo reasoning control
Gemma 4 31B Turbo262,144fp4Native tools parameterYeslow, medium, high; default high
GLM 4.7202,752fp4Native tools parameterNolow, medium, high; default high
GLM 5202,752fp4Native tools parameterNolow, medium, high; default high
GLM 5.1202,752fp4Native tools parameterNolow, medium, high; default high
GLM 5.21,048,576fp4Native tools parameterNolow, medium, high; default high
GLM 5.31,048,576fp4Native tools parameterNolow, medium, high, max; default high; no off switch
Nemotron 3 Super 120B262,144bfloat16Native tools parameterNolow, medium, high; default high
gpt-oss 20B131,072bfloat16Native tools parameterNolow, medium, high; default high
Kimi K2.6262,144fp4Native tools parameterYeslow, medium, high; default high
Kimi K2.7 Code262,144fp4Native tools parameterYeslow, medium, high; default high
MiniMax M2.7196,608fp8Native tools parameterNolow, medium, high; default high
MiniMax M3524,288fp8Native tools parameterYeslow, medium, high; default high
Mistral Small 3.2 24B128,000fp8Native tools parameterYesNo reasoning control

Still deciding?

The quota, privacy and billing questions everyone asks, answered.

Read the FAQ