HomeAboutModelsPricingSupport

Gemini 3.1 Flash-Lite Preview vs Qwen3-Max

Side-by-side specs, pricing and capabilities. Both models run on Clade under one subscription, so you can switch between them in the same conversation.

Gemini 3.1 Flash-Lite Preview comes from Google and Qwen3-Max from Alibaba. The practical difference for most people is context, price, and which input types each one accepts.

Gemini 3.1 Flash-Lite Preview holds more in a single conversation — 1.0M tokens against 256K — which matters for long documents and large codebases.

Gemini 3.1 Flash-Lite Preview is the cheaper of the two on input tokens at $0.25 / 1M tokens.

SpecificationGemini 3.1 Flash-Lite PreviewQwen3-Max
ProviderGoogleAlibaba
Model IDgemini-3.1-flash-lite-previewqwen3-max
Context window1.0M256K
Max output66K80K
Input price$0.25 / 1M tokens$1.20 / 1M tokens
Output price$1.50 / 1M tokens$6.00 / 1M tokens
Knowledge cutoff2025-01
Input typestext, image, video, audio, filetext
Output typestexttext
PlanFreePremium

Which should you use?

Pick Gemini 3.1 Flash-Lite Preview when…

  • You need the larger context window — 1.0M against 256K.
  • Cost matters: input runs at $0.25 / 1M tokens versus $1.20 / 1M tokens.
  • You need image and video and audio and file input, which the other model does not accept.
  • You are on the free plan — this model is included without an upgrade.
Gemini 3.1 Flash-Lite Preview details

Pick Qwen3-Max when…

  • You want longer single responses — up to 80K output tokens.
Qwen3-Max details

Related comparisons

Try both on Clade

You do not have to choose. One Clade subscription gives you Gemini 3.1 Flash-Lite Preview, Qwen3-Max, and every other model on the platform.