Qwen3 235B-A22B
Alibaba's flagship Qwen3. Competitive with GPT-4 class models.
Task Fit
Tool use, repo work, terminal workflows, and coding benchmarks.
Code generation, debugging, refactoring, and benchmark signal.
General writing, Q&A, and assistant use.
Document QA benefits from long context and instruction following.
Not marked for vision in the current library.
Not marked for image generation in the current library.
Not marked for video generation in the current library.
Not marked for voice in the current library.
Source Confidence
Variants and Quant Artifacts
Choose the artifact first; hardware fit follows from RAM, VRAM, format, and runtime.
| Quant | Format | Quality | Min RAM | Reco RAM | Runtime | Action |
|---|---|---|---|---|---|---|
| Q2_K | gguf | compact | 128GB | 192GB | ollama, llama.cpp, lm-studio | Plan with this |
| Q3_K_M | gguf | compact | 128GB | 192GB | ollama, llama.cpp, lm-studio | Plan with this |
| Q4_K_M | gguf | balanced | 145GB | 192GB | ollama, llama.cpp, lm-studio | Plan with this |
Recommended Hardware
Lowest estimated 5-year cost that can run this model.
Enough effective VRAM with a balanced 5-year cost.
Highest local performance signal among compatible hardware.
Benchmarks
Source and Review
Similar Models
Alibaba's open-weight Qwen3.6 27B. Strong coding-agent, reasoning, long-context, and vision-language model with 262K native context.
Alibaba's Qwen3.8. Watchlist entry: Qwen3.8 has not been verified as a public open-weight release yet. This preparation page uses provisional flagship-class specs so LocalAIRun can be updated quickly once Qwen publishes the official model card, weights, license, parameter count, context window, and runnable artifacts.
Qwen3 30B with MoE. Same family as Coder, general-purpose.