Articles Tagged "Qwen"

Bonsai 1.7B

Bonsai 1.7B

PrismML's smallest Bonsai checkpoint compresses Qwen3-1.7B under 0.5GB, and now powers the vision-language model behind Qualcomm's Snapdragon AR1 smart glasses demo.

Qwen3.8-Max

Qwen3.8-Max

Alibaba's 2.4 trillion parameter flagship ships with real pricing and a published benchmark table, but the open-weight release it promised for this week still hasn't shown up.

Qwen3-30B-A3B

Qwen3-30B-A3B

Qwen3-30B-A3B is Alibaba's efficient MoE model that activates 3.3B of 30.5B parameters per token, matching much larger dense models on reasoning and agent benchmarks under Apache 2.0.

POCKET-35B

POCKET-35B

VIDRAFT quantizes its Darwin-36B-Opus MoE model into a 35B GGUF that runs on stock llama.cpp with no GPU, trading GPQA Diamond score for CPU and phone portability.

Qwen3-VL-235B-A22B

Qwen3-VL-235B-A22B

Alibaba's flagship open-weight vision-language MoE beats every proprietary model on DocVQA at 96.5% and MathVista at 85.8%, but trails GPT-5.4 and Gemini 3.1 Pro on broad MMMU-Pro reasoning.

Qwen2.5-VL-72B-Instruct

Qwen2.5-VL-72B-Instruct

Alibaba's dense 72B vision-language model tops the open-weight DocVQA leaderboard at 96.4% and remains the default self-hosted choice for document and chart understanding.

Qwen3.8-Max-Preview

Qwen3.8-Max-Preview

Alibaba's 2.4 trillion parameter multimodal MoE model claims to trail only Claude Fable 5, but ships with no model card, no benchmark table, and no confirmed pricing.