TLDRocket
Sign in

Qwen

47 summarised stories about Qwen, each linking back to the original source. Browse all topics →

+ Follow this topic

Sunday, 19 July 2026

Alibaba Previews Qwen3.8-Max, a 2.4 Trillion-Parameter Multimodal Model, Days After Moonshot’s Kimi K3 Open-Weight Launch

MarkTechPost 1 month ago 51 7 sources

Alibaba previewed Qwen3.8-Max-Preview, described as a 2.4 trillion-parameter multimodal model, during Shanghai's World AI Conference on July 19, 2026, two days after Moonshot AI released its 2.8 trillion-parameter Kimi K3 model. The preview is available now at 10% standard pricing through Alibaba's Token Plan subscription, but the full model card, benchmark table, and open-weight release date remain unpublished. Without disclosure of active parameters per token—the actual number of parameters used per inference—the real serving cost and practical value of the model cannot be assessed, leaving developers dependent on unverified performance claims.

Fine-Tuning Qwen3 with LoRA Using NVIDIA NeMo AutoModel: A Complete Single-GPU Google Colab Workflow Tutorial

MarkTechPost 1 month ago 5 2 sources

NVIDIA NeMo AutoModel enables parameter-efficient fine-tuning of Qwen3-0.6B using LoRA on a single Google Colab GPU through a configuration-driven workflow. The tutorial adapts batch sizes, precision settings, and training steps to fit constrained hardware while maintaining the same distributed training architecture used for multi-GPU environments. The same YAML recipe-based approach scales from single-GPU experimentation to multi-node tensor-parallel and pipeline-parallel deployments without code changes.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.