Tag
14 articles
Alibaba's Qwen team introduces Qwen3.8-Flash-Next, a cost-efficient model that uses only 6% of its parameters per token, outperforming competitors like Claude Opus 4.6 and DeepSeek-V4-Flash.
Alibaba's Qwen team introduces Qwen3.8-Flash-Next, a 125B multimodal MoE model with only 6B active parameters per token, showcasing significant training efficiency and architectural innovation.
Alibaba is testing a new revenue-sharing model for its Qwen open-source AI platform, requiring larger companies that generate revenue from offering the model as a service to enter into commercial agreements with the company.
Learn about Alibaba's new AI model Qwen3.8-Max, a powerful 2.4 trillion parameter system that can understand text, images, and video.
Alibaba's Qwen Audio 3.0 TTS Plus has topped the Artificial Analysis Speech Arena leaderboard, showcasing advanced multilingual capabilities and expressive controls, though it lags in speed compared to competitors.
This article explores the limitations of hybrid thinking in AI models and why researchers like Junyang Lin are now advocating for agentic thinking as a more robust and scalable approach.
Alibaba's Qwen team launches Qwen3.7-Plus, a multimodal AI model on the Bailian platform, featuring vision understanding, deep reasoning, tool invocation, and autonomous iteration.
Alibaba integrates its Qwen AI assistant into Taobao, enabling users to shop through conversational AI rather than traditional search. This marks a major shift in how consumers interact with e-commerce platforms in China.
The Qwen team has released FlashQLA, a high-performance linear attention kernel library that achieves up to 3x speedup on NVIDIA Hopper GPUs, enhancing both pretraining and edge-side inference.
Alibaba's Qwen team has released Qwen3.6-27B, a dense open-weight model outperforming 397B MoE on agentic coding benchmarks. It introduces a Thinking Preservation mechanism and a hybrid attention architecture.
This article explains the advanced AI concepts behind Qwen 3.6-35B-A3B, a multimodal model that combines MoE routing, RAG, and session persistence for intelligent, context-aware AI applications.
Alibaba's Qwen team open-sources Qwen3.6-35B-A3B, a sparse MoE vision-language model with 3B active parameters and agentic coding capabilities.