Qwen grew out of Alibaba's Tongyi Qianwen. It first appeared through products and cloud services in 2023, then spread through open-weight 7B, 14B, 72B, and related releases.
Today / Thursday, August 13, 2026
limboData updated
Aug 13, 12:25 AM
Live sources
-
Ingestion status
Database first
Model Brief
Alibaba's model family spanning general, coding, multimodal, and open ecosystem releases.
Read full profile and timeline+
Qwen is Alibaba's model family spanning general language models, coding, multimodal models, mathematical reasoning, tool use, and open-weight releases. It is closely watched for Chinese, multilingual, and coding performance, and spreads through open communities and ecosystems such as ModelScope.
Qwen is especially important for Chinese users and developers because it combines cloud APIs, open models, and enterprise deployment options. Its updates often affect Chinese-language assistants, coding tools, agentic tool use, enterprise knowledge bases, and localized AI application stacks.
Timeline
The Qwen2 and Qwen2.5 phases expanded into coding, vision, audio, and mathematical reasoning, while QwQ-style releases pushed the family into reasoning-model competition.
From Qwen3 onward, Alibaba has advanced open weights, cloud APIs, specialized coding models, and multimodal models in parallel, making Qwen one of the most active model families in the Chinese developer ecosystem.
Heat / Trend
Discussion Heat Trend
Real community signals: HN points/comments + GitHub stars/forks
Total Heat
3106
Stories
8
Peak Day
07/01
Community / HN + GitHub
Community Discussion
Recent Hacker News and GitHub signals that add community context beyond the news feed.
HN
4
GitHub
4
Heat
3106
QwenLM/qwen-code
Score
834
Stars
25694
Forks
2584
QwenLM/Qwen3-VL
Score
794
Stars
19501
Forks
1798
QwenLM/Qwen3
Score
742
Stars
27346
Forks
1999
QwenLM/Qwen
Score
724
Stars
21366
Forks
1842
Show HN: Cline subscription plan to access GLM-5.2 at 2-5x discount
Score
7
Comments
2
Source
HN
Qwen3.5 2B burns all the output tokens while thinking
Score
2
Comments
0
Source
HN
Show HN: Apex-1-flash, 4B LLM finetuned on RTX 5070
Score
2
Comments
0
Source
HN
Show HN:Swarm intelligence without degradation using two Qwen models
Score
1
Comments
0
Source
HN
official / Qwen
Official Updates
Qwen3Guard: Real-time Safety for Your Token Stream
Tech Report GitHub Hugging Face ModelScope DISCORD Introduction We are excited to introduce Qwen3Guard, the first safety guardrail model in the Qwen family.
Qwen-Image-Edit: Image Editing with Higher Quality and Efficiency
QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD We are excited to introduce Qwen-Image-Edit, the image editing version of Qwen-Image. Built upon our 20B Qwen-Image model, Qwen-Image-Edit successfully extends Qwen-Image’s unique text rendering capabiliti...
media / Qwen
Media Coverage

Alibaba's new Qwen model is also taking your job, but this time it's great
Alibaba is marketing its new AI model Qwen 3.8 with a video that shows the AI working while a person enjoys their hobbies. It's a deliberate contrast to the job loss warnings from OpenAI and Anthropic. Of course, it's still just marketing.

Claude Code's complicated China problem involves bans on both sides of the Pacific
Anthropic is trying to block Chinese companies like ByteDance and Ant Financial from accessing Claude Code, but they're getting around the restrictions through VPNs and overseas subsidiaries.
research / Qwen
Research Papers

It's Not What You Say, It's How You Say It: Evaluating LLM Responses to Expressions of Belief
Users frequently express their beliefs to large language models (LLMs). In some situations, the LLM should accept these contextual beliefs as true. In others, they should stick to their prior knowledge.

Control Under Compression: Reliability Frontiers for Tool-Using Agents
Tool-using language-model agents are governed not only by task prompts but also by persistent system-side instructions that specify tools, arguments, policies, execution protocols, and recovery.

When Does Muon Help Agentic Reinforcement Learning?
Muon is competitive with AdamW in large-scale pre-training, but its value for reinforcement-learning (RL) post-training remains unclear. We study vanilla Muon in sparse-reward agentic RL through matched single-seed comparisons with AdamW on ALFWorld using Qwen...

ReToken: One Token to Improve Vision-Language Models for Visual Retrieval
Long visual context poses a challenge for vision-language models: performance degrades as the number of distractors grows, and processing all tokens at once is computationally infeasible under GPU memory constraints.

OpenThoughts-Agent: Data Recipes for Agentic Models
Agentic language models dramatically expand the applications of AI yet little is publicly known about how to curate training data for broadly capable agents.

Claude Code costs up to $200 a month. Goose does the same thing for free.
The artificial intelligence coding revolution comes with a catch: it's expensive. Claude Code , Anthropic's terminal-based AI agent that can write, debug, and deploy code autonomously, has captured the imagination of software developers worldwide.

Nous Research's NousCoder-14B is an open-source coding model landing right in the Claude Code moment
Nous Research , the open-source artificial intelligence startup backed by crypto venture firm Paradigm , released a new competitive programming model on Monday that it says matches or exceeds several larger proprietary systems — trained in just four days using...
community / Qwen
Community & Open Source

Alibaba's Qwen takes on Kimi K3 with open-weight Qwen 3.8, says model is "second only to Fable 5"
Alibaba has unveiled Qwen 3.8, a multimodal AI model with 2.4 trillion parameters that the Qwen team says rivals leading models and trails only Fable 5. A preview is available now. The article Alibaba's Qwen takes on Kimi K3 with open-weight Qwen 3.