/

Qwen

25 stories

Qwen

QwenAsiaQwen3Guard: Real-time Safety for Your Token Stream Tech Report GitHub Hugging Face ModelScope DISCORD Introduction We are excited to introduce Qwen3Guard, the first safety guardrail model in the Qwen family. Built upon the powerful Qwen3 foundation models and fine-tuned

September 22
Qwen

QwenAsiaQwen-Image-Edit: Image Editing with Higher Quality and Efficiency QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD We are excited to introduce Qwen-Image-Edit, the image editing version of Qwen-Image. Built upon our 20B Qwen-Image model, Qwen-Image-Edit successfully extends Qwen-Image’

August 18
Qwen

QwenAsiaQwen-Image: Crafting with Native Text Rendering GITHUB HUGGING FACE MODELSCOPE DEMO DISCORD We are thrilled to release Qwen-Image, a 20B MMDiT image foundation model that achieves significant advances in complex text rendering and precise image editing. To try the lat

August 4
Qwen

QwenAsiaGSPO: Towards Scalable Reinforcement Learning for Language Models PAPER DISCORD Introduction Reinforcement Learning (RL) has emerged as a pivotal paradigm for scaling language models and enhancing their deep reasoning and problem-solving capabilities. To scale RL, the foremost prerequi

July 27
Qwen

QwenAsiaQwen-MT: Where Speed Meets Smart Translation DEMO API DISCORD Introduction Here we introduce the latest update of Qwen-MT (qwen-mt-turbo) via Qwen API. This update builds upon the powerful Qwen3, leveraging trillions multilingual and translation tokens to comprehen

July 24
Qwen

QwenAsiaQwen3-Coder: Agentic Coding in the World GITHUB HUGGING FACE MODELSCOPE DISCORD Today, we’re announcing Qwen3-Coder, our most agentic code model to date. Qwen3-Coder is available in multiple sizes, but we’re excited to introduce its most powerful variant first:

July 22
Qwen

QwenAsiaTime to Speak Some Dialects, Qwen-TTS! API DISCORD Introduction Here we introduce the latest update of Qwen-TTS (qwen-tts-latest or qwen-tts-2025-05-22) through Qwen API . Trained on a large-scale dataset encompassing over millions of hours of speech, Qwen-TT

June 27
Qwen

QwenAsiaQwen VLo: From "Understanding" the World to "Depicting" It QWEN CHAT DISCORD Introduction The evolution of multimodal large models is continually pushing the boundaries of what we believe technology can achieve. From the initial QwenVL to the latest Qwen2.5 VL, we have made prog

June 26
Qwen

QwenAsiaQwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models GITHUB HUGGING FACE MODELSCOPE DISCORD We release Qwen3 Embedding series, a new proprietary model of the Qwen model family. These models are specifically designed for text embedding, retrieval, and reranking tasks, built

June 5
Qwen

QwenAsiaQwen3: Think Deeper, Act Faster QWEN CHAT GitHub Hugging Face ModelScope Kaggle DEMO DISCORD Introduction Today, we are excited to announce the release of Qwen3, the latest addition to the Qwen family of large language models. Our flagship model, Qwen3

April 28
Qwen

QwenAsiaQVQ-Max: Think with Evidence QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD Introduction Last December, we launched QVQ-72B-Preview as an exploratory model, but it had many issues. Today, we are officially releasing the first version of QVQ-Max, o

March 27
Qwen

QwenAsiaQwen2.5 Omni: See, Hear, Talk, Write, Do It All! QWEN CHAT HUGGING FACE MODELSCOPE DASHSCOPE GITHUB PAPER DEMO DISCORD We release Qwen2.5-Omni, the new flagship end-to-end multimodal model in the Qwen series. Designed for comprehensive multimodal perception, it seamles

March 26
Qwen

QwenAsiaQwen2.5-VL-32B: Smarter and Lighter QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD Introduction At the end of January this year, we launched the Qwen2.5-VL series of models, which received widespread attention and positive feedback from the community. Bu

March 23
Qwen

QwenAsiaQwQ-32B: Embracing the Power of Reinforcement Learning QWEN CHAT Hugging Face ModelScope DEMO DISCORD Scaling Reinforcement Learning (RL) has the potential to enhance model performance beyond conventional pretraining and post-training methods. Recent studies have demonstrate

March 5
Qwen

QwenAsia... QwQ-Max-Preview QWEN CHAT DISCORD This is a blog created by QwQ-Max-Preview. We hope you enjoy it! Introduction <think> Okay, the user wants me to create a title and introduction for their blog announcing the release of QwQ-Max-Preview.

February 24
Qwen

QwenAsiaQwen2.5-Max: Exploring the Intelligence of Large-scale MoE Model QWEN CHAT API DEMO DISCORD It is widely recognized that continuously scaling both data size and model size can lead to significant improvements in model intelligence. However, the research and industry community has limi

January 28
Qwen

QwenAsiaQwen2.5-1M: Deploy Your Own Qwen with Context Length up to 1M Tokens Tech Report HuggingFace ModelScope Qwen Chat HuggingFace Demo ModelScope Demo DISCORD Introduction Two months after upgrading Qwen2.5-Turbo to support context length up to one million tokens, we are back with the open-so

January 26
Qwen

QwenAsiaQwen2.5 VL! Qwen2.5 VL! Qwen2.5 VL! QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD We release Qwen2.5-VL, the new flagship vision-language model of Qwen and also a significant leap from the previous Qwen2-VL. To try the latest model, feel free to visit Q

January 26
Qwen

QwenAsiaGlobal-batch load balance almost free lunch to improve your MoE LLM training GITHUB HUGGING FACE MODELSCOPE DISCORD Background The Mixture-of-Experts (MoEs) architecture has become a popular model-parameter-scale-up technique. Typically, one MoE layer consists of a router (often parameterized as

January 20
Qwen

QwenAsiaTowards Effective Process Supervision in Mathematical Reasoning GITHUB HUGGING FACE MODELSCOPE DISCORD Introduction In recent years, Large Language Models (LLMs) have made remarkable advances in mathematical reasoning, yet they can make mistakes, such as miscalculations or logical er

January 13
Qwen

QwenAsiaQVQ: To See the World with Wisdom GITHUB HUGGING FACE MODELSCOPE KAGGLE DEMO DISCORD Language and vision intertwine in the human mind, shaping how we perceive and understand the world around us. Our ability to reason is deeply rooted in both linguistic t

December 24
Qwen

QwenAsiaQwQ: Reflect Deeply on the Boundaries of the Unknown GITHUB HUGGING FACE MODELSCOPE DEMO DISCORD Note: This is the pronunciation of QwQ: /kwju:/ , similar to the word “quill”. What does it mean to think, to question, to understand? These are the deep waters that QwQ (Qwen

November 27
Qwen

QwenAsiaExtending the Context Length to 1M Tokens! API Documentation (Chinese) HuggingFace Demo ModelScope Demo Introduction After the release of Qwen2.5, we heard the community’s demand for processing longer contexts. In recent months, we have made many optimizations fo

November 14
Qwen

QwenAsiaQwen2.5-Coder Series: Powerful, Diverse, Practical. GITHUB HUGGING FACE MODELSCOPE KAGGLE DEMO DISCORD Introduction Today, we are excited to open source the “Powerful”, “Diverse”, and “Practical” Qwen2.5-Coder series, dedicated to continuously promoting the development of

November 11