QwQ
QwQ is a family of open-weight reasoning models from the Qwen team at Alibaba Cloud, built to compete with OpenAI's o1 and DeepSeek-R1 at a fraction of their size.
Explore Chinese AI through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Chinese AI.
Showing 61-91 of 91 articles
QwQ is a family of open-weight reasoning models from the Qwen team at Alibaba Cloud, built to compete with OpenAI's o1 and DeepSeek-R1 at a fraction of their size.
Qwen is a family of large language models and multimodal models developed by the Qwen Team at Alibaba Cloud.
Qwen2 is the second major generation of open large language models developed by the Qwen team at Alibaba Cloud, released on 6 June 2024.
Qwen2-Math is a series of mathematics-specialized large language models released by the Qwen team at Alibaba on 8 August 2024.
Qwen2.5 is a family of open-weight large language models that Alibaba Cloud's Qwen team released on 19 September 2024, spanning seven dense sizes from 0.5 billion to 72 billion parameters, pretrained on…
Qwen2.5-Coder is the code-specialized series within the Qwen2.5 generation of large language models developed by the Qwen team at Alibaba.
Qwen2.5-Math is a family of mathematics-specialized large language models developed by the Qwen team at Alibaba Cloud and released in September 2024.
Qwen3 is the third-generation family of large language models developed by the Qwen Team at Alibaba Cloud (also known as Tongyi Qianwen lab).
Qwen3 Embedding is a family of open text embedding and reranking models released by Alibaba's Qwen team in June 2025.
Qwen3-Coder is a family of open-weight large language models specialized for software engineering, developed by Alibaba's Qwen team (Tongyi Lab) and released under the Apache 2.0 license.
Qwen3-Max is the flagship large language model in Alibaba's Qwen series and the first Qwen model to cross one trillion parameters, released in preview on September 5, 2025 and formally launched at the Apsara…
Qwen3-Omni is a natively end-to-end omni-modal foundation model developed by the Qwen team at Alibaba Cloud, capable of understanding text, images, audio, and video and generating both text and natural speech…
Qwen3-VL is a family of open-weight vision-language models built by the Qwen team at Alibaba Cloud, first released in September 2025.
Qwen3.5 is a family of open-weight large language models developed by the Qwen team at Alibaba, the successor to the Qwen3 series.
Qwen3.6 is a generation of large language models from the Qwen team at Alibaba, released in April 2026 as the successor to Qwen3.5.
Qwen3.7-Max is a closed-weight frontier large language model developed by Alibaba's Qwen team, announced in May 2026 as the flagship of the Qwen3.7 generation.
Qwen3.8 is the name Qwen uses for a model generation that includes hosted services and downloadable checkpoints.
Qwen3.8-Flash-Next is an experimental open-weight multimodal large language model released by Alibaba Group's Qwen team on August 26, 2026.
Qwen3.8-Max is a 2.4-trillion-parameter mixture-of-experts large language model developed by Alibaba's Qwen team and released on August 3, 2026 under the title "Qwen3.8-Max: A New Bar for Coding and Cowork".
StepFun (Chinese: 阶跃星辰, pinyin: Jiēyuè Xīngchén), formally Shanghai Jieyue Xingchen Intelligent Technology Co., Ltd., is a Shanghai-based Chinese artificial intelligence startup that builds the "Step" series…
Tang Jie (Chinese: 唐杰; born 1977), also published as Jie Tang, is a Chinese computer scientist, a chair professor at Tsinghua University, and the co-founder and chief scientist of Zhipu AI, the Beijing startup…
Tencent AI is the artificial intelligence research, products, and services developed by Tencent Holdings Ltd., one of the world's largest technology companies, built around the Hunyuan family of foundation…
Tencent Hunyuan Hy3 (marketed internationally as Tencent Hy3) is an open-weights large language model published by Tencent.
Tongyi Qianwen (通义千问), the brand often glossed in English as "seeking truth by asking a thousand questions," is Alibaba Group's flagship large language model and conversational AI brand, launched by Alibaba…
UltraChat is a large-scale synthetic multi-turn instructional conversation dataset released in May 2023 by the OpenBMB group at Tsinghua University, comprising approximately 1.5 million dialogues generated by…
Wu Dao (Chinese: 悟道, roughly "enlightenment" or "understanding the way") is a series of large pretrained models built by the Beijing Academy of Artificial Intelligence (BAAI), a non-profit research institute…
Yang Zhilin (Chinese: 杨植麟; pinyin: Yáng Zhílín) is the co-founder and chief executive officer of Moonshot AI (月之暗面), the Beijing startup that develops the Kimi chatbot and the Kimi family of large language…
Yi is a series of open, bilingual (English and Chinese) large language models developed by the Chinese startup 01.AI (Chinese: 零一万物, Lingyiwanwu), the company founded in March 2023 by Kai-Fu Lee.
Yi-Large is a closed-source large language model developed by Chinese artificial intelligence company 01.AI (零一万物, Língyi Wànwù), founded by Kai-Fu Lee.
Yi-Lightning is a closed-source large language model developed by Chinese artificial intelligence company 01.AI (零一万物, Língyī Wànwù), the company founded by Kai-Fu Lee.
Zhipu AI (智谱AI), now branded internationally as Z.ai, is a Chinese artificial intelligence company headquartered in Beijing.