Aquila (language model)
Aquila (Chinese: 悟道·天鹰, Wudao Tianying) is a series of open bilingual (Chinese and English) large language models developed by the Beijing Academy of Artificial Intelligence (BAAI).
Explore Chinese AI through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Chinese AI.
Showing 1-60 of 76 articles
Aquila (Chinese: 悟道·天鹰, Wudao Tianying) is a series of open bilingual (Chinese and English) large language models developed by the Beijing Academy of Artificial Intelligence (BAAI).
BGE (BAAI General Embedding) is a family of open-source text embedding and reranking models from the Beijing Academy of Artificial Intelligence (BAAI), first released in August 2023 and distributed through the…
BigBang-V1 is an open-weights large language model released on August 2, 2026 by The Endless Frontier, a Shanghai-based research team drawn from Shanghai Jiao Tong University's School of Artificial…
CodeGeeX is an open series of multilingual code generation models developed by the Knowledge Engineering Group (KEG) and Data Mining lab at Tsinghua University together with Zhipu AI.
CogVLM is an open vision language model developed by Zhipu AI and the Knowledge Engineering Group (KEG) at Tsinghua University.
CogVideoX is a family of open-source text-to-video generation models developed by Zhipu AI and Tsinghua University (THUDM) that generate up to 10-second continuous videos from a text prompt at 16 frames per…
DeepSeek-V3 is a 671-billion-parameter open-weights Mixture of Experts large language model from Chinese AI lab DeepSeek, released on December 26, 2024, that activates only 37 billion parameters per token and…
DeepSeek V3.1 is a large language model developed by DeepSeek, released on August 19, 2025 and made broadly available via the official API on August 21, 2025.
DeepSeek-V3.2 is an open-weight Mixture of Experts large language model family developed by DeepSeek that introduces DeepSeek Sparse Attention (DSA)
DeepSeek V4 is a family of open-weight Mixture of Experts large language models developed by DeepSeek, a Hangzhou-based AI research lab.
DeepSeek V4-Pro is the flagship model of the DeepSeek V4 family: a Mixture of Experts large language model with 1.6 trillion total parameters, 49 billion of them activated per token, and a one-million-token…
DeepSeek V4.1-Flash is an open-weight multimodal mixture-of-experts model released by DeepSeek on September 10, 2026. It accepts text and images and generates text.
DeepSeek-OCR is an open-source optical character recognition (OCR) and document-understanding system released by DeepSeek on 20 October 2025 that pioneers a contexts optical compression paradigm: it encodes…
DeepSeek-R1-Distill is a family of six open-weight reasoning language models released by DeepSeek on January 20, 2025, alongside the flagship DeepSeek-R1 reasoning model.
ERNIE 4.5 is a family of large language models released by the Chinese technology company Baidu, open-sourced on June 30, 2025 under the Apache 2.0 license .
GLM-130B is a 130-billion-parameter bilingual (English and Chinese) large language model released in August 2022 by the Knowledge Engineering Group (KEG) and the Data Mining research group at Tsinghua…
GLM-4.5 is an open-weights large language model released by Zhipu AI (operating internationally as Z.ai) on July 28, 2025, built on a 355-billion-parameter Mixture of Experts architecture that activates 32…
GLM-4.6 is a flagship open-weight large language model released by Zhipu AI under its international brand Z.ai on September 30, 2025, built on a sparse Mixture of Experts (MoE) architecture with roughly 357…
GLM-5 is an open-weight flagship large language model released by the Chinese AI company Zhipu AI, under its international brand Z.ai, on February 11, 2026.
GLM-5.1 is an open-weight large language model developed by the Chinese AI company Zhipu AI, which markets its products internationally under the brand Z.ai.
GLM-5.2 is an open-weight large language model developed by the Chinese company Zhipu AI, which sells its products internationally under the brand Z.ai.
GLM-5.3-Flash is an open-weight, natively multimodal mixture-of-experts large language model released by Z.ai on August 26, 2026.
HiDream most commonly refers to HiDream-I1, an open-source text-to-image generative foundation model released in April 2025 by the Chinese company HiDream.ai (Chinese: 智象未来).
Hunyuan is Tencent's brand for its foundation models, an umbrella covering large language models, video generation, image synthesis, 3D asset creation, and multimodal reasoning.
Hunyuan 3D is a family of open weight generative artificial intelligence models from Tencent that turn text prompts, single images, sketches, and other inputs into ready to use three dimensional assets…
Hunyuan-A13B is an open-weight mixture-of-experts large language model released by Tencent in late June 2025.
HunyuanVideo is an open-source video generation model developed by Tencent and released on December 3, 2024, with over 13 billion parameters
Hy-Embodied is a family of open-weight embodied-AI foundation models built by Tencent, jointly developed by the Tencent Hunyuan model team (the "Hy Team") and the Tencent Robotics X laboratory.
Hy4 Preview is an open-weight large language model released by the Tencent Hy Team on August 28, 2026.
inclusionAI is an open-source artificial general intelligence (AGI) research initiative established by Ant Group, the financial-technology affiliate of the Alibaba ecosystem.
InternLM (Chinese name Shusheng Puyu, 书生·浦语) is a family of open-weight large language models developed primarily by Shanghai AI Laboratory, with partners that include SenseTime, the Chinese University of Hong…
InternVL is a family of open-source multimodal large language models developed by the OpenGVLab research group at the Shanghai Artificial Intelligence Laboratory in collaboration with academic partners…
Kimi K2 is an open-weights Mixture of Experts language model from Moonshot AI, a Beijing startup, released on July 11, 2025 with 1.04 trillion total parameters and 32.6 billion activated per token
Kimi K2 Thinking is a reasoning and agentic large language model released by the Chinese startup Moonshot AI on November 6, 2025.
Kimi K2.5 is an open-weights, natively multimodal large language model developed by Moonshot AI and released on January 27, 2026 .
Kimi K2.6 is an open-weight, trillion-parameter mixture of experts (MoE) large language model released by Moonshot AI on 20 April 2026 for agentic coding and long-horizon autonomous execution.
Kimi K2.7-Code is an open-weight, coding-focused agentic model released by Moonshot AI on 2026-06-12.
Kimi K3 is an open-weight multimodal reasoning model developed by Moonshot AI. Moonshot made the model available through its hosted products on July 16, 2026 and published the weights, code, configuration…
Kimi Linear is a hybrid linear attention architecture published by Moonshot AI on October 30, 2025, together with a 48-billion-parameter mixture-of-experts model that activates 3 billion parameters per token.
Ling-1T is a trillion-parameter open-weight language model released by Ant Group through its Inclusion AI research group on October 9, 2025.
Ling-3.0-flash is an open-weight mixture-of-experts language model from inclusionAI, the open-source AI initiative of Ant Group, and the first model of the Ling 3.0 generation.
Marco-o1 is an open reasoning model released in November 2024 by the MarcoPolo team at Alibaba International Digital Commerce (AIDC).
MiniCPM is a family of compact, openly licensed language models published by OpenBMB, the shared open-source brand of Tsinghua University's natural language processing lab (THUNLP) and the Beijing company…
MiniCPM-V is a family of open-weights multimodal large language models published by OpenBMB, the shared open-source brand of Tsinghua University's Natural Language Processing lab (THUNLP) and the Beijing…
MiniCPM5-2B is an open-weights small language model published by OpenBMB on September 7, 2026 under the Apache License 2.0.
MiniMax H3 (also written MiniMax-H3) is a multimodal video generation model developed by MiniMax, the Shanghai-based AI company that operates the Hailuo AI video service.
MiniMax M2 is an open-weight large language model released on October 27, 2025 by the Shanghai-based AI company MiniMax, built as a Mixture of Experts model with 230 billion total parameters and roughly 10…
ModelScope is an open-source Model-as-a-Service (MaaS) platform developed by Alibaba Cloud and DAMO Academy and launched on November 3, 2022, that functions as China's largest AI model and dataset hub…
MoonEP is an open-source expert-parallel communication library for mixture-of-experts models, released by Moonshot AI on July 27, 2026 under the MIT License.
OpenBMB, short for "Open Lab for Big Model Base," is a Chinese open-source AI community and GitHub organization that builds foundation models, training and inference toolkits, agent frameworks, and…
OpenGVLab is the open-source organization and project hub for general-vision and multimodal foundation models run by the General Vision group at Shanghai AI Laboratory.
PaddlePaddle (Chinese name Feijiang, 飞桨) is an open-source deep learning framework developed by the Chinese technology company Baidu, and it is generally described as the first deep learning platform developed…
QwQ is a family of open-weight reasoning models from the Qwen team at Alibaba Cloud, built to compete with OpenAI's o1 and DeepSeek-R1 at a fraction of their size.
Qwen is a family of large language models and multimodal models developed by the Qwen Team at Alibaba Cloud.
Qwen-VL is the first family of open vision-language (multimodal) models from the Qwen team at Alibaba Cloud, able to take images, text, and bounding boxes as input and produce text and bounding boxes as output.
Qwen2 is the second major generation of open large language models developed by the Qwen team at Alibaba Cloud, released on 6 June 2024.
Qwen2-VL is a family of open-weight vision-language models released by the Qwen team at Alibaba Cloud between August and September 2024, in 2B, 7B, and 72B Instruct sizes.
Qwen2.5 is a family of open-weight large language models that Alibaba Cloud's Qwen team released on 19 September 2024, spanning seven dense sizes from 0.5 billion to 72 billion parameters, pretrained on…
Qwen2.5-Math is a family of mathematics-specialized large language models developed by the Qwen team at Alibaba Cloud and released in September 2024.
Qwen2.5-VL is a series of open-weight vision-language models released on 26 January 2025 by the Qwen team at Alibaba (Alibaba Cloud), succeeding the earlier Qwen2-VL family.