
Qwen2 7B is a transformer-based model that excels in language understanding, multilingual capabilities, coding, mathematics, and reasoning.
It features SwiGLU activation, attention QKV bias, and group query attention. It is pretrained on extensive data with supervised finetuning and direct preference optimization.
For more details, see this blog post(opens in new tab) and GitHub repo(opens in new tab).
Usage of this model is subject to Tongyi Qianwen LICENSE AGREEMENT(opens in new tab).
Modalities
Context
33K
Released
Jul 16, 2024
Knowledge Cutoff
Jun 2024
Qwen2 7B is a transformer-based model that excels in language understanding, multilingual capabilities, coding, mathematics, and reasoning. It features SwiGLU activation, attention QKV bias, and group query attention. It is pretrained on extensive data with supervised finetuning and direct preference optimization. For more details, see this blog post and GitHub repo.
Qwen 2 7B Instruct has a 32,768 token context window.
Qwen3.8 27B, Qwen3.8 2.4T A95B, Qwen3.8 Max and 47 more are other text models from Qwen.
Qwen 2 7B Instruct was released on July 16, 2024. Its knowledge cutoff is June 30, 2024.
Token volume and request traffic to this model over time.