Technology releaseSeptember 23, 2026

Qwen3.8-27B: an Apache 2.0 Open Model You Can Deploy Inside Your Perimeter

The Qwen team released Qwen3.8-27B, a 27-billion-parameter multimodal model with a 262,144-token context under the Apache 2.0 licence. What it means for companies that cannot send data to external clouds.

In short. Qwen3.8-27B is a 27-billion-parameter open model from the Qwen team, published under the Apache 2.0 licence. It understands text, images and video, natively handles a 262,144-token context extensible to 1 million, and deploys with SGLang and vLLM. For companies that cannot send data to external clouds, it is another strong candidate for an AI assistant inside their own perimeter.

What we know about the model

  • 27 billion parameters, Apache 2.0 licence.
  • Understanding of text, images and video.
  • A native context of 262,144 tokens, extensible to 1,000,000.
  • Recommended serving engines are SGLang and vLLM; an FP8 variant and quantised versions for llama.cpp, Ollama and LM Studio are available.

What it means for companies in Kazakhstan

The long context lets the model take regulations, contracts and technical documentation in full, and image understanding helps with scans and photos of documents. An open licence and deployment on your own server remove the main barrier for banks, manufacturers and the public sector: data never leaves the company's perimeter. Quality in Kazakh should be tested on your own tasks, as the model card does not list supported languages.

How 105 works with such models

105 deploys open language models inside the customer's perimeter: enterprise assistants and document search that cites its sources. More in the Enterprise LLMs section.

Sources

  1. Hugging Face: Qwen/Qwen3.8-27B — карточка модели

This material was prepared with AI from the sources listed above. Spotted an inaccuracy? Write to info@105.kz.

FAQ

Frequently asked questions

Short, straight answers to the questions we hear most often.

Can Qwen3.8-27B be used in commercial projects?

The model is published under the Apache 2.0 licence, which permits commercial use. Lawyers should review the licence terms before deployment.

Can Qwen3.8-27B run on our own server?

Yes, the model is open. The model card recommends SGLang and vLLM for serving; an FP8 variant and quantised versions for llama.cpp, Ollama and LM Studio are available.

Next step

Discuss a pilot project

Tell us about your task — we'll come back with a data audit plan and an effect estimate within two business days.

  1. 1A 30-minute call: the task, your current systems, who makes the decision
  2. 2A review of data and processes, a staged plan with time and cost ranges
  3. 3A 2–4 week prototype or an 8–12 week pilot with success criteria agreed upfront

NDA from the first contact

Reply within one business day