DOCUMENTATION
Supported models
These models have been tested with TypeLLM on a live SGLang GPU server.
| Model | Thinking |
|---|---|
Qwen/Qwen3.8-27B | On / off |
Qwen/Qwen3.5-0.8B / 4B / 9B | On / off |
openbmb/MiniCPM5-1B | On / off |
inclusionAI/Ling-mini-2.0 | Off only |
inclusionAI/Ring-mini-2.0 | Always on |
Image input has been tested with Qwen/Qwen3.8-27B. See Image input.
Other sizes in the Qwen3.5 and Qwen3.8 families are expected to be compatible. If the server's tokenizer path is unavailable locally, set tokenizer= to the matching Hugging Face ID or local tokenizer directory.