Alibaba · Released Mar 20, 2026
Qwen3.8 27B (xhigh)
A compact 27B parameter powerhouse capable of running locally or on single-GPU instances with near-frontier scores.
OpenRating Index
52.1
0–100 composite
Reasoning Index
71.4
0–100 composite
Cost / Task
$0.26
benchmark run
Output speed
129 tok/s
blended tok/s
Time to first token
0.5s
median latency
Input price
$0.12
per 1M tokens
Output price
$0.48
per 1M tokens
Context window
128K
tokens
Strengths
Runs on single-GPU hardware with high inference efficiency
Rapid 129 tok/s output speed
Cost-effective alternative for private enterprise deployments
How we rate Qwen3.8 27B (xhigh)
OpenRating scores every model on the same axes. The OpenRating Index is a weighted composite of reasoning, knowledge, and agentic benchmarks, normalised to a 0–100 scale. Speed is blended output measured in tokens per second under standardised load, with time-to-first-token reported separately. Prices are list input and output prices per million tokens, and Task Cost measures end-to-end workload spending.
OpenRating Index
52.1
Reasoning Index
71.4
Token Pricing
Input per 1M tokens
$0.12
Output per 1M tokens
$0.48
Cost of 1M input + 1M output
$0.6
Task Cost Economics
Empirical cost breakdown for a standard OpenRating Index benchmark task ($0.26 total).
Related models
Compare Qwen3.8 27B (xhigh) with the models it is most often measured against.